Advertisement
Sections
Researchers train neural network to generate realistic videos lip-synced to audio clips
In a video conference, the users can avoid interruptions from video feeds timing out because of connectivity problems.

Researchers from the University of Washington have developed an algorithm that generates realistic looking video based on audio input. An artificial neural network was first trained using hours of videos to generate realistic looking mouth shapes, that are synced to the audio files. The shapes are then superimposed seamlessly on an existing video of the person involved. The system was tested using audio and video files of Barack Obama available in the public domain.Steve Seitz, co-author of the paper says "When you watch Skype or Google Hangouts, often the connection is stuttery and low-resolution and really unpleasant, but often the audio is pretty good. So if you could use the audio to produce much higher-quality video, that would be terrific." The researchers made sure to get the mouth area absolutely right, to prevent the resulting video from looking unnatural.[caption id="attachment_388782" align="aligncenter" width="640"]
Image: University of Washington[/caption]There are a number of potential applications of the algorithm. Users can have realistic conversations with historical figures, in a virtual reality environment. In a video conference, the users can avoid interruptions from video feeds timing out because of connectivity problems. The system can also be used to generate high quality video feeds, even if the base video is not of a high resolution.
Image: University of Washington[/caption]There are a number of potential applications of the algorithm. Users can have realistic conversations with historical figures, in a virtual reality environment. In a video conference, the users can avoid interruptions from video feeds timing out because of connectivity problems. The system can also be used to generate high quality video feeds, even if the base video is not of a high resolution.First Published:Jul 12, 2017, 18:20:27 IST
Advertisement
Advertisement

Why AI notetakers are raising serious privacy and security concerns
AI notetakers promise effortless meeting summaries, but experts warn they could expose confidential conversations, corporate secrets and personal voiceprints. As businesses increasingly adopt AI-powered meeting assistants, questions over data storage, privacy, consent and legal risks are becoming impossible to ignore
5 min read
China's low-cost AI models are changing the global AI race. Here's why Silicon Valley is worried
As Chinese firms continue to improve performance while keeping prices low, the AI race is no longer just about building the smartest model—it is increasingly becoming a battle over who can deliver the best value
2 min read
China's Kimi K3 challenges US AI leaders with frontier-level performance at lower cost
Chinese artificial intelligence startup Moonshot AI has unveiled its latest open-weight AI model, Kimi K3, with early results suggesting it could compete with some of the world's most advanced AI systems developed by leading US companies
2 min read
How did Instagram run ads promoting child abuse in India?
India has issued a notice to Meta after an investigation alleged that Instagram displayed paid advertisements promoting child sexual abuse material. MeitY ordered Meta to remove such Instagram ads and explain within seven days how they were approved
8 min read
Why has India halted WhatsApp’s username feature before launch?
India has halted WhatsApp’s planned username feature, citing concerns about cybercrime, impersonation, and law enforcement challenges. As Meta races to address security concerns, here’s why MeitY has paused the rollout, what the feature does, and how it could influence privacy and online safety in India
3 min read
Advertisement
Advertisement
