arrow
返回

Audio-Driven Talking Video Frame Restoration

delete2024-01-01
delete1
PRE
AI
H
Harry H. Cheng
Y
Yangyang Guo
J
Jianhua Yin
H
Haonan Chen
J
Jiafang Wang
Liqiang Nie 封面图
Liqiang Nie (Liqiang Nie) *
DOI:10.1109/TMM.2021.3118287delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Talking video frames occasionally drop while streaming for reasons like network errors, which greatly hurts the online team collaboration and user experiences. Directly generating the dropped frames from the remaining ones is unfavorable since a person's lip motion is usually non-linear and thus hard to be restored when consecutive frames are missing. Nevertheless, the audio content provides strong signals for lip motion and is less likely to drop during transmitting. Inspired by this, as an initial attempt, we present the task of audio-driven talking video frame restoration in this paper, i.e., restoring dropped video frames by jointly leveraging the audio and remaining video frames. Towards the high-quality frame generation, we devise a cross-modal frame restoration network. This network aligns the complete audio content with video frames, precisely identifies and sequentially generates the dropped frames. To justify our model, we construct a new dataset, Talking Video Frames Drop, TVFD for short, consisting of 2.5K video and 144K frames in total. We conduct extensive experiments over TVFD and another publicly accessible dataset - Voxceleb2. Our model obtains significantly improved performance as compared to other state-of-the-art competitors.
Keyword:
Streaming media
Faces
Lips
Task analysis
Image restoration
Visualization
Synchronization
Frame Restoration
Frame-Dropped Video
Cross-Modal Learning
Dynamic Programming
Generative Adversial Network

期刊

IEEE Transactions on Multimedia 封面图
IEEE Transactions on Multimedia
IF:
9.7
论文数:
4.5K
被引数:
2.4W

机构

A
alibaba group
学者数:
1.1K
论文数: 789
被引数: 0
S
shandong university
学者数:
9.5W
论文数: 6.4W
被引数: 94
引用论文

引用论文

3D Face Reconstruction From A Single Image Assisted by 2D Face Images in the Wild
err2021-01-01
err72
errOAAI
errTu, Xiaoguang; Zhao, Jian; Xie, Mei; Jiang, Zihang; Balamurugan, Akshaya; Luo, Yao; Zhao, Yang; He, Lingxiao; Ma, Zheng; Feng, Jiashi
err分享
err收藏
Planar monomode optical couplers based on multimode interference effects
err1992-12-01
err0
errOAAI
errL.B. Soldano; F.B. Veerman; M.K. Smit; B.H. Verbeek; A.H. Dubost; E.C.M. Pennings
err分享
err收藏
Planning future studies based on the conditional power of a meta‐analysis
err2012-07-11
err0
errOAAI
errVerena Roloff; Julian P.T. Higgins; Alex J. Sutton
err分享
err收藏
Soft magnetic properties of nanocrystalline Ni3Fe and Fe75Al12.5Ge12.5
err1999-11-01
err0
PREAI
errH.N. Frase; R.D. Shull; L.-B. Hong; T.A. Stephens; Z.-Q. Gao; B. Fultz
err分享
err收藏
err分享
err收藏
学者 查看更多内容