arrow
返回

Audio-Driven Talking Face Video Generation With Dynamic Convolution Kernels

delete2023-01-01
delete19
delete
OA
AI
Z
Zipeng Ye
M
Mengfei Xia
R
Ran Yi *
张举勇 (Juyong Zhang)
Y
Yu‐Kun Lai
X
Xuwei Huang
G
Guoxin Zhang
Y
Yong‐Jin Liu *
DOI:10.1109/TMM.2022.3142387delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
In this paper, we present a dynamic convolution kernel (DCK) strategy for convolutional neural networks. Using a fully convolutional network with the proposed DCKs, high-quality talking-face video can be generated from multi-modal sources (i.e., unmatched audio and video) in real time, and our trained model is robust to different identities, head postures, and input audios. Our proposed DCKs are specially designed for audio-driven talking face video generation, leading to a simple yet effective end-to-end system. We also provide a theoretical analysis to interpret why DCKs work. Experimental results show that our method can generate high-quality talking-face video with background at 60 fps. Comparison and evaluation between our method and the state-of-the-art methods demonstrate the superiority of our method.
Keyword:
Dynamic kernel
convolutional neural network
multi-modal generation task
audio-driven talking-face generation

期刊

IEEE Transactions on Multimedia 封面图
IEEE Transactions on Multimedia
IF:
9.7
论文数:
4.5K
被引数:
2.4W

机构

S
shanghai jiao tong university
学者数:
15.7W
论文数: 11.7W
被引数: 159
T
tsinghua university
学者数:
11.9W
论文数: 10.0W
被引数: 137
U
university of science & technology of china, cas
学者数:
3.2W
论文数: 2.7W
被引数: 74
C
Cardiff University
学者数:
2.7W
论文数: 2.5W
被引数: 3.5W
C
chinese academy of sciences
学者数:
56.7W
论文数: 45.0W
被引数: 704
学者 查看更多机构
引用论文

引用论文

Planar monomode optical couplers based on multimode interference effects
err1992-12-01
err0
errOAAI
errL.B. Soldano; F.B. Veerman; M.K. Smit; B.H. Verbeek; A.H. Dubost; E.C.M. Pennings
err分享
err收藏
Text-based Editing of Talking-head Video基于文本的说话视频编辑
err2019-07-12
err171
errOAAI
errFried, Ohad; Tewari, Ayush; Zollhofer, Michael; Finkelstein, Adam; Shechtman, Eli; Goldman, Dan B.; Genova, Kyle; Jin, Zeyu; Theobalt, Christian; Agrawala, Maneesh
err分享
err收藏
err分享
err收藏
Effect of sildenafil on ocular hemodynamics in 3 months regular use
err2005-11-17
err0
errOAAI
errS O Dündar; Y Dayanir; A Topaloğlu; M Dündar; İ Koçak
err分享
err收藏
Viability of ram spermatozoa in relation to the abstinence period and successive ejaculations
err2008-06-28
err0
errOAAI
errM. OLLERO; T. MUIÑO‐BLANCO; M. J. LÓPEZ‐PÉREZ; J. A. CEBRIÁN‐PÉREZ
err分享
err收藏
Deep Video Portraits
err2018-07-30
err279
errOAAI
errKim, Hyeongwoo; Garrido, Pablo; Tewari, Ayush; Xu, Weipeng; Thies, Justus; Niessner, Matthias; Perez, Patrick; Richardt, Christian; Zollhofer, Michael; Theobalt, Christian
err分享
err收藏
Bringing Portraits to Life
err2017-11-20
err124
PREAI
errAverbuch-Elor, Hadar; Cohen-Or, Daniel; Kopf, Johannes; Cohen, Michael F.
err分享
err收藏
学者 查看更多内容