arrow
返回

Variational Learned Talking-Head Semantic Coded Transmission System

delete2024-07-01
delete0
PRE
AI
Y
Yue Weijie
Z
Zhongwei Si *
DOI:10.23919/JCC.fa.2024-0036.202407delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Video transmission requires considerable bandwidth, and current widely employed schemes prove inadequate when confronted with scenes featuring prominently. Motivated by the strides in talkinghead generative technology, the paper introduces a semantic transmission system tailored for talking-head videos. The system captures semantic information from talking-head video and faithfully reconstructs source video at the receiver, only one-shot reference frame and compact semantic features are required for the entire transmission. Specifically, we analyze video semantics in the pixel domain frame-by-frame and jointly process multi-frame semantic information to seamlessly incorporate spatial and temporal information. Variational modeling is utilized to evaluate the diversity of importance among group semantics, thereby guiding bandwidth resource allocation for semantics to enhance system efficiency. The whole endto-end system is modeled as an optimization problem and equivalent to acquiring optimal rate-distortion performance. We evaluate our system on both reference frame and video transmission, experimental results demonstrate that our system can improve the efficiency and robustness of communications. Compared to the classical approaches, our system can save over 90% of bandwidth when user perception is close.
Keyword:
semantic communications
source- channel coding
talking-head transmission
variational modeling

期刊

China Communications 封面图
China Communications
IF:
3.1
论文数:
1.9K
被引数:
5.0K

机构

B
beijing university of posts & telecommunications
学者数:
1.4W
论文数: 1.2W
被引数: 9