arrow
返回

A Unified Framework for Jointly Compressing Visual and Semantic Data

delete2024-05-15
delete0
delete
OA
AI
S
Shizhan Liu
林
林巍峣 (Weiyao Lin) *
Y
Yihang Chen
Y
Yufeng Zhang
W
Wenrui Dai
J
John See
熊
熊红凯 (Hongkai Xiong)
DOI:10.1145/3654800delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
The rapid advancement of multimedia and imaging technologies has resulted in increasingly diverse visual and semantic data. A large range of applications such as remote-assisted driving requires the amalgamated storage and transmission of various visual and semantic data. However, existing works suffer from the limitation of insufficiently exploiting the redundancy between different types of data. In this article, we propose a unified framework to jointly compress a diverse spectrum of visual and semantic data, including images, point clouds, segmentation maps, object attributes, and relations. We develop a unifying process that embeds the representations of these data into a joint embedding graph according to their categories, which enables flexible handling of joint compression tasks for various visual and semantic data. To fully leverage the redundancy between different data types, we further introduce an embedding-based adaptive joint encoding process and a Semantic Adaptation Module to efficiently encode diverse data based on the learned embeddings in the joint embedding graph. Experiments on the Cityscapes, MSCOCO, and KITTI datasets demonstrate the superiority of our framework, highlighting promising steps toward scalable multimedia processing.
Keyword:
Joint visual and semantic data compression
visual semantic data
multimedia processing
unified compression framework

期刊

ACM Transactions on Multimedia Computing Communications and Applications 封面图
ACM Transactions on Multimedia Computing Communications and Applications
IF:
6
论文数:
2.0K
被引数:
5.4K

机构

S
shanghai jiao tong university
学者数:
15.7W
论文数: 11.7W
被引数: 159
H
Heriot Watt University
学者数:
6.0K
论文数: 6.5K
被引数: 57
引用论文

引用论文

Unified Binary Generative Adversarial Network for Image Retrieval and Compression
err2020-02-18
err54
errOAAI
errSong, Jingkuan; He, Tao; Gao, Lianli; Xu, Xing; Hanjalic, Alan; Shen, Heng Tao
err分享
err收藏
Just Recognizable Distortion for Machine Vision Oriented Image and Video Coding
err2021-08-13
err13
PREAI
errZhang, Qi; Wang, Shanshe; Zhang, Xinfeng; Ma, Siwei; Gao, Wen
err分享
err收藏
Conceptual Compression via Deep Structure and Texture Synthesis
err2022-01-01
err19
errOAAI
errChang, Jianhui; Zhao, Zhenghui; Jia, Chuanmin; Wang, Shiqi; Yang, Lingbo; Mao, Qi; Zhang, Jian; Ma, Siwei
err分享
err收藏
err分享
err收藏
err分享
err收藏
err分享
err收藏
Multi-Modal Neural Feature Fusion for Automatic Driving Through Perception-Aware Path Planning
err2021-01-01
err10
errOAAI
errLi, Zhenyu; Zhou, Aiguo; Pu, Jiakun; Yu, Jiangyang
err分享
err收藏
学者 查看更多内容