arrow
返回

Learned Image Compression Using Cross-Component Attention Mechanism

delete2023-01-01
delete3
PRE
AI
W
Wenhong Duan
Z
Zheng Chang
C
Chuanmin Jia *
王苫社 封面图
王苫社 (Shanshe Wang)
马
马思伟 (Siwei Ma) *
李
李松 (Li Song)
高
高雯 (Wen Gao)
DOI:10.1109/TIP.2023.3319275delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Learned image compression methods have achieved satisfactory results in recent years. However, existing methods are typically designed for RGB format, which are not suitable for YUV420 format due to the variance of different formats. In this paper, we propose an information-guided compression framework using cross-component attention mechanism, which can achieve efficient image compression in YUV420 format. Specifically, we design a dual-branch advanced information-preserving module (AIPM) based on the information-guided unit (IGU) and attention mechanism. On the one hand, the dual-branch architecture can prevent changes in original data distribution and avoid information disturbance between different components. The feature attention block (FAB) can preserve the important information. On the other hand, IGU can efficiently utilize the correlations between Y and UV components, which can further preserve the information of UV by the guidance of Y. Furthermore, we design an adaptive cross-channel enhancement module (ACEM) to reconstruct the details by utilizing the relations from different components, which makes use of the reconstructed Y as the textural and structural guidance for UV components. Extensive experiments show that the proposed framework can achieve the state-of-the-art performance in image compression for YUV420 format. More importantly, the proposed framework outperforms Versatile Video Coding (VVC) with 8.37% BD-rate reduction on common test conditions (CTC) sequences on average. In addition, we propose a quantization scheme for context model without model retraining, which can overcome the cross-platform decoding error caused by the floating-point operations in context model and provide a reference approach for the application of neural codec on different platforms.
Keyword:
Image coding
Context modeling
Transforms
Decoding
Standards
Image reconstruction
Transform coding
Image compression
cross-component
information-guided unit
attention mechanism
information-preserving

期刊

IEEE Transactions on Image Processing 封面图
IEEE Transactions on Image Processing
IF:
13.7
论文数:
1.0W
被引数:
8.4W

机构

S
shanghai jiao tong university
学者数:
15.7W
论文数: 11.7W
被引数: 159
P
peking university
学者数:
11.9W
论文数: 8.7W
被引数: 146
I
institute of computing technology, cas
学者数:
1.0K
论文数: 878
被引数: 1
C
chinese academy of sciences
学者数:
56.7W
论文数: 45.0W
被引数: 704
学者 查看更多机构
引用论文

引用论文

Activator and Repressor Functions of the Mot3 Transcription Factor in the Osmostress Response of Saccharomyces cerevisiae
err2013-05-01
err0
errOAAI
errFernando Martínez-Montañés; Alessandro Rienzo; Daniel Poveda-Huertes; Amparo Pascual-Ahuir; Markus Proft
err分享
err收藏
Online condition monitoring of engine oil
err2005-12-01
err0
PREAI
errSaurabh Kumar; P.S. Mukherjee; N.M. Mishra
err分享
err收藏
Systemic GM-CSF Recruits Effector T Cells into the Tumor Microenvironment in Localized Prostate Cancer
err2016-10-31
err0
errOAAI
errXiao X. Wei; Stephen Chan; Serena Kwek; Jera Lewis; Vinh Dao; Li Zhang; Matthew R. Cooperberg; Charles J. Ryan; Amy M. Lin; Terence W. Friedlander; Brian Rini; Christopher Kane; Jeffry P. Simko; Peter R. Carroll; Eric J. Small; Lawrence Fong
err分享
err收藏
学者 查看更多内容