arrow
返回

A Super-Resolution-Based Feature Map Compression for Machine-Oriented Video Coding

delete2023-01-01
delete3
delete
OA
AI
J
JungHeum Kang
M
Muhammad Salman Ali
H
Hyewon Jeong
C
Chang-Kyun Choi
Y
Younhee Kim
S
Seyoon Jeong
S
Sung‐Ho Bae
H
Hui Yong Kim *
DOI:10.1109/ACCESS.2023.3260223delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Recently, video and image compression methods using neural networks have received much attention. In MPEG standardization, Video Coding for Machine (VCM) is a newly arising topic which attempts to compress features/images for the purpose of machine vision tasks. Especially, compressing features has advantages in terms of privacy protection and computation off-loading. In this paper, we propose an effective feature compression method equipped with a super-resolution (SR) module for features. Our main motivation comes from the observation that features are somewhat robust to spatial distortions (e.g., AWGN, blur, quantization distortions, coding artifacts), which leads us to integrating an SR module into the compression framework. We also further explore the best training strategy of the proposed method, i.e., finding the best combination of various losses and proper input feature shapes. Our comprehensive experiments show that the proposed method outperforms the baseline in the original VCM anchor scenario on various QP values with Versatile Video Coding (VVC). Specifically, the proposed framework achieved up to 50% BD-rate reduction compared to the conventional P-layer feature map compression method for the object detection task on the OpenImage dataset.
Keyword:
Versatile video codec
video coding for machine
feature compression
deep neural network
super resolution

期刊

IEEE Access 封面图
IEEE Access
IF:
3.6
论文数:
9.8W
被引数:
29.4W

机构

K
kyung hee university
学者数:
2.3W
论文数: 2.2W
被引数: 234