arrow
返回

The Interpretable Fast Multi-Scale Deep Decoder for the Standard HEVC Bitstreams

delete2020-07-01
delete10
PRE
AI
H
Huiguo He
T
Tingting Wang *
H
Hongyang Chao *
DOI:10.1109/TMM.2020.2978664delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
It is a research hotspot to restore decoded videos with existing bitstreams by applying deep neural network to improve compression efficiency at decoder-end. Existing research has verified that the utilization of redundancy at decoder-end, which is underused by the encoder, can bring an increase of compression efficiency. However, most existing research neglects the abundant multi-scale information among video frames as a typical type of such redundancy. It remains an interesting yet challenging topic how to build an effective, interpretable and fast deep neural network for the purpose of using the multi-scale similarity at decoder-end and further enhancing compression efficiency. To this end, this paper considers the use of underused inter multi-scale information and proposes the Fast Multi-Scale Deep Decoder (Fast MSDD) for the state-of-the-art video coding standard HEVC. The advantages of Fast MSDD are three-fold. First, it achieves a higher coding efficiency without modifying any encoding algorithm. Second, Fast MSDD is interpretable based on the framework of using the underused redundancy. Third, it guarantees the model's inference speed while fully using the multi-scale similarity among video frames. Extensive experimental results verify Fast MSDD's effectiveness, interpretability, and computational efficiency. Fast MSDD obtains averagely 14.3%, 10.8%, 8.5% and 7.6% BD gains for AI, LP, LB and RA respectively. Compared with our previous work MSDD, Fast MSDD achieves increases of 59.3%, 49.1%, 61.0% and 29.3%. Meanwhile, 16.9%, 11.2%, 9.2% and 8.3% BD gains are observed on videos with scale changes, which validate the interpretability of the proposed method. Furthermore, Fast MSDD can save at most 56.3% time compared to MSDD.
Keyword:
Videos
Encoding
Decoding
Standards
Redundancy
Computational efficiency
Neural networks
HEVC
multi-scale similarity
compression efficiency
deep learning
interpretability
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Transactions on Multimedia 封面图
IEEE Transactions on Multimedia
IF:
9.7
论文数:
4.5K
被引数:
2.4W

机构

S
Sun Yat Sen University
学者数:
9.9W
论文数: 7.2W
被引数: 95
引用论文

引用论文

Insights into rRNA processing and modification mapping in Archaea using Nanopore-based RNA sequencing
err
IF0
err2021-06-14
err0
errOAAI
errFelix Grünberger; Michael Jüttner; Robert Knüppel; Sébastien Ferreira-Cerca; Dina Grohmann
err分享
err收藏
Low-Area Active-Feedback Low-Noise Amplifier Design in Scaled Digital CMOS
err2008-11-01
err0
PREAI
errJonathan Borremans; Piet Wambacq; Charlotte Soens; Yves Rolain; Maarten Kuijk
err分享
err收藏
学者 查看更多内容