返回
Intra-Frame Coding Using a Conditional Autoencoder
DOI:10.1109/JSTSP.2020.3034768.png)
摘要
En 中文
Exploiting spatial redundancy in images is responsible for a large gain in the performance of image and video compression. The main tool to achieve this is called intra-frame prediction. In most state-of-the-art video coders, intra prediction is applied in a block-wise fashion. Up to now angular prediction was dominant, providing a low-complexity method covering a large variety of content. With deep learning, however, it is possible to create prediction methods covering a wider range of content, being able to predict structures which traditional modes can not predict accurately. Using the conditional autoencoder structure, we are able to train a single artificial neural network which is able to perform multi-mode prediction. In this paper, we derive the approach from the general formulation of the intra-prediction problem and introduce two extensions for spatial mode prediction and for chroma prediction support. Moreover, we propose a novel latent-space-based cross component prediction. We show the power of our prediction scheme with visual examples and report average gains of 1.13% in Bjontegaard delta rate in the luma component and 1.21% in the chroma component compared to VTM using only traditional modes.
Keyword:
Neural networks
Encoding
Image coding
Training
Tools
Decoding
Prediction methods
Video coder
intra prediction
conditional autoencoder
deep learning
期刊
IF:
13.7
论文数:
1.9K
被引数:
1.1W
机构
引用论文
Real-time and offline techniques for identifying obstructive sleep apnea patients用于识别阻塞性睡眠呼吸暂停患者的实时和离线技术
Red, green, and blue electrochromism in ambipolar poly(amine–amide–imide)s based on electroactive tetraphenyl‐p‐phenylenediamine units基于电活性四苯基 p-苯二胺单元的双极性聚 (胺-酰胺-酰亚胺) 中的红色,绿色和蓝色电致变色

