arrow
返回

Convolutional Neural Network-Based Occupancy Map Accuracy Improvement for Video-Based Point Cloud Compression

delete2022-01-01
delete17
delete
OA
AI
W
Wei Jia
李
李莉 (Li Li)
A
Anique Akhtar
朱
朱理 (Li, Zhu) *
S
Shan Liu
DOI:10.1109/TMM.2021.3079698delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
In video-based point cloud compression (V-PCC), a dynamic point cloud is projected onto geometry and attribute videos patch by patch for compression. In addition to the geometry and attribute videos, an occupancy map video is compressed into a V-PCC bitstream to indicate whether a two-dimensional (2D) point in the projected geometry video corresponds to any point in three-dimensional (3D) space. The occupancy map video is usually downsampled before compression to obtain a tradeoff between the bitrate and the reconstructed point cloud quality. Due to the accuracy loss in the downsampling process, some noisy points are generated, which leads to severe objective and subjective quality degradation of the reconstructed point cloud. To improve the quality of the reconstructed point cloud, we propose using a convolutional neural network (CNN) to improve the accuracy of the occupancy map video. We mainly make the following contributions. First, we improve the accuracy of the occupancy map video by formulating the problem as a binary segmentation problem since the pixel values of the occupancy map video are either 0 or 1. Second, in addition to the downsampled occupancy map video, we introduce a reconstructed geometry video as the other input of the CNN to provide more useful information in order to indicate the occupancy map video. To the best of our knowledge, this is the first learning-based work to improve the performance of V-PCC. Compared to state-of-the-art schemes, our proposed CNN-based approach achieves much more accurate occupancy map videos and significant bitrate savings.
Keyword:
Three-dimensional displays
Videos
Geometry
Heuristic algorithms
Noise measurement
Bit rate
Software algorithms
Convolutional neural network
high efficiency video coding
occupancy map
segmentation
video-based point cloud compression
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

IEEE Transactions on Multimedia 封面图
IEEE Transactions on Multimedia
IF:
9.7
论文数:
4.5K
被引数:
2.4W

机构

U
university of missouri kansas city
学者数:
3.9K
论文数: 3.3K
被引数: 3
C
chinese academy of sciences
学者数:
56.7W
论文数: 45.0W
被引数: 704
University of Missouri System 封面图
University of Missouri System
学者数:
3.0W
论文数: 2.7W
被引数: 75
学者 查看更多机构
引用论文

引用论文

A tutorial on the cross-entropy method关于交叉熵方法的教程
err2005-02-01
err2.2K
PREAI
errDe Boer, PT; Kroese, DP; Mannor, S; Rubinstein, RY
err分享
err收藏
Emerging MPEG Standards for Point Cloud Compression
err2019-03-01
err576
errOAAI
errSchwarz, Sebastian; Preda, Marius; Baroncini, Vittorio; Budagavi, Madhukar; Cesar, Pablo; Chou, Philip A.; Cohen, Robert A.; Krivokuca, Maja; Lasserre, Sebastien; Li, Zhu; Llach, Joan; Mammou, Khaled; Mekuria, Rufael; Nakagami, Ohji; Siahaan, Ernestasia; Tabatabai, Ali; Tourapis, Alexis M.; Zakharchenko, Vladyslav
err分享
err收藏
err分享
err收藏
err分享
err收藏
学者 查看更多内容