arrow
Return

Structure-Aware Generative Point Cloud Compression for Visual Perception

delete2025-01-01
delete0
PRE
AI
Y
Yichen Zhou
X
Xinfeng Zhang
Y
Yingzhan Xu
K
Kai Zhang
L
Li Zhang
Q
Qingming Huang
DOI:10.1109/TIP.2025.3607630delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
In recent years, there has been a rapid growth in applications that rely on point clouds to represent the 3D world, driven by the increasing demand for immersive and other related scenarios. However, compressing the large and high-precision point cloud data efficiently while maintaining high perceptual quality for human vision remains a challenge. To solve the problem, we propose a new structure-aware generative point cloud compression framework for human vision. In the encoder, we focus on information that is more sensitive to the human vision and obtain this type of information from different scale. This allows us to capture structural importance information from global scale and local scale, which are more difficult to reconstruct. For the decoder, we introduce a progressive generative reconstruction approach that utilizes acquired information from the encoder to guide the generation of point cloud surfaces. Moreover, we propose a novel probability cloud-based discriminator. Instead of directly assessing the authenticity of the generated point clouds, our discriminator evaluates the probability distribution of the existence of points within the generated point cloud. This approach reduces the difficulty of discrimination while effectively improving the accuracy of the generator in generating probability distributions. According to the correct probability, we can obtain a high accuracy point cloud by pruning the points with low probability. Through comprehensive experiments, we demonstrate the effectiveness and superiority of our proposed framework in terms of encoding efficiency, high perceptual quality, and generation quality.
Keywords:
Point cloud compression
generative adversarial network
geometry
visual perception

Journal

IEEE Transactions on Image Processing cover
IEEE Transactions on Image Processing
IF:
13.7
Papers:
1.0W
Citations:
8.4W

Organization

B
bytedance inc., beijing, china
Scholars:
5
Papers: 3
Citations: 0
U
University of Chinese Academy of Sciences
Scholars:
6.2K
Papers: 2.5K
Citations: 24.6W
B
bytedance inc., san diego, ca, usa
Scholars:
3
Papers: 2
Citations: 0
researcher View more organizations