返回
3D-MSFC: A 3D multi-scale features compression method for object detection☆
DOI:10.1016/j.displa.2024.102880.png)
摘要
En 中文
As machine vision tasks rapidly evolve, anew concept of compression, namely video coding for machines (VCM), has emerged. However, current VCM methods are only suitable for 2D machine vision tasks. With the popularization of autonomous driving, the demand for 3D machine vision tasks has significantly increased, leading to an explosive growth in LiDAR data that requires efficient transmission. To address this need, we propose a machine vision-based point cloud coding paradigm inspired by VCM. Specifically, we introduce a 3D multi-scale features compression (3D-MSFC) method, tailored for 3D object detection. Experimental results demonstrate that 3D-MSFC achieves less than a 3% degradation in object detection accuracy at a compression ratio of 2796x. Furthermore, its low-profile variant, 3D-MSFC-L, achieves less than a 2% degradation in accuracy at a compression ratio of 463x. The above results indicate that our proposed method can provide an ultra-high compression ratio while ensuring no significant drop inaccuracy, greatly reducing the amount of data required for transmission during each detection. This can significantly lower bandwidth consumption and save substantial costs in application scenarios such as smart cities.
Keyword:
Machine vision-based point cloud coding
3D multi-scale features compression
3D object detection
期刊
IF:
3.4
论文数:
2.2K
被引数:
3.2K
机构
引用论文
Deep-PCAC: An End-to-End Deep Lossy Compression Framework for Point Cloud AttributesDeep-pcac: 点云属性的端到端深度有损压缩框架
Thermal comfort, perceived air quality, and cognitive performance when personally controlled air movement is used by tropically acclimatized persons当热带适应的人使用个人控制的空气运动时,热舒适性,感知的空气质量和认知表现
Indoor Air
IF0

