arrow
返回

MSSA: Multi-Representation Semantics-Augmented Set Abstraction for 3D Object Detection

delete2024-10-29
delete0
delete
OA
AI
H
Huaijin Liu
J
Ji‐Xiang Du *
张勇 (Yong Zhang)
H
Hongbo Zhang
J
Jiandian Zeng *
DOI:10.1145/3686157delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Accurate recognition and localization of 3D objects is a fundamental research problem in 3D computer vision. Benefiting from transformation-free point cloud processing and flexible receptive fields, point-based methods have become accurate in 3D point cloud modeling, but still fall behind voxel-based competitors in 3D detection. We observe that the set abstraction module, commonly utilized by point-based methods for downsampling points, tends to retain excessive irrelevant background information, thus hindering the effective learning of features for object detection tasks. To address this issue, we propose MSSA, a Multi-representation Semantics- augmented Set Abstraction for 3D object detection. Specifically, we first design a backbone network to encode different representation features of point clouds, which extracts point-wise features through PointNet to preserve fine-grained geometric structure features, and adopts VoxelNet to extract voxel features and BEV features to enhance the semantic features of key points. Second, to efficiently fuse different representation features of keypoints, we propose a Point feature-guided Voxel feature and BEV feature fusion (PVB-Fusion) module to adaptively fuse multi-representation features and remove noise. At last, a novel Multi-representation Semantic-guided Farthest Point Sampling (MS-FPS) algorithm is designed to help set abstraction modules progressively downsample point clouds, thereby improving instance recall and detection performance with more important foreground points. We evaluate MSSA on the widely used KITTI dataset and the more challenging nuScenes dataset. Experimental results show that compared to PointRCNN, our method improves the AP of moderate level for three classes of objects by 7.02%, 6.76%, and 5.44%, respectively. Compared to the advanced point-voxel-based method PV-RCNN, our method improves the AP of moderate level by 1.23%, 2.84%, and 0.55% for the three classes, respectively.

期刊

ACM Transactions on Multimedia Computing Communications and Applications 封面图
ACM Transactions on Multimedia Computing Communications and Applications
IF:
6
论文数:
2.0K
被引数:
5.4K

机构

B
Beijing Normal University
学者数:
3.3W
论文数: 2.7W
被引数: 4.2W
H
huaqiao university
学者数:
1.1W
论文数: 7.1K
被引数: 131
Z
Zhejiang Chinese Medical University
学者数:
1.3W
论文数: 6.5K
被引数: 7.3K
学者 查看更多机构
引用论文

引用论文

The myeloma cell antigen syndecan‐1 is lost by apoptotic myeloma cells
err2001-12-25
err0
errOAAI
errMichel Jourdan; Martine Ferlin; Eric Legouffe; Mira Horvathova; Janny Liautard; Jean FranÇois Rossi; John Wijdenes; Jean Brochier; Bernard Klein
err分享
err收藏
err分享
err收藏
Key Aspects of Nucleic Acid Library Design for in Vitro Selection
err2018-02-05
err0
errOAAI
errMaria Vorobyeva; Anna Davydova; Pavel Vorobjev; Dmitrii Pyshnyi; Alya Venyaminova
err分享
err收藏
学者 查看更多内容