arrow
返回

Enhancing Point Cloud Feature Utilization for 3D Object Detection

delete2025-01-01
delete0
PRE
AI
Z
Zhou, Gongxiang
蒋峰 封面图
蒋峰 (Feng Jiang) *
R
Renjie Dong‬
M
Meng Wang
徐飘荣 (Piaorong Xu)
DOI:10.1109/LSP.2025.3626268delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Traditional voxelization methods use many artificial components to handle sparse and uneven point clouds, which may lead to spatial discretization distortion. In addition, the fixed convolution size limits the detector's ability to capture feature correlations, resulting in low feature utilization. To adderss this challenge, this paper proposes a two-stage LiDAR 3D object detector, Pillar-CT3D++. First, it introduces a self-attention encoding module that uses a Transformer to process voxelized point cloud information. By capturing the position information of point cloud objects within voxels through multi-dimensional position encoding and an encoder-decoder mechanism, the quality of feature suggestions is improved. Second, a feature weight-aware module is proposed, which designs multi-scale pooling and group attention mechanisms in the 2D backbone network to adaptively recalibrate the feature responses of channels, generating more representative feature information. The detection performance of Pillar-CT3D++ is validated through experiments on the KITTI dataset and compared with existing detectors. Specifically, on the KITTI dataset, the proposed model outperforms the baseline CT3D by 0.43%, 1.04%, and 0.87% in detecting simple, medium, and difficult-level objects, respectively.
Keyword:
Feature extraction
Point cloud compression
Encoding
Three-dimensional displays
Object detection
Convolution
Vectors
Detectors
Decoding
Attention mechanisms
3D object detection
autonomous driving
LiDAR
point cloud

期刊

I
IEEE Signal Processing Letters
IF:
3.9
论文数:
622
被引数:
0

机构

C
Central South University of Forestry & Technology
学者数:
601
论文数: 169
被引数: 0
引用论文

引用论文

DSVT: Dynamic Sparse Voxel Transformer with Rotated Sets
err2023-06-01
err0
errOAAI
errHaiyang Wang; Chen Shi; Shaoshuai Shi; Meng Lei; Sen Wang; Di He; Bernt Schiele; Liwei Wang
err分享
err收藏
Voxel R-CNN: Towards High Performance Voxel-based 3D Object Detection
err2021-05-18
err0
errOAAI
errJiajun Deng; Shaoshuai Shi; Peiwei Li; Wengang Zhou; Yanyong Zhang; Houqiang Li
err分享
err收藏
End-to-End Autonomous Driving: Challenges and Frontiers
err2024-12-01
err0
errOAAI
errLi Chen; Penghao Wu; Kashyap Chitta; Bernhard Jaeger; Andreas Geiger; Hongyang Li
err分享
err收藏
err分享
err收藏
Pillar-Based Object Detection for Autonomous Driving基于支柱的自动驾驶目标检测
err2020-11-17
err0
errOAAI
errYue Wang; Alireza Fathi; Abhijit Kundu; David A. Ross; Caroline Pantofaru; Tom Funkhouser; Justin Solomon
err分享
err收藏
学者 查看更多内容