arrow
返回

SCRTN: Enhancing multi-modal 3D object detection in complex environments

delete2026-01-13
delete0
PRE
AI
X
Xiufeng Zhu
Q
Qing Shen
Z
Zhenfang Liu
K
Kang Zhao
J
Jungang Lou
DOI:10.1016/j.patcog.2026.113068delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
• A multi-modal 3D object detection framework, the SCRTN, is proposed to fuse 3D point cloud data with two-dimensional (2D) image data, which enhances recognition performance of 3D objects in complex environments. • The ResTransfusion feature fusion technique is used to strengthen the global information connection between the point cloud features and enhanced point clouds, which improves the synergy between semantic and shape features, thus improving the model’s depth optimization capability. • A far-reaching voxel preservation sampling strategy is designed and combined with sparse convolutional technology to increase the specificity of feature extraction in 3D data, which increases the efficiency of 3D object detection. • The results of the extensive experiments demonstrate that the proposed method can achieve the state-of-the-art performance on the Hard KITTI dataset, with an accuracy of 85.75 and a mean average precision (mAP) of 89.67. Moreover, the proposed model also has excellent performance on the Nuscenes dataset and Waymo Open Dataset.

期刊

Pattern Recognition 封面图
Pattern Recognition
IF:
7.6
论文数:
1.3W
被引数:
4.5W

机构

H
Huzhou University
学者数:
4.1K
论文数: 3.5K
被引数: 6.7K
引用论文

引用论文

Parallel disentangling network for human-object interaction detection
err2024-02-01
err7
PREAI
errCheng, Yamin; Duan, Hancong; Wang, Chen; Chen, Zhijun
err分享
err收藏
err分享
err收藏
MCNet: Magnitude consistency network for domain adaptive object detection under inclement environments
err2024-01-01
err4
PREAI
errPang, Jian; Liu, Weifeng; Zhang, Bingfeng; Yang, Xinghao; Liu, Baodi; Tao, Dapeng
err分享
err收藏
没有更多内容