arrow
Return

PAN: Pillars-Attention-Based Network for 3D Object Detection

delete2026-05-12
delete0
delete
OA
AI
R
Ruan Bispo
D
Dane Mitrev
L
Letizia Mariotti
C
Clément Botty
D
Denver Humphrey
A
Anthony Scanlan
C
Ciarán Eising
DOI:10.1109/ojvt.2026.3692626delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Camera-radar fusion offers a robust and low-cost alternative to camera-lidar fusion for the 3D object detection task in real-time under adverse weather and lighting conditions. However, currently, in the literature, it is possible to find few works focusing on this modality and, most importantly, developing new architectures to explore the advantages of the radar point cloud, such as accurate distance estimation and speed information. Therefore, this work presents a novel and efficient 3D object detection algorithm using cameras and radars in the bird’s-eye-view (BEV). Our algorithm exploits the advantages of radar before fusing the features into a detection head. A new backbone is introduced, which maps the radar pillar features into an embedded dimension. A self-attention mechanism allows the backbone to model the dependencies between the radar points. We used a simplified convolutional layer to replace the FPN-based convolutional layers used in the PointPillars-based architectures with the main goal of reducing inference time. Experimental results show that our approach achieves strong performance on the nuScenes dataset, reaching an NDS of 58.2 with a ResNet-50 backbone, while maintaining real-time inference.
Keywords:
3D object detection
camera-radar
nuScenes
perception
sensor fusion

Journal

I
IEEE Open Journal of Vehicular Technology
IF:
4.8
Papers:
557
Citations:
987

Organization

U
university of limerick
Scholars:
1.2K
Papers: 569
Citations: 0
P
provizio
Scholars:
5
Papers: 1
Citations: 0