arrow
Return

6D Object Pose Estimation With Compact Generalized Non-Local Operation

delete2024-01-01
delete0
delete
OA
AI
C
Changhong Jiang
X
Xiaoqiao Mu
B
Bingbing Zhang
L
Liang Chao *
M
Mujun Xie *
DOI:10.1109/ACCESS.2024.3508772delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Real-time object detection and pose estimation are critical in practical applications such as virtual reality, scene understanding, and robotics. In this paper, we propose a compact generalized non-local pose estimation network capable of directly predicting the projection of an object's 3D bounding box vertices onto a 2D image, facilitating the estimation of the object's 6D pose. The network is constructed using the YOLOv5 model, with the integration of an improved non-local module termed the Compact Generalized Non-local Block. This module enhances feature representation by learning the correlations between the positions of all elements across channels, effectively capturing subtle feature cues. The proposed network is end-to-end trainable, producing accurate pose predictions without the need for any post-processing operations. Extensive validation on the LineMod dataset shows that our approach achieves a final accuracy of 46.1% on the average 3D distance of model vertices (ADD) metric, outperforming existing methods by 6.9% and our baseline model by 1.8%, thus underscoring the efficacy of the proposed network.
Keywords:
Pose estimation
Feature extraction
Three-dimensional displays
Training
Correlation
Predictive models
Computational modeling
Accuracy
Solid modeling
YOLO
Correlations
subtle feature
end-to-end
long-range spatiotemporal
fine-grained details
representational power

Journal

IEEE Access cover
IEEE Access
IF:
3.6
Papers:
9.8W
Citations:
29.4W

Organization

C
Changchun University of Technology
Scholars:
5.0K
Papers: 2.7K
Citations: 3.3K
D
Dalian Minzu University
Scholars:
2.0K
Papers: 1.7K
Citations: 2.6K