arrow
Return

Bounding Box Vectorization for Oriented Object Detection With Tanimoto Coefficient Regression

delete2024-01-01
delete4
PRE
AI
L
Linfei Wang
Y
Yibing Zhan
刘玮 cover
刘玮 (Wei Liu)
B
Baosheng Yu
陶大鹏 cover
陶大鹏 (Dapeng Tao) *
DOI:10.1109/TMM.2023.3330103delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Current oriented object detection methods mainly utilize a vanilla coordinate-angle representation for bounding box regression, which usually suffers from inconsistency between the bounding box regression losses and prediction errors induced with respect to different rotation angles, aspect ratios, and scales. Therefore, although the existing oriented object detectors have achieved very good performances under coarse evaluation metrics such as AP(50), their performance significantly degrades when using stricter evaluation metric such as AP(75). To address the abovementioned issues, we propose a new regression method with bounding box vectorization that implicitly represents the shape and orientation of an object with a set of orthogonal vectors. By doing this, the proposed method delicately avoids the inconsistency issues encountered in oriented bounding box regression. During training, we introduce the Tanimoto coefficient to evaluate the similarity of the bounding box vector in a shape- and orientation-aware manner, and we refer to the proposed box-to-vector loss as the B2V loss. In addition to 2D object detection, the proposed method can be easily generalized to 3D scenarios involving orientation estimation, such as autonomous driving. We evaluate the proposed method through extensive experiments conducted on four popular oriented object detection datasets, including both 2D and 3D datasets, where the proposed method significantly outperforms the recently developed state-of-the-art methods when using a more accurate evaluation metric.
Keywords:
Object detection
Detectors
Measurement
Three-dimensional displays
Task analysis
Shape
Convergence
Bounding box regression
implicit vector representation
Oriented object detection

Journal

IEEE Transactions on Multimedia cover
IEEE Transactions on Multimedia
IF:
9.7
Papers:
4.5K
Citations:
2.4W

Organization

U
University of Sydney
Scholars:
6.5W
Papers: 6.2W
Citations: 90
Y
Yunnan University
Scholars:
1.6W
Papers: 9.9K
Citations: 13