arrow
返回

Semi-Open Set Object Detection Algorithm Leveraged by Multi-Modal Large Language Models

delete2024-11-29
delete0
delete
OA
AI
K
Kewei Wu
Y
Yiran Wang
X
Xiaogang He
J
Jinyu Yan
Y
Yang Guo
Z
Zhuqing Jiang
X
Xing Zhang
W
Wei Wang
Y
Yongping Xiong
A
Aidong Men
X
Xiao, Li *
DOI:10.3390/bdcc8120175delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Currently, closed-set object detection models represented by YOLO are widely deployed in the industrial field. However, such closed-set models lack sufficient tuning ability for easily confused objects in complex detection scenarios. Open-set object detection models such as GroundingDINO expand the detection range to a certain extent, but they still have a gap in detection accuracy compared with closed-set detection models and cannot meet the requirements for high-precision detection in practical applications. In addition, existing detection technologies are also insufficient in interpretability, making it difficult to clearly show users the basis and process of judgment of detection results, causing users to have doubts about the trust and application of detection results. Based on the above deficiencies, we propose a new object detection algorithm based on multi-modal large language models that significantly improves the detection effect of closed-set object detection models for more difficult boundary tasks while ensuring detection accuracy, thereby achieving a semi-open set object detection algorithm. It has significant improvements in accuracy and interpretability under the verification of seven common traffic and safety production scenarios.
Keyword:
large-scale foundation model
computer vision
object detection

期刊

B
Big Data and Cognitive Computing
IF:
4.4
论文数:
1.3K
被引数:
2.4K

机构

B
beijing university of posts & telecommunications
学者数:
1.4W
论文数: 1.2W
被引数: 9
引用论文

引用论文

Comparative analysis of occlusion methods for artificial sphincters人工括约肌闭塞方法的比较分析
err2020-04-07
err0
PREAI
errLeonardo Marziale; Gioia Lucarini; Tommaso Mazzocchi; Leonardo Ricotti; Arianna Menciassi
err分享
err收藏
Selective Search for Object Recognition
err2013-04-02
err3.9K
PREAI
errUijlings, J. R. R.; van de Sande, K. E. A.; Gevers, T.; Smeulders, A. W. M.
err分享
err收藏
Oxidation‑reduction potential parameters worsen following intraarterial therapy in patients with reduced collateral circulation and middle cerebral artery occlusions
err2023-05-05
err0
errOAAI
errBenjamin Atchie; Stephanie Jarvis; Richard Bellon; Trevor Barton; Lauren Disalvo; Kristin Salottolo; Raphael Bar‑Or; David Bar‑or
err分享
err收藏
Haemostasis, Thrombosis, and Endothelium in Behcet's Disease
err1998-05-22
err0
PREAI
errIbrahim C. Haznedaroglu; İsmail Çelik; Yahya Büyükaşık; Ali Koşar; Şerafettin Kirazlı; Semra V. Dündar
err分享
err收藏
Decrease of regional cerebral blood flow in liver cirrhosis
err2000-09-01
err0
PREAI
errMotoh Iwasa; Kaname Matsumura; Masahiko Kaito; Jiro Ikoma; Yoshinao Kobayashi; Naoki Nakagawa; Shozo Watanabe; Kan Takeda; Yukihiko Adachi
err分享
err收藏
cIBR Effectively Targets Nanoparticles to LFA-1 on Acute Lymphoblastic T Cells
err2009-11-17
err0
errOAAI
errChuda Chittasupho; Prakash Manikwar; Jeffrey P. Krise; Teruna J. Siahaan; Cory Berkland
err分享
err收藏
Object Detection in 20 Years: A Survey20年来的目标检测: 一项调查
err2023-03-01
err812
errOAAI
errZou, Zhengxia; Chen, Keyan; Shi, Zhenwei; Guo, Yuhong; Ye, Jieping
err分享
err收藏
学者 查看更多内容