arrow
返回

Query by Example: Semantic Traffic Scene Retrieval Using LLM-Based Scene Graph Representation

delete2025-04-17
delete0
delete
OA
AI
Y
Yafu Tian *
A
Alexander Carballo
R
Ruifeng Li
S
Simon Thompson
K
Kazuya Takeda
DOI:10.3390/s25082546delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
In autonomous driving, retrieving a specific traffic scene in huge datasets is a significant challenge. Traditional scene retrieval methods struggle to cope with the semantic complexity and heterogeneity of traffic scenes and are unable to meet the variable needs of different users. This paper proposes Query-by-Example, a traffic scene retrieval approach based on Visual-Large Language Model (VLM)-generated Road Scene Graph (RSG) representation. Our method uses VLMs to generate structured scene graphs from video data, capturing high-level semantic attributes and detailed object relationships in traffic scenes. We introduce an extensible set of scene attributes and a graph-based scene description to quantify scene similarity. We also propose a RSG-LLM benchmark dataset containing 1000 traffic scenes, their corresponding natural language descriptions, and RSGs to evaluate the performance of LLMs in generating RSGs. Experiments show that our method can effectively retrieve semantically similar traffic scenes from large databases, supporting various query formats, including natural language, images, video clips, rosbag, etc. Our method provides a comprehensive and flexible framework for traffic scene retrieval, promoting its application in autonomous driving systems.
Keyword:
traffic scene retrieval
visual LLMs
subgraph isomorphism matching
scene graph
query by example
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Sensors 封面图
Sensors
IF:
3.5
论文数:
7.2W
被引数:
20.9W

机构

H
harbin institute of technology
学者数:
8.0W
论文数: 6.6W
被引数: 66
N
Nagoya University
学者数:
3.3W
论文数: 2.5W
被引数: 2.6W
引用论文

引用论文

暂无论文信息