arrow
返回

Task-Aware Semantic Map++: Cost-Efficient Task Assignment With Advanced Benchmark

delete2026-01-22
delete0
PRE
AI
D
Daewon Choi
S
Soeun Hwang
Y
Yoonseon Oh
DOI:10.1109/LRA.2026.3656794delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Enabling robots to perform diverse tasks autonomously requires a sophisticated semantic understanding of 3D scenes. However, conventional scene representations, which primarily rely on static attributes like visual information or object labels, have significant limitations in allowing robots to infer context-aware actions. We introduce Task-Aware Semantic Map++ (TASMap++), the framework that overcomes these limitations by constructing a map that assigns appropriate tasks to objects based on their holistic context. While prior work like TASMap pioneered this task-centric approach, it suffered from high computational costs and inaccuracies due to its reliance on single-frame analysis, which often fails to capture an object’s complete state. In contrast, TASMap++ resolves these issues with a multi-view synthesis pipeline that integrates multiple perspectives of an object for task assignment, resulting in significantly improved computational efficiency over its predecessor. Furthermore, to overcome biases in the existing TASMap evaluation, we established a reliable benchmark derived from the consensus of 32 participants across 231 cluttered scenes. On this benchmark, TASMap++ demonstrates superior accuracy over baselines. Finally, we introduce context-aware grounding, a paradigm distinct from conventional object grounding that relies on visual and spatial attributes. We present a downstream application of TASMap++ as a method to address this challenge and show experimentally that conventional grounding methods struggle in this setting, whereas TASMap++ is markedly more effective. To confirm these findings, the framework’s robustness and practicality were validated through extensive experiments on 3D indoor datasets, including real-world scan datasets.
Keyword:
Semantic scene understanding
mapping
AI-based methods

期刊

I
IEEE Robotics and Automation Letters
IF:
5.3
论文数:
1.7K
被引数:
3.9W

机构

H
Hanyang University
学者数:
439
论文数: 173
被引数: 0
引用论文

引用论文

ASHiTA: Automatic Scene-Grounded HIerarchical Task Analysis
err2025-06-10
err0
PREAI
errChang,Yun; Fermoselle,Leonor; Ta,Duy; Bucher,Bernadette; Carlone,Luca; Wang,Jiuguang
err分享
err收藏
An extensive experimental comparison of methods for multi-label learning
err2012-09-01
err554
PREAI
errMadjarov, Gjorgji; Kocev, Dragi; Gjorgjevikj, Dejan; Dzeroski, Saso
err分享
err收藏
LERF: Language Embedded Radiance Fields
err2023-10-01
err0
errOAAI
errJustin Kerr; Chung Min Kim; Ken Goldberg; Angjoo Kanazawa; Matthew Tancik
err分享
err收藏
ConceptFusion: Open-set multimodal 3D mapping
err2023-07-10
err0
errOAAI
errKrishna Jatavallabhula; Alihusein Kuwajerwala; Qiao Gu; Mohd Omama; Ganesh Iyer; Soroush Saryazdi; Tao Chen; Alaa Maalouf; Shuang Li; Nikhil Keetha; Ayush Tewari; Joshua Tenenbaum; Celso Melo; Madhava Krishna; Liam Paull; Florian Shkurti; Antonio Torralba
err分享
err收藏
学者 查看更多内容