arrow
Return

Task-Aware Semantic Map++: Cost-Efficient Task Assignment With Advanced Benchmark

delete2026-01-22
delete0
PRE
AI
D
Daewon Choi
S
Soeun Hwang
Y
Yoonseon Oh
DOI:10.1109/LRA.2026.3656794delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Enabling robots to perform diverse tasks autonomously requires a sophisticated semantic understanding of 3D scenes. However, conventional scene representations, which primarily rely on static attributes like visual information or object labels, have significant limitations in allowing robots to infer context-aware actions. We introduce Task-Aware Semantic Map++ (TASMap++), the framework that overcomes these limitations by constructing a map that assigns appropriate tasks to objects based on their holistic context. While prior work like TASMap pioneered this task-centric approach, it suffered from high computational costs and inaccuracies due to its reliance on single-frame analysis, which often fails to capture an object’s complete state. In contrast, TASMap++ resolves these issues with a multi-view synthesis pipeline that integrates multiple perspectives of an object for task assignment, resulting in significantly improved computational efficiency over its predecessor. Furthermore, to overcome biases in the existing TASMap evaluation, we established a reliable benchmark derived from the consensus of 32 participants across 231 cluttered scenes. On this benchmark, TASMap++ demonstrates superior accuracy over baselines. Finally, we introduce context-aware grounding, a paradigm distinct from conventional object grounding that relies on visual and spatial attributes. We present a downstream application of TASMap++ as a method to address this challenge and show experimentally that conventional grounding methods struggle in this setting, whereas TASMap++ is markedly more effective. To confirm these findings, the framework’s robustness and practicality were validated through extensive experiments on 3D indoor datasets, including real-world scan datasets.
Keywords:
Semantic scene understanding
mapping
AI-based methods

Journal

I
IEEE Robotics and Automation Letters
IF:
5.3
Papers:
1.8K
Citations:
3.9W

Organization

H
Hanyang University
Scholars:
439
Papers: 173
Citations: 0
Cited Papers

Cited Papers

ASHiTA: Automatic Scene-Grounded HIerarchical Task Analysis
err2025-06-10
err0
PREAI
errChang,Yun; Fermoselle,Leonor; Ta,Duy; Bucher,Bernadette; Carlone,Luca; Wang,Jiuguang
errShare
errSave
Habitat Synthetic Scenes Dataset (HSSD-200): An Analysis of 3D Scene Scale and Realism Tradeoffs for ObjectGoal Navigation
err2024-06-16
err0
PREAI
errMukul Khanna; Yongsen Mao; Hanxiao Jiang; Sanjay Haresh; Brennan Shacklett; Dhruv Batra; Alexander Clegg; Eric Undersander; Angel X. Chang; Manolis Savva
errShare
errSave
An extensive experimental comparison of methods for multi-label learning
err2012-09-01
err554
PREAI
errMadjarov, Gjorgji; Kocev, Dragi; Gjorgjevikj, Dejan; Dzeroski, Saso
errShare
errSave
Visual Programming for Zero-Shot Open-Vocabulary 3D Visual Grounding
err2024-06-16
err0
PREAI
errZhihao Yuan; Jinke Ren; Chun-Mei Feng; Hengshuang Zhao; Shuguang Cui; Zhen Li
errShare
errSave
LERF: Language Embedded Radiance Fields
err2023-10-01
err0
errOAAI
errJustin Kerr; Chung Min Kim; Ken Goldberg; Angjoo Kanazawa; Matthew Tancik
errShare
errSave
ConceptFusion: Open-set multimodal 3D mapping
err2023-07-10
err0
errOAAI
errKrishna Jatavallabhula; Alihusein Kuwajerwala; Qiao Gu; Mohd Omama; Ganesh Iyer; Soroush Saryazdi; Tao Chen; Alaa Maalouf; Shuang Li; Nikhil Keetha; Ayush Tewari; Joshua Tenenbaum; Celso Melo; Madhava Krishna; Liam Paull; Florian Shkurti; Antonio Torralba
errShare
errSave
researcher View more