REVIEW 2 cited by
Think, Act, and Ask: Open-World Interactive Personalized Robot Navigation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Zero-Shot Object Navigation (ZSON) enables agents to navigate towards open-vocabulary objects in unknown environments. The existing works of ZSON mainly focus on following individual instructions to find generic object classes, neglecting the utilization of natural language interaction and the complexities of identifying user-specific objects. To address these limitations, we introduce Zero-shot Interactive Personalized Object Navigation (ZIPON), where robots need to navigate to personalized goal objects while engaging in conversations with users. To solve ZIPON, we propose a new framework termed Open-woRld Interactive persOnalized Navigation (ORION), which uses Large Language Models (LLMs) to make sequential decisions to manipulate different modules for perception, navigation and communication. Experimental results show that the performance of interactive agents that can leverage user feedback exhibits significant improvement. However, obtaining a good balance between task completion and the efficiency of navigation and interaction remains challenging for all methods. We further provide more findings on the impact of diverse user feedback forms on the agents' performance. Code is available at https://github.com/sled-group/navchat.
Forward citations
Cited by 2 Pith papers
-
AROMA: Mixed-Initiative AI Assistance for Non-Visual Cooking by Grounding Multi-modal Information Between Reality and Videos
AROMA pairs a blind cook's spoken descriptions of what they feel, smell, and taste with a wearable camera and a video recipe to answer questions and issue proactive alerts, and eight participants rated it usable despi...
-
AmbiK: Dataset of Ambiguous Tasks in Kitchen Environment
AmbiK is a human-validated, text-only benchmark of 1000 ambiguous kitchen tasks paired with unambiguous counterparts, on which current ambiguity detection methods perform poorly.
Discussion (0). Sign in to comment.