Pith. sign in

REVIEW 1 cited by

DoRO: Disambiguation of referred object for embodied agents

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2207.14205 v1 pith:WYYM3PZB submitted 2022-07-28 cs.RO cs.AI

classification cs.ROcs.AI
keywords objectambiguitydororeferredrobottaskareagrounding
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Robotic task instructions often involve a referred object that the robot must locate (ground) within the environment. While task intent understanding is an essential part of natural language understanding, less effort is made to resolve ambiguity that may arise while grounding the task. Existing works use vision-based task grounding and ambiguity detection, suitable for a fixed view and a static robot. However, the problem magnifies for a mobile robot, where the ideal view is not known beforehand. Moreover, a single view may not be sufficient to locate all the object instances in the given area, which leads to inaccurate ambiguity detection. Human intervention is helpful only if the robot can convey the kind of ambiguity it is facing. In this article, we present DoRO (Disambiguation of Referred Object), a system that can help an embodied agent to disambiguate the referred object by raising a suitable query whenever required. Given an area where the intended object is, DoRO finds all the instances of the object by aggregating observations from multiple views while exploring & scanning the area. It then raises a suitable query using the information from the grounded object instances. Experiments conducted with the AI2Thor simulator show that DoRO not only detects the ambiguity more accurately but also raises verbose queries with more accurate information from the visual-language grounding.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. AmbiK: Dataset of Ambiguous Tasks in Kitchen Environment

    cs.LG 2025-06 conditional novelty 6.0 of 10

    AmbiK is a human-validated, text-only benchmark of 1000 ambiguous kitchen tasks paired with unambiguous counterparts, on which current ambiguity detection methods perform poorly.

Pith tools