Robots detect underspecified reward features via demonstration variation and query targeted natural language explanations to improve reward recovery from imperfect demos.
Correcting robot plans with natural language feedback
2 Pith papers cite this work. Polarity classification is still indexing.
2
Pith papers citing it