Pith. sign in

Learning to Play by Imitating Humans

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

Acquiring multiple skills has commonly involved collecting a large number of expert demonstrations per task or engineering custom reward functions. Recently it has been shown that it is possible to acquire a diverse set of skills by self-supervising control on top of human teleoperated play data. Play is rich in state space coverage and a policy trained on this data can generalize to specific tasks at test time outperforming policies trained on individual expert task demonstrations. In this work, we explore the question of whether robots can learn to play to autonomously generate play data that can ultimately enhance performance. By training a behavioral cloning policy on a relatively small quantity of human play, we autonomously generate a large quantity of cloned play data that can be used as additional training. We demonstrate that a general purpose goal-conditioned policy trained on this augmented dataset substantially outperforms one trained only with the original human data on 18 difficult user-specified manipulation tasks in a simulated robotic tabletop environment. A video example of a robot imitating human play can be seen here: https://learning-to-play.github.io/videos/undirected_play1.mp4

fields

cs.RO 1

years

2025 1

verdicts

CONDITIONAL 1

representative citing papers

Efficient Sensorimotor Learning for Open-world Robot Manipulation

cs.RO · 2025-05-07 · conditional · novelty 4.0

A PhD dissertation argues that object, spatial, and behavioral regularities, extracted with foundation models, enable data-efficient, generalizable robot manipulation, and presents seven systems and a benchmark built on that idea.

citing papers explorer

Showing 1 of 1 citing paper.

  • Efficient Sensorimotor Learning for Open-world Robot Manipulation cs.RO · 2025-05-07 · conditional · none · ref 67 · internal anchor

    A PhD dissertation argues that object, spatial, and behavioral regularities, extracted with foundation models, enable data-efficient, generalizable robot manipulation, and presents seven systems and a benchmark built on that idea.