REVIEW 5 cited by
Learning Robot Soccer from Egocentric Vision with Deep Reinforcement Learning
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
We apply multi-agent deep reinforcement learning (RL) to train end-to-end robot soccer policies with fully onboard computation and sensing via egocentric RGB vision. This setting reflects many challenges of real-world robotics, including active perception, agile full-body control, and long-horizon planning in a dynamic, partially-observable, multi-agent domain. We rely on large-scale, simulation-based data generation to obtain complex behaviors from egocentric vision which can be successfully transferred to physical robots using low-cost sensors. To achieve adequate visual realism, our simulation combines rigid-body physics with learned, realistic rendering via multiple Neural Radiance Fields (NeRFs). We combine teacher-based multi-agent RL and cross-experiment data reuse to enable the discovery of sophisticated soccer strategies. We analyze active-perception behaviors including object tracking and ball seeking that emerge when simply optimizing perception-agnostic soccer play. The agents display equivalent levels of performance and agility as policies with access to privileged, ground-truth state. To our knowledge, this paper constitutes a first demonstration of end-to-end training for multi-agent robot soccer, mapping raw pixel observations to joint-level actions, that can be deployed in the real world. Videos of the game-play and analyses can be seen on our website https://sites.google.com/view/vision-soccer .
Forward citations
Cited by 5 Pith papers
-
Perceiving and Acting in First-Person: A Dataset and Benchmark for Egocentric Human-Object-Human Interactions
Claimed first large-scale egocentric and multi-view dataset of human-object-human assistance (11.4 hours, 1.2M frames) with three benchmarks; only the abstract was assessable because the submitted body text is a diffe...
-
Toward Real-World Cooperative and Competitive Soccer with Quadrupedal Robot Teams
Hierarchical MARL with fictitious self-play trains quadruped soccer teams in simulation and transfers them zero-shot to real robots, enabling onboard, decentralized 1v1 and 2v1 soccer with emergent passing and role al...
-
Emergent Active Perception and Dexterity of Simulated Humanoids from Visual Reinforcement Learning
PDC trains a single egocentric-vision policy that lets a simulated humanoid search for, grasp, and place objects and open drawers without privileged state information.
-
SoccerDiffusion: Toward Learning End-to-End Humanoid Robot Soccer from Gameplay Recordings
An end-to-end transformer diffusion policy, distilled to one inference step, reproduces low-level humanoid soccer behaviors from real RoboCup game recordings but lacks high-level tactical behavior.
-
Zero-Shot Reinforcement Learning Under Partial Observability
Behavior foundation models with GRU memory outperform memory-free zero-shot RL baselines in most partially observable ExORL settings, but the advantage is inconsistent on Cheetah.
Discussion (0). Continue with ORCID to comment.