Pith. sign in

REVIEW 1 cited by

Reinforcement Learning Within the Classical Robotics Stack: A Case Study in Robot Soccer

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2412.09417 v2 pith:Z77DXAJE submitted 2024-12-12 cs.RO cs.AIcs.LG

classification cs.ROcs.AIcs.LG
keywords approacharchitecturechallengelearningrobotbehaviorclassicaldecision-making
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Robot decision-making in partially observable, real-time, dynamic, and multi-agent environments remains a difficult and unsolved challenge. Model-free reinforcement learning (RL) is a promising approach to learning decision-making in such domains, however, end-to-end RL in complex environments is often intractable. To address this challenge in the RoboCup Standard Platform League (SPL) domain, we developed a novel architecture integrating RL within a classical robotics stack, while employing a multi-fidelity sim2real approach and decomposing behavior into learned sub-behaviors with heuristic selection. Our architecture led to victory in the 2024 RoboCup SPL Challenge Shield Division. In this work, we fully describe our system's architecture and empirically analyze key design decisions that contributed to its success. Our approach demonstrates how RL-based behaviors can be integrated into complete robot behavior architectures.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Toward Real-World Cooperative and Competitive Soccer with Quadrupedal Robot Teams

    cs.RO 2025-05 conditional novelty 6.0 of 10

    Hierarchical MARL with fictitious self-play trains quadruped soccer teams in simulation and transfers them zero-shot to real robots, enabling onboard, decentralized 1v1 and 2v1 soccer with emergent passing and role al...

Pith tools