REVIEW 5 cited by
Body Transformer: Leveraging Robot Embodiment for Policy Learning
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
In recent years, the transformer architecture has become the de facto standard for machine learning algorithms applied to natural language processing and computer vision. Despite notable evidence of successful deployment of this architecture in the context of robot learning, we claim that vanilla transformers do not fully exploit the structure of the robot learning problem. Therefore, we propose Body Transformer (BoT), an architecture that leverages the robot embodiment by providing an inductive bias that guides the learning process. We represent the robot body as a graph of sensors and actuators, and rely on masked attention to pool information throughout the architecture. The resulting architecture outperforms the vanilla transformer, as well as the classical multilayer perceptron, in terms of task completion, scaling properties, and computational efficiency when representing either imitation or reinforcement learning policies. Additional material including the open-source code is available at https://sferrazza.cc/bot_site.
Forward citations
Cited by 5 Pith papers
-
UMI-on-Air: Embodiment-Aware Guidance for Embodiment-Agnostic Visuomotor Policies
Embodiment-Aware Diffusion Policy steers a UMI-trained diffusion policy with controller tracking-cost gradients at inference time, improving aerial manipulation success in simulation and real flights.
-
Arnold: a generalist muscle transformer policy
A single transformer policy with a compositional sensorimotor vocabulary achieves expert or super-expert performance on 14 musculoskeletal control tasks spanning four embodiments.
-
AnyBody: A Benchmark Suite for Cross-Embodiment Manipulation
AnyBody is a benchmark suite that tests cross-embodiment manipulation generalization along interpolation, extrapolation, and composition axes, and finds zero-shot generalization to unseen robot bodies remains difficult.
-
SPI-BoTER: Error Compensation for Industrial Robots via Sparse Attention Masking and Hybrid Loss with Spatial-Physical Information
SPI-BoTER, a physics-informed Transformer with sparse attention masks and a distance-matrix loss, reports 0.2515 mm mean 3D positioning error for UR5 robot arm error compensation, a 35.16% reduction over a standard DNN.
-
McARL:Morphology-Control-Aware Reinforcement Learning for Generalizable Quadrupedal Locomotion
A morphology-conditioned PPO policy trained only on the Unitree Go1 transfers zero-shot in simulation to Go2, A1 and Mini Cheetah, with the best variant reaching 3.5 m/s on the Go2.
Discussion (0). Sign in to comment.