REVIEW 4 cited by
SayTap: Language to Quadrupedal Locomotion
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Large language models (LLMs) have demonstrated the potential to perform high-level planning. Yet, it remains a challenge for LLMs to comprehend low-level commands, such as joint angle targets or motor torques. This paper proposes an approach to use foot contact patterns as an interface that bridges human commands in natural language and a locomotion controller that outputs these low-level commands. This results in an interactive system for quadrupedal robots that allows the users to craft diverse locomotion behaviors flexibly. We contribute an LLM prompt design, a reward function, and a method to expose the controller to the feasible distribution of contact patterns. The results are a controller capable of achieving diverse locomotion patterns that can be transferred to real robot hardware. Compared with other design choices, the proposed approach enjoys more than 50% success rate in predicting the correct contact patterns and can solve 10 more tasks out of a total of 30 tasks. Our project site is: https://saytap.github.io.
Forward citations
Cited by 4 Pith papers
-
GhostShell: Streaming LLM Function Calls for Concurrent Embodied Programming
A streaming XML function-token interface with multi-channel scheduling lets robots execute concurrent speech and motion while the LLM is still generating, reportedly beating native function calling 15/15 vs 6/15 on co...
-
Discovery of skill switching criteria for learning agile quadruped locomotion
A hierarchical reinforcement learning framework lets a quadruped robot automatically switch between trotting, bounding, galloping, and fall recovery based on distance to the goal, with switch distances tuned by CMA-ES.
-
Efficiently Generating Expressive Quadruped Behaviors via Language-Guided Preference Learning
A hybrid method uses LLM-generated candidate gaits and then refines them with a few human preference rankings, achieving quadruped behaviors aligned with user intent in as few as four queries.
-
Taking Flight with Dialogue: Enabling Natural Language Control for PX4-based Drone Agent
An open-source ROS2/PX4 framework using locally hosted LLMs and VLMs enables natural language drone commands, with the best simulated mission success rate at 40%.
Discussion (0). Continue with ORCID to comment.