REVIEW 2 cited by
VLM-driven Behavior Tree for Context-aware Task Planning
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
The use of Large Language Models (LLMs) for generating Behavior Trees (BTs) has recently gained attention in the robotics community, yet remains in its early stages of development. In this paper, we propose a novel framework that leverages Vision-Language Models (VLMs) to interactively generate and edit BTs that address visual conditions, enabling context-aware robot operations in visually complex environments. A key feature of our approach lies in the conditional control through self-prompted visual conditions. Specifically, the VLM generates BTs with visual condition nodes, where conditions are expressed as free-form text. Another VLM process integrates the text into its prompt and evaluates the conditions against real-world images during robot execution. We validated our framework in a real-world cafe scenario, demonstrating both its feasibility and limitations.
Forward citations
Cited by 2 Pith papers
-
A Generative Partially Specified Finite State Machine Approach to Complex Behaviour Planning
Generative FSM planning (GPSFSM/Fabric) lets LLMs write XML state-machine behaviour plans for ROS2 robots and outperforms BTGenBot on GPT models, but not on local models.
-
Automatic Robot Task Planning by Integrating Large Language Model with Genetic Programming
An LLM generates robot behavior trees that are filtered by fitness and then evolved by genetic programming, reaching good task plans in fewer generations than starting from random trees.
Discussion (0). Continue with ORCID to comment.