Pith. sign in

REVIEW 3 cited by

Hierarchical Few-Shot Imitation with Skill Transition Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2107.08981 v2 pith:PDKPA2N6 submitted 2021-07-19 cs.LG cs.AIcs.RO

classification cs.LGcs.AIcs.RO
keywords tasksunseenskillfistimitationagentsbehavioraldata
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

A desirable property of autonomous agents is the ability to both solve long-horizon problems and generalize to unseen tasks. Recent advances in data-driven skill learning have shown that extracting behavioral priors from offline data can enable agents to solve challenging long-horizon tasks with reinforcement learning. However, generalization to tasks unseen during behavioral prior training remains an outstanding challenge. To this end, we present Few-shot Imitation with Skill Transition Models (FIST), an algorithm that extracts skills from offline data and utilizes them to generalize to unseen tasks given a few downstream demonstrations. FIST learns an inverse skill dynamics model, a distance function, and utilizes a semi-parametric approach for imitation. We show that FIST is capable of generalizing to new tasks and substantially outperforms prior baselines in navigation experiments requiring traversing unseen parts of a large maze and 7-DoF robotic arm experiments requiring manipulating previously unseen objects in a kitchen.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Extracting Visual Plans from Unlabeled Videos via Symbolic Guidance

    cs.RO 2025-05 conditional novelty 6.0 of 10

    Vis2Plan extracts object symbols from unlabeled play videos with vision models, plans symbolically with A* search, and retrieves reachable real images as subgoals for a goal-conditioned robot policy.

  2. NBDI: A Simple and Effective Termination Condition for Skill Extraction from Task-Agnostic Demonstrations

    cs.LG 2025-01 conditional novelty 6.0 of 10

    NBDI learns skill termination from state-action novelty (ICM prediction error) on task-agnostic demonstrations, improving downstream RL performance in maze and manipulation benchmarks.

  3. Bridging Perception and Action: Spatially-Grounded Mid-Level Representations for Robot Generalization

    cs.RO 2025-06 conditional novelty 5.0 of 10

    A mixture-of-experts diffusion policy conditioned on object, pose, depth, and trajectory mid-level representations is reported to outperform language-only and representation-free baselines on bimanual dexterous tasks,...

Pith tools