Pith. sign in

Paper Citation Record · LEDGER

A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2405.17418.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.17418 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:02:23.963963Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T13:05:45.043767Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e0f07e5a-4754-434c-b3ad-c1f867011397 · inbound

HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model cites this paper.

HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T22:00:48.968379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T22:00:48.667428Z digest=sha256:53b078f8a8bec6edff4925f26d913be38c7dc0a5a8427601738cdbe1b582dfcf

Observation 2cc89cae-e33d-429f-b495-13b8f6e381fc · inbound

ManipLVM-R1: Reinforcement Learning for Reasoning in Embodied Manipulation with Large Vision-Language Models cites this paper.

ManipLVM-R1: Reinforcement Learning for Reasoning in Embodied Manipulation with Large Vision-Language Models A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:02:23.963963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:02:23.963963Z digest=sha256:e7d5234813840582956b41ef05d534c63a1901fac489c7a14faeff860a90f79a

Observation 2f0ebc89-1577-484e-b049-91abb1db6d05 · inbound

Fast-in-Slow: A Dual-System Foundation Model Unifying Fast Manipulation within Slow Reasoning cites this paper.

Fast-in-Slow: A Dual-System Foundation Model Unifying Fast Manipulation within Slow Reasoning A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T11:34:16.484483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:34:16.484483Z digest=sha256:70145135a6a1507bcbff596b9c9c5ad5957a48a9e2478e1d385568a32d359db0

Observation 5c472cac-20d5-4e4e-a927-03706a803155 · inbound

Reinforcement Learning for Flow-Matching Policies cites this paper.

Reinforcement Learning for Flow-Matching Policies A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T15:47:41.911320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:47:41.911320Z digest=sha256:f9a53c782d190ba0be1121803f5f763be07e58ed2817b5403f5faa8bef46b4f1

Observation fb69eac5-e136-44f9-8330-3e71a657ff1a · inbound

AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models cites this paper.

AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T21:30:18.227523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T21:28:18.630934Z digest=sha256:a17b5cae4d83588a1ae447170d9c35978c7fa2bb088a091b30ad93d3e28fdbbd

Observation f6d0d885-b373-4912-8df6-9478e458394b · inbound

CycleVLA: Proactive Self-Correcting Vision-Language-Action Models via Subtask Backtracking and Minimum Bayes Risk Decoding cites this paper.

CycleVLA: Proactive Self-Correcting Vision-Language-Action Models via Subtask Backtracking and Minimum Bayes Risk Decoding A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-03T12:37:27.492757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:37:27.492757Z digest=sha256:0df3c7d9e4a3474c06a87abbea4709d65c3c1ab4a039714604fc88217ab06b5d

Observation 3243e82d-d122-4007-a230-2589704edc41 · inbound

The Latent Color Subspace: Emergent Order in High-Dimensional Chaos cites this paper.

The Latent Color Subspace: Emergent Order in High-Dimensional Chaos A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-14T22:24:08.310224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T22:24:08.310224Z digest=sha256:b139f835deb99db3e7d109ca85d5c2e73c65621f335ac692e033061186f97d6f

Observation 241cb840-a15f-4c70-aad4-53589ef4e32a · inbound

Sentinel-VLA: A Metacognitive VLA Model with Active Status Monitoring for Dynamic Reasoning and Error Recovery cites this paper.

Sentinel-VLA: A Metacognitive VLA Model with Active Status Monitoring for Dynamic Reasoning and Error Recovery A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:46:04.903328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-09T15:14:03.865035Z digest=sha256:5c84558a80b458e927dade13cf28d96582aead4d7d252d17dc7c346b33c4d262

Observation f9291fae-bfa5-4287-87a8-e78dad21e8f8 · inbound

Sentinel-VLA: A Metacognitive VLA Model with Active Status Monitoring for Dynamic Reasoning and Error Recovery cites this paper.

Sentinel-VLA: A Metacognitive VLA Model with Active Status Monitoring for Dynamic Reasoning and Error Recovery A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T13:05:45.046082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-01T01:10:48.111387Z digest=sha256:d10870eac9d1cda00413ece49c16cfc84739e02ad6e8bfbf37b192987472e731

Observation b3f8ac2d-30d1-402a-824a-f9a9eea32c92 · inbound

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation cites this paper.

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:11:27.367232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T03:36:24.941205Z digest=sha256:b746687692b0031db5860ad3e7055fd895c08c2d94369a22103695a7704d9e99

Observation d92c9ae3-9b50-4c62-bab0-2f759da5d721 · inbound

VISOR: A Vision-Language Model-based Test Oracle for Testing Robots cites this paper.

VISOR: A Vision-Language Model-based Test Oracle for Testing Robots A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:21:31.544358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T05:17:03.803350Z digest=sha256:278ec23df5f7872d66ab27164437e0e299a10b5f509a201307b9759bf98c8635

Observation 54ab0ffb-ebc6-4b14-8e26-05d8b4256dd5 · inbound

VISOR: A Vision-Language Model-based Test Oracle for Testing Robots cites this paper.

VISOR: A Vision-Language Model-based Test Oracle for Testing Robots A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:29:09.871897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T22:24:25.871817Z digest=sha256:823ded27c2895033ffd9b56e8b1370a449583e1d29cc46449ae08232641d7a98

Observation 1cc93585-4968-42d8-88b2-8af218133987 · inbound

General Covariant Action Modeling: Constructing Generalized Manifolds via Spatio-Temporal Decoupling cites this paper.

General Covariant Action Modeling: Constructing Generalized Manifolds via Spatio-Temporal Decoupling A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 198

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:33:27.701922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-29T13:33:03.368006Z digest=sha256:343bf2d7ea52e5234b86e2177b9a60e631680caed2e6339c780b5382d93e3183

Observation 839ce8f3-afe6-41b9-9a6c-2f203e529d79 · inbound

Token-Wise Latent Streaming from Slow Reasoners to Fast Planners for Dynamic Vision Language Navigation cites this paper.

Token-Wise Latent Streaming from Slow Reasoners to Fast Planners for Dynamic Vision Language Navigation A Self-Correcting Vision-Language-Action Model for Fast and Slow System Manipulation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T19:56:56.910815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T19:56:56.910815Z digest=sha256:97e50cfc646c8b505c3c328d6d116b1a51946b1dba3c8f041e8ddfa13fbffdef