Pith. sign in

Paper Citation Record · LEDGER

VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2502.13508.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.13508 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T22:13:14.942425Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T17:30:00.525942Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 9e98d851-58a3-499e-ae3a-cf331b776022 · inbound

UP-VLA: A Unified Understanding and Prediction Model for Embodied Agent cites this paper.

UP-VLA: A Unified Understanding and Prediction Model for Embodied Agent VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T22:13:14.942425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T22:13:14.942425Z digest=sha256:5560f4797adb3897debc485de6830f5eb2d65a85a6ac3299e9f4d01a706d1e99

Observation 387c0b76-0387-4d93-8362-cc628d7755df · inbound

DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control cites this paper.

DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:48:48.994403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T19:48:48.725800Z digest=sha256:84b97445f0fa5c96479228e74fcfb094d45810af4701b9db7247399ba73ffcb8

Observation ef8e333e-3baf-4787-8619-ad432f15f447 · inbound

VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers cites this paper.

VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T21:07:25.936596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:07:25.936596Z digest=sha256:d37ccd7ff351bf55121975b49dd7dfc5800f1b549bf486448040d426f14d7e40

Observation 5e1a6cb3-5c40-453d-b33e-088a8634317a · inbound

Bring My Cup! Personalizing Vision-Language-Action Models with Visual Attentive Prompting cites this paper.

Bring My Cup! Personalizing Vision-Language-Action Models with Visual Attentive Prompting VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T14:34:25.882971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T14:34:25.882971Z digest=sha256:3a8bfdd9ab8176d8916e7a998c4510deecfa72c5c7fc9a81988ca636fdd655ac

Observation e7d6c07d-6368-4be3-8a11-877a77ba460e · inbound

PALM: Progress-Aware Policy Learning via Affordance Reasoning for Long-Horizon Robotic Manipulation cites this paper.

PALM: Progress-Aware Policy Learning via Affordance Reasoning for Long-Horizon Robotic Manipulation VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 149

Resolution
verified exact
arxiv_id, observed 2026-05-16T15:08:02.092746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T15:05:21.907878Z digest=sha256:41845530e48f9fcae26cf268dcc1d948ce746a3ecacd7f946a70aa4bbdddb60b

Observation 5bdac23c-6650-4273-aff4-9e98bbd7a717 · inbound

ThermoAct:Thermal-Aware Vision-Language-Action Models for Robotic Perception and Decision-Making cites this paper.

ThermoAct:Thermal-Aware Vision-Language-Action Models for Robotic Perception and Decision-Making VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:48:24.841267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T00:46:03.360770Z digest=sha256:ce828f6c7aedffdaaafa943d6c0c3f2d9130b78973d1381178e1fa7c5c9ff7e4

Observation 31c068bb-f719-4f21-997a-637e5a176494 · inbound

dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model cites this paper.

dWorldEval: Scalable Robotic Policy Evaluation via Discrete Diffusion World Model VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:31:09.582834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T11:45:18.081248Z digest=sha256:21d9d5b5e647c30f4d84f96ab9d9af9333534a3aab6f4098d104d9b0276c221a

Observation 34a23613-c686-48b8-b1c4-34a9376162dd · inbound

UniSteer: Unified Noise Steering for Efficient Human-Guided VLA Adaptation cites this paper.

UniSteer: Unified Noise Steering for Efficient Human-Guided VLA Adaptation VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:36:25.676152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T04:08:43.222818Z digest=sha256:25763e697c5c115bf0d3538d7504ff595701297218415db260d9dd1b51c626fd

Observation db8ae8a9-2adb-461a-a5bd-d466c803c1e2 · inbound

UniSteer: Unified Noise Steering for Efficient Human-Guided VLA Adaptation cites this paper.

UniSteer: Unified Noise Steering for Efficient Human-Guided VLA Adaptation VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T14:25:59.725499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:25:59.725499Z digest=sha256:c7075519c4f9db6f128198949361fddd45683b6cb1feee1d6ac09795673c823b

Observation 12c552b4-9e30-47a5-9f6e-02907289948b · inbound

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation cites this paper.

MuseVLA: An Adaptive Multimodal Sensing Vision-Language-Action Model for Robotic Manipulation VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:08:58.769029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T00:49:13.291897Z digest=sha256:eec3acc444e48f73e0dd51d788bbf7f8a8693469ace83068e2c94b8a23fb728b

Observation d9c2a461-1dfc-4374-b928-9b2f0ea7cf50 · inbound

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation cites this paper.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:30:00.527553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-25T23:39:14.779122Z digest=sha256:21c102956ad85412f6c4a9a64c9336f3c4954947bc1b18d460d8b3e8e302e80f

Observation d0f4b36f-b5a1-4bcb-ae7a-03fd37fc19de · inbound

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation cites this paper.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:d8299a954de907f19ed3a7376558219cb73bfd56465cb84e0db0977d34412812

Observation 2cda399c-9843-416e-91da-f3a99b59713c · inbound

Differential Amplifier-Inspired AmpAttention for Multi-View Robotic Manipulation cites this paper.

Differential Amplifier-Inspired AmpAttention for Multi-View Robotic Manipulation VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-12T06:39:16.303751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:39:16.303751Z digest=sha256:002253b6f3f799259a5d69b43bd0efeeafc5a94b8c7d78ddacd39604b5168d94