Pith. sign in

Paper Citation Record · LEDGER

LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2406.11815.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.11815 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:02:24.077240Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T15:18:33.791076Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 51d6c4cf-99bf-44fa-981d-a445625cb05c · inbound

TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies cites this paper.

TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 89

Resolution
verified exact
arxiv_id, observed 2026-05-15T18:27:22.981258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-15T18:27:22.760982Z digest=sha256:ed3d9e22b3540b3939f27b39bdd7ee37138cfa572974470ff443776f4363f701

Observation c6264bba-281b-41fd-b8b8-bb7f49518a39 · inbound

$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization cites this paper.

$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-22T18:05:00.919649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T18:02:23.305313Z digest=sha256:2defadefb216e2d9116957d743cd5a01899ddb5ede39f0f6d98baf1a3dfa4bf6

Observation c2df2264-8199-4720-ba16-74c35d79db47 · inbound

ManipLVM-R1: Reinforcement Learning for Reasoning in Embodied Manipulation with Large Vision-Language Models cites this paper.

ManipLVM-R1: Reinforcement Learning for Reasoning in Embodied Manipulation with Large Vision-Language Models LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-07T15:02:24.077240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:02:24.077240Z digest=sha256:716bcd1943623cbc357ad5cade75d0e3751c9e5857973821d9acef2a3d2712e0

Observation 9e66d2b2-3332-410e-b1d7-9f79c874c03c · inbound

Bridging Perception and Action: Spatially-Grounded Mid-Level Representations for Robot Generalization cites this paper.

Bridging Perception and Action: Spatially-Grounded Mid-Level Representations for Robot Generalization LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T06:04:16.013005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:04:16.013005Z digest=sha256:380c32d386964b80cb5608a745bade7175fea4da1e676c1d7230f62e6c04c058

Observation 0f393c74-1e1a-4441-9de2-6e647414a394 · inbound

ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning cites this paper.

ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-19T03:22:01.021432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T03:18:14.655384Z digest=sha256:b35573e50f16399d781d524a699691ebaaa1c4615125e4cb2c12c5222aec3135

Observation fe60effd-845e-43db-8772-d8a66a868c31 · inbound

RAGNet: Large-scale Reasoning-based Affordance Segmentation Benchmark towards General Grasping cites this paper.

RAGNet: Large-scale Reasoning-based Affordance Segmentation Benchmark towards General Grasping LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T10:32:25.140850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:32:25.140850Z digest=sha256:2d7b1979bd3386fe3b8f6edb970a3641145729288fe3c0551a24f699f9026cae

Observation dc6d773c-9e4a-4678-94a5-6d73163ec34e · inbound

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver cites this paper.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:45.390488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:45.390488Z digest=sha256:7ba5622b447edc5fddca4c00a8226892674a0221cceb92872462827eda4f08a9

Observation c674e204-38c9-4720-9e81-7012a735d468 · inbound

Ego-centric Predictive Model Conditioned on Hand Trajectories cites this paper.

Ego-centric Predictive Model Conditioned on Hand Trajectories LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T15:29:25.505592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:29:25.505592Z digest=sha256:a05ad0d8be54d9926b1c793e092b6e9a0932fcda846d10f3469f81ad5cc73143

Observation c72e8b14-500b-43d8-b202-a7f9555e8d68 · inbound

Towards Spatial Trace with Reasoning in Vision-Language Models for Robotics cites this paper.

Towards Spatial Trace with Reasoning in Vision-Language Models for Robotics LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-03T16:27:34.062828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:27:34.062828Z digest=sha256:a7fb0950b71bbce50c1ae76249c20e464c93d13c915873139b1b7bca43c41c8f

Observation 6beaaddc-ce38-4903-9f5a-de8fff1354f4 · inbound

PALM: Progress-Aware Policy Learning via Affordance Reasoning for Long-Horizon Robotic Manipulation cites this paper.

PALM: Progress-Aware Policy Learning via Affordance Reasoning for Long-Horizon Robotic Manipulation LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-05-16T15:08:01.994473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T15:05:21.907878Z digest=sha256:6ddc263cc0a8ff74ab1037cae03452dc38f59ac396bc9e77bf623ed289ad4f7c

Observation e2dbe256-8bf9-4d7e-8486-d813c3031b00 · inbound

Decompose and Recompose: Reasoning New Skills from Existing Abilities for Cross-Task Robotic Manipulation cites this paper.

Decompose and Recompose: Reasoning New Skills from Existing Abilities for Cross-Task Robotic Manipulation LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:56:06.795208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-09T14:30:00.890117Z digest=sha256:8089dde8587ade5bf2dcf8ecb4f3a85d9099e9443bbc2005312a4db3481d4bce

Observation 451547fd-6822-49fb-b9c4-ff8ecbc68ca4 · inbound

Beyond 3D VQAs: Injecting 3D Spatial Priors into Vision-Language Models for Enhanced Geometric Reasoning cites this paper.

Beyond 3D VQAs: Injecting 3D Spatial Priors into Vision-Language Models for Enhanced Geometric Reasoning LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:53:13.630877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T07:47:52.739735Z digest=sha256:286e78472634cf30651b16e8471827b3e67cc2a9cc0f3d05c92c13de34dd500d

Observation e064ffb5-c8a1-4853-8dcd-e25a536c8d55 · inbound

Grasp-Then-Plan with Failure Attribution: A Closed Two-Stage Framework for Precise and Generalizable Robotic Manipulation cites this paper.

Grasp-Then-Plan with Failure Attribution: A Closed Two-Stage Framework for Precise and Generalizable Robotic Manipulation LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 75

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T03:56:34.473901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T09:37:58.434897Z digest=sha256:fc39d2ed8fe01b5a927911947adfd2796e48b8dc9f49da20290689a8618b7282

Observation 53be3b28-1bf8-4a88-9266-7cde92c907ed · inbound

Trajectory-Level Redirection Attacks on Vision-Language-Action Models cites this paper.

Trajectory-Level Redirection Attacks on Vision-Language-Action Models LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:18:33.793178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T06:33:53.013076Z digest=sha256:510cad03b4ec7f7e03f6961d4246cf393697da7a922df104079d86383917a688

Observation e21013a9-57c6-4ceb-8be4-35d30bdc961b · inbound

Towards Human-like Physical Intelligence: Lifelong Vision-Language-Action Learning for Robotic Manipulation cites this paper.

Towards Human-like Physical Intelligence: Lifelong Vision-Language-Action Learning for Robotic Manipulation LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T00:57:42.768376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:57:42.768376Z digest=sha256:723ff5c150b03cca9dad23d090bff63208b1ebea3c5eebfd3d85b20aa9583e39

Observation 5335d7da-d38d-4ad2-b593-3db0b03e1cd0 · inbound

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation cites this paper.

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-01T14:39:40.274531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T14:39:40.274531Z digest=sha256:d62f94a80f8a472b413060091c154030b97766bbfc584768d846d79cc4f1a2ee