Pith. sign in

Paper Citation Record · LEDGER

LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2406.11815.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.11815 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:02:24.077240Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T15:18:33.791076Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 51d6c4cf-99bf-44fa-981d-a445625cb05c · inbound

TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies cites this paper.

TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 89

Resolution
verified exact
arxiv_id, observed 2026-05-15T18:27:22.981258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-15T18:27:22.760982Z digest=sha256:a3f60e3d5896e30d4b753db843299f1bff9a1c2b96e2334ac474554e31699317

Observation c6264bba-281b-41fd-b8b8-bb7f49518a39 · inbound

$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization cites this paper.

$\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-22T18:05:00.919649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T18:02:23.305313Z digest=sha256:bd13a9d4beef76f2398d0aa2e8133df06f2d18c72d17d8f69f8f2aedf4ffebc7

Observation c2df2264-8199-4720-ba16-74c35d79db47 · inbound

ManipLVM-R1: Reinforcement Learning for Reasoning in Embodied Manipulation with Large Vision-Language Models cites this paper.

ManipLVM-R1: Reinforcement Learning for Reasoning in Embodied Manipulation with Large Vision-Language Models LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-07T15:02:24.077240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:02:24.077240Z digest=sha256:189d29e5385bd89d61673b9c91f78c01f1a3eae0b3336f65ee4f4a1719ad3c1f

Observation 9e66d2b2-3332-410e-b1d7-9f79c874c03c · inbound

Bridging Perception and Action: Spatially-Grounded Mid-Level Representations for Robot Generalization cites this paper.

Bridging Perception and Action: Spatially-Grounded Mid-Level Representations for Robot Generalization LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T06:04:16.013005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:04:16.013005Z digest=sha256:380c32d386964b80cb5608a745bade7175fea4da1e676c1d7230f62e6c04c058

Observation 0f393c74-1e1a-4441-9de2-6e647414a394 · inbound

ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning cites this paper.

ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-19T03:22:01.021432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T03:18:14.655384Z digest=sha256:e354f464b888b06e761d448bc6393843e0d29cfb7483df43d4a88bce7b8bc9aa

Observation fe60effd-845e-43db-8772-d8a66a868c31 · inbound

RAGNet: Large-scale Reasoning-based Affordance Segmentation Benchmark towards General Grasping cites this paper.

RAGNet: Large-scale Reasoning-based Affordance Segmentation Benchmark towards General Grasping LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T10:32:25.140850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:32:25.140850Z digest=sha256:2d7b1979bd3386fe3b8f6edb970a3641145729288fe3c0551a24f699f9026cae

Observation dc6d773c-9e4a-4678-94a5-6d73163ec34e · inbound

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver cites this paper.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T20:32:45.390488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:32:45.390488Z digest=sha256:7ba5622b447edc5fddca4c00a8226892674a0221cceb92872462827eda4f08a9

Observation c674e204-38c9-4720-9e81-7012a735d468 · inbound

Ego-centric Predictive Model Conditioned on Hand Trajectories cites this paper.

Ego-centric Predictive Model Conditioned on Hand Trajectories LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T15:29:25.505592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:29:25.505592Z digest=sha256:49a5a0ee25906a71f2dc6a242f9687732a922a3db742dc31f421666c24c930cb

Observation c72e8b14-500b-43d8-b202-a7f9555e8d68 · inbound

Towards Spatial Trace with Reasoning in Vision-Language Models for Robotics cites this paper.

Towards Spatial Trace with Reasoning in Vision-Language Models for Robotics LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-03T16:27:34.062828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:27:34.062828Z digest=sha256:68c83207f39be612a20fe960694bfdb5db86c492d1759f69b141e989cb18b44d

Observation 6beaaddc-ce38-4903-9f5a-de8fff1354f4 · inbound

PALM: Progress-Aware Policy Learning via Affordance Reasoning for Long-Horizon Robotic Manipulation cites this paper.

PALM: Progress-Aware Policy Learning via Affordance Reasoning for Long-Horizon Robotic Manipulation LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-05-16T15:08:01.994473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T15:05:21.907878Z digest=sha256:2e218501276d1357e9bda76bf1949b76f93e385df71c681352ec6608d2fbd06e

Observation e2dbe256-8bf9-4d7e-8486-d813c3031b00 · inbound

Decompose and Recompose: Reasoning New Skills from Existing Abilities for Cross-Task Robotic Manipulation cites this paper.

Decompose and Recompose: Reasoning New Skills from Existing Abilities for Cross-Task Robotic Manipulation LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:56:06.795208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-09T14:30:00.890117Z digest=sha256:f5e7f4a289ecb96db28201c3ecad4859404149c26ebf28a7c076c7c3e9504610

Observation 451547fd-6822-49fb-b9c4-ff8ecbc68ca4 · inbound

Beyond 3D VQAs: Injecting 3D Spatial Priors into Vision-Language Models for Enhanced Geometric Reasoning cites this paper.

Beyond 3D VQAs: Injecting 3D Spatial Priors into Vision-Language Models for Enhanced Geometric Reasoning LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:53:13.630877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T07:47:52.739735Z digest=sha256:cfb0fdba093411e0cd6e583705ca24d37f7c40cbc472bf3528a06ac4e115fceb

Observation e064ffb5-c8a1-4853-8dcd-e25a536c8d55 · inbound

Grasp-Then-Plan with Failure Attribution: A Closed Two-Stage Framework for Precise and Generalizable Robotic Manipulation cites this paper.

Grasp-Then-Plan with Failure Attribution: A Closed Two-Stage Framework for Precise and Generalizable Robotic Manipulation LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 75

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T03:56:34.473901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T09:37:58.434897Z digest=sha256:c27277d149c128761e6ecabd2b42522fd04a9c0ab11525a648924b500a311442

Observation 53be3b28-1bf8-4a88-9266-7cde92c907ed · inbound

Trajectory-Level Redirection Attacks on Vision-Language-Action Models cites this paper.

Trajectory-Level Redirection Attacks on Vision-Language-Action Models LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:18:33.793178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T06:33:53.013076Z digest=sha256:4970ebb2133cf1f72ad7e0eee85df7221fd97f4aaf6f203b00df532cc06528d3

Observation e21013a9-57c6-4ceb-8be4-35d30bdc961b · inbound

Towards Human-like Physical Intelligence: Lifelong Vision-Language-Action Learning for Robotic Manipulation cites this paper.

Towards Human-like Physical Intelligence: Lifelong Vision-Language-Action Learning for Robotic Manipulation LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T00:57:42.768376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:57:42.768376Z digest=sha256:aff523bc4139f344b60a7db10589d8482daa27cfc528f6e5628a14c59753ba21

Observation 5335d7da-d38d-4ad2-b593-3db0b03e1cd0 · inbound

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation cites this paper.

RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation LLARVA: Vision-Action Instruction Tuning Enhances Robot Learning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-01T14:39:40.274531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T14:39:40.274531Z digest=sha256:3abbd86bb50c1b5b5fdf00a714210ec471f30e9f9e348b4b741c494f91c1cfba