Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T03:56:05.890768Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 3 inbound Pith citation observations for arXiv:2602.06575.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T03:56:05.890768Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T01:01:44.122151Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T23:07:27.134549Z
32 of 32 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e79c0e1f-5ab5-4967-bf62-3c2d1efe2343 · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca558d04-ed66-414f-9d4e-36c09a6f5fe3 · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 866479ce-3577-489f-b0ec-00297cb223a3 · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce5024b7-fe2a-47fe-90ea-ebddaf07717a · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies R., Finn, C., Kumar, A., and Levine, S
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a02d83f-7c01-4b53-8c32-d1abd795b7eb · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies Diffusion policy: Visuomotor policy learning via action diffusion
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df2cc7c7-9f52-487f-90cd-13cfab484e06 · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies The Ingredients for Robotic Diffusion Transformers
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2eacf72c-5def-46fe-b13f-20757e186e7d · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies U., Akram, W., Saoud, L
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fea78ca6-3221-45b2-8b6f-f2d5b3c5c48a · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies Latent Theory of Mind: A Decentralized Diffusion Architecture for Cooperative Manipulation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85b205a8-ea66-405b-94bf-339714a66c7e · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies Demystifying Diffusion Policies: Action Memorization and Simple Lookup Table Alternatives
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4026ade-74dc-4276-8eaf-1bf3f6496f0c · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies Dita: Scaling Diffusion Transformer for Generalist Vision-Language-Action Policy
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfd835a7-5f16-4087-a785-e8ed73593688 · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies Video prediction policy: A generalist robot policy with predictive visual representations
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90fb4eaf-421f-4151-95a3-dd0dd3999e1c · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies Otter: A vision-language-action model with text-aware visual feature extraction
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 822fa9cc-27c0-42aa-804c-b12f3cc2c5b0 · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68cc964f-af0e-415a-acf5-9449d98324b9 · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies The better you learn, the smarter you prune: Towards efficient vision-language-action models via differentiable token pruning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b503360-8bb5-4e77-bb98-2bb75e829740 · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad7a0ac2-e012-4c8d-a8dd-4998de472333 · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies J., Pertsch, K., Karamcheti, S., Xiao, T., Balakrishna, A., Nair, S., Rafailov, R., Foster, E
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cb37b32-e59c-4e28-904d-aa6761045a37 · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies CoA-VLA: Improving Vision-Language-Action Models via Visual-Textual Chain-of-Affordance
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6214642d-950d-46ac-a6d1-7c53307f2485 · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fca8ba7-75e2-492c-90b4-926308054eb2 · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies Vision-Language Foundation Models as Effective Robot Imitators
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c5d9faa-2b20-47fd-883d-0ef0312d7808 · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies Flow Matching for Generative Modeling
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6547da2f-54b9-4ce2-a21a-70f92dbd2a79 · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies Libero: Benchmarking knowledge transfer for lifelong robot learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc3bae6d-825b-4b3f-a816-02c0e955849d · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies Calvin: A benchmark for language-conditioned policy learning for long-horizon robot manipulation tasks
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a4daf34-bf94-4cfc-9186-dee12cc36bda · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies and Xie, S
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82f5e85d-e02a-40d5-a581-3a169114d44e · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies Dynamicvit: Efficient vision transformers with dynamic token sparsification
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15fbe715-5e6c-4768-9716-468af86a2bc5 · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies Multimodal Diffusion Transformer: Learning Versatile Behavior from Multimodal Goals
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e99d5310-da24-401b-bfe8-fdbc90974028 · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies E., Otto, F., and Lioutikov, R
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4b86630-081a-4e87-b794-675d41d62a52 · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies TokenLearner: What Can 8 Learned Tokens Do for Images and Videos?
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54748125-b756-47b8-9bf5-fe9c211be117 · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87a6f7a6-44de-44f9-8bce-769e41b39213 · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 537af961-80dc-4af7-90f5-d7c32efc8a9b · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies Unleashing large-scale video generative pre-training for visual robot manipulation
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddf23d8c-a1e1-482a-8f98-b654a88a821d · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies Deer-vla: Dynamic inference of multimodal large language models for efficient robot execution
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31281768-7a8c-4a6f-8d55-0dd8cdc96c43 · outbound
Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies Rt-2: Vision-language-action models transfer web knowledge to robotic control
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc31dfb5-636f-475a-9eb3-7fda664f660e · inbound
When Video Misreads: Closed-Loop Distillation of Reading Heuristics for Exploratory Manipulation Trace QA Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e54940ea-f858-4887-9295-6bd8d2cf92cc · inbound
How Should Vision-Language-Action Models Use Proprioceptive State? Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c5ad290-f9e0-4386-bc10-c1d1fc7b6ee2 · inbound
How Should Vision-Language-Action Models Use Proprioceptive State? Think Proprioceptively: State-Grounded Visual Token Selection for VLA Policies
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.