Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:31:31.854787Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 2 inbound Pith citation observations for arXiv:2506.13679.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:31:31.854787Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T01:01:44.128884Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-08T01:01:44.731686Z
46 of 46 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c814ad01-a6d5-4cb3-92ba-6f0b2c45367a · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b09b29b0-0459-461f-96db-ea6356a82b61 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Improved baselines with visual instruction tuning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2997546c-b778-430d-bf0c-43a9f82d0535 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment LLaVA-OneVision: Easy Visual Task Transfer
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 904717a8-69be-4bdd-a694-e1085e59ddd0 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Prismatic vlms: Investigating the design space of visually-conditioned language models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c58b818e-dfaf-4cd2-a62e-d3fdd6d3fdc1 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Flamingo: a visual language model for few-shot learning.Advances in neural information processing systems, 35:23716–23736, 2022
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc828413-b729-4b96-b211-90854633e3c7 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8407032-91d5-48bb-a752-8ba9c45b3204 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment OpenVLA: An Open-Source Vision-Language-Action Model
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fbb4476-14fa-406a-b8d6-a6499681c0fa · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Open x-embodiment: Robotic learning datasets and rt-x models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3c4898bc-9a6a-4a92-9291-2be204bd8d2f · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b56d11b-00c3-4ad2-aa8b-4ced69a563dd · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 165a2863-afd9-46ef-9a11-eeb195aa3bcd · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment ChatVLA: Unified Multimodal Understanding and Robot Control with Vision-Language-Action Model
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fac4feb-cba5-4cc3-b41a-23c6ed007222 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment RT-H: Action Hierarchies Using Language
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74dab451-b4d4-4257-b8c1-0d5ecffc4965 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Octo: An open-source generalist robot policy
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 59b08fe7-ffbe-492e-9fcc-56f8c2beba24 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Perceiver-actor: A multi-task transformer for robotic manipulation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b4db716-9f2c-4606-9d99-6282f0170e54 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Vision-Language Foundation Models as Effective Robot Imitators
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b992ffc5-e23c-4980-81b7-5aa8b2145abe · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Rvt: Robotic view transformer for 3d object manipulation
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d15c9dcd-c39e-4e06-8e8d-338192f51988 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment RVT-2: Learning Precise Manipulation from Few Demonstrations
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bfe38de-0259-45f7-b169-78caaec5f641 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Bc-z: Zero-shot task generalization with robotic imitation learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b281560-27ce-4836-b863-aace77d3cfc0 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Coarse-to-fine q- attention: Efficient learning for visual robotic manipulation via discretisation
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e2a4cdef-d831-43af-b315-087c08719fd5 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Robot learning with sensorimotor pre-training
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a70a004-3e9b-46d0-a405-26dc6d678029 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment VIMA: General Robot Manipulation with Multimodal Prompts
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd92e520-1295-4fd6-aee2-c3b5429fdcff · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment RT-1: Robotics Transformer for Real-World Control at Scale
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ae0e367-31f4-4e7b-b225-0d48e47ac061 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acb22942-725d-4b03-9e9d-ad6442942450 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment PaLM-E: An Embodied Multimodal Language Model
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1192e56-1c1f-429e-8e4e-dcdfc069c8d9 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Learning transferable visual models from natural language supervision
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8589d5cb-f1e4-4b80-b6c2-c6a8b42f89f1 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Blip: Bootstrapping language- image pre-training for unified vision-language understanding and generation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef2fa0cb-c995-4d5e-8dc1-51545a5300d9 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Robotic Control via Embodied Chain-of-Thought Reasoning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f6cc393-fa1c-4e30-bbae-f03991b8191f · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment What Matters in Building Vision-Language-Action Models for Generalist Robots
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41f07b63-3605-43c8-a2ac-81775e33644d · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Llarva: Vision-action instruction tuning enhances robot learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation dd50b737-8aac-407b-a9c3-55f3d821b0cf · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a53aae7b-3c62-416e-9a3a-fc5a03f3ecc1 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment LLaRA: Supercharging Robot Learning Data for Vision-Language Policy
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ced5fe00-6396-4a32-8e96-f20f57f39cb0 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Tinyvla: Towards fast, data-efficient vision-language-action models for robotic manipulation.IEEE Robotics and Automation Letters, 2025
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1090907d-ac6e-4a27-ad05-3b22f1083705 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Towards Fast, Memory-based and Data-Efficient Vision-Language Policy
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation fbe82f9c-440f-4a16-9397-b85d4f3eb02a · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment MoLe-VLA: Dynamic Layer-skipping Vision Language Action Model via Mixture-of-Layers for Efficient Robot Manipulation
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee6b8ece-75f1-4d92-8a42-ee1ae762e9fc · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment An Atomic Skill Library Construction Method for Data-Efficient Embodied Manipulation
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b74616f8-37a9-4c9f-91d3-09548dfd9de0 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Pose estimation for an autonomous vehicle using monocular vision
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 0e7fec20-bfcd-423a-8a42-f529c0193c35 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Pose estimation and map building with a time-of-flight-camera for robot navigation.International Journal of Intelligent Systems Technologies and Applications, 5(3-4):355–364, 2008
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 0a41b70d-1cae-4cc4-9308-da39c0a7c1d1 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Learning human-to-robot handovers from point clouds
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation a55ac823-aecf-4896-b0f8-5f943267b7d9 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Pose estima- tion and adaptive robot behaviour for human-robot interaction
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 7401f7d3-8dbf-40dc-9761-86f2867a01d5 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Privacy-Preserving Pose Estimation for Human-Robot Interaction
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ec4159cd-4cc7-4493-9087-514d9950cc0e · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment 3D Robot Pose Estimation from 2D Images
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 24dd54f8-a3d3-4790-a295-2b17553f9b5d · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Multi-objective convolutional neural networks for robot localisation and 3d position estimation in 2d camera images
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 688cad5f-3aba-4f4b-956b-ee7eb8eda90f · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Robots’ state estimation and observability analysis based on statistical motion models.IEEE Transactions on Control Systems Technology, 30(5):2030–2045, 2022
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation bc44f804-0131-4409-8eac-bd62b734ec53 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Real-time holistic robot pose estimation with unknown states
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 65460fd1-5104-4768-bee8-b9ebbc0aa0ab · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment Qwen2.5 Technical Report
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 534a41be-be97-4cec-aa30-2727b67db404 · outbound
ROSA: Harnessing Robot States for Vision-Language and Action Alignment push the maroon button, then push the green button
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 806cce81-a07d-4a0d-964b-55c4be1e767f · inbound
Grounded Semantic Re-Binding for Robust Instruction Generalization in Vision-Language-Action Models ROSA: Harnessing Robot States for Vision-Language and Action Alignment
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87fdd124-616f-4096-9bdc-3ba1d7801856 · inbound
How Should Vision-Language-Action Models Use Proprioceptive State? ROSA: Harnessing Robot States for Vision-Language and Action Alignment
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.