Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 100 inbound Pith citation observations for arXiv:2307.15818.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T10:24:59.855729Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T10:07:00.865372Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 96658184-8130-456b-a56a-c52d5da317bc · inbound
Cognitive Architectures for Language Agents RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ef194d76-9d83-44ad-b50e-1b4cf63ca8b8 · inbound
GPT-Driver: Learning to Drive with GPT RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f50b1ee2-ebe6-4ea7-b375-ed4d48be7db3 · inbound
Reinforcement Learning with Foundation Priors: Let the Embodied Agent Efficiently Learn on Its Own RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 649fbcd2-8058-433b-a436-24504b3c9aae · inbound
Learning Interactive Real-World Simulators RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 23bba126-a8fd-4f01-b6e9-09e8094c89de · inbound
Open X-Embodiment: Robotic Learning Datasets and RT-X Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 38c972c4-232c-4ccb-9a83-c2208992c99e · inbound
Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3fd271d6-e777-4d79-844f-8ab2ab44ba53 · inbound
TD-MPC2: Scalable, Robust World Models for Continuous Control RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 156
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 21e3ab03-a12d-4689-9be7-9794c12146e1 · inbound
Vision-Language Foundation Models as Effective Robot Imitators RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e5f92c29-2975-4e70-9586-65035dd4c01b · inbound
Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 73ce41a8-5064-48c6-80d9-0e4fd8e93486 · inbound
AppAgent: Multimodal Agents as Smartphone Users RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bef27680-7526-4a59-889c-cf316c3614fa · inbound
Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 916b2d1a-9167-419b-aad0-b5a2b05d2e1a · inbound
Agent AI: Surveying the Horizons of Multimodal Interaction RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 120
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f022c4e6-8fd8-4677-aec9-aeda741ec060 · inbound
DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b58cb528-c895-49c0-87a0-a7c36819cb45 · inbound
NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 30ab2bec-9c11-49af-ba4c-a82f90499660 · inbound
RT-H: Action Hierarchies Using Language RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fb41528e-1b26-46b0-ae0e-888efea28365 · inbound
3D-VLA: A 3D Vision-Language-Action Generative World Model RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 37a39ef4-f617-4a6d-abd2-6c000a0eaf42 · inbound
OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4882d3c3-7a9d-4410-9107-ade4035e05b6 · inbound
The Platonic Representation Hypothesis RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 219
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cdb17367-46f1-4777-8115-4007acb99935 · inbound
A Survey on Vision-Language-Action Models for Embodied AI RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 69f71b86-68fc-4236-957a-6ae74cb5af71 · inbound
OpenVLA: An Open-Source Vision-Language-Action Model RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 54bf625c-a741-4d60-81c8-4c46b7298ba1 · inbound
LongVILA: Scaling Long-Context Visual Language Models for Long Videos RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8158301a-499c-477b-ba24-38f401b22131 · inbound
ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 107
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6bb2cd43-b995-43ca-9603-ddb731b32873 · inbound
TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a130f85b-dd09-4f7b-9b58-01a4f0dcb0df · inbound
Training Language Models to Self-Correct via Reinforcement Learning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 150
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 668a0c42-cfd3-4db6-ab7e-484a94654fc8 · inbound
Towards Robust Surgical Automation via Digital Twin Representations from Foundation Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cb94fd1d-303c-4ee1-8144-d70b041977cc · inbound
GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f9ac1071-0598-4bd7-bc05-8314ca9f4cb9 · inbound
Language Conditioned Multi-Finger Dexterous Manipulation Enabled by Physical Compliance and Switching of Controllers RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9e164021-00cd-4c88-bc85-f7c0beeb88f3 · inbound
OS-ATLAS: A Foundation Action Model for Generalist GUI Agents RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0ead8dff-ba03-423b-94af-1257f7b9ddca · inbound
$\pi_0$: A Vision-Language-Action Flow Model for General Robot Control RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ea61dd21-5e13-46dc-98ab-88fb0eb5a5a4 · inbound
DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation afa4cfbc-ac71-4591-86be-8a12bd8caafb · inbound
VeriGraph: Scene Graphs for Execution Verifiable Robot Planning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 929fd749-adcc-45c6-a506-7b00056469fe · inbound
CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7da82f87-4651-4a8b-8fc2-a7b3a5b42c37 · inbound
Preference Goal Tuning: Post-Training as Latent Control for Frozen Policies RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation abfa9331-448a-47a8-98cf-831b0e642247 · inbound
RoboMD: Uncovering Robot Vulnerabilities through Semantic Potential Fields RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5457b199-a646-4e3c-8f33-a7ca63a66dbb · inbound
Uni-NaVid: A Video-based Vision-Language-Action Model for Unifying Embodied Navigation Tasks RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f09baac6-b13e-4ac8-9ffd-647ba224fe8e · inbound
TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8f70d6d8-e7a0-4264-823a-6d06bd4319d0 · inbound
What Matters in Building Vision-Language-Action Models for Generalist Robots RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 870883cb-c452-47dd-b7f6-86fa504fe8fa · inbound
Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e0b41df9-e0d6-4477-9494-b451fa7e1ab9 · inbound
FAST: Efficient Action Tokenization for Vision-Language-Action Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bf8d3fb4-5f69-4029-a1e2-c59cca2d2d0c · inbound
SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dc9505dc-693f-4cd5-98f0-83215b576692 · inbound
Large Language Models for Multi-Robot Systems: A Survey RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7eff8d05-1492-4e42-9627-d79c14e123fd · inbound
DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 95d55974-f9f6-4a4d-918f-621136103741 · inbound
Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e7ea5adf-e6d5-43bd-bb69-c4acd5fa0b3e · inbound
Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6aa27e55-540d-4995-8c9c-cc0503f68167 · inbound
SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ae82a16e-4262-4e53-90a8-bd0e1ea5d805 · inbound
Kaiwu: A Multimodal Manipulation Dataset and Framework for Robot Learning and Human-Robot Interaction RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cc8628dd-8783-4662-b625-90cf7d12897b · inbound
AlphaDrive: Unleashing the Power of VLMs in Autonomous Driving via Reinforcement Learning and Reasoning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 436be115-68ae-483b-9d86-890585998ca3 · inbound
HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6b34d820-1b21-4719-9585-b85c951bd20b · inbound
Growing a Multi-head Twig via Distillation and Reinforcement Learning to Accelerate Large Vision-Language Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 75830fd6-8b1a-463a-8de5-bd989145fc16 · inbound
GR00T N1: An Open Foundation Model for Generalist Humanoid Robots RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4fe36b50-369c-41ad-9dda-3fffae657425 · inbound
Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 94ee9ac4-a713-4272-a414-f9028c9d8af9 · inbound
NORA: A Small Open-Sourced Generalist Vision Language Action Model for Embodied Tasks RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 674c5f46-84b4-4d7c-99b7-cb893664acb8 · inbound
GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation da7ba968-5944-4f1c-8ea5-dad8ec899c31 · inbound
VLAs are Confined yet Capable of Generalizing to Novel Instructions RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4b1b22d7-c527-41f1-b267-fa64f50eaa04 · inbound
DreamGen: Unlocking Generalization in Robot Learning through Video World Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 50fff180-de6d-44fe-991c-7c4bbfc4cc06 · inbound
Policy Contrastive Decoding for Robotic Foundation Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6ec1a03f-3de6-410c-94b1-d6626c5d596a · inbound
FLARE: Robot Learning with Implicit World Modeling RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9ec798c0-283b-47df-b3c8-0e7e80bd1c3b · inbound
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d04b9ed3-4832-4b36-a696-5c5eb67e1d1c · inbound
DreamPolicy: A Unified World-model Policy for Scalable Humanoid Locomotion RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 09d7e6b4-a3a3-49cf-9678-444178251b0e · inbound
SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a1c0ce06-e65d-4ac7-91ef-ac9d61fd6c15 · inbound
Real-Time Execution of Action Chunking Flow Policies RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c6b04c45-8f01-4280-97ac-3488f0dd72d2 · inbound
V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4eaf2be8-1aa8-49b4-92eb-4ea2fc502cd1 · inbound
Block-wise Adaptive Caching for Accelerating Diffusion Policy RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f11fb299-de8e-418b-baec-7c9deb9959a6 · inbound
GRaD-Nav++: Vision-Language Model Enabled Visual Drone Navigation with Gaussian Radiance Fields and Differentiable Dynamics RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9d2cf2ee-fa5c-4141-afe1-6dd8233bd2d9 · inbound
RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3d6a7b7e-3f2a-4e8a-a616-d3b28e46cabd · inbound
WorldVLA: Towards Autoregressive Action World Model RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a89c0052-ad6d-47a9-b50b-dbf080a2c966 · inbound
A Careful Examination of Large Behavior Models for Multitask Dexterous Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 03dcca80-50b5-4ba6-9b3b-85dfc3804b4a · inbound
Preemptive Solving of Future Problems: Multitask Preplay in Humans and Machines RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 53bfa4d4-aa9d-4ef3-9eb8-5fa31b68ff16 · inbound
GR-3 Technical Report RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 402f1164-5b32-4826-a52b-22a45bf3ff43 · inbound
ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 24f98853-d3f5-47ab-b2f1-9d88375305fb · inbound
Genie Envisioner: A Unified World Foundation Platform for Robotic Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8b306df5-cd1a-445c-a1ac-a4a7c46ced50 · inbound
MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 34e88aee-4055-4ac5-99ea-018ca00f7785 · inbound
Aligning Perception, Reasoning, Modeling and Interaction: A Survey on Physical AI RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 225
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8518c6a5-a2ee-4014-98c3-1e5baad3ec48 · inbound
R2RGEN: Real-to-Real 3D Data Generation for Spatially Generalized Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f9786ea3-b935-4cfa-bd83-263aff97709f · inbound
Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4c6ec9b-26e4-4e2a-84d8-dfc3a381bddc · inbound
InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 28da32b0-07fa-4681-b657-bf4255060496 · inbound
QDepth-VLA: Quantized Depth Prediction as Auxiliary Supervision for Vision-Language-Action Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90529829-446a-46ba-b581-2aed5cffbc7e · inbound
Dual-Stream Diffusion for World-Model Augmented Vision-Language-Action Model RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5071b48-8ac2-4f67-a6f6-98e51df95086 · inbound
LACY: A Vision-Language Model-based Language-Action Cycle for Self-Improving Robotic Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a35b5dd0-f523-48e5-9031-b0825f01733f · inbound
AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bf80d64c-1d5d-4bbf-aa01-24c17c134a5f · inbound
DynaMimicGen: A Data Generation Framework for Robot Learning of Dynamic Tasks RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30cf9be9-17be-4f22-a6b1-808061085006 · inbound
DASIP: Dynamic Test-Time Compute Scaling for Robot Control with Stochastic Interpolant Policies RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bfd0808-ce72-4877-b94c-a3803f1582f2 · inbound
SimScale: Learning to Drive via Real-World Simulation at Scale RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e32151b2-88c8-4d91-aa2f-8c8a6e415793 · inbound
Training One Model to Master Cross-Level Agentic Actions via Reinforcement Learning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47318a33-5b9f-4103-9de0-5b38181a5ee9 · inbound
Large Video Planner Enables Generalizable Robot Control RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d8b81cd4-22d4-4f2b-a90c-d2801d265914 · inbound
VLA-Arena: An Open-Source Framework for Benchmarking Vision-Language-Action Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ec77fc6-5cf2-45eb-a555-a5c37206250b · inbound
Learning to Feel the Future: DreamTacVLA for Contact-Rich Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0a4bbb9e-f51d-4272-a56c-0a5d670723be · inbound
Learning to Feel the Future: DreamTacVLA for Contact-Rich Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62233dc9-936e-419b-a6f3-c38e0f58c4dc · inbound
VLM4VLA: Revisiting Vision-Language-Models in Vision-Language-Action Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e64fe6ad-2e91-488a-839a-5a187154cd2b · inbound
State Backdoor: Towards Stealthy Real-world Poisoning Attack on Vision-Language-Action Model in State Space RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcec3e9b-37a2-47e4-8bac-8cbafd3baa8f · inbound
Real2Sim via Active Perception with Behavior Trees Automatically Generated by VLMs RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7c612aea-1fe0-47d7-bedf-09063fbbef2c · inbound
Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6cba2652-a325-4b4d-acc6-69686cbb1b06 · inbound
Eval-Actions: Fine-Grained Execution Quality Evaluation for Robotic Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6e86cb1-2d70-4e36-b77e-547a39ef03ea · inbound
How Users Understand Robot Foundation Model Performance through Task Success Rates and Beyond RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f131e54-7d80-47a3-914f-60f2e8965678 · inbound
ST-BiBench: Benchmarking Multi-Stream Multimodal Coordination in Bimanual Embodied Tasks for MLLMs RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 52dc38f3-28cf-4a64-b69c-f63ec1c7d087 · inbound
TwinRL: Digital Twin-Driven Reinforcement Learning for Real-World Robotic Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 33d214d7-c209-4613-86ea-8c58d81925b5 · inbound
PRISM-XR: Empowering Privacy-Aware XR Collaboration with Multimodal Large Language Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7f0aa1ec-7573-4c8c-89bd-9d7c10351990 · inbound
World Action Models are Zero-shot Policies RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f2996fc3-11d9-4bfe-bdbd-062bea046405 · inbound
When Vision Overrides Language: Evaluating and Mitigating Counterfactual Failures in VLAs RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c20da66f-f8b7-4a19-a200-b6c04932699b · inbound
UniLACT: Depth-Aware RGB Latent Action Learning for Vision-Language-Action Models RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.