Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:33:55.651566Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 76 of 76 outbound references and 35 inbound Pith citation observations for arXiv:2505.21432.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:33:55.651566Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:21:09.315435Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
76 of 76 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 81007fef-4341-4caf-a768-4a4bb9bd60e8 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Vision-language foundation models as effective robot imitators
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a679ba6e-d086-424b-9449-b56e638d267f · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbb5c2de-20a4-4c51-9fb5-d85fe417456e · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Fastumi: A scalable and hardware-independent universal manipulation interface with dataset
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f7a06d05-3a29-4c1f-aa89-74229fc47b72 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model OpenVLA: An Open-Source Vision-Language-Action Model
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72bb9ad0-c3ad-4caf-92d2-0ae9f341887a · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Learning 2d invariant affordance knowledge for 3d affordance grounding
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 523ae9b6-ae5a-4258-84ba-9c07faaa7f70 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b16ca1ec-1d7f-44e0-94fd-da39d0af58cc · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Improving domain generalization in self-supervised monocular depth estimation via stabilized adversarial training
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4a1439db-2ae7-4131-a03b-830828a63ccd · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model MoMa-Kitchen: A 100K+ Benchmark for Affordance-Grounded Last-Mile Navigation in Mobile Manipulation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4c00cb3-77d4-4214-b80e-5fa8af747019 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model ORLA*: Mobile Manipulator-Based Object Rearrangement with Lazy A Star
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f130c18-80c5-4cd6-9397-eb6338f72955 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Decentralized Transformers with Centralized Aggregation are Sample-Efficient Multi-Agent World Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94eaf881-2020-449e-ab3a-2b66d3dc8480 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Gemini Robotics: Bringing AI into the Physical World
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b973670-3451-42d6-ada0-178115d937c5 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5d2e319-e561-438f-b073-5c14a9713627 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Thinking, fast and slow
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7005a5ee-e8a7-4078-b940-90adc3152e70 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model COHERENT: Collaboration of Heterogeneous Multi-Robot System with Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dad5dba8-74ae-4f28-932a-592062ea71e2 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Kinematic- aware prompting for generalizable articulated object manipulation with llms
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8c35e3b9-4873-4117-a821-d10f28fa2b16 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model MoRE: Mixture of Residual Experts for Humanoid Lifelike Gaits Learning on Complex Terrains
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6dddf33d-15cc-4694-939f-d600362909b7 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Skill-Nav: Enhanced Navigation with Versatile Quadrupedal Locomotion via Waypoint Interface
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 597e2a01-7034-4650-b63b-32cafef1ea61 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Robotic policy learning via human-assisted action preference optimization
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92e6060e-9b8b-44a7-a2c4-33d3bfb0e455 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Phoenix: A motion-based self-reflection framework for fine-grained robotic action correction
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 85cd5166-56d9-467d-ae2d-9aa866a98f5e · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Think Small, Act Big: Primitive Prompt Learning for Lifelong Robot Manipulation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1aa81f56-cba4-40c7-b47d-aabcf0641517 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Spatial-temporal graph diffusion policy with kinematic modeling for bimanual robotic manipulation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1af8d243-7c3a-4325-b55d-d0415085ca79 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Chain-of-thought prompting elicits reasoning in large language models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 33ca2d67-8cba-48d1-8c09-e441a74ddea4 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Robotic Control via Embodied Chain-of-Thought Reasoning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afbe1c2c-f490-4a80-a099-310c6b6a865f · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model A dual process vla: Efficient robotic manipulation leveraging vlm
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0593f14e-de49-4900-97f9-ced694cb1150 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model HiRT: Enhancing Robotic Control with Hierarchical Robot Transformers
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbc7f4d3-eef8-4e1c-9eb2-c8f2cbfe781d · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Towards Synergistic, Generalized, and Efficient Dual-System for Robotic Manipulation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ca8f132-157a-4cc0-a0b7-b44fbf22ab28 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c1db6a1-da02-49f4-a743-2c29061bc1a0 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation faf03009-fb2c-461a-bd4d-14ef462cf812 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3b374f1-6e31-40ef-983a-45e2b9a5fa7a · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Helix: A vision-language-action model for generalist humanoid control, 2025
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 32fb4798-d4e5-4b97-82c7-86ae6160cf74 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Flow Matching for Generative Modeling
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e36e8996-aef0-42d3-9da8-5e096f4a255c · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Beyond Optimal Transport: Model-Aligned Coupling for Flow Matching
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e2d2e85-f720-40dd-9f4e-71daf55ff675 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model LIBERO: Benchmarking Knowledge Transfer for Lifelong Robot Learning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed099548-4729-4806-a43b-4047d2b92ee5 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Evaluating real-world robot manipulation policies in simulation
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b6980a7b-23da-4fac-9da9-898329f2d91a · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4661f28f-b4fd-428c-8cab-dd0ea2467d6c · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 257b1298-91e4-4d6a-b99e-c8e12cacef87 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15194271-763a-4a9e-ae20-58bc0171772c · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model RoboMP2: A robotic multimodal perception-planning framework with multimodal large language models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6f71dcb0-3c70-43a6-b824-801f1038533a · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Universal Actions for Enhanced Embodied Foundation Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98106cf3-bab8-479b-bef4-aedf443dc267 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Learning causality-inspired representation consistency for video anomaly detection
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7987215a-b994-45c6-abc3-f9d0c0442d99 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Pali-x: On scaling up a multilingual vision and language model
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a686961f-ed9d-4b19-866a-bda44f39afa6 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Prismatic vlms: Investigating the design space of visually-conditioned language models
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f30ddaf8-0366-41b9-ac3c-a4e833fb91c2 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Open x-embodiment: Robotic learning datasets and rt-x models
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bb30c695-ea59-47f2-b80d-691d8ebc69ca · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Reflexion: Language agents with verbal reinforcement learning
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0c08af9-ddd5-4a01-aed7-1077f92b290c · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Reasoning with Language Model is Planning with World Model
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76800bd8-a57b-47d3-be37-85e5bc17e353 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 232a0add-3c70-40f2-85a8-dc5ba1c1ae8a · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Cross from Left to Right Brain: Adaptive Text Dreamer for Vision-and-Language Navigation
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c515e4aa-907b-4748-b0fe-ab70894abfb7 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Mllmguard: A multi- dimensional safety evaluation suite for multimodal large language models
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 51c50869-8ef4-42d7-9746-4530cff6eec3 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model SelfCheck: Using LLMs to Zero-Shot Check Their Own Step-by-Step Reasoning
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b14d58a8-8cfa-42de-b0fc-c8e86a6fd48f · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model MorphMark: Flexible Adaptive Watermarking for Large Language Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ade5a7e-5cda-44de-b5bf-de1ce2217f55 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Tree of thoughts: Deliberate problem solving with large language models
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 52dd79ae-1572-46ca-b6ee-a04dcf031ad7 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model AlignBot: Aligning VLM-powered Customized Task Planning with User Reminders Through Fine-Tuning for Household Robots
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56722f56-4188-4b25-aede-774f1ca05879 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Sets: Leveraging self-verification and self-correction for improved test-time scaling
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23c7015d-acb4-4e75-9d04-2bd052aeccba · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Interpretable Contrastive Monte Carlo Tree Search Reasoning
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51468219-21ee-4e85-a594-c758e0825966 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model O1-Pruner: Length-Harmonizing Fine-Tuning for O1-Like Reasoning Pruning
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0acd536-3231-4720-905d-cf6adbc16476 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Cascaded Diffusion Models for High Fidelity Image Generation
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55a1e608-4ed0-4fa6-bde0-4ddcb2a9a26a · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Revis- iting multi-agent world modeling from a diffusion-inspired perspective
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f73061fa-010c-4826-9dc2-865bbd20f8fd · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model f-DM: A Multi-stage Diffusion Model via Progressive Signal Transformation
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d2f7f3f1-10d6-4d08-ab54-bf8ea252f4b1 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Bring Metric Functions into Diffusion Models
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 50763b71-7e9d-4562-b70a-c7dee4706c6c · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Spectral-cascaded diffusion model for remote sensing image spectral super-resolution
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f51f89c3-d1cb-4027-84ce-d0d70f442a8f · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model High-resolution frame interpolation with patch-based cascaded diffusion
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c527b747-cc31-4b73-8b80-f6f1e271296e · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Cascaded diffusion models for virtual try-on: Improving control and resolution
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2c28b4f4-4442-4192-947e-bf9bac79cfd2 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05da854f-a948-4814-ad54-9ffdfb1cc9e1 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Pre-training for robots: Offline rl enables learning new tasks from a handful of trials
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 287d1e5a-a6e8-4633-a58d-c5f73ef9088f · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Cal-ql: Calibrated offline rl pre-training for efficient online fine-tuning
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f5826893-a8b4-4a18-848d-bfba099c12f1 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model RT-1: Robotics Transformer for Real-World Control at Scale
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b73c3ab7-a074-4ed8-bac6-3ee1f980e446 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Octo: An open-source generalist robot policy
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 949e21aa-1443-418a-a674-d4b10016708d · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Scaling proprioceptive-visual learning with het- erogeneous pre-trained transformers
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fb8bd160-8b34-48c3-9899-1b76a00db3e0 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model What Matters in Building Vision-Language-Action Models for Generalist Robots
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ee513b6-aba5-475c-9d35-14fb20a3afc8 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad6b665d-616d-44f5-bdca-262e55c9930f · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Diffusion policy: Visuomotor policy learning via action diffusion
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ac1ebce-6d32-4038-8412-5bb69cd72bc6 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae1bf471-dc67-454c-9dff-58f2d586f78c · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Improving Large Language Model Fine-tuning for Solving Math Problems
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1791c3be-837b-4f62-b6bf-1553d4807b86 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Continuous control with deep reinforcement learning
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b18831f-3bec-46ab-82ca-e17095e176ce · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Addressing function approximation error in actor-critic methods
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5d8e87df-c426-4323-8554-e9c8d0a2fea1 · outbound
Hume: Introducing System-2 Thinking in Visual-Language-Action Model Soft Actor-Critic Algorithms and Applications
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f957329-1602-4060-b1fb-d8301b5ef22a · inbound
FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9943f6b5-3f45-44ca-a277-c0385c11385e · inbound
Skill-Nav: Enhanced Navigation with Versatile Quadrupedal Locomotion via Waypoint Interface Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf1e6f74-a5fa-4853-a89d-9d6deb14871f · inbound
DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e6c42ed0-3a1e-41ec-be9f-2ac5e9d0b556 · inbound
F1: A Vision-Language-Action Model Bridging Understanding and Generation to Actions Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3edc250c-f891-48c5-987d-ad183462ea08 · inbound
InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c5a28893-c809-42c3-bf40-11c38ea3f232 · inbound
Efficient Vision-Language-Action Models for Embodied Manipulation: A Systematic Survey Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa07afbc-06fa-432a-b218-afd5a0eb7310 · inbound
VGAS: Value-Guided Action-Chunk Selection for Few-Shot Vision-Language-Action Adaptation Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 150d728d-7652-4a3c-bc24-46f9e1ba84ac · inbound
StemVLA:An Open-Source Vision-Language-Action Model with Future 3D Spatial Geometry Knowledge and 4D Historical Representation Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73a9281f-48a8-4fa0-a5d9-b79e4083dc45 · inbound
ZeroWBC: Learning Natural Whole-Body Humanoid Interaction from Human Egocentric Data Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e773579a-6a7d-4e0d-8a2f-5b3e8483c3fb · inbound
AR-VLA: True Autoregressive Action Expert for Vision-Language-Action Models Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8deee7bd-f6c2-4f84-b333-13298a69cef4 · inbound
Fine-tuning is Not Enough: A Parallel Framework for Collaborative Imitation and Reinforcement Learning in End-to-end Autonomous Driving Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 635e0017-9b9f-4d17-b2f7-0e961982ed91 · inbound
Spatial navigation in preclinical Alzheimer's disease: A review Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5241d33f-930f-4b74-8528-f7da22a0e8d4 · inbound
UAV-Track VLA: Embodied Aerial Tracking via Vision-Language-Action Models Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 90f67b5f-aacf-4d18-a407-f2dfa1a89af7 · inbound
Deep Image Clustering Based on Curriculum Learning and Density Information Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 28a5c88f-922d-4a9e-afe3-e79a16620ddd · inbound
Adaptive Action Chunking at Inference-time for Vision-Language-Action Models Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bce58944-e575-4e19-aeeb-2f96c4ab1f0d · inbound
Sentinel-VLA: A Metacognitive VLA Model with Active Status Monitoring for Dynamic Reasoning and Error Recovery Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c66e8685-deec-4791-a0d1-cde96d09ed9f · inbound
Sentinel-VLA: A Metacognitive VLA Model with Active Status Monitoring for Dynamic Reasoning and Error Recovery Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c8b0042f-638e-4de8-b0bf-1858489ed884 · inbound
VLA-ATTC: Adaptive Test-Time Compute for VLA Models with Relative Action Critic Model Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9237db13-56b0-4f9c-8a7b-61c5ff07d408 · inbound
Hyper-DP3: Frequency-Aware Right-Sizing of 3D Diffusion Policies for Visuomotor Control Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2913645b-6fa2-4e10-8c4d-453eb1512fbb · inbound
Hyper-DP3: Frequency-Aware Right-Sizing of 3D Diffusion Policies for Visuomotor Control Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 30972e20-377c-40dd-b05c-ff820df2d8c4 · inbound
From Imagined Futures to Executable Actions: Mixture of Latent Actions for Robot Manipulation Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0c5495cd-9e10-45be-8883-67a3eef00e56 · inbound
Q-VGM: Q-Value-Gradient Matching for Off-Policy Reinforcement Learning of Flow-Matching VLA Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation de9b08b3-5e48-45e1-9deb-259b4f49c5da · inbound
Q-VGM: Q-Value-Gradient Matching for Off-Policy Reinforcement Learning of Flow-Matching VLA Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e3fd18b-156e-424d-98fa-ba4770f928df · inbound
PearlVLA: Progressive Embodied Action-Plan Refinement in Latent Space Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 53e4a2ea-4c7a-4e32-b782-a1ccb57258a1 · inbound
UniFS: Unified Fast-to-Slow Hierarchical Architecture for Vision-Language-Action Models Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8cf1726e-96e8-4ac7-8326-8919eb2502a4 · inbound
FORCE: Efficient VLA Reinforcement Fine-Tuning via Value-Calibrated Warm-up and Self-Distillation Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a96ac403-2fb8-43bc-bf7f-951da40ff893 · inbound
E-TTS: A New Embodied Test-Time Scaling Framework for Robotic Manipulation Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1c1c081c-dc97-461f-81d4-34b61c42e28c · inbound
Recursive Self-Evolving Agents via Held-Out Selection Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6d0ce5b9-3a94-40f1-b56c-203336e675e1 · inbound
Position: Vision-Language-Action Models Cannot Be Verified to Perform Physical Reasoning Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 984e3d47-f327-4463-81e0-474103ed5288 · inbound
ROSA: A Robotics Foundation Model Serving System for Robot Factories Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 67504498-1254-428f-91dc-015b7465a512 · inbound
Look Before You Leap: Distilling Tree Search into Action Evaluation for Frozen VLA Models Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73f704b0-a5e4-469e-b3e9-0953aa79b3a5 · inbound
ABot-N1: Toward a General Visual Language Navigation Foundation Model Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5efd68fd-27c8-462e-ab83-5827f5d3a342 · inbound
ABot-N1: Toward a General Visual Language Navigation Foundation Model Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c83eae8-90ec-454b-8d38-30cba65811db · inbound
CosFly-VLA: A Spatially Aware Vision-Language-Action Model for UAV Tracking Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8f48e30-f03c-4298-aa2d-b3bd694feb13 · inbound
Token-Wise Latent Streaming from Slow Reasoners to Fast Planners for Dynamic Vision Language Navigation Hume: Introducing System-2 Thinking in Visual-Language-Action Model
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.