Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:33:12.777054Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 3 inbound Pith citation observations for arXiv:2505.12744.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:33:12.777054Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T20:32:29.151581Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T21:00:09.870570Z
50 of 50 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c9437be8-6633-42fe-93ba-a99950227dc9 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Do As I Can, Not As I Say: Grounding Language in Robotic Affordances
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 054d140b-ea7f-4dca-a1bc-a6c150433fb9 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Qwen2.5-VL Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e16e946-49ef-4832-850b-1edf634bf5c5 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da33bb5f-c837-4fb8-a835-2512447dfc05 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation RT-1: Robotics Transformer for Real-World Control at Scale
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5cfd8230-3e87-42a8-9973-5e58e278c02c · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 606d6697-a9b2-4e10-8309-64266fbc35a7 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 111b15a2-ae6e-41b3-8f13-6e8d87168424 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Moto: Latent motion token as the bridging language for robot manipulation.arXiv preprint arXiv:2412.04445, 2024
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f5ea101-12cc-4bdc-906f-f89c6caf2e26 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Diffusion Policy: Visuomotor Policy Learning via Action Diffusion
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e276d1f4-9344-43a9-b333-59525faf413a · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Dynamo: In-domain dynamics pretraining for visuo.Motor Control, 2024
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 25f9d2c0-de3e-4921-9b46-32627781030a · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Fineclip: Self-distilled region-based clip for better fine-grained understanding.Advances in Neural Information Processing Systems, 2024
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 2310d7c0-ac87-4591-9314-9448a7f5b668 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation PaLM-E: An Embodied Multimodal Language Model
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58f55d5e-666f-48ef-849e-e3f2976303b2 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Bridge Data: Boosting Generalization of Robotic Skills with Cross-Domain Datasets
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a626588-9adc-4552-b445-32193d6d50aa · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b14d0579-2e9c-457a-aec7-578d4479316b · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Deep residual learning for image recognition
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9bbd22c5-1ada-438a-81e9-17f19e201937 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Lora: Low-rank adaptation of large language models.ICLR, 1(2):3, 2022
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50121185-7d97-45e8-9ee5-c5b68b326891 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 986b2236-b0fb-41ab-87a0-c89adc56dad2 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation An Embodied Generalist Agent in 3D World
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46f8da5e-a78c-424f-a55f-ea76046ce8cd · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 303ff475-e242-476a-85db-5a372d88be0a · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a5dfefd-cce6-435a-b488-0ad208ff35e0 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation OpenAI o1 System Card
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fef3789-952a-4476-8729-aec616554738 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation macmillan, 2011
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e4172b8-3e57-4640-bdb3-cf3dc4763842 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation OpenVLA: An Open-Source Vision-Language-Action Model
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f107e40-f37e-45ca-825c-956417c6268c · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Lisa: Reasoning segmentation via large language model
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb1253bf-5e96-4e42-8e72-246894153c22 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Vision-Language Foundation Models as Effective Robot Imitators
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 337af9f0-41c6-439d-950f-33fc041554ac · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Evaluating Real-World Robot Manipulation Policies in Simulation
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8dd7c1a0-a8b9-41dd-9eb4-2746304ca2a7 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Open-vocabulary semantic segmentation with mask-adapted clip
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d25ab36c-4505-44c6-b94d-37d449624964 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Code as policies: Language model programs for embodied control
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f31f9f2-cb43-4c08-a1ca-2e79ef004195 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Swin transformer: Hierarchical vision transformer using shifted windows
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20cc83de-0d40-4efb-ad29-71bdc6ae1a2f · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation SGDR: Stochastic Gradient Descent with Warm Restarts
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c920773-b5c1-4072-be56-67ce2a73af87 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Decoupled Weight Decay Regularization
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92af7a2e-7f78-43c7-899f-37ac6c3d585c · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Open X-Embodiment: Robotic Learning Datasets and RT-X Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21cc9d64-dee7-4f4a-8b34-c857a59247c3 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Proximal Policy Optimization Algorithms
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79013272-2282-41c1-b221-9800cec958a6 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fe58cd1-3f50-4324-92e2-8239cfe36c68 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5edd807-34a1-45e9-b295-35521afbcc3a · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Computing euler angles from a rotation matrix.Retrieved on August, 6(2000):39–63, 1999
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 28c2bee9-4443-448e-a252-fb7b397b257f · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation GeoManip: Geometric Constraints as General Interfaces for Robot Manipulation
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c746766d-ba61-47ec-9324-b55f220a4ce6 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Overcoming Support Dilution for Robust Few-shot Semantic Segmentation
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81f954b5-859b-49ca-abc9-faf06bfd477e · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Octo: An Open-Source Generalist Robot Policy
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6b118d3-7c43-4508-b74f-24c7dfd2b14e · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f1396e7-c4ec-4a2d-b5c2-051ccc2a59a8 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Bridgedata v2: A dataset for robot learning at scale
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c61ae86-8fcc-43e6-bd73-3bcb10a1334d · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation DART-LLM: Dependency-Aware Multi-Robot Task Decomposition and Execution using Large Language Models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 256228c2-1388-40e6-a0b1-5a3249fde096 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Q-learning.Machine learning, 8:279–292, 1992
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3871ed7-281b-4b1f-9f3f-5ee4d1a88593 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Chain-of-thought prompting elicits reasoning in large language models.Advances in neural information processing systems, 35:24824–24837, 2022
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10a290b1-691d-4070-bb64-2b506ace1192 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Qwen2.5 Technical Report
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ad22a85-1b4c-49bd-8103-98a1074fff65 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation DeepCritic: Deliberate Critique with Large Language Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dad9383-d3d0-4063-95be-fb7038cdaa81 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Latent Action Pretraining from Videos
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1b71996-a53e-46b1-aec7-a7ac379ab7bd · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bfb273f-f52f-427f-a7e4-fd445edbdc30 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5dada62-7421-440c-9c94-90048fbaa2e1 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation Robotic Control via Embodied Chain-of-Thought Reasoning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3ad2b89-8e9e-4b9e-9570-0e14eb8bb404 · outbound
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97ee8c20-695a-4f66-bf59-d38a765773a6 · inbound
Mixture of Horizons in Action Chunking Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d9088c2-a2cb-4738-b39a-86e8688c1840 · inbound
When Vision Overrides Language: Evaluating and Mitigating Counterfactual Failures in VLAs Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc62dda3-bc48-4f55-87eb-adba98446f84 · inbound
Learning Action Priors for Cross-embodiment Robot Manipulation Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.