Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:22:26.851204Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 6 inbound Pith citation observations for arXiv:2505.22050.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:22:26.851204Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T04:50:32.142169Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-17T20:28:15.942007Z
65 of 65 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 24e9227a-ffd2-414c-8a7c-fc9cb3a023e7 · outbound
Reinforced Reasoning for Embodied Planning URL: https://openai.com/index/ gpt-4o-mini-advancing-cost-efficient-intelligence/
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 65f69c35-873e-4102-b65b-9e0a69ef96ee · outbound
Reinforced Reasoning for Embodied Planning URL: https://openai.com/index/hello-gpt-4o/
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3b59e599-9cd7-4bf1-9ca2-cdcc395a60a3 · outbound
Reinforced Reasoning for Embodied Planning URL: https://www.anthropic.com/news/ claude-3-5-sonnet
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3eb58784-d7ba-4e28-ad95-8d9006b35e7f · outbound
Reinforced Reasoning for Embodied Planning URL: https://blog
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 94ffbd43-4381-4b69-acce-c470cc9f74ba · outbound
Reinforced Reasoning for Embodied Planning URL: https://ai.meta.com/blog/ llama-3-2-connect-2024-vision-edge-mobile-devices/
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6d72c2fe-222c-4599-935a-6740ea66b12a · outbound
Reinforced Reasoning for Embodied Planning Do As I Can, Not As I Say: Grounding Language in Robotic Affordances
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1466b7fc-2907-4e6d-a50a-285ceff2fbb6 · outbound
Reinforced Reasoning for Embodied Planning Qwen2.5-VL Technical Report
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15d6ded9-411f-4a3c-8807-c8451aa847fe · outbound
Reinforced Reasoning for Embodied Planning RoboGPT: an intelligent agent of making embodied long-term decisions for daily instruction tasks
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab1fe195-2f7d-4b5c-bb0e-a44cd8a1818e · outbound
Reinforced Reasoning for Embodied Planning Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7eaaf131-0b02-4d53-bea1-345e82ad6be2 · outbound
Reinforced Reasoning for Embodied Planning Process Reinforcement through Implicit Rewards
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba721db9-318d-48eb-bfe2-050a9a589979 · outbound
Reinforced Reasoning for Embodied Planning A survey of embodied ai: From simulators to research tasks
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 144e999e-c563-48ce-84b8-313f1e88c37d · outbound
Reinforced Reasoning for Embodied Planning Open r1: A fully open reproduction of deepseek-r1, January 2025
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 521e9aaa-7842-462c-aa30-ced4c358632e · outbound
Reinforced Reasoning for Embodied Planning What can vlms do for zero-shot embodied task planning? In ICML 2024 Workshop on LLMs and Cognition, 2024
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 486f722f-2e98-4710-994c-fcb43486f642 · outbound
Reinforced Reasoning for Embodied Planning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b35a1a89-f812-4f18-8a4e-ba084039e568 · outbound
Reinforced Reasoning for Embodied Planning Lora: Low-rank adaptation of large language models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5814911-feb8-457b-8a69-998ae1b6c810 · outbound
Reinforced Reasoning for Embodied Planning OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 306cb825-6fc5-49d0-b7ac-c2cf156d7404 · outbound
Reinforced Reasoning for Embodied Planning Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 953bad54-af75-45d8-a8f4-18ea783b6597 · outbound
Reinforced Reasoning for Embodied Planning Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e8cbed3-1426-4168-89c8-6ca6e30aecaa · outbound
Reinforced Reasoning for Embodied Planning RoboBrain: A Unified Brain Model for Robotic Manipulation from Abstract to Concrete
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb0633e3-f0b5-4c7c-8c62-12fad8a46a5a · outbound
Reinforced Reasoning for Embodied Planning Context-Aware Planning and Environment-Aware Memory for Instruction Following Embodied Agents
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 015d696c-6ac0-4202-80c8-9d98520e52c0 · outbound
Reinforced Reasoning for Embodied Planning Openvla: An open-source vision-language-action model
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 59e6988c-2cf4-4ed4-b91d-b3763ddf060e · outbound
Reinforced Reasoning for Embodied Planning AI2-THOR: An Interactive 3D Environment for Visual AI
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a9a00ae-3fd0-46c2-be4e-a04520f0240e · outbound
Reinforced Reasoning for Embodied Planning VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd364a8c-639e-4932-8fb5-0d90a571eece · outbound
Reinforced Reasoning for Embodied Planning Let’s verify step by step
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 068979f8-27f0-45ce-abe9-af07d4cc13f9 · outbound
Reinforced Reasoning for Embodied Planning Visual-RFT: Visual Reinforcement Fine-Tuning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75ed4314-1218-4b26-9220-0e431c3a42f3 · outbound
Reinforced Reasoning for Embodied Planning A Survey on Vision-Language-Action Models for Embodied AI
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 594ba930-4cde-4303-9a9a-16a9b0917d39 · outbound
Reinforced Reasoning for Embodied Planning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4064ed7-2bcd-4381-970d-e4a9ac66cc24 · outbound
Reinforced Reasoning for Embodied Planning Composi- tional chain-of-thought prompting for large multimodal models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e6233108-f2b8-4000-acc6-16e29d635747 · outbound
Reinforced Reasoning for Embodied Planning Kam-cot: Knowledge augmented multimodal chain-of-thoughts reasoning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a379916d-3750-496e-bd5e-d1b9bb0b6812 · outbound
Reinforced Reasoning for Embodied Planning EmbodiedGPT: Vision-Language Pre-Training via Embodied Chain of Thought
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2893fc7d-e8dd-4666-bbd9-bb90b0675797 · outbound
Reinforced Reasoning for Embodied Planning Training language models to follow instructions with human feedback
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6b71b1c-ea92-4cd9-9494-f9eafb0d7505 · outbound
Reinforced Reasoning for Embodied Planning Reasoning with large language models, a survey
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d138363b-8b9b-4f20-9d31-e6b264d3761b · outbound
Reinforced Reasoning for Embodied Planning Direct preference optimization: Your language model is secretly a reward model
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ae79c2a-828f-4a90-8707-f0c4835ee5d8 · outbound
Reinforced Reasoning for Embodied Planning Say- Plan: Grounding large language models using 3d scene graphs for scalable robot task planning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2673d062-04d3-4f5a-b57a-7fce5fcb0cac · outbound
Reinforced Reasoning for Embodied Planning Habitat: A platform for embodied ai research
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d065253d-040a-4b08-a27e-ec1e2567bcfe · outbound
Reinforced Reasoning for Embodied Planning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82715e8d-d392-48f3-ab1c-027a5694d301 · outbound
Reinforced Reasoning for Embodied Planning VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 695f2968-25c7-4f31-8cea-ac6239232391 · outbound
Reinforced Reasoning for Embodied Planning Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccb339cc-995d-455f-be9d-1b86e9eb32e4 · outbound
Reinforced Reasoning for Embodied Planning Socratic Planner: Self-QA-Based Zero-Shot Planning for Embodied Instruction Following
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e46ce3b6-0470-4a02-a8e7-14900943e03c · outbound
Reinforced Reasoning for Embodied Planning Alfred: A benchmark for interpreting grounded instructions for everyday tasks
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9082b7a0-ca47-426e-8226-73f21312d533 · outbound
Reinforced Reasoning for Embodied Planning Tenenbaum, Leslie Kaelbling, and Michael Katz
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c45aed3-41bc-4ada-bbe9-090799f28ba2 · outbound
Reinforced Reasoning for Embodied Planning ProgPrompt: Generating Situated Robot Task Plans using Large Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5083a2e4-4759-4062-869c-bc31d46e0826 · outbound
Reinforced Reasoning for Embodied Planning Trial and Error: Exploration-Based Trajectory Optimization for LLM Agents
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 276bf8bd-f071-4bbf-b57f-362978007a35 · outbound
Reinforced Reasoning for Embodied Planning Reason-rft: Reinforcement fine-tuning for visual reasoning
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a330e41-731c-4acf-98b7-a7e1e58d8eb7 · outbound
Reinforced Reasoning for Embodied Planning Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ef661ec-05e2-4e3e-be61-65c46d171ec2 · outbound
Reinforced Reasoning for Embodied Planning World Modeling Makes a Better Planner: Dual Preference Optimization for Embodied Task Planning
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea7a803f-cf6c-450e-8f42-ca62b9577c92 · outbound
Reinforced Reasoning for Embodied Planning Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41ebaff6-6b3f-4a3f-b640-c6982e3f7ce5 · outbound
Reinforced Reasoning for Embodied Planning Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc5ff3e5-e4fe-45b7-9a14-d9e94a21e443 · outbound
Reinforced Reasoning for Embodied Planning RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15fc3c6d-8b35-4c25-836b-2832a0b31c7f · outbound
Reinforced Reasoning for Embodied Planning Chain-of-thought prompting elicits reasoning in large language models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a3083d7-dbd6-4f84-9679-685ec544d6eb · outbound
Reinforced Reasoning for Embodied Planning Embodied Task Planning with Large Language Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b9d0635-0905-4b8f-a6d7-e53cb7270b75 · outbound
Reinforced Reasoning for Embodied Planning The rise and potential of large language model based agents: A survey
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf71a2cf-9922-41d1-8c85-ac16f9273565 · outbound
Reinforced Reasoning for Embodied Planning A Survey on Robotics with Foundation Models: toward Embodied AI
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33ed233e-2815-4edb-9741-68178e0e3971 · outbound
Reinforced Reasoning for Embodied Planning EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0ed911a-61c7-49ab-87f3-87547d71b64f · outbound
Reinforced Reasoning for Embodied Planning LIMO: Less is More for Reasoning
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3d90c58-7546-4690-99e6-d95fd96565d2 · outbound
Reinforced Reasoning for Embodied Planning Robotic control via embodied chain-of-thought reasoning
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7b573dbf-8950-48cd-911b-78c2857b79dd · outbound
Reinforced Reasoning for Embodied Planning HiRT: Enhancing Robotic Control with Hierarchical Robot Transformers
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e896157e-17c9-47f5-bca2-da733487fbde · outbound
Reinforced Reasoning for Embodied Planning Vision-language models for vision tasks: A survey
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2401c5a6-982c-49ee-81e2-e35b5de658ad · outbound
Reinforced Reasoning for Embodied Planning R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3365ba74-cedf-4a8a-bb6a-1d6134f17680 · outbound
Reinforced Reasoning for Embodied Planning Embodied-Reasoner: Synergizing Visual Search, Reasoning, and Action for Embodied Interactive Tasks
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba02ac77-4754-4ff3-8aeb-e459fef570f7 · outbound
Reinforced Reasoning for Embodied Planning Multimodal Chain-of-Thought Reasoning in Language Models
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72e92544-ffec-44c3-805d-e83ce60e3698 · outbound
Reinforced Reasoning for Embodied Planning Embodied-R: Collaborative Framework for Activating Embodied Spatial Reasoning in Foundation Models via Reinforcement Learning
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2b29aa4-dbbc-4fdb-8da3-7385faca2cff · outbound
Reinforced Reasoning for Embodied Planning LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f57eff8-5fb6-4f00-974e-911afd275ee8 · outbound
Reinforced Reasoning for Embodied Planning reasoning_and_reflection\
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9bede566-deae-4e2d-a0ba-ca8adacd5eec · outbound
Reinforced Reasoning for Embodied Planning Unresolved cited work
Reference 224
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dc9d4a05-26d6-4f64-b7e6-b01053167dd8 · inbound
Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey Reinforced Reasoning for Embodied Planning
Reference 150
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8355abd1-3ad3-4a41-b036-c0dc8ba2d1d3 · inbound
Large Language Models for Next-Generation Wireless Network Management: A Survey and Tutorial Reinforced Reasoning for Embodied Planning
Reference 168
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 228d765b-1279-480a-b96e-82f25ef54e93 · inbound
RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Reinforced Reasoning for Embodied Planning
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0aefcf2-b090-463b-a156-431113c66b6c · inbound
RoboAgent: Chaining Basic Capabilities for Embodied Task Planning Reinforced Reasoning for Embodied Planning
Reference 115
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 29bc1a9f-64c8-42e9-b5c6-1f94a5eab6b9 · inbound
Odysseus: Scaling VLMs to 100+ Turn Decision-Making in Games via Reinforcement Learning Reinforced Reasoning for Embodied Planning
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d7ee1b53-19af-41f1-91ac-da99f4e9d0b8 · inbound
RoboEvolve: Co-Evolving Planner-Simulator for Robotic Manipulation with Limited Data Reinforced Reasoning for Embodied Planning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.