Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:55:06.821160Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 3 inbound Pith citation observations for arXiv:2506.23127.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:55:06.821160Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T09:33:38.485506Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-15T11:55:33.536071Z
64 of 64 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6cddb675-a768-4624-9741-83ba498cb8ce · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92f03b00-374b-4a14-9c7c-068dcff9429b · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning FireAct: Toward Language Agent Fine-tuning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44fdcf0d-5700-463d-831e-9d8541d497c4 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning RoboGPT: an intelligent agent of making embodied long-term decisions for daily instruction tasks
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46e7171b-ec46-4bee-b76d-f05f0060db78 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Process Reward Models for LLM Agents: Practical Framework and Directions
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30bfaac9-0107-4ba4-9eac-9899aff3e25f · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning DeepSeek-V3 Technical Report
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 291326ee-5722-4bd0-baa2-1e87c5d81ee1 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning A survey of embodied AI: from simulators to research tasks
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b2be25ea-7c01-4204-a2c6-394fa2d1dabf · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 99b878b9-8cba-4204-9b4c-e37587f6fd7a · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Chawla, Olaf Wiest, and Xiangliang Zhang
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0311ce09-d251-483a-b835-3ca4da948942 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11797ada-e311-4259-bce7-bb4206757cc2 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Controlling Large Language Model with Latent Actions
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e10d8081-b76b-49ce-8a53-2cfd8c19ab6c · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 349398af-d142-4647-8926-ff5acfa64ff4 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Search-o1: Agentic Search-Enhanced Large Reasoning Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f53b722-1eae-4f6b-a001-79151293c5cb · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning From System 1 to System 2: A Survey of Reasoning Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7654a8ce-2bd2-4e0f-b883-0fd401de6487 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Let's verify step by step
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c80cf3bd-6b7e-4a89-9abc-fc7703a8b8a9 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning QLASS: Boosting Language Agent Inference via Q-Guided Stepwise Search
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdef8133-077c-4f9b-82d3-ff157b6b2c48 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Aligning Cyber Space with Physical World: A Comprehensive Survey on Embodied AI
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a76f25c-b45c-473e-ba67-e1b922a9e1ea · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa62b6a0-3e2f-4f47-9455-ecab2d1090bc · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69dcfd01-d48e-4d21-85dd-be3372de2929 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f85de20-9000-4cc6-82a2-35398d56805f · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning FILM: following instructions in language with modular methods
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b199611a-4214-4d9e-8e4f-d1d02f70d8b6 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Skill set optimization: Reinforcing language model behavior via transferable skills
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 372079f1-79b5-4900-9ebb-f38615bd7ee4 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning GPT-4 Technical Report
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbde38d3-4b95-4ad0-9163-63e5af4152cf · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edfb6089-6ce5-4682-bddf-b787eb2e14f1 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 928ca13b-4213-46a8-af18-cdbadaa4124d · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Agent planning with world knowledge model
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 773b8528-02e7-4ffa-a46f-d9bbb0d3cf37 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Tarr, William W
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cc6260ec-5258-4a64-8d1a-6880233e3d7f · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b84dae45-9bbd-4e05-a1fb-2582714e1b90 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Reflexion: language agents with verbal reinforcement learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 88014892-07ad-4018-b6ef-a7cb5a6f0324 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Hausknecht
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f57c87d5-c1e2-472e-a7fe-5c3815e962c6 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Agentbank: Towards generalized LLM agents via fine-tuning on 50000+ interaction trajectories
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 30ea569b-757d-4043-b333-015c580c87eb · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Trial and Error: Exploration-Based Trajectory Optimization for LLM Agents
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b9603a2-5159-4b2b-a17b-79b7fc87cc1c · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Adaplanner: Adaptive planning from feedback with language models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 82d41c2d-4246-44ec-bca0-328c26073f9f · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning A Survey of Reasoning with Foundation Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4237dab2-8f17-4de4-8124-79efe558c357 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning InternLM2 Technical Report
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70f59dfc-0d62-4f21-802e-c5b1c646d098 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c6d6211-4907-4e1e-afef-1f97b687be2d · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning STeCa: Step-level Trajectory Calibration for LLM Agent Learning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c787b02-2301-4e54-b8de-08a800880ea4 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Offline Reinforcement Learning for LLM Multi-Step Reasoning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73074153-2e5e-44f9-81ac-73f5cfc3c85f · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Learning From Failure: Integrating Negative Examples when Fine-tuning Large Language Models as Agents
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 779b1c05-6328-4b05-ab3b-ac7f085a7e96 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Jansen, Marc - Alexandre C \^ o t \' e , and Prithviraj Ammanabrolu
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f8e557f-9eeb-46fc-bf09-8e1b78caf713 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning World Modeling Makes a Better Planner: Dual Preference Optimization for Embodied Task Planning
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb23b17a-7f74-443f-85ac-367da61f1473 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Describe, Explain, Plan and Select: Interactive Planning with Large Language Models Enables Open-World Multi-Task Agents
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36fc8b99-1a2c-46e7-ad87-53556fddc0e6 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Chi, Quoc V
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9312c6b6-0e23-4cd2-bc03-6d4bc44f2aa1 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning AgentGym: Evolving Large Language Model-based Agents across Diverse Environments
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3f216b5-4b4e-4209-a147-8f13ab461ca9 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning The rise and potential of large language model based agents: a survey
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbba7441-fa9f-4444-943d-2d1116594347 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a401f42-8a84-4a96-85aa-d93346523a38 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Watch every step! LLM agent learning via iterative step-level process refinement
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61f5e7dc-8d9f-4914-9e8b-0a5c6114f23a · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Watch every step! LLM agent learning via iterative step-level process refinement
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5ea0e976-666d-420d-a67b-5d90bd1ddfdc · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Qwen2.5 Technical Report
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c0eeca8-6703-4643-aacd-fdc1955a67d2 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning CoPS: Empowering LLM Agents with Provable Cross-Task Experience Sharing
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4c4ce4d-1dc0-4c3b-9edb-93ffd23c3513 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Narasimhan, and Yuan Cao
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31beb251-bf4e-4270-a5cf-b1ac54fdd52f · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning N., Zeyuan Chen, Jianguo Zhang, Devansh Arpit, Ran Xu, Phil Mui, Huan Wang, Caiming Xiong, and Silvio Savarese
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3848eb40-f0f0-4ab3-af49-cebe1ec2661c · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Agent lumos: Unified and modular training for open-source language agents
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8084ff0b-4e90-42fa-bc50-6cd0ad92309c · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning InternLM-Math: Open Math Large Language Models Toward Verifiable Reasoning
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60ac31f0-aca5-4678-9ad2-e093e0cffafa · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd366988-da8b-4428-9b9e-d5d54b3591c0 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Agenttuning: Enabling generalized agent abilities for llms
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac810c3d-7994-44d2-a01e-d81a1b9d02c9 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Enhancing decision-making for LLM agents via step-level q-value models
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cfa7755b-703c-4908-98a1-9002c37c780e · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Large language models as commonsense knowledge for large-scale task planning
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c298c002-3bf0-4527-bdd5-9f85043486d1 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning DeepResearcher: Scaling Deep Research via Reinforcement Learning in Real-world Environments
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 606cad4c-3865-416f-8d17-7281ef795196 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Agents: An Open-source Framework for Autonomous Language Agents
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5861ae12-b48c-443d-b637-12da6f2d9901 · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning K now A gent: Knowledge-augmented planning for LLM -based agents
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 61970ca3-7b6b-4135-9bf8-8e1a78af80fc · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning write newline
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb9523f3-12e8-4d73-a115-2a0433ca4c7a · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning @esa (Ref
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 048ecf1a-75c5-496b-94d8-7d7281645b8f · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning Unresolved cited work
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f087a48-75c0-4512-99f8-8e14c8b6bd0f · outbound
Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f8873d8-e2c1-4d93-b663-d1c42fe201b3 · inbound
RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d509ff7e-1e52-432c-8d78-a450afbf3bec · inbound
RoboAgent: Chaining Basic Capabilities for Embodied Task Planning Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6ce41588-75c6-48de-8f13-a86b0b9fbce5 · inbound
BrainMem: Brain-Inspired Evolving Memory for Embodied Agent Task Planning Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.