Pith. sign in

Paper Citation Record · LEDGER

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning

As of 11 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 2 inbound Pith citation observations for arXiv:2510.14828.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2510.14828 v3

Coverage vector

measured 68 of 68 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T09:33:42.644462Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-14T17:54:50.325820Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-14T17:57:33.302151Z

Reference resolution

68 of 68 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved68
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9f77073e-0b7e-4236-8853-8e31202b484e · outbound

This paper cites an unresolved cited work.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:37.164763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:37.164763Z digest=sha256:cc7330cf845d7e2be8330a5917419e126ef27ea7de18c23844d52cbaa509a156

Observation f6499bf0-acd5-4286-ad8a-4f31473a4172 · outbound

This paper cites an unresolved cited work.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:37.248595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:37.248595Z digest=sha256:8e97f6bfdebab244e45a87207d22d79d6d52ba19a6912545591dcd38343804b6

Observation a5cdd6ee-f146-4902-ab57-b49252306109 · outbound

This paper cites Do As I Can, Not As I Say: Grounding Language in Robotic Affordances.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Do As I Can, Not As I Say: Grounding Language in Robotic Affordances

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:37.334499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:37.334499Z digest=sha256:43b79b87a0547367154d8499948502d1aa4e0ffe8fe90283bbe33fbd998c794c

Observation 1c1435f5-dac0-425b-a2bd-3c3b35d214b8 · outbound

This paper cites an unresolved cited work.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:37.421925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:37.421925Z digest=sha256:aba1bc102ef0f97aa371b1c5816c093506eb4b97097df63a64fee43dc88af1ac

Observation 47da6e38-1052-4529-9f20-6bc6fe3661f4 · outbound

This paper cites DigiRL: Training In-The-Wild Device-Control Agents with Autonomous Reinforcement Learning.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning DigiRL: Training In-The-Wild Device-Control Agents with Autonomous Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:37.516249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:37.516249Z digest=sha256:0e87c5dd672fe009db081b30a5a792b389d6d66c472c844321d1f32502a37e68

Observation cc88a67d-9a4b-462d-a8cf-84457bdf72d5 · outbound

This paper cites an unresolved cited work.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:37.596870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:37.596870Z digest=sha256:d11b8b25cee94cc054ed498de91787736d718773014b7b222fff430a38001d9a

Observation 76637523-0550-43f2-a75f-7c0ca3167490 · outbound

This paper cites VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:37.673101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:37.673101Z digest=sha256:524eb3b2c574e12b5089840ad3867ddc0e7e84d63800c4c72b6ceffd680e8aec

Observation ddd89d9e-e6a6-480b-8d46-62ef500fe2c9 · outbound

This paper cites an unresolved cited work.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:37.732719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:37.732719Z digest=sha256:527e2471ab1183a29ad4555f2f2a238fcc9f60fcd60ffd2bbe1365db90eb6941

Observation 3c3a89a6-d496-4f2a-ac87-f3a88b90769f · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:37.824221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:37.824221Z digest=sha256:cec6c7456a813109031d355e9af6e9daea73fe593df1fb7308a075acc2b68741

Observation 287e2341-46b4-44fd-b14e-f4aaa56b3342 · outbound

This paper cites SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:37.915402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:37.915402Z digest=sha256:c2cd7e9eeb27b0a2f0167c9e2c15c1950025dc408078f4ff48923dd49c54accd

Observation 0c97f4fe-f96b-4e6f-88b6-1daac8b70f17 · outbound

This paper cites an unresolved cited work.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:37.979590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:37.979590Z digest=sha256:9e094799c87b70f729eb5e3b3768c20feee43e1f11e7afe4e02f9ea85c61f4af

Observation 76db443e-2e45-4aa5-9f20-d86165ee7018 · outbound

This paper cites Qwen2.5-VL Technical Report.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Qwen2.5-VL Technical Report

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:38.035048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:38.035048Z digest=sha256:264f6710388578599faafe1205b414bfa286315bfcd2b318a70fb2f04eaf5b2d

Observation 007b2173-e4b4-4ab0-aefa-1d67fda17d32 · outbound

This paper cites an unresolved cited work.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:38.122980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:38.122980Z digest=sha256:0671affb08d38222e365819406daafe25f65fb68e7e96cbae11728fa3fee5218

Observation 663cc9ae-d62e-491c-b3c8-513ccbff3950 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:38.225037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:38.225037Z digest=sha256:5adfbbf5bb58f0e622faa4854775c133932b04325f6e3148fc16b5164f0e95ce

Observation 03f72b09-5350-4862-bdd3-4180241fe0f5 · outbound

This paper cites A Survey of Embodied AI: From Simulators to Research Tasks.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning A Survey of Embodied AI: From Simulators to Research Tasks

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:38.311414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:38.311414Z digest=sha256:a87a51255689ba475a3ced0582752c51014a9a73283edcf376c566c56737aab2

Observation c3228154-aca9-4dec-9b7a-b40389d2f0c8 · outbound

This paper cites Agent AI: Surveying the Horizons of Multimodal Interaction.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Agent AI: Surveying the Horizons of Multimodal Interaction

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:38.396393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:38.396393Z digest=sha256:84c1216159f3792cad3fa160986ce892475e1e35732fccb1b13b85a6fae2544b

Observation 6f8873d8-e2c1-4d93-b663-d1c42fe201b3 · outbound

This paper cites Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:38.485506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:38.485506Z digest=sha256:7c588cc6cec8db03c39d4066764dc1d2b798b6a0427549f817b86c90a926d52b

Observation 3c6aa1f6-31f5-42e5-b836-6ce307f2176e · outbound

This paper cites an unresolved cited work.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:38.547736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:38.547736Z digest=sha256:1e26c5fd56bc3c8e837b76c3edbb388a83c9b276f58c4157090d9a7b41ba03b3

Observation a1e41c29-913e-4819-a24b-0f3b5294da57 · outbound

This paper cites Reflective Planning: Vision-Language Models for Multi-Stage Long-Horizon Robotic Manipulation.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Reflective Planning: Vision-Language Models for Multi-Stage Long-Horizon Robotic Manipulation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:38.644459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:38.644459Z digest=sha256:fa59fe45b911e929a91dcee9918afca53caef38a3b543f39fdeed571db34d21d

Observation 95a5c8a6-8a44-46ff-b8ce-25da8195f68b · outbound

This paper cites SRFT: A Single-Stage Method with Supervised and Reinforcement Fine-Tuning for Reasoning.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning SRFT: A Single-Stage Method with Supervised and Reinforcement Fine-Tuning for Reasoning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:38.707015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:38.707015Z digest=sha256:9263b46bf62ef6ad7459a2982998f964f6eb047b3003a591a83973bae6901724

Observation 9ad1ad25-bd18-4fa8-8088-c594776ae649 · outbound

This paper cites an unresolved cited work.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:38.742058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:38.742058Z digest=sha256:33494519a618b937824b8fe741a4d4308cb2e552d4baa5190a3db48906fdaf70

Observation 90be904b-6709-402a-8fce-10bcad633495 · outbound

This paper cites an unresolved cited work.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:38.828513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:38.828513Z digest=sha256:d105288a12705aaffa821d93f0d27baa33bf175d37af7ce107042bd149e6026f

Observation 20cb3071-f4eb-4fde-bb10-bec73bd737e6 · outbound

This paper cites RoboBrain: A Unified Brain Model for Robotic Manipulation from Abstract to Concrete.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning RoboBrain: A Unified Brain Model for Robotic Manipulation from Abstract to Concrete

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:38.976923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:38.976923Z digest=sha256:d3f72ca9593f64ac6aab79f6c4f79cddb53c8da60c205fc0ede1a9e9726f6a5c

Observation 3e9004e9-3fa5-45a2-8261-efb517b7d826 · outbound

This paper cites AlphaDrive: Unleashing the Power of VLMs in Autonomous Driving via Reinforcement Learning and Reasoning.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning AlphaDrive: Unleashing the Power of VLMs in Autonomous Driving via Reinforcement Learning and Reasoning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:39.054571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:39.054571Z digest=sha256:2048acc2b58bf8635b69399fe9f47159803595113cc0a1a9de026df52764b98f

Observation 7894e5b7-0ba8-48a8-a727-e74e6c2bce8e · outbound

This paper cites AI2-THOR: An Interactive 3D Environment for Visual AI.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning AI2-THOR: An Interactive 3D Environment for Visual AI

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:39.183849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:39.183849Z digest=sha256:0c068ff771bb18bf39ceb7dc10993bfc905b6908fa194bf11c01c4b795da0de0

Observation 239f54c4-f565-4212-99fa-7e479d963c5c · outbound

This paper cites VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:39.319465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:39.319465Z digest=sha256:f2b3b10bac888437401d190b0e9e586929c95dedab3800c58042267682ba8979

Observation 47e90a23-9977-4161-9e0f-227a905396ff · outbound

This paper cites Large Model Empowered Embodied AI: A Survey on Decision-Making and Embodied Learning.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Large Model Empowered Embodied AI: A Survey on Decision-Making and Embodied Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:39.414690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:39.414690Z digest=sha256:ce07a763398d96a90cf6bb8877776f489f1dc29bfb695ec90de89e0ebb1a3a07

Observation 053d6d53-f0ab-48d1-8353-563cfd4304aa · outbound

This paper cites Aligning Cyber Space with Physical World: A Comprehensive Survey on Embodied AI.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Aligning Cyber Space with Physical World: A Comprehensive Survey on Embodied AI

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:39.470901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:39.470901Z digest=sha256:dd31db170d899a9e320bc4298d1f5ab4fc8b6ce79d956a29866f7476a5d25cc9

Observation e972aef4-4091-42b8-9514-11782e3d20d8 · outbound

This paper cites Visual-RFT: Visual Reinforcement Fine-Tuning.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Visual-RFT: Visual Reinforcement Fine-Tuning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:39.545675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:39.545675Z digest=sha256:a88b2379eb4ef97263b3ccf5dfc46274395736c2c51869bba990ffa05b42c367

Observation 9d1e0453-2504-4012-ac5c-c00a62dbcb46 · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:39.629890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:39.629890Z digest=sha256:3a2e114ed05effc08220bd1645c69c374eb2785b47d303fd0e7a5512075be93e

Observation 60901f06-4fbb-470e-9d54-7cda08d151cb · outbound

This paper cites an unresolved cited work.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:39.735062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:39.735062Z digest=sha256:b227b4b0cb7b66f1fe8690c7e3e78024215cd3ed883837492766060c973cac9e

Observation 6c37f363-c13a-411c-99cd-3b7ced36bfd9 · outbound

This paper cites OpenAI o1 System Card.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning OpenAI o1 System Card

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:39.811735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:39.811735Z digest=sha256:c315a546afa0e794e7cb95c2b94a35ab320cbff8f52a8b0f8d62368b6481cde5

Observation ff9aac9d-6f3d-4b0b-8b05-791516e27ebd · outbound

This paper cites an unresolved cited work.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:39.875355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:39.875355Z digest=sha256:2b561b8446ecf554c0ddb69aa2718bfe50ee66c3a64f40e115a1f5cda4870f66

Observation b976c30e-d0f1-475c-9a13-8673b6cd41e2 · outbound

This paper cites an unresolved cited work.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:39.950444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:39.950444Z digest=sha256:35f53736b1487426c3da9bbfc66a34f9dab916307b34f7e81668bc403f086e18

Observation 9f7ee5c6-177d-49ff-9d5c-58239ead11f2 · outbound

This paper cites an unresolved cited work.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:40.030018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:40.030018Z digest=sha256:16573e5308db5aa873a67ece7f00591d88c21bcf37e43037c8f8ed1f05559d55

Observation 356fe00a-db9c-4419-bb9a-22d65ee941ed · outbound

This paper cites Habitat: A Platform for Embodied AI Research.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Habitat: A Platform for Embodied AI Research

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:40.107534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:40.107534Z digest=sha256:fe9ed21de377a656b33125f7772971c06557c9f622939589592fae6a592d450d

Observation e5524ead-982f-40f3-8021-cfc34f213871 · outbound

This paper cites RoboVQA: Multimodal Long-Horizon Reasoning for Robotics.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning RoboVQA: Multimodal Long-Horizon Reasoning for Robotics

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:40.163623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:40.163623Z digest=sha256:640046303856b31b66940c41715e23ecab28165eabbcbdab39a159d7f4e4eeb0

Observation ae0cad2e-052a-4f96-9b80-abbe5e7545ba · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:40.222514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:40.222514Z digest=sha256:9daa6fdcfef106441cb8063c56e2fc295f3db2d06596b857ccf84e9fac72deba

Observation 4b282ff1-f0d4-4fb3-982e-3019f96ae079 · outbound

This paper cites VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:40.274255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:40.274255Z digest=sha256:10fb43e785c9b8d01339238ff09738026ce307e6f4e33c07e32aff339be3c5cb

Observation 070e2f59-2fb9-4760-a14c-4a030752f44b · outbound

This paper cites Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:40.382324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:40.382324Z digest=sha256:82f5ca46c02bb7843459c79fa8ff076550b513c0383bb4aca388a904c1134cad

Observation 6841738b-d346-445d-9af7-c50c12831ce4 · outbound

This paper cites an unresolved cited work.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:40.470672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:40.470672Z digest=sha256:164dca229a52c5cac6e13ed992ea9365ac5091bbded9ee336075f695448648b9

Observation 23505add-31dc-4721-b5af-dc5c2eb429a2 · outbound

This paper cites ALFRED: A Benchmark for Interpreting Grounded Instructions for Everyday Tasks.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning ALFRED: A Benchmark for Interpreting Grounded Instructions for Everyday Tasks

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:40.630287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:40.630287Z digest=sha256:d721d887b7d7e7374982decfe7406097ec59fd0c6485bfac6f759cebb70a8a97

Observation e829214c-6955-4ef7-a539-c8a444a7be2b · outbound

This paper cites LLM-Planner: Few-Shot Grounded Planning for Embodied Agents with Large Language Models.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning LLM-Planner: Few-Shot Grounded Planning for Embodied Agents with Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:40.714342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:40.714342Z digest=sha256:d4bf1dfcffffbdb2a070245bfb2629f4351f1691d326b1fe6937bc193bf1ae5a

Observation 29e3a671-a61a-4d50-bdc4-781527c4b06f · outbound

This paper cites Gemma 3 Technical Report.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Gemma 3 Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:40.778430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:40.778430Z digest=sha256:b0643b2497c18338c5045658104b21701142ca1805182646cde46cd8d23d98e9

Observation e38770a1-69dd-46fe-9f6b-2e2af7817e33 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:40.825945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:40.825945Z digest=sha256:5913bd509e670bd13b1e496f18e50b1f9fea85202c76451d18f1184f0d3ead04

Observation 62801cd4-ce5a-4f1e-abfb-8c526f5297b7 · outbound

This paper cites Enhancing LLM Reasoning with Iterative DPO: A Comprehensive Empirical Investigation.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Enhancing LLM Reasoning with Iterative DPO: A Comprehensive Empirical Investigation

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:40.888746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:40.888746Z digest=sha256:adb02f340ff5964213d0b394334a53b59e558e620b592a6daf36f5d0fdf3f473

Observation 951575bc-7608-422b-8ab0-c162ca620349 · outbound

This paper cites an unresolved cited work.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:40.948686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:40.948686Z digest=sha256:d6a881c3d673b89e203b1228a0880db36d61f4fbc1ebf2c2417df63d63e81c69

Observation 3efd8b4e-e54e-4b28-8dc2-3bce5785d955 · outbound

This paper cites Online Preference-based Reinforcement Learning with Self-augmented Feedback from Large Language Model.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Online Preference-based Reinforcement Learning with Self-augmented Feedback from Large Language Model

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:41.033717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:41.033717Z digest=sha256:2491d44e7ec18f36ef54adc8bde8fb18fce8387ac66a5d3f48e03c6f813a7e14

Observation 87f8af6c-ef6f-4912-bfd4-c5f1fef05f77 · outbound

This paper cites Voyager: An Open-Ended Embodied Agent with Large Language Models.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Voyager: An Open-Ended Embodied Agent with Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:41.105277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:41.105277Z digest=sha256:ad455f60ee2fc5538fde4cbab7679330499ab01565a3b9392e4e5e3ffaa21b19

Observation abbbf8a1-ad4d-4672-9847-8cecb376950b · outbound

This paper cites an unresolved cited work.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:41.165390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:41.165390Z digest=sha256:bc8f29e644a09931ec4332c5671840a9365a3704748f2c9a53c01cb0ee0d6753

Observation cfd41063-37b0-4fa7-b3d8-dd85676fa80c · outbound

This paper cites Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:41.282310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:41.282310Z digest=sha256:eb6ba0af9e6a3648c304d0c8f4f0ced731f2eb3e4cb9be3a75355225fc6a9ee8

Observation e7824725-ce5f-4f52-b8f4-37a09d7b1c71 · outbound

This paper cites arXiv:2509.24494 [cs.CL] https://arxiv.org/abs/2509.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning arXiv:2509.24494 [cs.CL] https://arxiv.org/abs/2509

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:41.226056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:41.226056Z digest=sha256:31ba1223dd4cf6f14abf8aff1bd35bd6952f841def86bb6ac5b3ec6ca392f9d4

Observation cd9f6d05-1b05-40a7-99fa-612ad0a20a1a · outbound

This paper cites SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:41.412380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:41.412380Z digest=sha256:cb92fc4eaa5c1fca3574fd503bcc23bc94619e63c168fe294ba6dfbb0899c1d2

Observation 08d3e24a-16de-4ef9-9daf-916074e2ae47 · outbound

This paper cites ScienceWorld: Is your Agent Smarter than a 5th Grader?.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning ScienceWorld: Is your Agent Smarter than a 5th Grader?

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:41.370086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:41.370086Z digest=sha256:c89a542ad52d260e4ba5a0d8714f9781e313e01d2965d2cfd92a153db02b8af9

Observation 228d765b-1279-480a-b96e-82f25ef54e93 · outbound

This paper cites Reinforced Reasoning for Embodied Planning.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Reinforced Reasoning for Embodied Planning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:41.474386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:41.474386Z digest=sha256:813aee2fc73460fd1d8bb7ee6fac35e995a76a0b98cfbeadb1b1194ce5d2b581

Observation 68a20794-c4fb-400f-b5e3-ded744db4baa · outbound

This paper cites The Rise and Potential of Large Language Model Based Agents: A Survey.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning The Rise and Potential of Large Language Model Based Agents: A Survey

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:41.700343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:41.700343Z digest=sha256:ffe72158f5d3e159931707570d89039eb9f8a9bd299eefb32902c21077a785ac

Observation f1accbf5-2d65-4d33-890c-c145f07b63a5 · outbound

This paper cites Embodied Task Planning with Large Language Models.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Embodied Task Planning with Large Language Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:41.629369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:41.629369Z digest=sha256:0bca24b7d080c1ef898065e16681c42c769555a9297a3d4b7f91fa629584c6ef

Observation 2ce6b19d-f7b0-40cc-bf15-783a3356c17d · outbound

This paper cites EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:41.839773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:41.839773Z digest=sha256:cc3cfa4c2d3e7ff9be54818af7739494cb25a66d30cefbb3687e6a354a9cdfb7

Observation 7759031b-650e-4aaf-97af-23bb67c10b09 · outbound

This paper cites A Survey on Robotics with Foundation Models: toward Embodied AI.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning A Survey on Robotics with Foundation Models: toward Embodied AI

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:41.788473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:41.788473Z digest=sha256:762e5e3027feb92fa8f029c576a9f9e9157b3c9d6a2261ba743738501400d27e

Observation 3c9ae41d-8adc-41e7-b0f8-0f50a4da9c7d · outbound

This paper cites Robotic Control via Embodied Chain-of-Thought Reasoning.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Robotic Control via Embodied Chain-of-Thought Reasoning

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:42.041223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:42.041223Z digest=sha256:dd6ad15e2c36851303d00216df607c01d5767b90aaf66ad391794393fea95264

Observation e8e7f942-9ec8-44d1-a197-56410fa1ff20 · outbound

This paper cites Demystifying Long Chain-of-Thought Reasoning in LLMs.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Demystifying Long Chain-of-Thought Reasoning in LLMs

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:41.924759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:41.924759Z digest=sha256:f6a6244a229e933d1f43f550aaa0bb6eb4e4e9231b4ce5ea534b6fd3304954c9

Observation 15308678-5de8-4ed6-83b7-e44781a14490 · outbound

This paper cites ACECODER: Acing Coder RL via Automated Test-Case Synthesis.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning ACECODER: Acing Coder RL via Automated Test-Case Synthesis

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:42.262857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:42.262857Z digest=sha256:325bbe97aaf3744324862704c438b54ec9102b45e3ead120d0908ca3295afde9

Observation a041484a-e109-4286-b7a9-6420c0b5cf8e · outbound

This paper cites AgentTuning: Enabling Generalized Agent Abilities for LLMs.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning AgentTuning: Enabling Generalized Agent Abilities for LLMs

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:42.145467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:42.145467Z digest=sha256:bc55733b0efb508398ef08b726461dd8d8c981024348d35ea6ab1c57776d494d

Observation 088f09f5-4de6-462f-879c-1e007c5d2c71 · outbound

This paper cites R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:42.482668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:42.482668Z digest=sha256:59ec92ea551aae6bcfe43f7eb902515693d8532f4228fda16665a25ca5ab5149

Observation 06b7ccd8-d487-4846-ac50-b8239a562185 · outbound

This paper cites Vision-Language Models for Vision Tasks: A Survey.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Vision-Language Models for Vision Tasks: A Survey

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:42.375221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:42.375221Z digest=sha256:81e91d63f48e3152bb9dda58d858447f2405e79cce3a312d6e646619cefcc63e

Observation f88f39fd-5a20-47bc-aadf-f41e12188cba · outbound

This paper cites RLVMR: Reinforcement Learning with Verifiable Meta-Reasoning Rewards for Robust Long-Horizon Agents.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning RLVMR: Reinforcement Learning with Verifiable Meta-Reasoning Rewards for Robust Long-Horizon Agents

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:42.644462Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:42.644462Z digest=sha256:dbcf275f71fe7a3b9f39b140a3fb74da00b6a31bd4c08824cafee428466efb8d

Observation eba6f6a2-bc86-409c-ae91-36b052a70dc3 · outbound

This paper cites Embodied-Reasoner: Synergizing Visual Search, Reasoning, and Action for Embodied Interactive Tasks.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Embodied-Reasoner: Synergizing Visual Search, Reasoning, and Action for Embodied Interactive Tasks

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:42.587307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:42.587307Z digest=sha256:b51e399608a523a993314c9858b6857c72ebdf38aeddf601196183164cdaf870

Observation e1099645-a944-49a6-9be7-50fa2d184a26 · outbound

This paper cites Socratic Planner: Self-QA-Based Zero-Shot Planning for Embodied Instruction Following.

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning Socratic Planner: Self-QA-Based Zero-Shot Planning for Embodied Instruction Following

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-04T09:33:40.546288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:33:40.546288Z digest=sha256:a5f04f3a1946967727bc4bda0af9deecb18932375b6b22566af1889c19d98c9f

Pith citing papers

Observation 5929a515-634a-42c3-bd96-e4cbe4d82e70 · inbound

RoboAgent: Chaining Basic Capabilities for Embodied Task Planning cites this paper.

RoboAgent: Chaining Basic Capabilities for Embodied Task Planning RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-06-10T02:12:36.666320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T18:15:08.727921Z digest=sha256:b7928fd2e4d55744ab553039b9c3f3faa0654e6f964cb53d3f5d8e07322cf975

Observation 6a75c318-e443-42ec-9013-e5b95c201d67 · inbound

RoboEvolve: Co-Evolving Planner-Simulator for Robotic Manipulation with Limited Data cites this paper.

RoboEvolve: Co-Evolving Planner-Simulator for Robotic Manipulation with Limited Data RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-10T02:12:36.666320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-14T17:54:50.325820Z digest=sha256:388b1c5059ba91b267bacec9f530e58a58c7d6cf539feaac6d9e6b24b0bbadc5