Pith. sign in

Paper Citation Record · LEDGER

Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2303.16563.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2303.16563 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:45.947629Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-25T08:35:32.392777Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d995584e-6831-432c-a216-595dddd88057 · inbound

Reasoning with Language Model is Planning with World Model cites this paper.

Reasoning with Language Model is Planning with World Model Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 102

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T01:49:29.086602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-17T01:49:28.796581Z digest=sha256:99a9dc2dd4e6a80b265c67a2a98379ce47e3f708a8c6f57767969dabd4cfe17b

Observation 2d60cd34-4d75-404a-a7e3-415ef3a64af6 · inbound

Voyager: An Open-Ended Embodied Agent with Large Language Models cites this paper.

Voyager: An Open-Ended Embodied Agent with Large Language Models Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:11:41.063113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T13:11:40.995345Z digest=sha256:6befe3fc79cbed7e4c7af88a41c2a0be9528a3a3ce854e608cb5dbdff405f028

Observation 68864cdc-33fc-4197-9577-297aaf79c30f · inbound

Ghost in the Minecraft: Generally Capable Agents for Open-World Environments via Large Language Models with Text-based Knowledge and Memory cites this paper.

Ghost in the Minecraft: Generally Capable Agents for Open-World Environments via Large Language Models with Text-based Knowledge and Memory Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T18:37:20.919338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T18:37:20.889942Z digest=sha256:24cf7f636fe695daa372e1606bf8e9eb2e3f42a134a57161d0048635b26261a0

Observation dc17a8b5-22b0-4654-8e65-859ec19b6b2f · inbound

VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models cites this paper.

VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-13T08:57:22.545092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T08:57:22.299028Z digest=sha256:25110be3456d7367de9e229b5b8346db82aab14b6471f21c9d75ae1c090aa367

Observation 3054bffe-a5a6-46b1-a2b2-ea7ed92c861d · inbound

SkillTree: Explainable Skill-Based Deep Reinforcement Learning for Long-Horizon Control Tasks cites this paper.

SkillTree: Explainable Skill-Based Deep Reinforcement Learning for Long-Horizon Control Tasks Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-25T08:35:32.396566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-25T08:33:45.689151Z digest=sha256:392f6e5f019b234043f7e1918962906baf1f1b28318e2ce3db5f174162996273

Observation 17da2ae2-b907-4130-ab33-a253df18366f · inbound

EvolvingAgent: Curriculum Self-evolving Agent with Continual World Model for Long-Horizon Tasks cites this paper.

EvolvingAgent: Curriculum Self-evolving Agent with Continual World Model for Long-Horizon Tasks Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:32:28.134736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T03:31:13.180660Z digest=sha256:f23a7a69c953c365cb2ae18e5ab57ee7d5ac3480e3a50a8969d85ac12581ef91

Observation 747ba4ea-cbc0-41db-bfef-e65b1e614064 · inbound

BAR: A Backward Reasoning based Agent for Complex Minecraft Tasks cites this paper.

BAR: A Backward Reasoning based Agent for Complex Minecraft Tasks Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:45.947629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:45.947629Z digest=sha256:b0a4b3a702a1032e904918a0d9dfb549a6e322fdb30b01490680830d33bdc89b

Observation 05b6a540-17c4-4c2c-9c15-8ac7645b314a · inbound

SwitchVLA: Execution-Aware Task Switching for Vision-Language-Action Models cites this paper.

SwitchVLA: Execution-Aware Task Switching for Vision-Language-Action Models Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:05:38.066482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:05:38.066482Z digest=sha256:1cb01260739883395f17fb05cb9d57fed1e6c51f0e2809a28c4c6d536ddda47f

Observation 32b89ad3-01d7-4694-a7af-8cc8b92086b1 · inbound

Automated Skill Discovery for Language Agents through Exploration and Iterative Feedback cites this paper.

Automated Skill Discovery for Language Agents through Exploration and Iterative Feedback Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:29.691987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:29.691987Z digest=sha256:5ac007110013ddf1877fe53f3882a668d125b54fb5819181f24168cd21d25348

Observation fe9b2a5e-ecb2-4902-9605-207c538bea61 · inbound

UI-TARS-2 Technical Report: Advancing GUI Agent with Multi-Turn Reinforcement Learning cites this paper.

UI-TARS-2 Technical Report: Advancing GUI Agent with Multi-Turn Reinforcement Learning Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-13T10:13:58.973256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T10:13:58.774968Z digest=sha256:ceff873caaecc37d2beb25b42429b996db1c647845215e30bd92a39cf167fdb8

Observation a00e646c-7f14-4ccd-be5e-a0c359350c86 · inbound

KGLAMP: Knowledge Graph-guided Language model for Adaptive Multi-robot Planning and Replanning cites this paper.

KGLAMP: Knowledge Graph-guided Language model for Adaptive Multi-robot Planning and Replanning Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:10:45.594012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T08:07:48.508022Z digest=sha256:cd66d0f37e5f4772f5fb46c1b00af26d89b632593a70461559052e774de8bfa9

Observation c40154a7-df32-49fc-825e-72e47625bd24 · inbound

Gated Coordination for Efficient Multi-Agent Collaboration in Minecraft Game cites this paper.

Gated Coordination for Efficient Multi-Agent Collaboration in Minecraft Game Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:16:10.548887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T02:03:01.533652Z digest=sha256:72e88a1857e40a1658dd4045c84839011f40d2a0d664ba4d1c2b4dc6b274d281

Observation 1c47d1f8-3594-461d-81fb-0ef69326c1f5 · inbound

ReCAPA: Hierarchical Predictive Correction to Mitigate Cascading Failures cites this paper.

ReCAPA: Hierarchical Predictive Correction to Mitigate Cascading Failures Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T22:44:15.176273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-09T22:21:54.577215Z digest=sha256:7c13911a39200dc925140d7b1b7d7fb3857e357f399538a62dcb5cf5692710ff

Observation 6ed80ed1-a138-4aea-834d-4f35e832b0a2 · inbound

ReCAPA: Hierarchical Predictive Correction to Mitigate Cascading Failures cites this paper.

ReCAPA: Hierarchical Predictive Correction to Mitigate Cascading Failures Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T03:11:19.327198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T03:08:25.891080Z digest=sha256:683d56f849f2a0a69567159e04f72b6e77144c5d5b881079dbf31e2ab6154daa