Pith. sign in

Paper Citation Record · LEDGER

Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2303.16563.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2303.16563 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:45.947629Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-25T08:35:32.392777Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d995584e-6831-432c-a216-595dddd88057 · inbound

Reasoning with Language Model is Planning with World Model cites this paper.

Reasoning with Language Model is Planning with World Model Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 102

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T01:49:29.086602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T01:49:28.796581Z digest=sha256:4cf6658e7687642fb2f44dfdf9bb6d569aefcd21673a0e2434304af8a341daf4

Observation 2d60cd34-4d75-404a-a7e3-415ef3a64af6 · inbound

Voyager: An Open-Ended Embodied Agent with Large Language Models cites this paper.

Voyager: An Open-Ended Embodied Agent with Large Language Models Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:11:41.063113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T13:11:40.995345Z digest=sha256:d98290ce5e12241269d49426dcbbd06a3904c42c1c84103ec874e0f357fb0da9

Observation 68864cdc-33fc-4197-9577-297aaf79c30f · inbound

Ghost in the Minecraft: Generally Capable Agents for Open-World Environments via Large Language Models with Text-based Knowledge and Memory cites this paper.

Ghost in the Minecraft: Generally Capable Agents for Open-World Environments via Large Language Models with Text-based Knowledge and Memory Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T18:37:20.919338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T18:37:20.889942Z digest=sha256:c14b2d09a6103bda6d1dce99831293699088030f0f64d678b3cb5aa404cf5a94

Observation dc17a8b5-22b0-4654-8e65-859ec19b6b2f · inbound

VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models cites this paper.

VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-13T08:57:22.545092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T08:57:22.299028Z digest=sha256:42172a9c16fbdf9f3d75d33cf763d6f3abf35ac0a3d33f74a9c4d59dcb06bedf

Observation 3054bffe-a5a6-46b1-a2b2-ea7ed92c861d · inbound

SkillTree: Explainable Skill-Based Deep Reinforcement Learning for Long-Horizon Control Tasks cites this paper.

SkillTree: Explainable Skill-Based Deep Reinforcement Learning for Long-Horizon Control Tasks Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-25T08:35:32.396566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-25T08:33:45.689151Z digest=sha256:0cc0d113dfe0bbe561da66595d08f6ee868535432edc59a5d287ecad1df594c2

Observation 17da2ae2-b907-4130-ab33-a253df18366f · inbound

EvolvingAgent: Curriculum Self-evolving Agent with Continual World Model for Long-Horizon Tasks cites this paper.

EvolvingAgent: Curriculum Self-evolving Agent with Continual World Model for Long-Horizon Tasks Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:32:28.134736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T03:31:13.180660Z digest=sha256:c26ef8da9ca705916811e7f06b0eb0b41626aa04616d5958a4f2e29cdca9538a

Observation 747ba4ea-cbc0-41db-bfef-e65b1e614064 · inbound

BAR: A Backward Reasoning based Agent for Complex Minecraft Tasks cites this paper.

BAR: A Backward Reasoning based Agent for Complex Minecraft Tasks Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:45.947629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:45.947629Z digest=sha256:b0a4b3a702a1032e904918a0d9dfb549a6e322fdb30b01490680830d33bdc89b

Observation 05b6a540-17c4-4c2c-9c15-8ac7645b314a · inbound

SwitchVLA: Execution-Aware Task Switching for Vision-Language-Action Models cites this paper.

SwitchVLA: Execution-Aware Task Switching for Vision-Language-Action Models Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:05:38.066482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:05:38.066482Z digest=sha256:1cb01260739883395f17fb05cb9d57fed1e6c51f0e2809a28c4c6d536ddda47f

Observation 32b89ad3-01d7-4694-a7af-8cc8b92086b1 · inbound

Automated Skill Discovery for Language Agents through Exploration and Iterative Feedback cites this paper.

Automated Skill Discovery for Language Agents through Exploration and Iterative Feedback Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:29.691987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:29.691987Z digest=sha256:5ac007110013ddf1877fe53f3882a668d125b54fb5819181f24168cd21d25348

Observation fe9b2a5e-ecb2-4902-9605-207c538bea61 · inbound

UI-TARS-2 Technical Report: Advancing GUI Agent with Multi-Turn Reinforcement Learning cites this paper.

UI-TARS-2 Technical Report: Advancing GUI Agent with Multi-Turn Reinforcement Learning Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-13T10:13:58.973256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T10:13:58.774968Z digest=sha256:aadd79e9e7d46e4755ae02a1b408d8a1dc2b2a2357c146da570796cb877a202c

Observation a00e646c-7f14-4ccd-be5e-a0c359350c86 · inbound

KGLAMP: Knowledge Graph-guided Language model for Adaptive Multi-robot Planning and Replanning cites this paper.

KGLAMP: Knowledge Graph-guided Language model for Adaptive Multi-robot Planning and Replanning Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:10:45.594012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T08:07:48.508022Z digest=sha256:d87b7069884cdb9ae25f5fbeb9aa046e19365b2f3d42f36a1cac4462ae9433cc

Observation c40154a7-df32-49fc-825e-72e47625bd24 · inbound

Gated Coordination for Efficient Multi-Agent Collaboration in Minecraft Game cites this paper.

Gated Coordination for Efficient Multi-Agent Collaboration in Minecraft Game Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:16:10.548887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T02:03:01.533652Z digest=sha256:0ffaf62129622d4defd544e02fab8da8586151c260f6a72d6c4ca41a54168ff6

Observation 1c47d1f8-3594-461d-81fb-0ef69326c1f5 · inbound

ReCAPA: Hierarchical Predictive Correction to Mitigate Cascading Failures cites this paper.

ReCAPA: Hierarchical Predictive Correction to Mitigate Cascading Failures Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T22:44:15.176273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-09T22:21:54.577215Z digest=sha256:1f755da0de695ac1badd7abaa0a811ebb4a6b3e55441991f1e5ae1e2169e4c16

Observation 6ed80ed1-a138-4aea-834d-4f35e832b0a2 · inbound

ReCAPA: Hierarchical Predictive Correction to Mitigate Cascading Failures cites this paper.

ReCAPA: Hierarchical Predictive Correction to Mitigate Cascading Failures Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T03:11:19.327198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T03:08:25.891080Z digest=sha256:6b8e261d950d2772c2ef8d30209269fa5a75c83de0393f0dde637ebfa6e6363c