Pith. sign in

Paper Citation Record · LEDGER

OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2010.13611.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2010.13611 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:15:42.139432Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T23:57:29.181327Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e67ef3e2-5dc7-4102-b3b9-af35d619c232 · inbound

Decision Transformer: Reinforcement Learning via Sequence Modeling cites this paper.

Decision Transformer: Reinforcement Learning via Sequence Modeling OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T15:11:11.314700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T15:11:11.056013Z digest=sha256:e954ec9f6d801f468122fc3e52eb4d0415cda1f126d0c987ee1d944b506fae41

Observation 24ac5f17-dd9e-4fd0-82fd-598970556c8f · inbound

What Matters in Learning from Offline Human Demonstrations for Robot Manipulation cites this paper.

What Matters in Learning from Offline Human Demonstrations for Robot Manipulation OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-13T08:51:55.958219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T08:51:55.826747Z digest=sha256:a6ebdf2f64c1771c5ea3be98c36ea8944da55282e4336e991592ce9268d49df5

Observation ff9dbc7f-779e-41fd-89b2-78b7f97a0a6c · inbound

Behavioral Exploration: Learning to Explore via In-Context Adaptation cites this paper.

Behavioral Exploration: Learning to Explore via In-Context Adaptation OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:42.139432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:42.139432Z digest=sha256:ea5d39efe9a265bda5de62efa100d1261ad47aeeeeb4d7c185ec784590506e74

Observation 1192f49f-ab8d-4644-805a-7744d51c4ed3 · inbound

Learning Temporal Abstractions via Variational Homomorphisms in Option-Induced Abstract MDPs cites this paper.

Learning Temporal Abstractions via Variational Homomorphisms in Option-Induced Abstract MDPs OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:34.847846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:17:34.847846Z digest=sha256:7c8d9a5ddb836d152d7f819668cb4959f4f97ed1df79b184247321868afa1b6b

Observation fbbd8599-f619-4da0-873b-2e186e9b76db · inbound

Learning Upper Lower Value Envelopes to Shape Online RL: A Principled Approach cites this paper.

Learning Upper Lower Value Envelopes to Shape Online RL: A Principled Approach OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T08:45:47.978112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T08:45:47.978112Z digest=sha256:bbe909cbea87bc924b2650ddbd0eeff941380e7f50c80b124e72a17bbc2d84ae

Observation 384aca80-181a-45a0-93e7-fffcd6613991 · inbound

Learning Semantic Atomic Skills for Multi-Task Robotic Manipulation cites this paper.

Learning Semantic Atomic Skills for Multi-Task Robotic Manipulation OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T15:05:01.788849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:05:01.788849Z digest=sha256:cd92ca28a9496a059e6729fc9666546e2285319e02967c9df52abcf442e097a2

Observation 22e07cd7-df74-4651-be1c-1af7f93bbba8 · inbound

Implicit Safety Alignment from Crowd Preferences cites this paper.

Implicit Safety Alignment from Crowd Preferences OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-22T08:31:16.907176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-22T08:28:41.652865Z digest=sha256:c2069add2bef3371352d54b7890b590c69944e6ea726844333ef276d46e561c1

Observation 1fa937b6-3149-4cd3-89a4-eddf17901ccd · inbound

Counterfactual Transport Flows for Offline Conservative Trajectory Refinement cites this paper.

Counterfactual Transport Flows for Offline Conservative Trajectory Refinement OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-07-02T23:57:29.182689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T17:33:35.857240Z digest=sha256:98b74310f88fe63a263966bd030c1fd89a6887f8b678f454cd597b88716aa30a

Observation 097a2308-2c05-487d-91ef-4ae79db223c1 · inbound

Behavior Uncloning: Distilling Mode Redirection into Policy Weights without Inference-Time Steering cites this paper.

Behavior Uncloning: Distilling Mode Redirection into Policy Weights without Inference-Time Steering OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:54:22.212408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T07:49:04.825693Z digest=sha256:7e03dce1b5f1e6ecd270e7c83d020e5323e2c654bd45bfbd07e121398a4fce0e

Observation 01c0d124-f73a-4f69-979b-8e84014c9005 · inbound

Adapting Generalist Robot Policies with Semantic Reinforcement Learning cites this paper.

Adapting Generalist Robot Policies with Semantic Reinforcement Learning OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:45:42.729562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-01T05:09:29.625066Z digest=sha256:925901c2fdaa7f750680f8500f105e5e4886a1f07843a37eae50541afa7265fd

Observation ed133a94-6119-455a-9f1d-3b4e59acae50 · inbound

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills cites this paper.

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T19:45:26.462169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:45:26.462169Z digest=sha256:1ab3d90c97e969bdc77e94af93ffa4d716c7a210eb13918d60ed709ee5cdcc25