Pith. sign in

Paper Citation Record · LEDGER

Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2104.07749.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2104.07749 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:21:25.121008Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T06:53:12.435264Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5e376256-7564-4fe6-8cf9-1554d12f1bb0 · inbound

Do As I Can, Not As I Say: Grounding Language in Robotic Affordances cites this paper.

Do As I Can, Not As I Say: Grounding Language in Robotic Affordances Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

Reference 84

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:24:06.320155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T22:24:05.999350Z digest=sha256:31f1c03c37033c479935ff9eff1b155472e27107bfc18aae88a5ed5e8a2f8d0f

Observation 1e636b8c-b707-462d-a13f-1bbbb93a1da9 · inbound

VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training cites this paper.

VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T04:42:52.677145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T04:42:52.627166Z digest=sha256:d2d5cbf64463eb96cf22799f266cedb348542481a859200284be45a1c0db6480

Observation 933a4504-06f1-4922-9b93-77336d51342f · inbound

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies cites this paper.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.121008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.121008Z digest=sha256:19507f7914db03dd7ba15b46e84083d99a0cb1f0f1989a566b06739945237e7a

Observation 70300d74-6b40-4983-932a-2c0b16915638 · inbound

Reachability Weighted Offline Goal-conditioned Resampling cites this paper.

Reachability Weighted Offline Goal-conditioned Resampling Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-07T11:24:59.006613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:24:59.006613Z digest=sha256:302ec172178bf3be51bdc1735a147c74bbc9de4241f37a231528b2bf4c69bfed

Observation cb939337-9d3a-4bff-99dd-4dab5c33a910 · inbound

V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning cites this paper.

V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:33:50.741002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T00:33:50.471804Z digest=sha256:2a950fc7d62f559c3b659f2f9059c1fa76c6da0f493b0fb289385368f039b2a8

Observation 98b7e95c-a3b6-4886-ab6d-1702aa029fe1 · inbound

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement cites this paper.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.043894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.043894Z digest=sha256:2b96fca296c83be10c2a846136833bace91f32571826caeb5bdcb048d4b6a760

Observation eb5c0813-1f02-4401-9cbc-e656adcb021b · inbound

Chain-of-Goals Hierarchical Policy for Long-Horizon Offline Goal-Conditioned RL cites this paper.

Chain-of-Goals Hierarchical Policy for Long-Horizon Offline Goal-Conditioned RL Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

Reference 318

Resolution
unresolved
no resolver link, observed 2026-08-03T05:02:04.408498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:02:04.408498Z digest=sha256:7ce27e72020e0a58c383a92d9977a8fb74442cd5360383a7d8920717319fd304

Observation 8331ad51-66f8-4152-aeb1-89aa3e6113a6 · inbound

Mollified Value Learning cites this paper.

Mollified Value Learning Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T20:29:19.303477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:29:19.303477Z digest=sha256:3e6f3edcb810830fc0140139ffea3359397e9bab12d1364c46a75d853f652cb4

Observation 02bd0867-dcb5-4319-a41d-e1de17d958aa · inbound

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL cites this paper.

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:30:58.049416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-11T01:17:48.643521Z digest=sha256:09647d57537e0c7dd501a21780fb4ce795354a48018de9e33703e5c033bacbf3

Observation 1da7c0c8-fb2e-4d32-81d6-5bc91d480ade · inbound

Behavior-Consistent Deep Reinforcement Learning cites this paper.

Behavior-Consistent Deep Reinforcement Learning Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

Reference 182

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:39:40.716545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-21T05:37:29.919862Z digest=sha256:29daa3567a502bd7fe24f356944f55836769f52f2ffe5b45658d5b5cec51b623

Observation 34bb6813-18c5-4ef8-9969-85b719fd8cb9 · inbound

Behavior-Consistent Deep Reinforcement Learning cites this paper.

Behavior-Consistent Deep Reinforcement Learning Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

Reference 182

Resolution
verified exact
arxiv_id, observed 2026-05-22T10:06:21.792760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-22T10:06:04.478006Z digest=sha256:afad2e6427c2526b8fde2bf24f45267219277e344387d27ec86bff159037b170

Observation d3409602-ef58-4535-b591-0baf217837f6 · inbound

Physics-informed Goal-Conditioned Reinforcement Learning under Hybrid Contact Dynamics cites this paper.

Physics-informed Goal-Conditioned Reinforcement Learning under Hybrid Contact Dynamics Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-29T06:53:12.436430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T06:51:33.077793Z digest=sha256:83b06e7525b3b4e89fc27976bb6c5ca02807e19e282d8152323f025b1456ae66

Observation a99b93e8-fc91-484b-9fd1-59171dabe709 · inbound

Relative Value Learning cites this paper.

Relative Value Learning Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

Reference 114

Resolution
unresolved
no resolver link, observed 2026-08-01T08:32:00.373209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T08:32:00.373209Z digest=sha256:e5d9b10c55ed3bd7dae493d8f2af69e521a07267a09062fee02ad90a17e85a65