Pith. sign in

Paper Citation Record · LEDGER

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement

As of 11 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 0 inbound Pith citation observations for arXiv:2507.06701.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.06701 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:06:06.951316Z

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a27a6172-f7a0-46cf-b531-e4a07e871936 · outbound

This paper cites Wish you were here: Hindsight Goal Selection for long-horizon dexterous manipulation.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Wish you were here: Hindsight Goal Selection for long-horizon dexterous manipulation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.116005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.116005Z digest=sha256:7652313faa858b35ebb9673c865af7dbab67cae2191108d5551acd098d635ba1

Observation 35e2ba22-e5ea-4da3-aa74-c9507fab7691 · outbound

This paper cites A Generalist Agent.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement A Generalist Agent

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.602130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.602130Z digest=sha256:3f7c884a5af44fbc16107ca16828eb2c253af34b6f3352df8f9c480a91b4897d

Observation 1d3d755d-d3b6-4273-a0a9-604db3f09052 · outbound

This paper cites an unresolved cited work.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:06:07.444819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-06T19:06:06.743130Z digest=sha256:926e0b3ec5a3dc3053d44dd91db8f2a00b91c28b5c575104a9857bc68a0e6626

Observation 99c2d47e-7fff-4ee7-b1fd-7501c86f79c6 · outbound

This paper cites Offline Learning from Demonstrations and Unlabeled Experience.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Offline Learning from Demonstrations and Unlabeled Experience

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.902037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.902037Z digest=sha256:b25f16cce94de77f0d8e5a85097eec4b249cf34d35b125df41a61b3263826581

Observation 9c306700-d951-45f0-8fea-32d6a439cce9 · outbound

This paper cites an unresolved cited work.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:06:07.318510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-06T19:06:06.951316Z digest=sha256:3f84c22666850d00055a71d97a2b167560eb02a8681766815417b55705354021

Observation 379b2b3e-48a3-4756-a055-b3f92a1131a7 · outbound

This paper cites Large Language Models Can Self-Improve.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Large Language Models Can Self-Improve

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.248583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.248583Z digest=sha256:b068da3d3e6bc5f095a80a78232931246f739fadb5caff5a0856f7071dcd68ed

Observation 29bfdde9-336a-4003-b743-3b7accb25f0d · outbound

This paper cites Maximum Entropy Deep Inverse Reinforcement Learning.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Maximum Entropy Deep Inverse Reinforcement Learning

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.805435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.805435Z digest=sha256:bc786d18307c239ffe1f1a11ac5a387ec1064827fdb2482ad403e1443d15324a

Observation 4a047633-9609-4812-8932-74bf9035e92d · outbound

This paper cites Kwon, T., Palo, N.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Kwon, T., Palo, N

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.454510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.454510Z digest=sha256:62d81a5a3829be7ebf60d5c7bbc1323cd2237cd0a8c4929c38fcf4e4d009df9b

Observation 98b7e95c-a3b6-4886-ab6d-1702aa029fe1 · outbound

This paper cites Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.043894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.043894Z digest=sha256:0083513df1785f488b35d179f69ce498655b1a1a057bb2717304b1beedf18390

Observation a695a859-fe68-432b-9b3f-443714294e47 · outbound

This paper cites Imitation Learning via Off-Policy Distribution Matching.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Imitation Learning via Off-Policy Distribution Matching

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.332378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.332378Z digest=sha256:6b23d92d519696afe57b7d525d7449f74f611944ffe65e0b2ec666d5b6acd62f

Observation d3dbb16b-69ea-4a80-9af1-4462da4e71b3 · outbound

This paper cites Imitating Language via Scalable Inverse Reinforcement Learning.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Imitating Language via Scalable Inverse Reinforcement Learning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.864052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.864052Z digest=sha256:66829b1155dcd08f45f810d1c67b8e4a664d78d27172a49429f5adb20232dcf7

Observation a681f6c3-2c41-460e-9cd3-9279e8e30a02 · outbound

This paper cites org/CorpusID:269605913.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement org/CorpusID:269605913

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.677382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.677382Z digest=sha256:123ef25ba94464230ab9e8c59def3e8294057725fe5980b1124078b5c970f9ad

Pith citing papers

No inbound Pith citation observations are available.