Pith. sign in

Paper Citation Record · LEDGER

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement

As of 10 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 0 inbound Pith citation observations for arXiv:2507.06701.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.06701 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:06:06.951316Z

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a27a6172-f7a0-46cf-b531-e4a07e871936 · outbound

This paper cites Wish you were here: Hindsight Goal Selection for long-horizon dexterous manipulation.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Wish you were here: Hindsight Goal Selection for long-horizon dexterous manipulation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.116005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.116005Z digest=sha256:fb0f5c0bea05dd74e3bc5e6bb527ded77369e75123dc72d97b451c00796e6b85

Observation 35e2ba22-e5ea-4da3-aa74-c9507fab7691 · outbound

This paper cites A Generalist Agent.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement A Generalist Agent

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.602130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.602130Z digest=sha256:b56798a4e48d8b065497cf256beeb91a167a8368fd6367321823f523e651cf30

Observation 1d3d755d-d3b6-4273-a0a9-604db3f09052 · outbound

This paper cites an unresolved cited work.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:06:07.444819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:06:06.743130Z digest=sha256:780f18a7004dea1bc54e85a0f21c5c2a8779cd7f4a1085f0490e50f59b5c2fc3

Observation 99c2d47e-7fff-4ee7-b1fd-7501c86f79c6 · outbound

This paper cites Offline Learning from Demonstrations and Unlabeled Experience.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Offline Learning from Demonstrations and Unlabeled Experience

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.902037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.902037Z digest=sha256:2a2036116df18e3dfccb714736da9cf49c39267c6de61a85317f99522d40aff3

Observation 9c306700-d951-45f0-8fea-32d6a439cce9 · outbound

This paper cites an unresolved cited work.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:06:07.318510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T19:06:06.951316Z digest=sha256:b117dcfe72ea9f6ca3c070ea32f8a48102c0063b81b35f5aadb70f2fc8c69867

Observation 379b2b3e-48a3-4756-a055-b3f92a1131a7 · outbound

This paper cites Large Language Models Can Self-Improve.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Large Language Models Can Self-Improve

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.248583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.248583Z digest=sha256:671246ce75690f2c15fd2ccfea6d58977bf0f6f76c1013c2960cbd5dda6f7b77

Observation 29bfdde9-336a-4003-b743-3b7accb25f0d · outbound

This paper cites Maximum Entropy Deep Inverse Reinforcement Learning.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Maximum Entropy Deep Inverse Reinforcement Learning

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.805435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.805435Z digest=sha256:30ecb6f8cc8f55b6f6b69252c8972e45d13d23d402d9376732b0b0853c5d7338

Observation 4a047633-9609-4812-8932-74bf9035e92d · outbound

This paper cites Kwon, T., Palo, N.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Kwon, T., Palo, N

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.454510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.454510Z digest=sha256:31f9adc00de649997e433f75818d7e4b2c48f38e8282d4e7d98317fc0c21b1df

Observation 98b7e95c-a3b6-4886-ab6d-1702aa029fe1 · outbound

This paper cites Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.043894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.043894Z digest=sha256:9b088000b0899751dfd8a81544430933d8100f37828f0e9609ce0bda5b09a341

Observation a695a859-fe68-432b-9b3f-443714294e47 · outbound

This paper cites Imitation Learning via Off-Policy Distribution Matching.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Imitation Learning via Off-Policy Distribution Matching

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.332378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.332378Z digest=sha256:394e2f1a86465604dad9edb1aabccad8a3cb23479248296e46ee533b09c36046

Observation d3dbb16b-69ea-4a80-9af1-4462da4e71b3 · outbound

This paper cites Imitating Language via Scalable Inverse Reinforcement Learning.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Imitating Language via Scalable Inverse Reinforcement Learning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.864052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.864052Z digest=sha256:891d01a7a8af260b068b9d8068987269fc8e7a15b55078a8b779edff0af4163d

Observation a681f6c3-2c41-460e-9cd3-9279e8e30a02 · outbound

This paper cites org/CorpusID:269605913.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement org/CorpusID:269605913

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.677382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.677382Z digest=sha256:65da32f845fe38fed30183b87da3d8f6fa629c1dc8d251b82482fa75f705d278

Pith citing papers

No inbound Pith citation observations are available.