Pith. sign in

Paper Citation Record · LEDGER

Imitation Learning via Off-Policy Distribution Matching

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:1912.05032.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1912.05032 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:43:10.659776Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T22:09:07.605152Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0bd3cad1-c695-485d-b965-0ab957be8df1 · inbound

From Screens to Scenes: A Survey of Embodied AI in Healthcare cites this paper.

From Screens to Scenes: A Survey of Embodied AI in Healthcare Imitation Learning via Off-Policy Distribution Matching

Reference 226

Resolution
unresolved
no resolver link, observed 2026-08-10T20:43:10.659776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:43:10.659776Z digest=sha256:9a87eaebfb0fc5a9afc5c8539596c360bf61731c84789d4300f01071862a329b

Observation 7418e6c8-5dc5-4ab3-98e0-36583c608810 · inbound

Exploration from a Primal-Dual Lens: Value-Incentivized Actor-Critic Methods for Sample-Efficient Online RL cites this paper.

Exploration from a Primal-Dual Lens: Value-Incentivized Actor-Critic Methods for Sample-Efficient Online RL Imitation Learning via Off-Policy Distribution Matching

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T22:16:40.577492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:16:40.577492Z digest=sha256:2019b83fc976dea3d5331b782a9b4ab03a0eaf570e4e504396f91aaaaa98df8c

Observation a695a859-fe68-432b-9b3f-443714294e47 · inbound

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement cites this paper.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Imitation Learning via Off-Policy Distribution Matching

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.332378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.332378Z digest=sha256:6b23d92d519696afe57b7d525d7449f74f611944ffe65e0b2ec666d5b6acd62f

Observation d040edbd-238b-4cf2-a3a2-2d88181b49fc · inbound

Diffusion-Modeled Reinforcement Learning for Carbon and Risk-Aware Microgrid Optimization cites this paper.

Diffusion-Modeled Reinforcement Learning for Carbon and Risk-Aware Microgrid Optimization Imitation Learning via Off-Policy Distribution Matching

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T15:22:26.488059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:22:26.488059Z digest=sha256:5ce57303a9f907be3d9c7141bd18a9d6837a57c46d443b41f11b984f2748603f

Observation 0eb30226-101e-49d8-b78c-4e8d28c65da1 · inbound

Distributional Inverse Reinforcement Learning cites this paper.

Distributional Inverse Reinforcement Learning Imitation Learning via Off-Policy Distribution Matching

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T12:42:58.311175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:42:58.311175Z digest=sha256:b4d4624487fb6a8fd61a130c8ecb91f7840e7d8cbe5c9cda554013707230d896

Observation 22983097-7847-4451-b390-d3a2f1062ce0 · inbound

Learning When to Stop: Selective Imitation Learning Under Arbitrary Dynamics Shift cites this paper.

Learning When to Stop: Selective Imitation Learning Under Arbitrary Dynamics Shift Imitation Learning via Off-Policy Distribution Matching

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:21:25.285855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T03:27:41.845716Z digest=sha256:6a81e8126a332d74fdf4f31d77df0f071c3b9b3dd2b03e875d82bd6361f5087e

Observation 66db76f1-5f70-4dee-bd68-3c0d4e711c0c · inbound

Learning When to Stop: Selective Imitation Learning Under Arbitrary Dynamics Shift cites this paper.

Learning When to Stop: Selective Imitation Learning Under Arbitrary Dynamics Shift Imitation Learning via Off-Policy Distribution Matching

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:09:07.608511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-20T22:04:57.245024Z digest=sha256:c82f70a196e59517d0d23355fb322c07f82ea5346fc2b84f6c985ff60be74106