Pith. sign in

Paper Citation Record · LEDGER

RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2002.12292.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2002.12292 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T18:11:48.357673Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

34
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c865aed2-9004-4d0b-bfe8-64996daabd1c · inbound

Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution cites this paper.

Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments

Reference 267

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T08:12:31.148831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-16T08:12:30.984870Z digest=sha256:5b592f26afcf09243feb37aa0c1adf2195cdf74fc8b7e134fa7b7d891f43ad0e

Observation 94a24c8d-85ac-4ea5-9b53-cf2da92f5e3f · inbound

The impact of intrinsic rewards on exploration in Reinforcement Learning cites this paper.

The impact of intrinsic rewards on exploration in Reinforcement Learning RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T18:11:48.357673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:11:48.357673Z digest=sha256:2b414fa4938b6171c285c869ce5e54f8e2b3300ad03888ebb47d69169ea4237d

Observation 814bafc7-d09e-4e09-880e-5d93a4c7f339 · inbound

Episodic Novelty Through Temporal Distance cites this paper.

Episodic Novelty Through Temporal Distance RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T14:25:58.977311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T14:25:58.977311Z digest=sha256:b00b04f061c34396d777a24d5b45f3c67384afd51557b74fa19ff3977699d9e9

Observation d8cb284d-c321-4eff-9481-1f696a007576 · inbound

ManiTaskGen: A Comprehensive Task Generator for Benchmarking and Improving Vision-Language Agents on Embodied Decision-Making cites this paper.

ManiTaskGen: A Comprehensive Task Generator for Benchmarking and Improving Vision-Language Agents on Embodied Decision-Making RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T13:53:57.551508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:53:57.551508Z digest=sha256:11ee2feffdf1e409a2783b1e9a7a977421a69b1e3b8340e365b903a0ebbeb9b9

Observation d859f02f-2ab9-4cc4-b7b2-20e0ccb04bec · inbound

Beyond Noisy-TVs: Noise-Robust Exploration Via Learning Progress Monitoring cites this paper.

Beyond Noisy-TVs: Noise-Robust Exploration Via Learning Progress Monitoring RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:51:20.179492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-18T11:50:46.257332Z digest=sha256:c3bb2eb85898b234f3fc21dedc9bcbdbce136000091d91a1c788aeb12eb627b5

Observation 92993aed-0cd0-4596-9926-ddde69b56c67 · inbound

Beyond Single-Model Optimization: Preserving Plasticity in Continual Reinforcement Learning cites this paper.

Beyond Single-Model Optimization: Preserving Plasticity in Continual Reinforcement Learning RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:05:08.940023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T11:01:17.325738Z digest=sha256:500cc60b072ce084b235585658efc96fc9bfbe69bd885a92ad00968f84deded1

Observation 38ebcaba-2444-46e8-8ddd-7b7172f98ee2 · inbound

Beyond Single-Model Optimization: Preserving Plasticity in Continual Reinforcement Learning cites this paper.

Beyond Single-Model Optimization: Preserving Plasticity in Continual Reinforcement Learning RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-12T19:46:39.624903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T19:46:39.624903Z digest=sha256:285f222677cefac6eee315730f09530062a4f1bfd2afce3b199819458b94cd97

Observation e96c06e7-53f4-4751-8899-5b4665cb798f · inbound

Goal-Conditioned Agents that Learn Everything All at Once cites this paper.

Goal-Conditioned Agents that Learn Everything All at Once RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments

Reference 83

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:00:21.466068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-25T04:59:48.867927Z digest=sha256:06598cf4ad2b49e2b57a3138fe057a7f190e7b2da5b4e61da52590ff366254df

Observation c74f7ca1-cf00-4cef-9c71-7b6cc29f80e6 · inbound

Right Makes Might: Aligning Verified Hidden States Empowers RL Reasoning cites this paper.

Right Makes Might: Aligning Verified Hidden States Empowers RL Reasoning RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:06:27.252391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-28T11:13:21.565082Z digest=sha256:30fa05bb4c1c42333e51ba9226f86391b5fa9b9a648cba150257bd995db98394

Observation e39b4dfb-3766-4988-bd8f-80d31b93802a · inbound

Progress-conditioned Group Policy Optimization for Long-Horizon Agentic Tasks cites this paper.

Progress-conditioned Group Policy Optimization for Long-Horizon Agentic Tasks RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T11:54:58.489272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:54:58.489272Z digest=sha256:f39d282e9cac9d0d0876a2cb7d4ea9860da954790b3e112ed0fc99bf76530ca8