Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2002.12292.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-10T18:11:48.357673Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
34
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation c865aed2-9004-4d0b-bfe8-64996daabd1c · inbound
Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments
Reference 267
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 94a24c8d-85ac-4ea5-9b53-cf2da92f5e3f · inbound
The impact of intrinsic rewards on exploration in Reinforcement Learning RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 814bafc7-d09e-4e09-880e-5d93a4c7f339 · inbound
Episodic Novelty Through Temporal Distance RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8cb284d-c321-4eff-9481-1f696a007576 · inbound
ManiTaskGen: A Comprehensive Task Generator for Benchmarking and Improving Vision-Language Agents on Embodied Decision-Making RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d859f02f-2ab9-4cc4-b7b2-20e0ccb04bec · inbound
Beyond Noisy-TVs: Noise-Robust Exploration Via Learning Progress Monitoring RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 92993aed-0cd0-4596-9926-ddde69b56c67 · inbound
Beyond Single-Model Optimization: Preserving Plasticity in Continual Reinforcement Learning RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 38ebcaba-2444-46e8-8ddd-7b7172f98ee2 · inbound
Beyond Single-Model Optimization: Preserving Plasticity in Continual Reinforcement Learning RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e96c06e7-53f4-4751-8899-5b4665cb798f · inbound
Goal-Conditioned Agents that Learn Everything All at Once RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c74f7ca1-cf00-4cef-9c71-7b6cc29f80e6 · inbound
Right Makes Might: Aligning Verified Hidden States Empowers RL Reasoning RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e39b4dfb-3766-4988-bd8f-80d31b93802a · inbound
Progress-conditioned Group Policy Optimization for Long-Horizon Agentic Tasks RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.