Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2006.05990.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T04:36:12.619468Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
105
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 11bf0bf6-346d-4a68-8682-77a61759fe70 · inbound
What Matters in Learning from Offline Human Demonstrations for Robot Manipulation What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation edd1cefa-a79c-401e-a686-fd5a82942d12 · inbound
The Effects of Reward Misspecification: Mapping and Mitigating Misaligned Models What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5e7bc265-bb22-4746-9f1c-6b5223fd8b4b · inbound
Mastering Diverse Domains through World Models What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0c67352e-c4d3-4b5f-b882-fb53c819e4ef · inbound
Simulating Errors in Touchscreen Typing What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87b35a59-608e-447d-9477-3e97bb1daa7f · inbound
MuJoCo Playground What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4b230da-84bc-4768-b4c3-268d49f2c8c0 · inbound
RedRFT: A Light-Weight Benchmark for Reinforcement Fine-Tuning-Based Red Teaming What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58df91f7-f6d0-43f2-914c-f803515f43e7 · inbound
Magistral What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9bbad98f-46c2-43b4-bc7a-bd421c171fc0 · inbound
Communicating Smartly in Molecular Communication Environments: Neural Networks in the Internet of Bio-Nano Things What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05e8926f-c64c-4013-be31-a7adcfd16cf7 · inbound
On the Effect of Regularization in Policy Mirror Descent What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0922a89a-ec0c-4f0d-984e-0897a8430c28 · inbound
Learning human-to-robot handovers through 3D scene reconstruction What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0bd07b4-47d3-45fb-8dd6-c316ec3a8241 · inbound
Online Training and Pruning of Deep Reinforcement Learning Networks What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d1af3aa-8b4b-4dea-865c-f087b1bfce4d · inbound
Differentiable Weightless Controllers: Learning Logic Circuits for Continuous Control What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16ea3926-2720-442f-ae65-02428ba04317 · inbound
Preventing Learning Stagnation in PPO by Scaling to 1 Million Parallel Environments What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9348a80-ac48-439d-9a34-652d31809979 · inbound
From $P(y|x)$ to $P(y)$: Investigating Reinforcement Learning in Pre-train Space What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0897607c-7adc-46df-8bc8-f1172f451575 · inbound
Low-Rank Adaptation for Critic Learning in Off-Policy Reinforcement Learning What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 810aace4-d58a-49c6-85b6-0fb17d21ba16 · inbound
Application of Deep Reinforcement Learning to Event-Triggered Control for Networked Artificial Pancreas Systems What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ec7d7e36-c90a-4688-b726-47c03e42b11c · inbound
Application of Deep Reinforcement Learning to Event-Triggered Control for Networked Artificial Pancreas Systems What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e5544742-6a92-4df8-8429-0affd16ea488 · inbound
Hint Tuning: Less Data Makes Better Reasoners What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2306338e-254c-4b5e-a92e-c19e3e962391 · inbound
Hint Tuning: Less Data Makes Better Reasoners What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fd05a4da-3bbf-4276-a8e5-fcb36ddeedbf · inbound
TuniQ: Autotuning Compilation Passes for Quantum Workloads at Scale for Effectiveness and Efficiency What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ed702539-4593-4c4c-9d30-38c779010c53 · inbound
TOPPO: Rethinking PPO for Multi-Task Reinforcement Learning with Critic Balancing What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e8ccc653-d395-4959-978f-9d755b9a10bb · inbound
Randomized Advantage Transformation (RAT): Computing Natural Policy Gradients via Direct Backpropagation What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4a2c263d-a5cb-48b1-aca4-b3c0ef4abb7b · inbound
When are LLMs Sufficient Policy Optimizers for Sequential RL Tasks? What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4249dc2e-3a72-431a-b874-9af188867e0a · inbound
When are LLMs Sufficient Policy Optimizers for Sequential RL Tasks? What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ea047044-d8ca-484c-88d2-a6d70d7ccaf6 · inbound
PowerOPD: Stabilizing On-Policy Distillation with Bounded Power Transformation What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 69add4d6-dddd-4560-a164-723cc1443895 · inbound
Learning to Trigger: Reinforcement Learning at the Large Hadron Collider What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ae2dde90-abce-4a0e-ae84-b3abf3babac2 · inbound
Learning to Trigger: Reinforcement Learning at the Large Hadron Collider What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c078e5ea-f243-46fc-a542-efdd9bed50b0 · inbound
The Importance of Encoder Choice:A Tabular-Image Study What Matters In On-Policy Reinforcement Learning? A Large-Scale Empirical Study
Reference 168
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.