Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:1906.10306.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-14T10:27:42.130654Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-21T02:29:25.160173Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 5305a79e-1dd9-4bdd-a258-3c531cf10641 · inbound
Neural Policy Gradient Methods: Global Optimality and Rates of Convergence Neural Proximal/Trust Region Policy Optimization Attains Globally Optimal Policy
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ece489f0-733e-4085-8d4b-11e9f78920b1 · inbound
Adaptive Trust Region Policy Optimization: Global Convergence and Faster Rates for Regularized MDPs Neural Proximal/Trust Region Policy Optimization Attains Globally Optimal Policy
Reference 2002
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64d04a35-fb84-4010-b8a6-5d133077290e · inbound
Fast Convergence of Softmax Policy Mirror Ascent Neural Proximal/Trust Region Policy Optimization Attains Globally Optimal Policy
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1cc9075-7b04-4ef7-8916-3276575b5799 · inbound
On the Sample Complexity of Differentially Private Policy Optimization Neural Proximal/Trust Region Policy Optimization Attains Globally Optimal Policy
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1e26dc60-2e16-4f69-bc8e-957516cb4382 · inbound
Concentration of General Stochastic Approximation Under Heavy-Tailed Markovian Noise Neural Proximal/Trust Region Policy Optimization Attains Globally Optimal Policy
Reference 154
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.