Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2209.10579.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:54:11.213481Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-29T10:33:18.772880Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation ad6ae68f-dc42-47b6-84b5-6ddabd41e6d9 · inbound
Near-Optimal Policy Identification in Robust Constrained Markov Decision Processes via Epigraph Form First-order Policy Optimization for Robust Markov Decision Process
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6b2800b2-64ec-4e46-b8c4-4e056478a7aa · inbound
Efficient Q-Learning and Actor-Critic Methods for Robust Average-Reward Reinforcement Learning First-order Policy Optimization for Robust Markov Decision Process
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02e14e7c-de26-479c-bdd9-1439190d22ba · inbound
Sample Complexity for Markov Decision Processes and Stochastic Optimal Control with Static Risk Measures First-order Policy Optimization for Robust Markov Decision Process
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ca97eebf-6bc5-4ee6-9fb9-206d4d7de8ec · inbound
Value Mirror Descent for Reinforcement Learning First-order Policy Optimization for Robust Markov Decision Process
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6f144f27-1a45-4766-a00a-be694147343a · inbound
Revisiting Subgradient Dominance in Robust MDPs: Counterexamples, Hardness, and Sufficient Conditions First-order Policy Optimization for Robust Markov Decision Process
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 748102b9-a20c-4ab4-8810-057dbc871299 · inbound
Natural Policy Gradient as Doubly Smoothed Policy Iteration: A Bellman-Operator Framework First-order Policy Optimization for Robust Markov Decision Process
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c6747c55-653f-44c3-8ec3-2862ce2e3095 · inbound
Robust Markov Decision Processes on Continuous State Spaces First-order Policy Optimization for Robust Markov Decision Process
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9e666d68-b161-47fe-8196-cd229dfebfd1 · inbound
Non-Asymptotic Convergence of Stochastic Iterative Algorithms: A Lyapunov Framework First-order Policy Optimization for Robust Markov Decision Process
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.