Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T18:34:55.781594Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 1 inbound Pith citation observation for arXiv:2603.08956.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T18:34:55.781594Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-10T18:33:50.296933Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T00:25:50.044346Z
42 of 42 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 71a340cb-2ad4-4f7f-af90-c4e18d27a645 · outbound
A Survey of Reinforcement Learning For Economics Thompson Sampling for Dynamic Pricing
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c5f1fdb-cba4-4d90-897b-6322d092428b · outbound
A Survey of Reinforcement Learning For Economics Deep Reinforcement Learning from Self-Play in Imperfect-Information Games
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 927b2e5c-82e4-448e-bd24-903a4f5f2053 · outbound
A Survey of Reinforcement Learning For Economics Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc626079-6674-4405-84d6-62ddfef4638c · outbound
A Survey of Reinforcement Learning For Economics Is q-learning minimax optimal? a tight sample complexity analysis.Operations Research, 72(1), 2024a
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f92ec82-d915-44e7-9f10-cd44bb304454 · outbound
A Survey of Reinforcement Learning For Economics Clare Lyle, Mark Rowland, and Will Dabney
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4669983a-3fea-4362-8fcd-17d4fc19d61b · outbound
A Survey of Reinforcement Learning For Economics Empirical Design in Reinforcement Learning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 119e6291-060f-4239-901b-42e80bf3633a · outbound
A Survey of Reinforcement Learning For Economics Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f2cd22e-e456-4a26-8948-b6b7bd4decee · outbound
A Survey of Reinforcement Learning For Economics Proximal Policy Optimization Algorithms
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7167921-8f51-4979-bdf3-c24f93511b12 · outbound
A Survey of Reinforcement Learning For Economics Residual Policy Learning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64a11181-fc75-40e7-8191-253dfb935df7 · outbound
A Survey of Reinforcement Learning For Economics Solving Large Imperfect Information Games Using CFR+
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f82c125-154d-45af-aa97-ec4c8389fe68 · outbound
A Survey of Reinforcement Learning For Economics Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5d4fc0e-c118-4390-83a9-0a0b68b9bc99 · outbound
A Survey of Reinforcement Learning For Economics Near-Optimal Sample Complexity for Iterated CVaR Reinforcement Learning with a Generative Model
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f426e85-277b-48ca-a7f0-571f93d4a738 · outbound
A Survey of Reinforcement Learning For Economics Self-Rewarding Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bac77cc-6aac-4a84-ac73-212adbbfeb8b · outbound
A Survey of Reinforcement Learning For Economics A Deeper Look at Experience Replay
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f67ffb0-3c62-4ccc-916e-d369ab71e3b1 · outbound
A Survey of Reinforcement Learning For Economics Fine-Tuning Language Models from Human Preferences
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98777ef0-3492-46fb-a102-4aa0c37129cb · outbound
A Survey of Reinforcement Learning For Economics Santos and John Rust
Reference 1959
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3aa4ea9d-4aa4-4680-b70c-b827d8e4e9b5 · outbound
A Survey of Reinforcement Learning For Economics RL with KL penalties is better viewed as Bayesian inference
Reference 1960
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22f541ec-4d3b-49de-87f0-d9c7e45661cd · outbound
A Survey of Reinforcement Learning For Economics Classifying fermionic states via many-body correlation measures
Reference 1982
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bcf9018-48f8-4f49-94c2-872f28177592 · outbound
Reference 1988
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91d8cf4c-8412-4a2f-9c12-9dd68b69c670 · outbound
A Survey of Reinforcement Learning For Economics Markov games as a framework for multi-agent reinforcement learning
Reference 1992
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07be33d8-1cc9-4968-873e-45541c0561bc · outbound
A Survey of Reinforcement Learning For Economics First-order methods for Wasserstein distributionally robust MDP
Reference 1993
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00c241db-d497-4aa5-a9f6-3c4a36166254 · outbound
A Survey of Reinforcement Learning For Economics Unresolved cited work
Reference 1994
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 09a6e2d1-766a-4fb1-abcc-f68be9b8c1ed · outbound
A Survey of Reinforcement Learning For Economics Learning to Solve Constraint Satisfaction Problems with Recurrent Transformer
Reference 1997
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f93d637a-29e8-444d-bf6e-0117c7a7071c · outbound
A Survey of Reinforcement Learning For Economics Contextual Dynamic Pricing with Strategic Buyers
Reference 2001
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d0a0f84-62ca-4e8e-8e6d-8a31da115225 · outbound
A Survey of Reinforcement Learning For Economics Core equality of real sequences
Reference 2003
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4c7435d-f718-4e02-8e66-00a09d950a12 · outbound
A Survey of Reinforcement Learning For Economics Assessing Game Balance with AlphaZero: Exploring Alternative Rule Sets in Chess
Reference 2006
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91e31647-01c9-4f35-8e62-6adecfc4f605 · outbound
A Survey of Reinforcement Learning For Economics Peter Arcidiacono and Robert A
Reference 2008
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c587db8a-935b-4326-807c-d59b0b5ba6eb · outbound
A Survey of Reinforcement Learning For Economics Deep Reinforcement Learning and the Deadly Triad
Reference 2010
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6402ceee-ba27-4ed3-b4ff-d00b2dd5e9da · outbound
A Survey of Reinforcement Learning For Economics Insight from the elliptic flow of identified hadrons measured in relativistic heavy-ion collisions
Reference 2011
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e98d37c-c83d-4e75-a175-3f619ef22804 · outbound
A Survey of Reinforcement Learning For Economics Fleming and William M
Reference 2012
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15c6b402-2baf-4b60-8858-5b6bff1233a9 · outbound
A Survey of Reinforcement Learning For Economics Asynchronous methods for deep reinforce- ment learning
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc756144-df6d-4d3c-914e-b05d665be10e · outbound
A Survey of Reinforcement Learning For Economics Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Review
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 446d8768-e16b-46ff-a6dd-6144eb50ec26 · outbound
A Survey of Reinforcement Learning For Economics Music Source Separation with Band-split RNN
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cecaf6e8-7df4-4e35-8a21-a56aa9cfac76 · outbound
A Survey of Reinforcement Learning For Economics Off-policy deep reinforcement learning with- out exploration
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a5f69ed-2c80-46e3-b39d-11e9649ee6c5 · outbound
A Survey of Reinforcement Learning For Economics Policy Optimization for Constrained MDPs with Provable Fast Global Convergence
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35f2d2d3-0679-4b3b-becf-104b1167add1 · outbound
A Survey of Reinforcement Learning For Economics Deep reinforcement learning: Emerging trends in macroe- conomics and future prospects
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23280bfc-bf70-48a0-b34a-71d88ffa8a98 · outbound
A Survey of Reinforcement Learning For Economics Generalizing across Temporal Domains with Koopman Operators
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 560b2954-1697-4888-95b7-bb4c44c3845e · outbound
A Survey of Reinforcement Learning For Economics Unifying causal reinforcement learning: Survey, taxonomy, algorithms and applications.arXiv preprint arXiv:2512.18135,
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f95a455-8d94-49ca-aa7f-ae9c74fce8f0 · outbound
A Survey of Reinforcement Learning For Economics Lili Chen, Kevin Lu, Aravind Rajeswaran, Kimin Lee, Aditya Grover, Michael Laskin, Pieter Abbeel, Aravind Srinivas, and Igor Mordatch
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62fa3c21-54ae-41a9-b35c-980a993046a9 · outbound
A Survey of Reinforcement Learning For Economics Strategic classifi- cation
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70ce1f1a-f9f1-4941-ab16-f50dfa0933c0 · outbound
A Survey of Reinforcement Learning For Economics Jonas Mueller, Vasilis Syrgkanis, and Matt Taddy
Reference 2025
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3e7c990f-f04f-48fe-b847-7e44fa88a3c9 · outbound
A Survey of Reinforcement Learning For Economics Some asymptotic formulae for torsion in homotopy groups
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d15b34f2-af80-4e1c-8936-36870312e240 · inbound
The Theorems of Dr. David Blackwell and Their Contributions to Artificial Intelligence A Survey of Reinforcement Learning For Economics
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.