Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T11:22:20.728645Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2607.19928.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T11:22:20.728645Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
41 of 41 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d06da8c5-3bad-4d2f-825d-5c77a6105fc0 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Başar and G.J
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7aafdae3-4a94-40f8-a31b-b0fb9ff89a67 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Bouchard and N
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b31b335-27b2-43b6-ba67-000bd3af2580 · outbound
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d10ffc7-8672-4e90-8abe-68e0012e73d7 · outbound
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87893005-eae9-471a-9676-6fb86519d5bb · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Carmona and F
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b2649f5-97c9-4d13-b215-a83db45c365d · outbound
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08040e2d-72c3-4f18-ba10-2695434055f3 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Fleming and H.M
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68e41935-8c6d-4748-86ee-6e460ff9db31 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f149baea-b9c0-43bc-a52d-81f1d526e381 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Springer, 2001
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45546477-f6d9-44ec-9b83-a1dee15030a8 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7330bed7-adf5-4f46-a209-96a5dae68633 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Anα-potential game framework forn-player dynamic games.SIAM Journal on Control and Optimization, 63(4):2964–3005, 2025
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 944c4818-b030-4d57-804c-5689ca96aa46 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Entropy Regularized Reinforcement Learning for Zero-Sum Stochastic Differential Games in a Regime-Switching Jump-Diffusion Process
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 717eb36c-0c83-489d-b9b4-c6cf628148f9 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Entropy-Regularized Reinforcement Learning for Linear-Quadratic Stackelberg Differential Games in Regime-Switching Diffusion Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 842a9b81-0c06-4678-9a1b-da69c8cd85bd · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Recent developments in machine learning meth- ods for stochastic control and games.Numerical Algebra, Control and Optimization, 14(3):435–525, 2024
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1eceffb0-0050-45fb-9408-7358beba4edc · outbound
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4888772-7b01-4209-aa7b-cd21a5fdc2f2 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Continuous-time reinforcement learning for optimal switching over multiple regimes
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bba2d55-4741-4b6d-a2d9-b02a49d9c8ad · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Sublinear regret for a class of continuous- time linear-quadratic reinforcement learning problems.SIAM Journal on Control and Optimization, 63(5):3452–3474, 2025
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a8504d5-6aa5-4220-9480-8210bd94ac96 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 517acc6d-724f-4675-aca9-d448b4f2a479 · outbound
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62e0d2ed-5d8d-4da5-9163-1ce0b9d106a3 · outbound
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c322874c-fb4d-4158-a533-e533105f53d0 · outbound
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1daf5aba-d84e-46d7-bde3-c5a56bcf4254 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Karatzas and S.E
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f941b0b6-b5eb-4916-bf0e-6bb32294040f · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Krylov.Nonlinear Elliptic and Parabolic Equations of the Second Order
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fa5b3e9-a62c-4746-8aff-9f0693063227 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Ladyzhenskaya, V.A
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4092ac45-138c-46ca-89b6-82329c8fd0bd · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Lasry and P.-L
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea73c6d1-45b6-4829-8fae-809b1166cf53 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Actor-critic reinforcement learning al- gorithms for mean field games in continuous time, state and action spaces.Applied Mathematics & Optimization, 89(3):73, 2024
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c660c46a-1bee-4ba7-9cad-0bb6e4f720b8 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d2ba461-18db-48a1-87da-365ca2336570 · outbound
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4d7c324-fc34-4202-a87c-90b113d4bbf9 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Unresolved cited work
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b5bd82f-5a43-43f3-ba20-0c9ff50cd4be · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7109b10-09f6-4381-af13-db03688f3ba7 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Learning distributed equilibria in linear-quadratic stochastic differential games: Anα-potential approach.arXiv:2602.16555, 2026
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da386cc4-6a4a-4eb1-b4eb-0910641b873c · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Continuous-time q-learning for mean-field control with common noise, part-I: Theoretical foundations
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24211381-ebd3-4cb3-ab6b-2c4b3f964e1f · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Reinforcement learning for exploratory linear- quadratic two-person zero-sum stochastic differential games.Applied Mathematics and Computation, 442:127763, 2023
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea6fbfea-33a6-4975-bd56-ce3ab6ecf508 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Sutton and A.G
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c764ff7-65ef-41ed-ac5a-2d826cecf34a · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Regret of exploratory policy improvement and $q$-learning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18d40f3f-e0c1-4120-a2be-fe64d20e4678 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e576da72-ae9e-4f0b-a119-3bdd74ebd4e8 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Unresolved cited work
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b61a4e65-e746-4664-ae2a-7035de25d74d · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Unresolved cited work
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39132cc4-3f89-4c08-9e10-d14b113930d9 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Continuoustimeq-learningformean-fieldcontrolproblems.Applied Mathematics & Optimization, 91(1):10, 2025
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63e59b80-236e-4bb3-bdd5-36974a329498 · outbound
Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Unified continuous-time q-learning for mean-field game and mean-field control problems
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f5886ff-a829-4fe8-a48e-340def7ae7cd · outbound
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.