Pith. sign in

Paper Citation Record · LEDGER

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies

As of 9 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2607.19928.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.19928 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T11:22:20.728645Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d06da8c5-3bad-4d2f-825d-5c77a6105fc0 · outbound

This paper cites Başar and G.J.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Başar and G.J

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.565114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.565114Z digest=sha256:ef457ff54c4489be6b2eb7c6f7174022aab7911bf7345615f48b2f0948fc1557

Observation 7aafdae3-4a94-40f8-a31b-b0fb9ff89a67 · outbound

This paper cites Bouchard and N.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Bouchard and N

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.570220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.570220Z digest=sha256:6f8be2fe7c70cf972c4921ace65caf161a779ada319f9f33b8f1eb34998c65bd

Observation 7b31b335-27b2-43b6-ba67-000bd3af2580 · outbound

This paper cites Buckdahn, P.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Buckdahn, P

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.574696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.574696Z digest=sha256:1c5a3c476ba1048b10fd1de82958601be43e0fce737b83d0dc7edb409c58bd4e

Observation 5d10ffc7-8672-4e90-8abe-68e0012e73d7 · outbound

This paper cites Buşoniu, R.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Buşoniu, R

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.579306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.579306Z digest=sha256:ae594e2d3df4dac7074f3c46196af0d1cc57bdefc2376e30ac435bf92adc2e7c

Observation 87893005-eae9-471a-9676-6fb86519d5bb · outbound

This paper cites Carmona and F.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Carmona and F

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.583484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.583484Z digest=sha256:de56031b2290914ff09c377e785f2dce0d3ba35ac39e46160674bf71bb8f53f3

Observation 3b2649f5-97c9-4d13-b215-a83db45c365d · outbound

This paper cites Dockner, S.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Dockner, S

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.587571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.587571Z digest=sha256:f72e99dde125c8adc763542bf9dbdb28891801f6877e9acb7affc8500b5bd09c

Observation 08040e2d-72c3-4f18-ba10-2695434055f3 · outbound

This paper cites Fleming and H.M.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Fleming and H.M

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.592027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.592027Z digest=sha256:62b2a1332df6ed037fdb8d5c55ad6276a16933e06373345cd0d13ea19b94e9ff

Observation 68e41935-8c6d-4748-86ee-6e460ff9db31 · outbound

This paper cites an unresolved cited work.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.595905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.595905Z digest=sha256:5766f5aeef9acb906d31dba2affcd21dfbf5c14576c8a345bbb2262465c4f293

Observation f149baea-b9c0-43bc-a52d-81f1d526e381 · outbound

This paper cites Springer, 2001.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Springer, 2001

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.600048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.600048Z digest=sha256:93ff44722be8d8bb0040a4a211b6adf823b629d9915e23ae1540b237bc541341

Observation 45546477-f6d9-44ec-9b83-a1dee15030a8 · outbound

This paper cites an unresolved cited work.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.603807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.603807Z digest=sha256:085c98814f37faec8c15e2c6b46979c05d47f07e598ad264fd8e5254970fd6fc

Observation 7330bed7-adf5-4f46-a209-96a5dae68633 · outbound

This paper cites Anα-potential game framework forn-player dynamic games.SIAM Journal on Control and Optimization, 63(4):2964–3005, 2025.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Anα-potential game framework forn-player dynamic games.SIAM Journal on Control and Optimization, 63(4):2964–3005, 2025

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.607761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.607761Z digest=sha256:24d46823573cbda5cf273ef0e70f86cd5622d981b3df16dce910d876da99763c

Observation 944c4818-b030-4d57-804c-5689ca96aa46 · outbound

This paper cites Entropy Regularized Reinforcement Learning for Zero-Sum Stochastic Differential Games in a Regime-Switching Jump-Diffusion Process.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Entropy Regularized Reinforcement Learning for Zero-Sum Stochastic Differential Games in a Regime-Switching Jump-Diffusion Process

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.611862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.611862Z digest=sha256:ef14be0cde42cf9dd1e1363cda73c96bb13977d7bcdb4b13ba231a22e1b1f877

Observation 717eb36c-0c83-489d-b9b4-c6cf628148f9 · outbound

This paper cites Entropy-Regularized Reinforcement Learning for Linear-Quadratic Stackelberg Differential Games in Regime-Switching Diffusion Models.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Entropy-Regularized Reinforcement Learning for Linear-Quadratic Stackelberg Differential Games in Regime-Switching Diffusion Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.616066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.616066Z digest=sha256:9170810f1781bdc29078c52e2ceb8177ba928e5c1d7534b9721f9166c602749e

Observation 842a9b81-0c06-4678-9a1b-da69c8cd85bd · outbound

This paper cites Recent developments in machine learning meth- ods for stochastic control and games.Numerical Algebra, Control and Optimization, 14(3):435–525, 2024.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Recent developments in machine learning meth- ods for stochastic control and games.Numerical Algebra, Control and Optimization, 14(3):435–525, 2024

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.620044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.620044Z digest=sha256:315a422231d19ec013dd280159fd68e4cc801524522dce343a68ac67ac1650ee

Observation 1eceffb0-0050-45fb-9408-7358beba4edc · outbound

This paper cites Huang, P.E.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Huang, P.E

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.623927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.623927Z digest=sha256:2ba9e595d00bae2efffbf89609faae238fd48b80d897a6f820ad3f9a6eaa5701

Observation e4888772-7b01-4209-aa7b-cd21a5fdc2f2 · outbound

This paper cites Continuous-time reinforcement learning for optimal switching over multiple regimes.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Continuous-time reinforcement learning for optimal switching over multiple regimes

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.627811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.627811Z digest=sha256:70af40efd9eb43667fa54ec6ea2251215e78d6af3769c1f7f17f04150871d2c5

Observation 0bba2d55-4741-4b6d-a2d9-b02a49d9c8ad · outbound

This paper cites Sublinear regret for a class of continuous- time linear-quadratic reinforcement learning problems.SIAM Journal on Control and Optimization, 63(5):3452–3474, 2025.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Sublinear regret for a class of continuous- time linear-quadratic reinforcement learning problems.SIAM Journal on Control and Optimization, 63(5):3452–3474, 2025

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.631946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.631946Z digest=sha256:2de19a9e9f3da46aea3547d74c53f9f2a1d89488357e4e2c116adfe210179a9f

Observation 2a8504d5-6aa5-4220-9480-8210bd94ac96 · outbound

This paper cites Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Data-Driven Exploration for a Class of Continuous-Time Indefinite Linear--Quadratic Reinforcement Learning Problems

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.635669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.635669Z digest=sha256:9a7ae04b51483e77e1d1d3d0fe6439f89759f08b21dcb8033f8bfe414d4ab5cc

Observation 517acc6d-724f-4675-aca9-d448b4f2a479 · outbound

This paper cites Jia and X.Y.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Jia and X.Y

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.639819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.639819Z digest=sha256:0d1c5c7d3a5fdbd1161c12cb731afe7794f32353cc86b4f219dfb95d865cbfa8

Observation 62e0d2ed-5d8d-4da5-9163-1ce0b9d106a3 · outbound

This paper cites Jia and X.Y.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Jia and X.Y

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.643824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.643824Z digest=sha256:f4e7ae87ec7260851895647ff948fa0adcb1fdcc3ed486907af4a52bb3ed63f5

Observation c322874c-fb4d-4158-a533-e533105f53d0 · outbound

This paper cites Jia and X.Y.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Jia and X.Y

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.647639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.647639Z digest=sha256:a9fbd21ffb988377c3b13cafbad368f30510d47e7d9b01286f607e72b773de82

Observation 1daf5aba-d84e-46d7-bde3-c5a56bcf4254 · outbound

This paper cites Karatzas and S.E.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Karatzas and S.E

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.651537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.651537Z digest=sha256:f4ea8a720e7ad84072244c903173d4dbddc866d12051ee01ac6e81d7b629eb2d

Observation f941b0b6-b5eb-4916-bf0e-6bb32294040f · outbound

This paper cites Krylov.Nonlinear Elliptic and Parabolic Equations of the Second Order.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Krylov.Nonlinear Elliptic and Parabolic Equations of the Second Order

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.655449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.655449Z digest=sha256:68845151c3a4322d7e7f27e59bcdd837ff62f501d2b48139515ca76b624d4fe5

Observation 3fa5b3e9-a62c-4746-8aff-9f0693063227 · outbound

This paper cites Ladyzhenskaya, V.A.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Ladyzhenskaya, V.A

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.659370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.659370Z digest=sha256:677a3e941c528ec29782bec40e93ae8e50d1665defe2ecec8f7cdbdeaf3ff9ae

Observation 4092ac45-138c-46ca-89b6-82329c8fd0bd · outbound

This paper cites Lasry and P.-L.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Lasry and P.-L

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.663451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.663451Z digest=sha256:d366b9ac77778e7cb25fa894e0b17d818fa4d8d2c9e25b9edcdc3e4da4be9b24

Observation ea73c6d1-45b6-4829-8fae-809b1166cf53 · outbound

This paper cites Actor-critic reinforcement learning al- gorithms for mean field games in continuous time, state and action spaces.Applied Mathematics & Optimization, 89(3):73, 2024.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Actor-critic reinforcement learning al- gorithms for mean field games in continuous time, state and action spaces.Applied Mathematics & Optimization, 89(3):73, 2024

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.667751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.667751Z digest=sha256:1a0c38c2d9a293556136296e093382ca5469ce5452ce0aad2894ba1102ca55fb

Observation c660c46a-1bee-4ba7-9cad-0bb6e4f720b8 · outbound

This paper cites an unresolved cited work.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.672034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.672034Z digest=sha256:8a9ae42202751b642c97234af3f0e9be09da1348d63cb8aa8bdaa176468cf8f7

Observation 4d2ba461-18db-48a1-87da-365ca2336570 · outbound

This paper cites Mnih et al.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Mnih et al

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.676452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.676452Z digest=sha256:9fc5d4c7befaab92f4c918e9e0c9f3ea2a1c64c7fa966439fbc68c80bb091f87

Observation a4d7c324-fc34-4202-a87c-90b113d4bbf9 · outbound

This paper cites an unresolved cited work.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.680518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.680518Z digest=sha256:19138eda72b1cf7d68ad297f70daa114e2781f3442e65165004a8a53a7f81103

Observation 6b5bd82f-5a43-43f3-ba20-0c9ff50cd4be · outbound

This paper cites an unresolved cited work.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.684619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.684619Z digest=sha256:da664fe5003232c04717ef79dfb6ee7a41713401eff8019272dc6b4b85c83792

Observation d7109b10-09f6-4381-af13-db03688f3ba7 · outbound

This paper cites Learning distributed equilibria in linear-quadratic stochastic differential games: Anα-potential approach.arXiv:2602.16555, 2026.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Learning distributed equilibria in linear-quadratic stochastic differential games: Anα-potential approach.arXiv:2602.16555, 2026

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.688624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.688624Z digest=sha256:031455ca0173125a053f82b5b64792ba0a0466c97197f29b19a02a8341cecf3e

Observation da386cc4-6a4a-4eb1-b4eb-0910641b873c · outbound

This paper cites Continuous-time q-learning for mean-field control with common noise, part-I: Theoretical foundations.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Continuous-time q-learning for mean-field control with common noise, part-I: Theoretical foundations

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.692590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.692590Z digest=sha256:7e0869fb814a840ccbd65f48f41528716a6b28cc5d9b7a4e2fdcc6552ca0aa60

Observation 24211381-ebd3-4cb3-ab6b-2c4b3f964e1f · outbound

This paper cites Reinforcement learning for exploratory linear- quadratic two-person zero-sum stochastic differential games.Applied Mathematics and Computation, 442:127763, 2023.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Reinforcement learning for exploratory linear- quadratic two-person zero-sum stochastic differential games.Applied Mathematics and Computation, 442:127763, 2023

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.696871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.696871Z digest=sha256:c0f4d7daa88035a1f678771251d4a47ff3d03dd97c29374e8b7eb6807e73545e

Observation ea6fbfea-33a6-4975-bd56-ce3ab6ecf508 · outbound

This paper cites Sutton and A.G.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Sutton and A.G

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.700770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.700770Z digest=sha256:2c26534906e0ca62ca7a480d68bcab6b802f84b4b01e9a27aa508ff28b095cc1

Observation 4c764ff7-65ef-41ed-ac5a-2d826cecf34a · outbound

This paper cites Regret of exploratory policy improvement and $q$-learning.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Regret of exploratory policy improvement and $q$-learning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.704553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.704553Z digest=sha256:d92aa798b137dec79944024213991ba5aa8aedd837ddca70d95dc41826b56837

Observation 18d40f3f-e0c1-4120-a2be-fe64d20e4678 · outbound

This paper cites an unresolved cited work.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.709348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.709348Z digest=sha256:0c98bae7eb704bd6d8e2eee705eb98266179f50507691c976731f8fbb02cb901

Observation e576da72-ae9e-4f0b-a119-3bdd74ebd4e8 · outbound

This paper cites an unresolved cited work.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.713281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.713281Z digest=sha256:2ec9d6565770f264356981f15cdc6a1546d4f6d79f6b48b14b6e305f90e79852

Observation b61a4e65-e746-4664-ae2a-7035de25d74d · outbound

This paper cites an unresolved cited work.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.717121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.717121Z digest=sha256:ff51b0f3c62ebb89f7d4030eea01c8ce550ac6a4f393e1b3e59dc63f30e98d92

Observation 39132cc4-3f89-4c08-9e10-d14b113930d9 · outbound

This paper cites Continuoustimeq-learningformean-fieldcontrolproblems.Applied Mathematics & Optimization, 91(1):10, 2025.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Continuoustimeq-learningformean-fieldcontrolproblems.Applied Mathematics & Optimization, 91(1):10, 2025

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.720868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.720868Z digest=sha256:3d144ddebc8f13b4ef928a56e7152ebb29d00db7972b2efc4d7820ac81e486ee

Observation 63e59b80-236e-4bb3-bdd5-36974a329498 · outbound

This paper cites Unified continuous-time q-learning for mean-field game and mean-field control problems.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Unified continuous-time q-learning for mean-field game and mean-field control problems

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.724649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.724649Z digest=sha256:e2b849cafc865e85058988bc9f3a2aa809828beebc09503422586c0c66dca0d7

Observation 0f5886ff-a829-4fe8-a48e-340def7ae7cd · outbound

This paper cites Zhang, Z.

Continuous-Time Reinforcement Learning for $N$-Player Stochastic Differential Games with Exploratory Policies Zhang, Z

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T11:22:20.728645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:22:20.728645Z digest=sha256:314f7824de035fb65bda8bff944a4fc0e819acd77d329bafda8bbdb5b2e2a9b2

Pith citing papers

No inbound Pith citation observations are available.