Pith. sign in

Paper Citation Record · LEDGER

Policy Gradient for Continuous-Time Mean-Field Control

As of 10 August 2026, this Paper Citation Record lists 82 of 82 outbound references and 1 inbound Pith citation observation for arXiv:2605.20718.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.20718 v1

Coverage vector

measured 82 of 82 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-21T04:15:32.156380Z

measured 83 of 83 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T07:42:38.840826Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

82 of 82 outbound references displayed

  • verified exact6
  • verified fuzzy70
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 05de4c07-224a-4306-8b9a-8bf16bf4161c · outbound

This paper cites Mean field type control with congestion.Applied Mathematics & Optimization, 73(3):393–418.

Policy Gradient for Continuous-Time Mean-Field Control Mean field type control with congestion.Applied Mathematics & Optimization, 73(3):393–418

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.490060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:a348e0727a064155932e32d41594d6ba4c0c59c2d76f5ad8c4b80621f8fc2e5f

Observation 3dd960af-38e9-4a6d-aaf1-1032088a9090 · outbound

This paper cites Mean field type control with congestion II: An augmented lagrangian method.

Policy Gradient for Continuous-Time Mean-Field Control Mean field type control with congestion II: An augmented lagrangian method

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.517720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:32c2ad438e9c9cf6723a42e73424c9973d67dfe5fb5be9b80c3c6c094609f127

Observation 6fa60bf4-8ce6-44a1-a761-ee849b12035e · outbound

This paper cites A maximum principle for SDEs of mean-field type.Applied Mathematics and Optimization, 63(3):341–356.

Policy Gradient for Continuous-Time Mean-Field Control A maximum principle for SDEs of mean-field type.Applied Mathematics and Optimization, 63(3):341–356

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.499691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:9872a50e437511875a13e7b421713b4a85b81ec299e2dfea3331f6da7f3162f9

Observation ef6e21f7-309a-4258-be77-d44777e1c7d4 · outbound

This paper cites an unresolved cited work.

Policy Gradient for Continuous-Time Mean-Field Control Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-05-21T10:04:59.497337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:3425c6ea2e733ba457cd8831609c4bfd3548e6950b531adf91043fa934b65cd9

Observation 5d337347-a7c9-4a61-ab40-a7df3afb61e2 · outbound

This paper cites an unresolved cited work.

Policy Gradient for Continuous-Time Mean-Field Control Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-05-21T10:04:59.506991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:46142e761c0ce320140a5e61c170c8a6a4565e401069f12c54d1f334dcdb3f95

Observation cde67c11-0658-408a-8749-49c08e5e5535 · outbound

This paper cites A weak martingale approach to linear-quadratic McKean–Vlasov stochastic control problems.Journal of Optimization Theory and Applications, 181(2):347–382.

Policy Gradient for Continuous-Time Mean-Field Control A weak martingale approach to linear-quadratic McKean–Vlasov stochastic control problems.Journal of Optimization Theory and Applications, 181(2):347–382

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.510647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:19aae337d9ea124881d779b54bf1410b420eccd0ffa08cdbc0a51e6921b1009a

Observation ec294a9e-112a-4809-a0db-58f70fa30209 · outbound

This paper cites Viscosity solutions of fully second-order hjb equations in the wasserstein space.SIAM Journal on Control and Optimization.

Policy Gradient for Continuous-Time Mean-Field Control Viscosity solutions of fully second-order hjb equations in the wasserstein space.SIAM Journal on Control and Optimization

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.515366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:f9d0e46ba40aa9cdefeb66e0e82c63ffad6bbb5d65fb80bfa477280a2e3f0a22

Observation e692bfda-9e9b-44a6-beee-57e0867ade51 · outbound

This paper cites an unresolved cited work.

Policy Gradient for Continuous-Time Mean-Field Control Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-05-21T10:04:59.478246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:c5d0091d355a715facb069a312efec7e17c75c24b2dab3aad3ac8427a0ec05d5

Observation df652d8d-1a42-4a1a-813f-aaa5787aefe4 · outbound

This paper cites Comparison for semi-continuous viscosity solutions for second order pdes on the wasserstein space.Journal of Differential Equations, 455:1–31.

Policy Gradient for Continuous-Time Mean-Field Control Comparison for semi-continuous viscosity solutions for second order pdes on the wasserstein space.Journal of Differential Equations, 455:1–31

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.394800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:014a14b2f0f4d07a8074f5ce085bfe84cd8c66906e05670e55570e89f6567f48

Observation b4866fb1-ecc5-4341-9b5a-1471cad57481 · outbound

This paper cites Comparison of viscosity solutions for a class of second-order pdes on the wasserstein space.Communications in Partial Differential Equations, 50(4):570–613.

Policy Gradient for Continuous-Time Mean-Field Control Comparison of viscosity solutions for a class of second-order pdes on the wasserstein space.Communications in Partial Differential Equations, 50(4):570–613

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.378667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:4d3e458eb6bf11743f2c2f98ebe59d960ce505693091ea9f900d497da0e7d7f0

Observation d168fa9b-0f8d-4a70-840a-3b6981181431 · outbound

This paper cites Convergence rate of particle system for second-order pdes on wasserstein space.SIAM Journal on Control and Optimization, 63(3):1515–1782.

Policy Gradient for Continuous-Time Mean-Field Control Convergence rate of particle system for second-order pdes on wasserstein space.SIAM Journal on Control and Optimization, 63(3):1515–1782

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.520327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:292e7dce72815713fbd8a9e44f1bbee1797907f5bd9e3302b26f451f34439c5c

Observation 4dec7e8d-8736-4516-aa37-1f5df937b1ee · outbound

This paper cites Mean-field phibe: Continuous-time mean-field reinforcement learning from discrete-time data.Preprint.

Policy Gradient for Continuous-Time Mean-Field Control Mean-field phibe: Continuous-time mean-field reinforcement learning from discrete-time data.Preprint

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.392078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:e42120a481d235362224fb9849aa141b4bb1983968fcc7c09406df6d5e990f62

Observation bcad8097-7337-4ea7-b363-fbe4ad755c8e · outbound

This paper cites Ergodicity and turnpike properties of linear-quadratic mean field control problems.

Policy Gradient for Continuous-Time Mean-Field Control Ergodicity and turnpike properties of linear-quadratic mean field control problems

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.399715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:68c37a4543bab5b6afc098b885eed0be3cfbb3d01cb5a5cdc06f7aa6844c1c07

Observation 225173af-5163-41b6-b458-e17f4b74a673 · outbound

This paper cites Convergence and turnpike properties of linear-quadratic mean field control problems with common noise.arXiv preprint arXiv.

Policy Gradient for Continuous-Time Mean-Field Control Convergence and turnpike properties of linear-quadratic mean field control problems with common noise.arXiv preprint arXiv

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.470720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:247b504f4e60e69e5f5e7b67e7ed2caad6349906abe725fb7a7fbd5c6cdcd5f2

Observation 39ab7bfb-4943-4faf-a496-2c4a4d101bc8 · outbound

This paper cites Learning with linear function approximations in mean-field control.Journal of Machine Learning Research, 26(192):1–53.

Policy Gradient for Continuous-Time Mean-Field Control Learning with linear function approximations in mean-field control.Journal of Machine Learning Research, 26(192):1–53

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.527805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:df786dde092f18b673ca9d4ce47d764c95710a46695bfb01f6f4f51b554dee45

Observation d7f1b6b7-27fb-4d45-a2ca-6e6e7c752ad2 · outbound

This paper cites Propagation of chaos of forward-backward stochastic differential equa- tions with graphon interactions.Applied Mathematics and Optimization, 88(1):Article 25.

Policy Gradient for Continuous-Time Mean-Field Control Propagation of chaos of forward-backward stochastic differential equa- tions with graphon interactions.Applied Mathematics and Optimization, 88(1):Article 25

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.539194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:1e28571f9898c396c5310978b4e6c352b5038d5371283c001b4225921d039f1f

Observation d207f5f3-abf6-4604-883a-e28d5d600a40 · outbound

This paper cites SpringerBriefs in Mathematics.

Policy Gradient for Continuous-Time Mean-Field Control SpringerBriefs in Mathematics

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.541373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:c1f78d87c434d18a0fdceae9a704ebc1130be89c81bfacf98362a71cdd9074fa

Observation 8808a4bd-8df7-4f47-9962-32a1d05c70bc · outbound

This paper cites The pontryagin maximum principle in the wasserstein space.Calculus of Varia- tions and Partial Differential Equations, 58:1–36.

Policy Gradient for Continuous-Time Mean-Field Control The pontryagin maximum principle in the wasserstein space.Calculus of Varia- tions and Partial Differential Equations, 58:1–36

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.543446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:4d1b22a9dc42303374301b6edca982b3b2a561e3402664dd9720c5718e069a28

Observation 82a92baa-4dc1-41c9-a578-9f3ef272a291 · outbound

This paper cites A general maximum principle for SDEs of mean-field type.Applied Mathematics and Optimization, 64(2):197–216.

Policy Gradient for Continuous-Time Mean-Field Control A general maximum principle for SDEs of mean-field type.Applied Mathematics and Optimization, 64(2):197–216

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.545572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:a70302e42ce2add415d553ee059e883651ac8c5d7fb3ca3f7632916788be6fd9

Observation ba044ed8-9990-4128-b5b3-6aef188d1762 · outbound

This paper cites Mean-field stochastic differential equations and associated pdes.Annals of Probability, 45(2):824–878.

Policy Gradient for Continuous-Time Mean-Field Control Mean-field stochastic differential equations and associated pdes.Annals of Probability, 45(2):824–878

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.552625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:0636f6cc59c7edf6657157bed3ac04024e27443a2e295c2e5f83a82f4cdb6286

Observation 96513dbb-d7f9-4d98-9e53-619747b34459 · outbound

This paper cites Max Reppen, and H.

Policy Gradient for Continuous-Time Mean-Field Control Max Reppen, and H

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.554409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:6fea5326a9fba1ef64c713b42014d0eaeaf1c1235430b8098eec5e219ad3e474

Observation 28402ac9-17e9-4ff0-a63d-9e00d9396c1c · outbound

This paper cites Society for Industrial and Applied Mathematics (SIAM), Philadelphia, PA.

Policy Gradient for Continuous-Time Mean-Field Control Society for Industrial and Applied Mathematics (SIAM), Philadelphia, PA

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.492668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:1bda7bb972048da6ac248652d6ef6fa8bdce961701159ad97864a86cbfc0f9b5

Observation cd547b01-babd-46d0-954b-242c9ce3612d · outbound

This paper cites Forward–backward stochastic differential equations and controlled McKean– Vlasov dynamics.Annals of Probability, 43(5):2647–2700.

Policy Gradient for Continuous-Time Mean-Field Control Forward–backward stochastic differential equations and controlled McKean– Vlasov dynamics.Annals of Probability, 43(5):2647–2700

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.512969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:3696025cc21fb6ad0cebb06e4752dfd64e0f4e6dd77506399160201dde805fad

Observation b49508fa-a66b-437d-9356-dfcd4c8e5a95 · outbound

This paper cites I, volume 83 of Probability Theory and Stochastic Modelling.

Policy Gradient for Continuous-Time Mean-Field Control I, volume 83 of Probability Theory and Stochastic Modelling

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.564121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:1f9a81edc37e435193febfa932dfda8ba8b8eed50f7bf4ca1b6b5cbe5480264a

Observation 9c8ac905-bf30-40a9-8e0b-c0e701503751 · outbound

This paper cites II, volume 84 of Probability Theory and Stochastic Modelling.

Policy Gradient for Continuous-Time Mean-Field Control II, volume 84 of Probability Theory and Stochastic Modelling

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.460554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:76f5fa2fdf4d97c4e366b93e21584a611ad35677bb617374583a325dd04ac939

Observation 33e76a54-c660-4cfc-81fa-eb33ed096991 · outbound

This paper cites Control of McKean–Vlasov dynamics versus mean field games.

Policy Gradient for Continuous-Time Mean-Field Control Control of McKean–Vlasov dynamics versus mean field games

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.462811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:9280fefa5740c122ccd8d2be667298adb7fe372e99e7c9cc1652c3c436f58b6b

Observation 3ec279bb-67a8-4c4e-bb19-ac77d2674b18 · outbound

This paper cites Mean field games and systemic risk.Communications in Mathematical Sciences, 13(4):911–933.

Policy Gradient for Continuous-Time Mean-Field Control Mean field games and systemic risk.Communications in Mathematical Sciences, 13(4):911–933

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.467872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:512dae616703372312c8b6aa96c45cc9a1a8d80555763295d5732a2e20b66e69

Observation 5db4fbbe-c53a-4e30-916c-e4fb5ee5636b · outbound

This paper cites an unresolved cited work.

Policy Gradient for Continuous-Time Mean-Field Control Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-05-21T10:04:59.536369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:667ff0589b5f2c599617afba0597f3eae67524414dc4e747a9de30a76157fae5

Observation 48c6c7de-978a-48f0-b6ba-e4f829410715 · outbound

This paper cites Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods.

Policy Gradient for Continuous-Time Mean-Field Control Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:19:33.507409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:d536ddf065dcf031abd0c768b595938f9f4ef7d988061d8718bd349709a59364

Observation d99aaeb9-1669-40bc-9057-5e59ec113a97 · outbound

This paper cites Carrillo, Young-Pil Choi, and Maxime Hauray.

Policy Gradient for Continuous-Time Mean-Field Control Carrillo, Young-Pil Choi, and Maxime Hauray

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.397202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:5b051b6c10f1758f9023f3b4be15d1c05dc633c6a02559f9afc36c5833f063b5

Observation 0e36a81f-ea3d-4d28-808f-6a37742f26e7 · outbound

This paper cites Carrillo, Massimo Fornasier, Giuseppe Toscani, and Francesco Vecil.

Policy Gradient for Continuous-Time Mean-Field Control Carrillo, Massimo Fornasier, Giuseppe Toscani, and Francesco Vecil

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.457906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:4284f6fc6190a19ad99c6f9f90981dae15c525a1ac9b959284cfc46f9cd9920c

Observation c4f9cfa8-aaae-4a02-b634-f9e9ab96c9d1 · outbound

This paper cites Propagation of chaos: A review of models, methods and applications.

Policy Gradient for Continuous-Time Mean-Field Control Propagation of chaos: A review of models, methods and applications

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.559847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:462b65ec932d5db9423bf2203af8a1edc40d5eaabb2765815624d23736134472

Observation 4f1a21e8-3e36-4c01-9bb5-76742afaa769 · outbound

This paper cites Numerical method for FBSDEs of McKean–Vlasov type.Annals of Applied Probability, 29(3):1640–1684.

Policy Gradient for Continuous-Time Mean-Field Control Numerical method for FBSDEs of McKean–Vlasov type.Annals of Applied Probability, 29(3):1640–1684

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.473218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:7929025a9bca3cf1dc0a594c3c1fd7cfa28d1a123bf20e5c47e9caf047159580

Observation 9cdb784c-2560-41da-8ed4-3fca87cd6654 · outbound

This paper cites Emergent behavior in flocks.IEEE Transactions on Automatic Control, 52(5):852–862.

Policy Gradient for Continuous-Time Mean-Field Control Emergent behavior in flocks.IEEE Transactions on Automatic Control, 52(5):852–862

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.447406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:b74001fc4cfe102e1b6add309d5782e15501b307713fb676fa90eb6783d66bd4

Observation 53120db2-2bdb-42f9-935b-c6334b7ed650 · outbound

This paper cites Mckean–vlasov optimal control: Limit theory and equivalence between different formulations.Mathematics of Operations Research, 47(4):2891–2930.

Policy Gradient for Continuous-Time Mean-Field Control Mckean–vlasov optimal control: Limit theory and equivalence between different formulations.Mathematics of Operations Research, 47(4):2891–2930

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.452538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:fa46cf55c8d7658dd9c4c0df9652ea03987c0914cb9bfa0ffb1ca04967492e75

Observation e6713b11-f19b-481a-a0dc-f6ed3517fff9 · outbound

This paper cites Martingale measures and stochastic calculus.Probability Theory and Related Fields, 84(1–2):83–101.

Policy Gradient for Continuous-Time Mean-Field Control Martingale measures and stochastic calculus.Probability Theory and Related Fields, 84(1–2):83–101

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.444803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:a4192f54fc6aebda06efabb951ba1ff490eb91be49bf3014b860b25c246be04d

Observation e0e10495-ee5a-4664-8447-ddaa699b2b09 · outbound

This paper cites Actor-critic learning for mean-field control in continuous time.J.

Policy Gradient for Continuous-Time Mean-Field Control Actor-critic learning for mean-field control in continuous time.J

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.449933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:28ba0b1997b879cbba85c0b909e5ccf19bf508990688db7aab7046c858b151b3

Observation b91cba99-3aef-4121-adbf-7a019cae5bc3 · outbound

This paper cites Full error analysis of policy gradient learning algorithms for exploratory linear quadratic mean-field control problem in continuous time with common noise.

Policy Gradient for Continuous-Time Mean-Field Control Full error analysis of policy gradient learning algorithms for exploratory linear quadratic mean-field control problem in continuous time with common noise

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:19:33.522026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:0c18b2435cf5f9ebcee0b1528a44124f1fa306e50921088c2ce054bfb784d348

Observation 4ab26cd6-e292-4b09-871d-98e5010e1d6e · outbound

This paper cites Hamilton-Jacobi equations in the Wasserstein space.Methods Appl.

Policy Gradient for Continuous-Time Mean-Field Control Hamilton-Jacobi equations in the Wasserstein space.Methods Appl

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.561911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:a2ba6a97c72507e0a719ec11c299a5fb5442d9b0772a9d2362652cd83613ab1d

Observation 59dcfc00-79c1-44c8-a0ed-aab8e9abf1d5 · outbound

This paper cites Opinion dynamics and bounded confidence: Models, analysis and simulation.

Policy Gradient for Continuous-Time Mean-Field Control Opinion dynamics and bounded confidence: Models, analysis and simulation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.436610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:299e85fa13c53ed62c8cd4772ae8fedc67efbcb7645b0dac294d0fe91d7d9000

Observation 498872b3-5e19-41c8-a718-7b7eff381ecf · outbound

This paper cites Howard.Dynamic programming and Markov processes.

Policy Gradient for Continuous-Time Mean-Field Control Howard.Dynamic programming and Markov processes

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.404624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:7c9fb6a65e24c3faa6d43d6cc301b8e7a8f9a3bc7e516d8a507e895874b20efa

Observation b2d33492-3963-43e9-b498-24bc11a715c3 · outbound

This paper cites A linear-quadratic optimal control problem for mean-field stochastic differential equations in infinite horizon.Math.

Policy Gradient for Continuous-Time Mean-Field Control A linear-quadratic optimal control problem for mean-field stochastic differential equations in infinite horizon.Math

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.475956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:e09c19b2a6a75ce6839ba9dd071842126921adf277ba4780d18dffe08d27efea

Observation a31369c5-5175-4a8b-a28a-a19960545c0d · outbound

This paper cites Infinite horizon value functions in the Wasserstein spaces.J.

Policy Gradient for Continuous-Time Mean-Field Control Infinite horizon value functions in the Wasserstein spaces.J

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.434217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:5ca380b67a5d922df45f0a190a33c7f00b1ad8572d05ac89ce591ae62741bd71

Observation d9d8dd3c-e052-4017-aee4-6d552d7b7668 · outbound

This paper cites Accuracy of discretely sampled stochastic policies in continuous-time reinforcement learning.

Policy Gradient for Continuous-Time Mean-Field Control Accuracy of discretely sampled stochastic policies in continuous-time reinforcement learning

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:19:33.511413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:a5f9cb1134d233fe7bc649239be23b0aa02dd95e3e9589d368f7e764d764db8a

Observation 99f04930-10d8-40bc-bc74-82556801903d · outbound

This paper cites Policy evaluation and temporal-difference learning in continuous time and space: A martingale approach.Journal of Machine Learning Research, 23(1):6918–6972.

Policy Gradient for Continuous-Time Mean-Field Control Policy evaluation and temporal-difference learning in continuous time and space: A martingale approach.Journal of Machine Learning Research, 23(1):6918–6972

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.522899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:c3f3fc6336e3c086f47f2aab98270ad9a6a7c62ae126ceaaaaae9e417bbe1dc4

Observation 1c89e2b0-8c49-498e-b000-4347767a5658 · outbound

This paper cites Policy gradient and actor-critic learning in continuous time and space.Journal of Machine Learning Research, 23(1):12603–12652.

Policy Gradient for Continuous-Time Mean-Field Control Policy gradient and actor-critic learning in continuous time and space.Journal of Machine Learning Research, 23(1):12603–12652

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.480561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:6a42cf52a5f0ec44c8aae7350936f4363bc41c81763539fea8937e8259a1a9d3

Observation be968192-42a4-443c-be7d-a8692dc2f708 · outbound

This paper cites an unresolved cited work.

Policy Gradient for Continuous-Time Mean-Field Control Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-05-21T10:04:59.428864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:140d28e36d84077f07e6d6a1461ed9f31996960dc53cba6048c7499035d10ae5

Observation 08a394e0-e2db-4bbf-b050-766fcd1d206e · outbound

This paper cites an unresolved cited work.

Policy Gradient for Continuous-Time Mean-Field Control Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-05-21T10:04:59.529917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:0d4c321d50af7bfeb15b1b69d775255ecc7048a9181466f07615992256317f66

Observation 17085b17-2973-44a8-b454-a1f6ccb74323 · outbound

This paper cites Limit theory for controlled McKean–Vlasov dynamics.SIAM Journal on Control and Optimization, 55(3):1641–1672.

Policy Gradient for Continuous-Time Mean-Field Control Limit theory for controlled McKean–Vlasov dynamics.SIAM Journal on Control and Optimization, 55(3):1641–1672

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.487899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:fdff9fbe50c0645ad43482dd8f875f7a877baa74fbdfd60d6a636165bc652bde

Observation b1a2b330-50da-41ab-8bb1-ddb4cc4f77d2 · outbound

This paper cites Mean field games.Japanese Journal of Mathematics, 2(1):229–260.

Policy Gradient for Continuous-Time Mean-Field Control Mean field games.Japanese Journal of Mathematics, 2(1):229–260

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.482972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:d7253db7e9b5f5c5cf90dd5a7828fcef0252017b870264b3d09b91ab08b8883f

Observation 6871f50e-e74e-466b-bab9-4bbf6a2e86d8 · outbound

This paper cites Numerical methods for mean field games and mean field type control.Proceedings of Symposia in Applied Mathematics, 78:221–282.

Policy Gradient for Continuous-Time Mean-Field Control Numerical methods for mean field games and mean field type control.Proceedings of Symposia in Applied Mathematics, 78:221–282

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.534428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:99e15c24ce71b40618aedb389846d4e7b6790bf3870bcd7684dd639f70057b16

Observation 8b2ef5be-d50f-4397-9b32-d6569ec2a66c · outbound

This paper cites Dynamic programming for mean-field type control.Comptes Rendus Math´ ematique.

Policy Gradient for Continuous-Time Mean-Field Control Dynamic programming for mean-field type control.Comptes Rendus Math´ ematique

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.532000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:e564b7e5e7ff32d8d46cc3f053749ac3827264c18f0f37df98061dcbbb3d5c96

Observation 750f29c2-a508-46ea-a5f9-d833186900ac · outbound

This paper cites Mean-field stochastic linear quadratic optimal control problems: Closed-loop solvability.Probability, Uncertainty and Quantitative Risk, 1(1):2.

Policy Gradient for Continuous-Time Mean-Field Control Mean-field stochastic linear quadratic optimal control problems: Closed-loop solvability.Probability, Uncertainty and Quantitative Risk, 1(1):2

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.485361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:b94cd73584a9abc44d52818170aed19427fd5c701873873fd385ddb497b2515d

Observation a896a42a-f91f-4b0f-974c-5f53b29afb26 · outbound

This paper cites Cours au Coll` ege de France: Th´ eorie des jeux ` a champ moyen.

Policy Gradient for Continuous-Time Mean-Field Control Cours au Coll` ege de France: Th´ eorie des jeux ` a champ moyen

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.426546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:f73cd06e00616555a33c31cf370ff761497129931d4e4a21ad2e5ebc2bcd7332

Observation c6f34b89-825b-4fcc-8308-0c01065439f3 · outbound

This paper cites Linear quadratic optimal control of conditional McKean–Vlasov equation with random coefficients and applications.Journal of Mathematical Economics, 66:7–26.

Policy Gradient for Continuous-Time Mean-Field Control Linear quadratic optimal control of conditional McKean–Vlasov equation with random coefficients and applications.Journal of Mathematical Economics, 66:7–26

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.431615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:bef07ccb59b82bfe0dbdd9af0ada9bf4cd892ca9e684e76ced93d03a081c8b53

Observation a0d208c2-9953-4632-96ef-429bee3168c7 · outbound

This paper cites Actor-critic learning algorithms for mean-field control with moment neural networks.

Policy Gradient for Continuous-Time Mean-Field Control Actor-critic learning algorithms for mean-field control with moment neural networks

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.439243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:9ff1f9147e528733319e493dd6796b4a945a818f4365dfe6901e4669aa307d9f

Observation 73951ded-2a0f-448b-a4c4-607421a6fb57 · outbound

This paper cites Dynamic programming for optimal control of stochastic McKean–Vlasov dynamics.

Policy Gradient for Continuous-Time Mean-Field Control Dynamic programming for optimal control of stochastic McKean–Vlasov dynamics

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.495075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:f9e7a2f95720f2000de4f28feb6032a03e547a8005877045af4da26e0ce5574c

Observation 33d2da7e-64cc-45b0-996d-4e149cbfe98a · outbound

This paper cites Bellman equation and viscosity solutions for mean-field stochastic control problem.

Policy Gradient for Continuous-Time Mean-Field Control Bellman equation and viscosity solutions for mean-field stochastic control problem

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.419115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:6d7add3bdd2caf3c97cb25a96569267d8f66f970798b14485e4cee052fdbf039

Observation ee922660-e942-4f26-9215-b07d7eaf58b9 · outbound

This paper cites Mean-field neural networks: Learning mappings on wasserstein space.Neural Net- works, 168:380–393.

Policy Gradient for Continuous-Time Mean-Field Control Mean-field neural networks: Learning mappings on wasserstein space.Neural Net- works, 168:380–393

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.424023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:5dde4686e5acd0622ca6ed6d40c5c4cfc676fe7dc5403195d0df786eb9bea936

Observation 5d6cec1d-165e-412e-b94c-f66e66f1fec7 · outbound

This paper cites Continuous-time q-learning for mean-field control with common noise, part-i: Theoretical foundations.

Policy Gradient for Continuous-Time Mean-Field Control Continuous-time q-learning for mean-field control with common noise, part-i: Theoretical foundations

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.416581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:47ac3ab3a9fb0238c98863ee5e03976e83a6495a5eec491ddc75a24f6783510d

Observation 5c65b131-2f3d-47b2-a3ec-500ec902646c · outbound

This paper cites Continuous-time q-learning for mean-field control with common noise, part-ii: q-learning algorithms.

Policy Gradient for Continuous-Time Mean-Field Control Continuous-time q-learning for mean-field control with common noise, part-ii: q-learning algorithms

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.421629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:b0ca93a49f48f832f10e896750ed3cb5812f774a6b5c0670c6ddef009841939d

Observation 7f8605af-8b72-47ce-80d7-d89af6336aec · outbound

This paper cites Osher, Wuchen Li, Levon Nurbekyan, and Samy Wu Fung.

Policy Gradient for Continuous-Time Mean-Field Control Osher, Wuchen Li, Levon Nurbekyan, and Samy Wu Fung

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.441828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:a558a381b813c6cd3e89e9277a5a4622f330ef72ba88ec7b09d80ac20300af7f

Observation 288cf3e9-8ef0-43b6-bcb8-7d44fc1e9bb3 · outbound

This paper cites Learning algorithms for mean field optimal control.

Policy Gradient for Continuous-Time Mean-Field Control Learning algorithms for mean field optimal control

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:19:33.514765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:ad8eecbc14688bd135b29fd0f0e78f9853801417abc130bffc252bdc9de9109f

Observation 695a9fc3-c832-493b-bb48-f36ca5edef3e · outbound

This paper cites Mete Soner and Qinxin Yan.

Policy Gradient for Continuous-Time Mean-Field Control Mete Soner and Qinxin Yan

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.550478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:660cef0151ea688fe24867fa2ce6373425fea2393fab1837a4d42683e239f2cc

Observation a27f048f-3d6b-469f-95d9-e39eb45c394e · outbound

This paper cites Mete Soner and Qinxin Yan.

Policy Gradient for Continuous-Time Mean-Field Control Mete Soner and Qinxin Yan

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.409253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:884e91798a62ca7a146ade93a841773246f956b74b9a70c70a8fbab7fbdfc964

Observation 215e1ce4-f58b-40b7-81f5-bba0820c2935 · outbound

This paper cites Mean-field stochastic linear quadratic optimal control problems: Open-loop solvabilities.ESAIM: Control, Optimisation and Calculus of Variations, 23(3):1099–1127.

Policy Gradient for Continuous-Time Mean-Field Control Mean-field stochastic linear quadratic optimal control problems: Open-loop solvabilities.ESAIM: Control, Optimisation and Calculus of Variations, 23(3):1099–1127

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.384181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:e3197cbd6fbdd64b322974d76c50083644411aad1692fc3e4376d655aa9e9a47

Observation 8eeaf5bf-84c3-4b26-96da-ee7a43ae9c64 · outbound

This paper cites Mean-field stochastic linear-quadratic optimal control problems: Weak closed-loop solvability.Mathematical Control and Related Fields, 11(1):47–71.

Policy Gradient for Continuous-Time Mean-Field Control Mean-field stochastic linear-quadratic optimal control problems: Weak closed-loop solvability.Mathematical Control and Related Fields, 11(1):47–71

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.381332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:7252e6eb1ff41158b427647889a3bad01f9e2cb83acaf90b348bc10f2c552df1

Observation 97ec6ec2-e7f9-4206-81c6-2dd49cdf60c4 · outbound

This paper cites Springer Nature.

Policy Gradient for Continuous-Time Mean-Field Control Springer Nature

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.455576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:9914fd6a9c7d8a77f74e6497e3b0583d31e340e449af27f717b3fd7dd4f140a7

Observation 695e3a34-7ebb-4324-b980-ef4580092137 · outbound

This paper cites The exact law of large numbers via fubini extension and characterization of insurable risks.Journal of Economic Theory, 126(1):31–69.

Policy Gradient for Continuous-Time Mean-Field Control The exact law of large numbers via fubini extension and characterization of insurable risks.Journal of Economic Theory, 126(1):31–69

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.504533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:555489ce43a3dcc3103f158e2ede67149a7e3713d8f215d141a3d50abcae7579

Observation 27c5ee06-cd2e-46ce-b259-15c9c6894deb · outbound

This paper cites Sutton and Andrew G.

Policy Gradient for Continuous-Time Mean-Field Control Sutton and Andrew G

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.414273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:cc08e0ec8e27770d63f54c9f6f745c962678bf1eafe12d193c9ff56ff3fafe23

Observation ee42fb72-20b3-4147-ae61-4e6c5795b0c1 · outbound

This paper cites Synthesis Lectures on Artificial Intelligence and Machine Learning.

Policy Gradient for Continuous-Time Mean-Field Control Synthesis Lectures on Artificial Intelligence and Machine Learning

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.502141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:91b5a0721317e745581a715967c7bf0965d5a59d7191bcdf9408dcba81f7ca21

Observation cb9dc937-8511-442d-901c-1fa2ce1b82d5 · outbound

This paper cites Topics in propagation of chaos.

Policy Gradient for Continuous-Time Mean-Field Control Topics in propagation of chaos

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.525348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:be792a134fd69e3e113aa615c257c0e2b3a36e5620b7cf5e49f16242b94bc6ee

Observation db9aa6b9-6bc6-4496-bcfd-b45ad55f4b41 · outbound

This paper cites Optimal scheduling of entropy regularizer for continuous- time linear-quadratic reinforcement learning.SIAM Journal on Control and Optimization, 62(1):135–166.

Policy Gradient for Continuous-Time Mean-Field Control Optimal scheduling of entropy regularizer for continuous- time linear-quadratic reinforcement learning.SIAM Journal on Control and Optimization, 62(1):135–166

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.402198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:4f8cd59de4e4148949e323dea60ec4a696fdd7891d434883698de95942daf3be

Observation b5b60a16-c453-4f04-9944-54b1f1d33506 · outbound

This paper cites Making deep Q-learning methods robust to time discretization.

Policy Gradient for Continuous-Time Mean-Field Control Making deep Q-learning methods robust to time discretization

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.407104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:d02d58cf5823a038de94e7842036ba1f5f77760323a268e46a41aba2e7ecf484

Observation e7fbaa1d-bd78-42ed-af19-367c06530e3c · outbound

This paper cites Kinetic models of opinion formation.Communications in Mathematical Sciences, 4(3):481–496.

Policy Gradient for Continuous-Time Mean-Field Control Kinetic models of opinion formation.Communications in Mathematical Sciences, 4(3):481–496

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.389495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:269b525beffcf32cc42a3bdee0627e584b61590ce1b41932df5eaf6ef2fff5e3

Observation f5f4cf72-5e45-4958-adc9-34d24c3fe061 · outbound

This paper cites Reinforcement learning in continuous time and space: A stochastic control approach.Journal of Machine Learning Research, 21(198):1–34.

Policy Gradient for Continuous-Time Mean-Field Control Reinforcement learning in continuous time and space: A stochastic control approach.Journal of Machine Learning Research, 21(198):1–34

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.386913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:5025fd131943cc431ebf1064d4635f81edd9d3ed705068925415f20466998700

Observation b1047761-c87e-4955-a12a-fac5d9b565e0 · outbound

This paper cites Global convergence of policy gradient for linear- quadratic mean-field control/game in continuous time.

Policy Gradient for Continuous-Time Mean-Field Control Global convergence of policy gradient for linear- quadratic mean-field control/game in continuous time

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.411820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:dd5df01c264f104eb0bf4a1907481d0b9fff186652f5885ccef9db5bb43040e9

Observation 0093bcd8-b17e-4bc6-a2cb-298d91ca5470 · outbound

This paper cites Continuous time q-learning for mean-field control problems.Applied Mathematics & Opti- mization, 91(1):10.

Policy Gradient for Continuous-Time Mean-Field Control Continuous time q-learning for mean-field control problems.Applied Mathematics & Opti- mization, 91(1):10

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.547797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:fb7abd4ad75d723b97dfef7d8e53f095b8363308547c217a039fa4aad1b067b4

Observation e5db062e-971d-4d5a-86c4-ebd664ea8d69 · outbound

This paper cites Unified continuous-time q-learning for mean-field game and mean-field control problems.

Policy Gradient for Continuous-Time Mean-Field Control Unified continuous-time q-learning for mean-field game and mean-field control problems

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-08-03T01:10:22.407760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:c03476d0d28107aa2291b71af6a44b162d562c63bd767701c8b255c8082c9908

Observation 77da7314-7111-4e8d-a5b9-4924946a5b33 · outbound

This paper cites Efficient local planning with linear function approximation.

Policy Gradient for Continuous-Time Mean-Field Control Efficient local planning with linear function approximation

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.465245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:713efc2032cdf99b04789f5d71d63c0cf347c788cf5cb7b19e44b5ddae1ace84

Observation 50d101de-2164-4cc4-b593-91e8ba4da176 · outbound

This paper cites Linear-quadratic optimal control problems for mean-field stochastic differential equations.SIAM Journal on Control and Optimization, 51(4):2809–2838.

Policy Gradient for Continuous-Time Mean-Field Control Linear-quadratic optimal control problems for mean-field stochastic differential equations.SIAM Journal on Control and Optimization, 51(4):2809–2838

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.557643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:a26b69052cbc264627f887b4b82952b343c476f47ed589f39cc665f64e016919

Observation 5c5ca38c-7cdb-48f8-adcb-87334f300f85 · outbound

This paper cites PhiBE: A PDE-based Bellman equation for continuous time policy evaluation.

Policy Gradient for Continuous-Time Mean-Field Control PhiBE: A PDE-based Bellman equation for continuous time policy evaluation

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:19:33.525856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:dd4a478f8be043e3ceab951ae45c4461a51974e1607858a49b0359b3e767321f

Pith citing papers

Observation 4ad6043e-c0ec-4375-9534-e26bdd39c1c0 · inbound

Actor-Critic Learning for Extended Mean Field Control with Deterministic Policies cites this paper.

Actor-Critic Learning for Extended Mean Field Control with Deterministic Policies Policy Gradient for Continuous-Time Mean-Field Control

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-14T07:42:38.840826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T07:42:38.840826Z digest=sha256:fdd250dd753109372b741d1388a6644099e64eb24b4a5c45cea1bb928fc6a00f