Pith. sign in

Paper Citation Record · LEDGER

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games

As of 9 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2506.05894.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.05894 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:25:56.452058Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

63 of 63 outbound references displayed

  • verified exact2
  • verified fuzzy48
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 785a7eff-dc7e-424c-afed-7efebc62578f · outbound

This paper cites Value iteration algorithm for mean- field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Value iteration algorithm for mean- field games

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.559605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:49.554219Z digest=sha256:7088622184c18f2e440596ea8524c0541c320f4ef4886041f0b55d08dd44aa8e

Observation 7db36181-eddf-4575-8802-3f5cdfca4e99 · outbound

This paper cites Unified reinforcement q-learning for mean field game and control problems.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Unified reinforcement q-learning for mean field game and control problems

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.551580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:49.639583Z digest=sha256:c64601f09148fb9795180f7cfa5a56a88513b412d4afe9083d212e6b35c80fe8

Observation efcfd009-787b-483b-b5fe-2e6069da6fb5 · outbound

This paper cites Stochastic graphon games: II.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Stochastic graphon games: II

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.543561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:49.790981Z digest=sha256:fafed03f39462b8ca5cdc12f1ab2a4d78906d202f1508750a7ca3916ee868cef

Observation a2c3fda2-815f-492e-a65d-f853df8b6ba4 · outbound

This paper cites Berahas, Liyuan Cao, Krzysztof Choromanski, and Katya Scheinberg.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Berahas, Liyuan Cao, Krzysztof Choromanski, and Katya Scheinberg

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.535643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:49.928649Z digest=sha256:d0626a6d5edbc445f01b29469f08f6875cb07c05c6eec267409ab7c141df0ec0

Observation a0e11cf2-e17c-418f-8c40-8cd5655da34e · outbound

This paper cites An Lp theory of sparse graph convergence i: Limits, sparse random graph models, and power law distributions.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games An Lp theory of sparse graph convergence i: Limits, sparse random graph models, and power law distributions

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.527738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:50.035037Z digest=sha256:ac75ca8e66bd4cbd52ba615a5f52aaffcc4f07782fa3edb5a1afe9226b0f6899

Observation 123f7447-30c8-42d1-97e2-3bb0285ef568 · outbound

This paper cites Chayes, Henry Cohn, and Yufei Zhao.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Chayes, Henry Cohn, and Yufei Zhao

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.519563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:50.119793Z digest=sha256:a048b7fb11dc53542eff384d0b0a90c057f24df0ded8fda2da11f1f1fc1f4c0e

Observation 3542a397-2415-4f04-afee-fc70ea509587 · outbound

This paper cites Functional Analysis, Sobolev Spaces and Partial Differential Equations.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Functional Analysis, Sobolev Spaces and Partial Differential Equations

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.511383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:50.175222Z digest=sha256:c570cd4f8b9b0d893beccbd36fcb359687203ad39dda0a54847775888264ab80

Observation 268cc0c8-0e71-4c4f-8d2b-eff23765fff7 · outbound

This paper cites Policy Gradient-based Algorithms for Continuous-time Linear Quadratic Control.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Policy Gradient-based Algorithms for Continuous-time Linear Quadratic Control

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:50.242646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:50.242646Z digest=sha256:fa7289eafc80dc995b169fb8ea97de56c35c90350936cd9527bb1b9c813195e0

Observation 7175448f-8be3-4a07-9c46-f35e44628ed0 · outbound

This paper cites Caines and Minyi Huang.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Caines and Minyi Huang

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.503030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:50.336317Z digest=sha256:3f1b26947fbb5287d8677fab95b4b3fce752fadb4dbc44a260eaa4013cd0a0b2

Observation ebf7009f-a1af-48e6-a049-0a496c937c73 · outbound

This paper cites Financial Mathematics.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Financial Mathematics

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.495330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:50.443526Z digest=sha256:6c0956834b9d8c5f668412f1fb45ba573bb6a637c82069ecdfbb63c2c8832109

Observation 07a7c107-9089-4b83-9fbb-b66acd672b52 · outbound

This paper cites Cooney, Christy V.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Cooney, Christy V

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.487337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:50.539531Z digest=sha256:613c9b6d0607fe038e36a09b41751f4ba578cbce274fe242f4a9eb5a79cd0446

Observation e669e92f-5b9c-46d4-bd91-2aff43417c13 · outbound

This paper cites Deep learning for mean field games and mean field control with applications to finance.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Deep learning for mean field games and mean field control with applications to finance

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.479292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:50.619486Z digest=sha256:ea1a5611c77dc286b82d951e00a37eaeec0c3108dbefedd3fe8a572aa2e3a4f8

Observation eec5f744-7b5d-4f39-9d94-85ca914afc08 · outbound

This paper cites Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:50.732625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:50.732625Z digest=sha256:29274c8fdfad053235e700774316b9e89a732249a8e9862022990ca10018c4f7

Observation 1004d101-8bb8-45d1-be4c-5d1933abb5e1 · outbound

This paper cites Approximately solving mean field games via entropy-regularized deep reinforcement learning.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Approximately solving mean field games via entropy-regularized deep reinforcement learning

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.471147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:50.868633Z digest=sha256:75aa0e080cc58429100643f40d9578a481eb92c51d9d8ab7a74496e8633339e8

Observation 78277913-cbb5-471b-b07c-d73702318517 · outbound

This paper cites Learning graphon mean field games and approximate Nash equilibria.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Learning graphon mean field games and approximate Nash equilibria

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.463392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:51.023614Z digest=sha256:cb13f30158a2753208a68519d2e43c405cec9f85f1cab90efaea6cfc6f8e3761

Observation d9107452-ab73-4e50-975e-d078c9c77215 · outbound

This paper cites Dechevski and Lars Erik Persson.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Dechevski and Lars Erik Persson

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.455466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:51.198586Z digest=sha256:7c2629373c4b721ed431ab645a892df4405b86b86e0414f9a4f8ba246dae27f3

Observation 11d9b855-377c-420a-828b-ad632b70dc71 · outbound

This paper cites On the convergence of model free learning in mean field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games On the convergence of model free learning in mean field games

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.446518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:51.317948Z digest=sha256:7f5ff602169f8523d7aaddc4dc367e8ef092e557f9dc2db7833297ab22b115a8

Observation 50967019-6c8a-4d3c-b346-641383314127 · outbound

This paper cites Learning sparse graphon mean field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Learning sparse graphon mean field games

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.438140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:51.437786Z digest=sha256:ae3726310ebc94ac524979211051b3dc1092d811d48dbc81d9e908d4406c8cb5

Observation 5b96362f-5e25-461f-9226-f9eaae412486 · outbound

This paper cites Global convergence of policy gradient methods for the linear quadratic regulator.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Global convergence of policy gradient methods for the linear quadratic regulator

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.430093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:51.590918Z digest=sha256:e0f120c2c7980a8b592b045f49aaf130c85360a463e95c6d390773b818bea4f9

Observation 3f191497-aaaa-4169-b700-ae468df838aa · outbound

This paper cites Full error analysis of policy gradient learning algorithms for exploratory linear quadratic mean-field control problem in continuous time with common noise.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Full error analysis of policy gradient learning algorithms for exploratory linear quadratic mean-field control problem in continuous time with common noise

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:51.664862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:51.664862Z digest=sha256:35df9642a966d7eada72c5b2d81e2b76066356c3df2df4c0a2ad62b447066ae8

Observation ed498024-705a-41c7-9624-ad7968ee1dc1 · outbound

This paper cites Caines, and Minyi Huang.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Caines, and Minyi Huang

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.421695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:51.764041Z digest=sha256:165e0e0499c40ba706b055cd5f9cbbf9d2117da39a92d5795fbab14948007439

Observation 6e630b54-f70a-4e26-bdb3-8a924e38c07d · outbound

This paper cites an unresolved cited work.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:25:58.413846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:51.831285Z digest=sha256:3ac78eea279e2db5bfe2cb158395a8986c321b4c544007cdf4cdacde3b8ac193

Observation e710deb9-6db8-4643-b121-a361319ae50d · outbound

This paper cites Convergence of policy gradient methods for finite-horizon exploratory linear-quadratic control problems.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Convergence of policy gradient methods for finite-horizon exploratory linear-quadratic control problems

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.405448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:51.925261Z digest=sha256:ad8d909c685b97e7648090bcde2bc6268ff23c74566a3482212781577fc647c8

Observation 4428bf96-8c76-4a93-b366-ee0bc5c299ba · outbound

This paper cites Volterra Integral and Functional Equations.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Volterra Integral and Functional Equations

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.397411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:52.040532Z digest=sha256:de6b33c6e65abf9ce312b6b43bb9d3f6e89bbb0d2f0ff8a4354dc3754d2bffc1

Observation d7e87767-315b-4220-ae25-0eb85ec081a4 · outbound

This paper cites Learning mean-field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Learning mean-field games

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.389457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:52.121901Z digest=sha256:526fee0490e7b44490aad347571c3cbabb98d77e7536e6ea4ac4c792dd266b9a

Observation 7367bb92-5693-49f8-989d-86d41753bcb4 · outbound

This paper cites Policy gradient methods for the noisy lin- ear quadratic regulator over a finite horizon.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Policy gradient methods for the noisy lin- ear quadratic regulator over a finite horizon

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.381270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:52.207116Z digest=sha256:a08f3c2652521b6ddc06b16d2c7d4ac99e1ceb7bc0738aa07ee181f9e757b431

Observation 7b08fa9a-7303-454a-82f9-e396cc4d831c · outbound

This paper cites Policy gradient methods find the nash equilib- rium in n-player general-sum linear-quadratic games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Policy gradient methods find the nash equilib- rium in n-player general-sum linear-quadratic games

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.372977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:52.308626Z digest=sha256:24b8fd9352a7737fe1821cba228ec04d266af1806348963c6943c37f96029292

Observation 7ab5e38b-98e3-41a0-b27d-789487e3e0b5 · outbound

This paper cites MF-OML: Online Mean-Field Reinforcement Learning with Occupation Measures for Large Population Games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games MF-OML: Online Mean-Field Reinforcement Learning with Occupation Measures for Large Population Games

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:25:56.883606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:52.453170Z digest=sha256:648b09dc0c04619f0d4ad1a61e2911ab613afe0b16246cb6c6d07e45f5f53b8d

Observation f9743ba9-84a0-4d46-a4d1-769d8d2caa33 · outbound

This paper cites Model-based RL for mean-field games is not statistically harder than single-agent RL.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Model-based RL for mean-field games is not statistically harder than single-agent RL

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.364277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:52.547369Z digest=sha256:7e7b7d4c1bb9e5a5c32c4f47b0ed792ca5db40d1c8a7bbc6eebafbb068dda7fd

Observation f5a3cbd0-5d59-42d6-8dde-eb4278c8ae55 · outbound

This paper cites On the statistical efficiency of mean-field reinforcement learning with general function approximation.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games On the statistical efficiency of mean-field reinforcement learning with general function approximation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.355754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:52.656118Z digest=sha256:7dd47b9686900ef66ef4ab56ea44c5eebebbe6b5fb8cb67b0400d15c9204f207

Observation 95c6d6a1-48f8-4431-a192-d11babe7cdce · outbound

This paper cites Malham´ e, and Peter E.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Malham´ e, and Peter E

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.347534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:52.806221Z digest=sha256:7fb880368f0acdd92966588984ccdcf1292692753467f44eda7ff793bbf5a8c0

Observation d97d0940-69cf-4b85-b549-70b4643ccd2c · outbound

This paper cites Reinforcement Learning for SBM Graphon Games with Re-Sampling.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Reinforcement Learning for SBM Graphon Games with Re-Sampling

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:25:56.722759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:52.929133Z digest=sha256:20e18356d32b7740a6aff5c300963561db9a5c4e87ce2a4da9683b5f2a62d01f

Observation f49b81b1-5b32-4582-a3a9-742c8bc91cd0 · outbound

This paper cites Real-time bidding with multi-agent reinforcement learning in display advertising.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Real-time bidding with multi-agent reinforcement learning in display advertising

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.339344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:53.031245Z digest=sha256:f576122feefeafb49edc1aa89801c1660af7e711cca7ddd08c89204c8a3f609b

Observation 7ed1f384-d10b-4eb2-b5cf-2add752210e1 · outbound

This paper cites A natural policy gradient.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games A natural policy gradient

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.331341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:53.182249Z digest=sha256:cdbc664a5b927a7287bb35b7723e0327ddf66951f39b3b1992e45f7c42ab675e

Observation 6de26ac7-ccc6-42f6-924e-b4f45d1856cd · outbound

This paper cites A Fisher-Rao gradient flow for entropy-regularised Markov decision processes in Polish spaces.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games A Fisher-Rao gradient flow for entropy-regularised Markov decision processes in Polish spaces

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:53.289607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:53.289607Z digest=sha256:5edfac1a52384fc8b35b5fcb6c9b19627a0e65e0214be7227f288e3056d506f1

Observation 62c2cc9f-ef5f-495d-ac4e-0dee4c654e42 · outbound

This paper cites Actor-critic algorithms.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Actor-critic algorithms

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.323352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:53.374843Z digest=sha256:75e5992272ee23cb01b2138a59e268e8acaf200b293ee55f5a9e4f72e95f1a67

Observation 858b1adf-c7f9-4d50-8379-7af58a46ea6c · outbound

This paper cites A label-state formulation of stochastic graphon games and approximate equilibria on large networks.Mathematics of Operations Research, 48:1811–2382, 2022.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games A label-state formulation of stochastic graphon games and approximate equilibria on large networks.Mathematics of Operations Research, 48:1811–2382, 2022

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.314990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:53.459363Z digest=sha256:0958e50c8555a3e9158f3d00ec3a8db8e7b91f2fbb783ae376491b94a13e0acb

Observation 13e0aa9e-5bce-43cd-8a38-1ac86aa8ce23 · outbound

This paper cites Mean field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Mean field games

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.306776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:53.577904Z digest=sha256:cd76b30b5a74a0d481f8814c17f7d0e1872d4855e6bb93667ec387b0167f0b68

Observation a1055e9f-55a4-4063-b2be-b4f652e0f15e · outbound

This paper cites Learning in Mean Field Games: A Survey.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Learning in Mean Field Games: A Survey

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:53.713354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:53.713354Z digest=sha256:4a767af213d159e9b29aab4606a45c070b5f6792f448d1f0eb1eaf85f96ecb9c

Observation acc9d4f3-cd17-48f8-8725-88b75e370f2f · outbound

This paper cites American Mathematical Society Col- loquium Publications.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games American Mathematical Society Col- loquium Publications

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.296721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:53.816578Z digest=sha256:aca59f4208dc044ba2067e13c660f4481169b15fce794d89a07492077d27afef

Observation 1efd3b1b-a6e1-41c3-b455-ca7b01153109 · outbound

This paper cites On the global con- vergence rates of softmax policy gradient methods.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games On the global con- vergence rates of softmax policy gradient methods

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.288192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:53.946606Z digest=sha256:3cb99c1f32ba2c200b3a425c4b9d2975f90653d9fd41af9d68ece0855b00bfd8

Observation 21362dd4-02e6-4ad6-a1d4-04b54f59df70 · outbound

This paper cites Stochastic Graphon Games with Memory.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Stochastic Graphon Games with Memory

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:54.024854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:54.024854Z digest=sha256:8fb06253598b236a926fa40925fc4002dc5a13d0262110d09e523e5cac5757f1

Observation 4bb99848-22c8-4cac-90d6-99b885720120 · outbound

This paper cites Dota 2 with Large Scale Deep Reinforcement Learning.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Dota 2 with Large Scale Deep Reinforcement Learning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:54.194986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:54.194986Z digest=sha256:bd052d4c5047df0f80da3a89a5c2a654131aa2e6ea0bc76aa4324adcccabc05b

Observation c916ef8f-1cbe-44b6-98c1-b40aead4983d · outbound

This paper cites Graphon games: A statistical framework for network games and interventions.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Graphon games: A statistical framework for network games and interventions

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.280044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:54.367333Z digest=sha256:f7ca57767c4e45907ea89969478b173966197927b595e0655233f0c68bfc4af4

Observation 06ae224f-521a-4a7a-af7f-51273bf7b88e · outbound

This paper cites Time discretization-invariant safe action re- petition for policy gradient methods.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Time discretization-invariant safe action re- petition for policy gradient methods

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.271175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:54.597645Z digest=sha256:ec0e57712b8ca1c668afb146de75c55342508f0f0a49c8fc50865c2c9517a1a6

Observation 8086cd2c-b1bb-41a8-b175-f2c040dfb26c · outbound

This paper cites Entropy annealing for policy mirror descent in continuous time and space.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Entropy annealing for policy mirror descent in continuous time and space

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:54.770111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:54.770111Z digest=sha256:c936cd027ce8264e8da98bd9f62791d81cd10b9c1aec9bf81c70e31439b22865

Observation e8ae0642-33bf-4a24-be32-f7021115328a · outbound

This paper cites Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:54.893054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:54.893054Z digest=sha256:f2df359e483023429172a89e9ebb2771fcc0fb333ae80fffc2e066385c2fb00c

Observation 12d16d47-64f4-4f3d-be60-99f0e6d88012 · outbound

This paper cites an unresolved cited work.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:25:58.262557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:54.993284Z digest=sha256:71926e4732ff9c7e96f884eb35f3e6414fac258eafa1e91a83703aef03c5cd48

Observation 38e1ed21-e657-4664-abba-9b397c27886b · outbound

This paper cites Deterministic policy gradient algorithms.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Deterministic policy gradient algorithms

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.254608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:55.118458Z digest=sha256:0f5b6525c72a0e29917e55d0d7ff3b93e9a306ed64c79df20e1d4051bb271644

Observation 3c8ff049-5a9f-4a34-b805-ee9926f1a9fe · outbound

This paper cites Sutton and Andrew G.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Sutton and Andrew G

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.246089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:55.217096Z digest=sha256:c31d0c129d3fad92bf24dda5519c059ce5455a68815216c10b76169d5bd14cd8

Observation 67698907-189d-49f5-9b7e-8f15530d809c · outbound

This paper cites Policy gradient methods for reinforcement learning with function approximation.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Policy gradient methods for reinforcement learning with function approximation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.237861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:55.288745Z digest=sha256:669c3e887f9ab25bcea73e09a9fd93a175337624acde562dbbf1a9c484f2f1fb

Observation 83bd1352-65d0-40a7-8246-706c1de4042e · outbound

This paper cites Caines, and Minyi Huang.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Caines, and Minyi Huang

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.229525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:55.380324Z digest=sha256:c2fa7b8de41e1874a132a8302b1f6f6e2b37396e5bc878986f72be23b6a06dec

Observation 72bcac30-3a04-4b6b-8cae-edf56c682601 · outbound

This paper cites Ordinary Differential Equations and Dynamical Systems, volume 140 ofGradu- ate Studies in Mathematics.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Ordinary Differential Equations and Dynamical Systems, volume 140 ofGradu- ate Studies in Mathematics

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.220709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:55.502383Z digest=sha256:219cb9f35d6830bbd0678b1cc50c9312c75ab3662f09e2f06a23c57543f9d7c5

Observation 76ebe283-1d2d-43dd-afb8-5bd04ea4373a · outbound

This paper cites Global convergence of policy gradient for linear-quadratic mean-field control/game in continuous time.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Global convergence of policy gradient for linear-quadratic mean-field control/game in continuous time

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.211525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:55.600933Z digest=sha256:6d3b8d12950677a2f9a5120689253e272e2ed36e2d6cb3e57b50d58e2c8e2a1b

Observation 24536580-ed8b-408c-ad84-9579b75df894 · outbound

This paper cites Learning while playing in mean-field games: Convergence and optimality.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Learning while playing in mean-field games: Convergence and optimality

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.202300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:55.699378Z digest=sha256:5608ddd214253818933f39499ca5dbb4147e33e2eca96ec9bf92320cc002630d

Observation 20296ccf-e8b7-4bc3-9617-ba463e04787a · outbound

This paper cites Linear-Quadratic Graphon Mean Field Games with Common Noise.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Linear-Quadratic Graphon Mean Field Games with Common Noise

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:55.825786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:55.825786Z digest=sha256:65bc5fca9b29d5d961bf2e484db3b8b73479eab0a12b3acdbc824551a0ff3320

Observation 4b5344fa-0afc-4adc-a7d2-c3c2938c86ad · outbound

This paper cites Policy mirror ascent for efficient and independent learning in mean field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Policy mirror ascent for efficient and independent learning in mean field games

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.194028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:55.922529Z digest=sha256:3d1bb01c041dee01f7ba8e5a94cf8a3157233bfc9a595f6efc2b04d47c0889dd

Observation ff24be89-593a-455b-8716-f8dcd5f494b8 · outbound

This paper cites When is mean-field reinforcement learning tractable and relevant? In Proceedings of the 23rd International Conference on Autonomous Agents and Multiagent Systems , page 2038–2046.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games When is mean-field reinforcement learning tractable and relevant? In Proceedings of the 23rd International Conference on Autonomous Agents and Multiagent Systems , page 2038–2046

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.184842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:55.991887Z digest=sha256:0f945f7aec394c397dc713b3e1d0a8bef95a18f096c17b2121df9b7941f40dd3

Observation c9ce61b0-55e5-4ec9-8000-0390774079cd · outbound

This paper cites Stochastic Controls: Hamiltonian Systems and HJB Equations.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Stochastic Controls: Hamiltonian Systems and HJB Equations

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.041054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:56.087561Z digest=sha256:0658a4d6fff963708c15962642e3c08f8ed60736890171ec43bf6f54c176d688

Observation aeac5648-6171-4005-80f5-15a06d9647db · outbound

This paper cites Learning regularized monotone graphon mean-field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Learning regularized monotone graphon mean-field games

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:57.702449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:56.207569Z digest=sha256:53eebb694355f562c94b259fba5086a01df8322c8b36b2599e815327e1f563cc

Observation 63aba5c9-8a33-473d-ae08-5410e8861d43 · outbound

This paper cites an unresolved cited work.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:25:57.444249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:56.268960Z digest=sha256:7d7d263a1a8aa1cf19eae2a03081258caaff0f118ead28c6103e0cda123e4deb

Observation bef945bd-b82f-4aea-b744-4dc9ce4270aa · outbound

This paper cites Convergence of policy gradient for stochastic linear quad- ratic optimal control problems in infinite horizon.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Convergence of policy gradient for stochastic linear quad- ratic optimal control problems in infinite horizon

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:57.137627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:56.358470Z digest=sha256:f77d74eb99d0e9a06c5b8b0b5cfd98fc186f4a67e5143af450304533cd271b6c

Observation dcf4c283-7e19-4b07-8b49-44d3999aefae · outbound

This paper cites Graphon mean field games with a representative player: Analysis and learning algorithm.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Graphon mean field games with a representative player: Analysis and learning algorithm

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:57.055423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T10:25:56.452058Z digest=sha256:34aaf8754b6b9644c2e855176abfe4cc7544fead08079de128268f46777cad57

Pith citing papers

No inbound Pith citation observations are available.