Pith. sign in

Paper Citation Record · LEDGER

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games

As of 16 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2506.05894.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.05894 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:25:56.452058Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

63 of 63 outbound references displayed

  • verified exact2
  • verified fuzzy48
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 785a7eff-dc7e-424c-afed-7efebc62578f · outbound

This paper cites Value iteration algorithm for mean- field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Value iteration algorithm for mean- field games

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.559605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:49.554219Z digest=sha256:259b1fe6d8bf483bcd5f880b3b51464211d8fd813a16e6d11a8fde34a2a4ea10

Observation 7db36181-eddf-4575-8802-3f5cdfca4e99 · outbound

This paper cites Unified reinforcement q-learning for mean field game and control problems.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Unified reinforcement q-learning for mean field game and control problems

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.551580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:49.639583Z digest=sha256:1895a65a183b5f64e9c234156974b7ff389c56dd9b731c93f3157c6c06dfb039

Observation efcfd009-787b-483b-b5fe-2e6069da6fb5 · outbound

This paper cites Stochastic graphon games: II.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Stochastic graphon games: II

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.543561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:49.790981Z digest=sha256:2f17097d855283c311ec1b55e0f168b7a784e2052c23b59bda17db3412eb26a5

Observation a2c3fda2-815f-492e-a65d-f853df8b6ba4 · outbound

This paper cites Berahas, Liyuan Cao, Krzysztof Choromanski, and Katya Scheinberg.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Berahas, Liyuan Cao, Krzysztof Choromanski, and Katya Scheinberg

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.535643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:49.928649Z digest=sha256:d1ef4c69125bab70e8827f9dec030e7e665db3f9eb216a4c99de58c21da5774e

Observation a0e11cf2-e17c-418f-8c40-8cd5655da34e · outbound

This paper cites An Lp theory of sparse graph convergence i: Limits, sparse random graph models, and power law distributions.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games An Lp theory of sparse graph convergence i: Limits, sparse random graph models, and power law distributions

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.527738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:50.035037Z digest=sha256:f7c3da9dbd7d34ca0b17436cbbf0254bad0ea895d175393fb38b9b932cc928b3

Observation 123f7447-30c8-42d1-97e2-3bb0285ef568 · outbound

This paper cites Chayes, Henry Cohn, and Yufei Zhao.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Chayes, Henry Cohn, and Yufei Zhao

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.519563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:50.119793Z digest=sha256:19ef25d3510dffd8f086550b89315b95e9720139e958799be81dee7687605265

Observation 3542a397-2415-4f04-afee-fc70ea509587 · outbound

This paper cites Functional Analysis, Sobolev Spaces and Partial Differential Equations.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Functional Analysis, Sobolev Spaces and Partial Differential Equations

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.511383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:50.175222Z digest=sha256:d73bba88913caae45d7f3d1c6f00210b2fb7db0da141f1965a01bfe643b255b6

Observation 268cc0c8-0e71-4c4f-8d2b-eff23765fff7 · outbound

This paper cites Policy Gradient-based Algorithms for Continuous-time Linear Quadratic Control.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Policy Gradient-based Algorithms for Continuous-time Linear Quadratic Control

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:50.242646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:50.242646Z digest=sha256:621a7c615a0b05292ebbd1e6297f0cdf9b391db3ad9ae5ab7693435a9bf07ef5

Observation 7175448f-8be3-4a07-9c46-f35e44628ed0 · outbound

This paper cites Caines and Minyi Huang.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Caines and Minyi Huang

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.503030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:50.336317Z digest=sha256:1be3c6ac39af0e9aef78f05bc18e4144cbd2618e32674ef896aa48fb9ceb05d9

Observation ebf7009f-a1af-48e6-a049-0a496c937c73 · outbound

This paper cites Financial Mathematics.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Financial Mathematics

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.495330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:50.443526Z digest=sha256:d0fee0edff6556e4ede864f364efa895f1140746d5249565c0434439b6660ee9

Observation 07a7c107-9089-4b83-9fbb-b66acd672b52 · outbound

This paper cites Cooney, Christy V.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Cooney, Christy V

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.487337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:50.539531Z digest=sha256:316083842fc9e060a3909629bd246b72b3641c6b8166c644c2e2555b6d930c8c

Observation e669e92f-5b9c-46d4-bd91-2aff43417c13 · outbound

This paper cites Deep learning for mean field games and mean field control with applications to finance.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Deep learning for mean field games and mean field control with applications to finance

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.479292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:50.619486Z digest=sha256:38bd38f2f7fe83e2abf5e8be69e144a64271758eead2b1d5cb76854166ec22d2

Observation eec5f744-7b5d-4f39-9d94-85ca914afc08 · outbound

This paper cites Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:50.732625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:50.732625Z digest=sha256:fcf68edf13223f22f988739404b5570311dc0b2d73e278bff9d71a772be84f54

Observation 1004d101-8bb8-45d1-be4c-5d1933abb5e1 · outbound

This paper cites Approximately solving mean field games via entropy-regularized deep reinforcement learning.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Approximately solving mean field games via entropy-regularized deep reinforcement learning

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.471147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:50.868633Z digest=sha256:bd45c94b2262937c4bdfc1f0c075c36fa92ba54f17f2cb651d3145ab52efefa6

Observation 78277913-cbb5-471b-b07c-d73702318517 · outbound

This paper cites Learning graphon mean field games and approximate Nash equilibria.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Learning graphon mean field games and approximate Nash equilibria

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.463392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:51.023614Z digest=sha256:72e32b4c88d1722244a89295d3492dd16a3a268daa17b7881c70fb9bc512980e

Observation d9107452-ab73-4e50-975e-d078c9c77215 · outbound

This paper cites Dechevski and Lars Erik Persson.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Dechevski and Lars Erik Persson

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.455466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:51.198586Z digest=sha256:6a2debee66b8e60015206bdb26125d7655faf7c9e81cd24654b8e854d36cb171

Observation 11d9b855-377c-420a-828b-ad632b70dc71 · outbound

This paper cites On the convergence of model free learning in mean field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games On the convergence of model free learning in mean field games

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.446518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:51.317948Z digest=sha256:a17a8c65964ddecd3375f802a4282af86ab2abd40cc960873ef2fd6206e1a84a

Observation 50967019-6c8a-4d3c-b346-641383314127 · outbound

This paper cites Learning sparse graphon mean field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Learning sparse graphon mean field games

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.438140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:51.437786Z digest=sha256:cb5b7a5ffb6414b88f264d9e5a99fc272968421cb5190488d1731e2caaf01eba

Observation 5b96362f-5e25-461f-9226-f9eaae412486 · outbound

This paper cites Global convergence of policy gradient methods for the linear quadratic regulator.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Global convergence of policy gradient methods for the linear quadratic regulator

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.430093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:51.590918Z digest=sha256:48bcaa38fdd5557f6b6c141abc93c34514a6937451f5cf33fc583b4feec8fecb

Observation 3f191497-aaaa-4169-b700-ae468df838aa · outbound

This paper cites Full error analysis of policy gradient learning algorithms for exploratory linear quadratic mean-field control problem in continuous time with common noise.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Full error analysis of policy gradient learning algorithms for exploratory linear quadratic mean-field control problem in continuous time with common noise

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:51.664862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:51.664862Z digest=sha256:b69ed866c5c1f20e38b9e3009e0bbd29a6ad437b68fd789c7edca4d041f9bac4

Observation ed498024-705a-41c7-9624-ad7968ee1dc1 · outbound

This paper cites Caines, and Minyi Huang.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Caines, and Minyi Huang

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.421695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:51.764041Z digest=sha256:8306e53408f41dffa9c8aff936d4f1cf5d5d418965a25567903bf3894331234c

Observation 6e630b54-f70a-4e26-bdb3-8a924e38c07d · outbound

This paper cites an unresolved cited work.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:25:58.413846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:51.831285Z digest=sha256:f5e47be39e6e8ee5d8ed277528cfdb43c11f03104b4eeb00dacbc845ce380072

Observation e710deb9-6db8-4643-b121-a361319ae50d · outbound

This paper cites Convergence of policy gradient methods for finite-horizon exploratory linear-quadratic control problems.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Convergence of policy gradient methods for finite-horizon exploratory linear-quadratic control problems

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.405448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:51.925261Z digest=sha256:1ec49dac4095c222704ec4f9dc0778aa4fc27ab7fe4b52189d819710cf3eb687

Observation 4428bf96-8c76-4a93-b366-ee0bc5c299ba · outbound

This paper cites Volterra Integral and Functional Equations.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Volterra Integral and Functional Equations

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.397411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:52.040532Z digest=sha256:a9b885cdea2b68e60418ac957d1b6ececb6bcc67f36d1f8451d3b9876e1f3767

Observation d7e87767-315b-4220-ae25-0eb85ec081a4 · outbound

This paper cites Learning mean-field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Learning mean-field games

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.389457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:52.121901Z digest=sha256:1cfca2e001b2375b67469c904dd0c7936bc935751a834f9370a571b58860a272

Observation 7367bb92-5693-49f8-989d-86d41753bcb4 · outbound

This paper cites Policy gradient methods for the noisy lin- ear quadratic regulator over a finite horizon.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Policy gradient methods for the noisy lin- ear quadratic regulator over a finite horizon

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.381270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:52.207116Z digest=sha256:9470f8e803c92ea36ad42cba0e6282ac9041d773d4a60153aaa864903ec7b2d7

Observation 7b08fa9a-7303-454a-82f9-e396cc4d831c · outbound

This paper cites Policy gradient methods find the nash equilib- rium in n-player general-sum linear-quadratic games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Policy gradient methods find the nash equilib- rium in n-player general-sum linear-quadratic games

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.372977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:52.308626Z digest=sha256:0d728c1676733f82936fdcea5131062275e0f3f1b8ec70a7ba67eb78eab8e417

Observation 7ab5e38b-98e3-41a0-b27d-789487e3e0b5 · outbound

This paper cites MF-OML: Online Mean-Field Reinforcement Learning with Occupation Measures for Large Population Games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games MF-OML: Online Mean-Field Reinforcement Learning with Occupation Measures for Large Population Games

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:25:56.883606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:52.453170Z digest=sha256:448c3941c7eb30795db30a45638a47345a12eb39272852998242fc265e336db5

Observation f9743ba9-84a0-4d46-a4d1-769d8d2caa33 · outbound

This paper cites Model-based RL for mean-field games is not statistically harder than single-agent RL.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Model-based RL for mean-field games is not statistically harder than single-agent RL

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.364277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:52.547369Z digest=sha256:4fbc0de297b60d4fcd55307b1e78ed0bf648d071b4f8d6ae8f6b43c6925c2a60

Observation f5a3cbd0-5d59-42d6-8dde-eb4278c8ae55 · outbound

This paper cites On the statistical efficiency of mean-field reinforcement learning with general function approximation.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games On the statistical efficiency of mean-field reinforcement learning with general function approximation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.355754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:52.656118Z digest=sha256:dd385f0e17db2d59e143255f2c790d6b331d99349f23a2f172808571c1d4ccf3

Observation 95c6d6a1-48f8-4431-a192-d11babe7cdce · outbound

This paper cites Malham´ e, and Peter E.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Malham´ e, and Peter E

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.347534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:52.806221Z digest=sha256:32e9fe1966697385f99b672779b978cb0b70c6b845f7d519979f33a79465a448

Observation d97d0940-69cf-4b85-b549-70b4643ccd2c · outbound

This paper cites Reinforcement Learning for SBM Graphon Games with Re-Sampling.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Reinforcement Learning for SBM Graphon Games with Re-Sampling

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:25:56.722759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:52.929133Z digest=sha256:9d36afca22f7af56650cbcfd9b0e1a6654c35bd72ddfe0ad70574e3eb151ae88

Observation f49b81b1-5b32-4582-a3a9-742c8bc91cd0 · outbound

This paper cites Real-time bidding with multi-agent reinforcement learning in display advertising.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Real-time bidding with multi-agent reinforcement learning in display advertising

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.339344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:53.031245Z digest=sha256:e31ab3bb41abcd37a35deb489735730f6da8aa9113a5b07b918e023b29b99521

Observation 7ed1f384-d10b-4eb2-b5cf-2add752210e1 · outbound

This paper cites A natural policy gradient.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games A natural policy gradient

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.331341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:53.182249Z digest=sha256:312f4414c82350e3e1eb99df60300590b7a8a7e55c9df16161f7e8f7cf0ddcb8

Observation 6de26ac7-ccc6-42f6-924e-b4f45d1856cd · outbound

This paper cites A Fisher-Rao gradient flow for entropy-regularised Markov decision processes in Polish spaces.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games A Fisher-Rao gradient flow for entropy-regularised Markov decision processes in Polish spaces

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:53.289607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:53.289607Z digest=sha256:d0a348d7aa6752640b9d165bb11bda0f675a24a1e44f529b29b89239c2d1533b

Observation 62c2cc9f-ef5f-495d-ac4e-0dee4c654e42 · outbound

This paper cites Actor-critic algorithms.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Actor-critic algorithms

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.323352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:53.374843Z digest=sha256:8052a33878f9cb3dfdcafcec317d4378b60dcb3252f91f97ca988e60fcb3854b

Observation 858b1adf-c7f9-4d50-8379-7af58a46ea6c · outbound

This paper cites A label-state formulation of stochastic graphon games and approximate equilibria on large networks.Mathematics of Operations Research, 48:1811–2382, 2022.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games A label-state formulation of stochastic graphon games and approximate equilibria on large networks.Mathematics of Operations Research, 48:1811–2382, 2022

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.314990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:53.459363Z digest=sha256:f83b8046ee6c51cf033ebbae01183dd660a6b9fb6fbf975346d9569867a52865

Observation 13e0aa9e-5bce-43cd-8a38-1ac86aa8ce23 · outbound

This paper cites Mean field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Mean field games

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.306776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:53.577904Z digest=sha256:06bc3673817abdf69312d7b9605af707e883e82b4fd98f7fa7ad0973cf73ec04

Observation a1055e9f-55a4-4063-b2be-b4f652e0f15e · outbound

This paper cites Learning in Mean Field Games: A Survey.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Learning in Mean Field Games: A Survey

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:53.713354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:53.713354Z digest=sha256:b727efccf824ce9ab0c278149307524bf851df09b7c40d5f76c1056c0d4eb0d9

Observation acc9d4f3-cd17-48f8-8725-88b75e370f2f · outbound

This paper cites American Mathematical Society Col- loquium Publications.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games American Mathematical Society Col- loquium Publications

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.296721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:53.816578Z digest=sha256:17f6033b1f65636191d06714c56766a10632954c42b91e78756aff074ccd0a1f

Observation 1efd3b1b-a6e1-41c3-b455-ca7b01153109 · outbound

This paper cites On the global con- vergence rates of softmax policy gradient methods.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games On the global con- vergence rates of softmax policy gradient methods

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.288192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:53.946606Z digest=sha256:6b0ebb8625e6a31440b0412df6a027f03fe241a759c45761e1ff15c73046ee6d

Observation 21362dd4-02e6-4ad6-a1d4-04b54f59df70 · outbound

This paper cites Stochastic Graphon Games with Memory.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Stochastic Graphon Games with Memory

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:54.024854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:54.024854Z digest=sha256:6af8e521e95862d7cfed39dae8f6e9ce03fae3652034e53e54c726a4fcaa28e5

Observation 4bb99848-22c8-4cac-90d6-99b885720120 · outbound

This paper cites Dota 2 with Large Scale Deep Reinforcement Learning.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Dota 2 with Large Scale Deep Reinforcement Learning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:54.194986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:54.194986Z digest=sha256:6bb59e849b0940f21607bb7bf035d5b409fe1823f99443f83d2c1d977196f577

Observation c916ef8f-1cbe-44b6-98c1-b40aead4983d · outbound

This paper cites Graphon games: A statistical framework for network games and interventions.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Graphon games: A statistical framework for network games and interventions

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.280044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:54.367333Z digest=sha256:362122be142790782425ca728b48ae524ac6113a92baa93abf0811b879dd7fd0

Observation 06ae224f-521a-4a7a-af7f-51273bf7b88e · outbound

This paper cites Time discretization-invariant safe action re- petition for policy gradient methods.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Time discretization-invariant safe action re- petition for policy gradient methods

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.271175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:54.597645Z digest=sha256:4af591329ff83dfbe1ee99c0650925a6cc1d1ce052a36eb1c54576d02cd3f361

Observation 8086cd2c-b1bb-41a8-b175-f2c040dfb26c · outbound

This paper cites Entropy annealing for policy mirror descent in continuous time and space.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Entropy annealing for policy mirror descent in continuous time and space

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:54.770111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:54.770111Z digest=sha256:8f21bcc0c35fbdbe37d684038113c9827820d9d6fc3acc77a77903c91198925d

Observation e8ae0642-33bf-4a24-be32-f7021115328a · outbound

This paper cites Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:54.893054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:54.893054Z digest=sha256:ea9ce6a06ca825baf3c0d6363b354e19c00deddc30a4ec02ab870cb2e1605b63

Observation 12d16d47-64f4-4f3d-be60-99f0e6d88012 · outbound

This paper cites an unresolved cited work.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:25:58.262557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:54.993284Z digest=sha256:b936090178ef3a46583a1b62e3800306f00ecfb2d74dd827c44f3f517d19e7ec

Observation 38e1ed21-e657-4664-abba-9b397c27886b · outbound

This paper cites Deterministic policy gradient algorithms.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Deterministic policy gradient algorithms

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.254608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:55.118458Z digest=sha256:8401849fa3db124afdfb9373f601c88134ca848389c26dd416fe09dbd27118f1

Observation 3c8ff049-5a9f-4a34-b805-ee9926f1a9fe · outbound

This paper cites Sutton and Andrew G.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Sutton and Andrew G

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.246089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:55.217096Z digest=sha256:d8bd1cdc3dc218a61555f8c8b8a48dc2847cf6b2ae34408ec8eea2d6d154e4f9

Observation 67698907-189d-49f5-9b7e-8f15530d809c · outbound

This paper cites Policy gradient methods for reinforcement learning with function approximation.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Policy gradient methods for reinforcement learning with function approximation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.237861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:55.288745Z digest=sha256:a842cb8185ecb4006f97d86b0c0549bc9661f028b507d4ac9657f16af5ec4ffe

Observation 83bd1352-65d0-40a7-8246-706c1de4042e · outbound

This paper cites Caines, and Minyi Huang.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Caines, and Minyi Huang

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.229525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:55.380324Z digest=sha256:e9bb435bca285bbaae25f775b48b45c7ccad212566dd9b93b86f2252cae42c46

Observation 72bcac30-3a04-4b6b-8cae-edf56c682601 · outbound

This paper cites Ordinary Differential Equations and Dynamical Systems, volume 140 ofGradu- ate Studies in Mathematics.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Ordinary Differential Equations and Dynamical Systems, volume 140 ofGradu- ate Studies in Mathematics

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.220709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:55.502383Z digest=sha256:e05eb38a9aafc39fac596a744a7c7cd9eddbd85b11b09aefd4d7d41176ea3737

Observation 76ebe283-1d2d-43dd-afb8-5bd04ea4373a · outbound

This paper cites Global convergence of policy gradient for linear-quadratic mean-field control/game in continuous time.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Global convergence of policy gradient for linear-quadratic mean-field control/game in continuous time

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.211525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:55.600933Z digest=sha256:7af9e05cdc0505f730fc7d750de352e8acd21123285cfe86afabca8b06bb5fda

Observation 24536580-ed8b-408c-ad84-9579b75df894 · outbound

This paper cites Learning while playing in mean-field games: Convergence and optimality.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Learning while playing in mean-field games: Convergence and optimality

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.202300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:55.699378Z digest=sha256:fd0d3d05a9ee1647d5f8129944e056731d8a8bf7220d05f5377bc2cbd4574ac9

Observation 20296ccf-e8b7-4bc3-9617-ba463e04787a · outbound

This paper cites Linear-Quadratic Graphon Mean Field Games with Common Noise.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Linear-Quadratic Graphon Mean Field Games with Common Noise

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:55.825786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:55.825786Z digest=sha256:1dae14698c3bb20b72b58cd4366b37c4d92cd4cf982277383251f351dc2d5128

Observation 4b5344fa-0afc-4adc-a7d2-c3c2938c86ad · outbound

This paper cites Policy mirror ascent for efficient and independent learning in mean field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Policy mirror ascent for efficient and independent learning in mean field games

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.194028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:55.922529Z digest=sha256:f5940f1571f503f58aee0442036e816d9b2055d797490c0d284e0a4f2e72a59b

Observation ff24be89-593a-455b-8716-f8dcd5f494b8 · outbound

This paper cites When is mean-field reinforcement learning tractable and relevant? In Proceedings of the 23rd International Conference on Autonomous Agents and Multiagent Systems , page 2038–2046.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games When is mean-field reinforcement learning tractable and relevant? In Proceedings of the 23rd International Conference on Autonomous Agents and Multiagent Systems , page 2038–2046

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.184842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:55.991887Z digest=sha256:00f72e2b3756787d04b93a55d1da3933095108cd4a22b000afb526aef406c89d

Observation c9ce61b0-55e5-4ec9-8000-0390774079cd · outbound

This paper cites Stochastic Controls: Hamiltonian Systems and HJB Equations.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Stochastic Controls: Hamiltonian Systems and HJB Equations

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.041054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:56.087561Z digest=sha256:2b508c9cdbd9daa9fef3afc64dc968917ac679a32fdccae290959befd4863e50

Observation aeac5648-6171-4005-80f5-15a06d9647db · outbound

This paper cites Learning regularized monotone graphon mean-field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Learning regularized monotone graphon mean-field games

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:57.702449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:56.207569Z digest=sha256:9b09d2f3fdf1add69d331485130c74ce2d218e96538372eeaf80a3b5b7cbeaca

Observation 63aba5c9-8a33-473d-ae08-5410e8861d43 · outbound

This paper cites an unresolved cited work.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:25:57.444249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:56.268960Z digest=sha256:29a1beba67effcf81b019e6978656a13eda2323194f041d57ae8e42b74969116

Observation bef945bd-b82f-4aea-b744-4dc9ce4270aa · outbound

This paper cites Convergence of policy gradient for stochastic linear quad- ratic optimal control problems in infinite horizon.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Convergence of policy gradient for stochastic linear quad- ratic optimal control problems in infinite horizon

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:57.137627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:56.358470Z digest=sha256:f0fff6799307c6abd65ebcc3308fc0e0b295bc1ab1fb7beaa406c13d3bb24fbd

Observation dcf4c283-7e19-4b07-8b49-44d3999aefae · outbound

This paper cites Graphon mean field games with a representative player: Analysis and learning algorithm.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Graphon mean field games with a representative player: Analysis and learning algorithm

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:57.055423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T10:25:56.452058Z digest=sha256:5b0182255b471a249ca466d9ebb8de57cfe24390c64cb50ad8ec100f5a61b3a6

Pith citing papers

No inbound Pith citation observations are available.