Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-20T22:13:00.107205Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 2 inbound Pith citation observations for arXiv:2605.18809.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-20T22:13:00.107205Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T00:45:45.078825Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T00:45:45.612357Z
20 of 20 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 431d1d1c-af84-48bb-8dd3-dc82d76142c2 · outbound
Metric-Gradient Projection for Stable Multi-Agent Policy Learning Markov games as a framework for multi-agent reinforcement learning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5f6b9916-7d09-45dc-8f47-b98d89f3e164 · outbound
Metric-Gradient Projection for Stable Multi-Agent Policy Learning Learning to collaborate with unknown agents in the absence of reward
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cc9f3ea4-086c-4107-b949-0480238b3f36 · outbound
Metric-Gradient Projection for Stable Multi-Agent Policy Learning Modeling Other Players with Bayesian Beliefs for Games with Incomplete Information
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cf33e756-3e32-44b9-aa43-1f5a4befbeda · outbound
Metric-Gradient Projection for Stable Multi-Agent Policy Learning Lipschitz Lifelong Monte Carlo Tree Search for Mastering Non-Stationary Tasks
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7a72cee9-0251-4351-a65b-890e3433b57d · outbound
Metric-Gradient Projection for Stable Multi-Agent Policy Learning Geometry of drifting mdps with path-integral stability certificates
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d1cf2afa-db78-4c34-b2ee-0f2738b3b7f6 · outbound
Metric-Gradient Projection for Stable Multi-Agent Policy Learning High-Dimensional Continuous Control Using Generalized Advantage Estimation
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 71da555e-6cdf-4216-9200-4bdc9f925f19 · outbound
Metric-Gradient Projection for Stable Multi-Agent Policy Learning Value-Decomposition Networks For Cooperative Multi-Agent Learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d59117f1-7687-457b-a8a9-845e873dae32 · outbound
Metric-Gradient Projection for Stable Multi-Agent Policy Learning A Variational Inequality Perspective on Generative Adversarial Networks
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 00259185-6ade-4ca5-a807-c3042e6f934a · outbound
Metric-Gradient Projection for Stable Multi-Agent Policy Learning Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 62bb4fd6-ead0-4037-a239-64bbf3ad5d45 · outbound
Metric-Gradient Projection for Stable Multi-Agent Policy Learning neurips.cc/paper/2005/hash/9752d873fa71c19dc602bf2a0696f9b5-Abstract.html
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7fe04ec0-4364-45d5-9891-d4b31e0832d8 · outbound
Metric-Gradient Projection for Stable Multi-Agent Policy Learning Optimistic mirror descent in saddle-point problems: Going the extra (gradient) mile
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 76497578-d2c3-4c27-9e54-8c5b78121797 · outbound
Metric-Gradient Projection for Stable Multi-Agent Policy Learning Training GANs with Optimism
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c1475e4b-3f28-4bc2-bfe8-f5199a3f1f71 · outbound
Metric-Gradient Projection for Stable Multi-Agent Policy Learning Algorithms, graph theory, and linear equations in laplacian matrices
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8ba9d677-3c5b-4b4e-9b5d-219586756a36 · outbound
Metric-Gradient Projection for Stable Multi-Agent Policy Learning Discrete calculus: Applied analysis on graphs for computational science, leo grady, jonathan polimeni, springer (2010), $129.00, isbn: 978-1-84996-289-6
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d321031c-eb48-4861-b417-ee43ed033c99 · outbound
Metric-Gradient Projection for Stable Multi-Agent Policy Learning Cochain perspectives on temporal-difference signals for learning beyond markov dynamics.arXiv preprint arXiv:2602.06939, 2026b
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d6bf5b80-299b-4bec-804d-3ff3c5d4d63a · outbound
Metric-Gradient Projection for Stable Multi-Agent Policy Learning Melting Pot 2.0
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8be89070-a4b1-43b4-9bfb-a0c50b7d81b2 · outbound
Metric-Gradient Projection for Stable Multi-Agent Policy Learning Agent alpha: Tree search unifying generation, exploration and evaluation for computer-use agents.arXiv preprint arXiv:2602.02995
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bf6547c0-7d3a-4068-8e95-5c9bc24a9ec3 · outbound
Metric-Gradient Projection for Stable Multi-Agent Policy Learning NonZero: Interaction-Guided Exploration for Multi-Agent Monte Carlo Tree Search
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7fe68f37-42e7-4e88-93bd-679eaf6c85e7 · outbound
Metric-Gradient Projection for Stable Multi-Agent Policy Learning Lisfc-search: Lifelong search for network sfc optimization under non-stationary drifts.arXiv preprint arXiv:2602.14360, 2026a
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 24334fc3-f4fa-41d3-9a93-1f7445636fe1 · outbound
Metric-Gradient Projection for Stable Multi-Agent Policy Learning First, cyclic matrix games, including Rock–Paper–Scissors and generalized cyclic games on∆K ×∆ K, expose cyclic interaction dynamics in a familiar simplex geometry
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 32377c4b-0ede-4256-9a68-3ab688476fde · inbound
FlowEdit: Information-Theoretic Control of LLM Reasoning Flows for Ill-posed Problems Involving Conflicts Metric-Gradient Projection for Stable Multi-Agent Policy Learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41ada73a-9ebb-4a48-bad2-b28bc9ef4772 · inbound
Learning Not to Optimize: Physics-Informed Action-Space Reshaping for Intent-Based Network Control Metric-Gradient Projection for Stable Multi-Agent Policy Learning
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.