Pith. sign in

Paper Citation Record · LEDGER

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization

As of 17 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 0 inbound Pith citation observations for arXiv:2607.17924.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.17924 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T16:42:04.090374Z

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation dfb80bfe-04ff-4c44-bc48-6d234cb14935 · outbound

This paper cites Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:02.462350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:02.462350Z digest=sha256:ca134b50a3507608424c3d8d641a28aad09bb4fdf4c193722dd8f2829c7a6945

Observation e86fa2a0-4660-40d3-8c4f-a290712c33d6 · outbound

This paper cites Bridging MARL to SARL: An Order-Independent Multi-Agent Transformer via Latent Consensus.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization Bridging MARL to SARL: An Order-Independent Multi-Agent Transformer via Latent Consensus

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:03.341501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:03.341501Z digest=sha256:30062d2554f12b728fc3379b734380dfc6f8d37db8d36340c6ca594bcac5c8f3

Observation 7f9ffecb-a468-438e-b0bc-786abaf1f2c7 · outbound

This paper cites an unresolved cited work.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:03.917531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:03.917531Z digest=sha256:4bfff16d5a997b83e2975357fe25f102b1cf6599c97c8800ad733e9fbabef5a6

Observation 3d652191-670e-4ac6-beae-fb61a0474d0f · outbound

This paper cites an unresolved cited work.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:04.090374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:04.090374Z digest=sha256:2043f4f740892d6e8cca04d13db26af26cbac04aaa445c3211b944e4447efde0

Observation 70fece1c-6a00-458e-a4cf-a8cf22f55e76 · outbound

This paper cites High-Dimensional Continuous Control Using Generalized Advantage Estimation.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization High-Dimensional Continuous Control Using Generalized Advantage Estimation

Reference 1994

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:03.179272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:03.179272Z digest=sha256:aecbf7d367d76084c5751d8b43199cd83d56d7d9736456d6dafbdf6ac92c0dd0

Observation 98b8ac4a-bf9b-41ac-bb3d-c29bff3ef3e2 · outbound

This paper cites an unresolved cited work.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization Unresolved cited work

Reference 2011

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:03.687268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:03.687268Z digest=sha256:4af099b42f1d50aae396c370384a3db3d23a96c29989e00441ca20d7852b6abe

Observation 28c4cdc4-dd85-4fcf-88e9-e25e415ddba4 · outbound

This paper cites A Comprehensive Survey on Multi-Agent Cooperative Decision-Making: Scenarios, Approaches, Challenges and Perspectives.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization A Comprehensive Survey on Multi-Agent Cooperative Decision-Making: Scenarios, Approaches, Challenges and Perspectives

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:02.557564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:02.557564Z digest=sha256:7ddee64ec068300ae4bd81879ebb81cd736863637df263a06c20981e28e68783

Observation a830c2a8-6882-4e8d-9f3e-8dd3eb6a4478 · outbound

This paper cites Sumo–simulation of urban mobility: an overview.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization Sumo–simulation of urban mobility: an overview

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:02.386172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:02.386172Z digest=sha256:9e05b262a007107faef54b80161c5fd4a9cc21145acf279002797dfa4914eeb7

Observation 9cdcdc7e-db66-4795-8f60-5af32579eef5 · outbound

This paper cites Markov games as a framework for multi-agent reinforcement learning.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization Markov games as a framework for multi-agent reinforcement learning

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:02.957872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:02.957872Z digest=sha256:6b61eef51093966dc34f2dc87cadb19784dc901edc273f1f481dfac335bca4c4

Observation 647be801-6234-4bea-b8bc-1ab307dcc323 · outbound

This paper cites 13 B.2 Proof of Lemma 2 (multiplicative variance).

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization 13 B.2 Proof of Lemma 2 (multiplicative variance)

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:03.520863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:03.520863Z digest=sha256:5e2ea1294fe8c2a8aa756910b7220a8e8af9b1abb4d6ae76ff110abf562eab10

Observation 1326ed37-126f-4953-9039-bc2868288590 · outbound

This paper cites Generalized per-agent advantage estimation for multi-agent policy optimization.arXiv preprint arXiv:2603.02654,.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization Generalized per-agent advantage estimation for multi-agent policy optimization.arXiv preprint arXiv:2603.02654,

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:02.678810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:02.678810Z digest=sha256:617ba21438bc83815fd47524bcece4232f8ecdfdbac9634e0f861d8534256fa2

Observation 45d08777-a3d6-49ab-accc-75a350072ee6 · outbound

This paper cites Trust region policy optimisation in multi-agent reinforcement learning.

Aggregate in the Advantage, Not the Ratio: A Canonical-Form Analysis of Cooperative Multi-Agent Policy Optimization Trust region policy optimisation in multi-agent reinforcement learning

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-01T16:42:02.792432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:42:02.792432Z digest=sha256:3cf864829f2425df571aa966116ddeecf610e5a1314c7a6331abe7c6a6c1d6b2

Pith citing papers

No inbound Pith citation observations are available.