Pith. sign in

Paper Citation Record · LEDGER

An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2409.03052.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2409.03052 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 28 of 28 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:06:35.898809Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

14
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 782ae683-f35f-4269-8e58-665ee6edf90f · inbound

A Survey on Large Language Model-Based Social Agents in Game-Theoretic Scenarios cites this paper.

A Survey on Large Language Model-Based Social Agents in Game-Theoretic Scenarios An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T21:58:45.695716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T21:58:45.695716Z digest=sha256:15b80f640b1239ff3f2df617dc6db1556e9c947ad2e7d2f4ecd5ed3f697ed0ff

Observation 75c72700-4ee4-4fc1-90df-51bcb9033149 · inbound

Low-Rank Agent-Specific Adaptation (LoRASA) for Multi-Agent Policy Learning cites this paper.

Low-Rank Agent-Specific Adaptation (LoRASA) for Multi-Agent Policy Learning An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T18:50:49.747276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T18:50:49.747276Z digest=sha256:c67cc88186206539b6707f28f5d360561e09f94253a4536cb09c88b5b77b87ae

Observation 89ca6b7a-aacf-4940-b807-c6b33978a579 · inbound

Multi-Agent Reinforcement Learning in Wireless Distributed Networks for 6G cites this paper.

Multi-Agent Reinforcement Learning in Wireless Distributed Networks for 6G An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-08T17:54:46.259862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:54:46.259862Z digest=sha256:a0064966f8365c89ddd3e131de132fd8ad030a3ccd543e799653c35c94458404

Observation a79c9291-51df-4bc2-8aea-f6f4abf8281d · inbound

Multi-Agent Reinforcement Learning Scheduling to Support Low Latency in Teleoperated Driving cites this paper.

Multi-Agent Reinforcement Learning Scheduling to Support Low Latency in Teleoperated Driving An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T23:52:40.956366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:52:40.956366Z digest=sha256:61edb6cf5658f9d17ccfbb0f79f755b6b8ef070523ff57f56797852998957a1a

Observation 665d4dda-d0fc-450b-905f-a9f7bf7f5ceb · inbound

Multi-agent Embodied AI: Advances and Future Directions cites this paper.

Multi-agent Embodied AI: Advances and Future Directions An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T23:16:15.318064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:16:15.318064Z digest=sha256:243ba8878c80e0d958bd0ecf82e60d41f315ae4cddd5d720e87c2cbe3e349e01

Observation 14b5a24a-cb95-4c3c-9211-ae08216c2c98 · inbound

A Multi-Agent Reinforcement Learning Approach for Cooperative Air-Ground-Human Crowdsensing in Emergency Rescue cites this paper.

A Multi-Agent Reinforcement Learning Approach for Cooperative Air-Ground-Human Crowdsensing in Emergency Rescue An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T22:34:11.213822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:34:11.213822Z digest=sha256:78be86f9731fe153ae2097e098af2581327e84c89e49d86750bd3db93174b36b

Observation 1099799c-3f0b-4b0c-9b86-453c0e31561f · inbound

Overcoming Environmental Meta-Stationarity in MARL via Adaptive Curriculum and Counterfactual Group Advantage cites this paper.

Overcoming Environmental Meta-Stationarity in MARL via Adaptive Curriculum and Counterfactual Group Advantage An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:12:15.521735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-19T11:08:51.426575Z digest=sha256:a94754ab25257b1ec0051df30f4db8e3bc3dead67d39f6baf780c3288bc0ecc2

Observation a2a49b79-ac74-41a4-9e13-1cc06653ec66 · inbound

ReCoDe: Reinforcement Learning-based Dynamic Constraint Design for Multi-Agent Coordination cites this paper.

ReCoDe: Reinforcement Learning-based Dynamic Constraint Design for Multi-Agent Coordination An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T18:05:58.333520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:05:58.333520Z digest=sha256:dbb5cbc4f522bb5a78098c8084c2adfe199803778b46ad83b99d6f2740107c3e

Observation d8fa179c-fde7-4f44-b1b1-ab79d3212837 · inbound

Hierarchical Message-Passing Policies for Multi-Agent Reinforcement Learning cites this paper.

Hierarchical Message-Passing Policies for Multi-Agent Reinforcement Learning An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T10:40:38.686052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:40:38.686052Z digest=sha256:69814b7135f61f1f5a0122279a91ebaa49cd66694f9f5f6fe842f6e1992da510

Observation 44742ba5-51e9-4b7f-932b-1efb0f5ff205 · inbound

Real-time adaptive quantum error correction by model-free multi-agent learning cites this paper.

Real-time adaptive quantum error correction by model-free multi-agent learning An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 109

Resolution
unresolved
no resolver link, observed 2026-08-05T10:37:10.405790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:37:10.405790Z digest=sha256:0feb45f8a9987b22dbb720a40439d9debd0031e52d2281ad1099499332e7feb8

Observation dd073fbe-2a90-42f6-972e-65120b48f8ff · inbound

Coupling Smoothed Particle Hydrodynamics with Multi-Agent Deep Reinforcement Learning for Cooperative Control of Point Absorbers cites this paper.

Coupling Smoothed Particle Hydrodynamics with Multi-Agent Deep Reinforcement Learning for Cooperative Control of Point Absorbers An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T11:28:05.024515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T11:28:05.024515Z digest=sha256:821b0c836dc7840d756d495cb8abb6ae1ce14ae93180cd51d79ade51bfcda8b1

Observation eb300f35-c264-4348-8be0-7189985268ae · inbound

AGMARL-DKS: An Adaptive Graph-Enhanced Multi-Agent Reinforcement Learning for Dynamic Kubernetes Scheduling cites this paper.

AGMARL-DKS: An Adaptive Graph-Enhanced Multi-Agent Reinforcement Learning for Dynamic Kubernetes Scheduling An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-15T11:59:59.397066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T11:57:24.591538Z digest=sha256:77f6b9e71bf585a66c65bbf3a449bd1fccc0f50592f8b587c39c7e6a6997044d

Observation cb1ac08b-b1da-4f0b-a87a-c074f88c0bf8 · inbound

Plasticity-Enhanced Multi-Agent Mixture of Experts for Dynamic Objective Adaptation in UAVs-Assisted Emergency Communication Networks cites this paper.

Plasticity-Enhanced Multi-Agent Mixture of Experts for Dynamic Objective Adaptation in UAVs-Assisted Emergency Communication Networks An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:55:58.907161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T16:55:19.978358Z digest=sha256:005c6d8a388c3b3031ec7e9682838b0e86575d73ef39b4502a57faf2e4c6877d

Observation f127e8cc-0c51-4bce-8faf-5cbd518ea800 · inbound

Do LLM-derived graph priors improve multi-agent coordination? cites this paper.

Do LLM-derived graph priors improve multi-agent coordination? An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:11:20.221104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T06:09:08.100187Z digest=sha256:5b38f30556de5262fa3da2521f4b47b5df11beebcf46af880d5fcf265c7ca783

Observation 48714f95-240c-41bd-9963-1c7687bce540 · inbound

Cross-Modal Navigation with Multi-Agent Reinforcement Learning cites this paper.

Cross-Modal Navigation with Multi-Agent Reinforcement Learning An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:31:12.757224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-08T08:43:54.882467Z digest=sha256:101b434241b48ddab7a7cd182ff705921c6acb4eee7f63b0ea0b5bc880180672

Observation 630fcd5f-cbcd-4d5f-92f8-603fd71df216 · inbound

ERPPO: Entropy Regularization-based Proximal Policy Optimization cites this paper.

ERPPO: Entropy Regularization-based Proximal Policy Optimization An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 66

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:32:51.887848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-14T19:32:24.940497Z digest=sha256:2db859167f0a6fc7faa8c512aaf082c3e96b34c465481c9922687399ff792a82

Observation e0b1f735-c955-477c-9042-6794e1f3754a · inbound

Probabilistic Verification of Recurrent Neural Networks for Single and Multi-Agent Reinforcement Learning cites this paper.

Probabilistic Verification of Recurrent Neural Networks for Single and Multi-Agent Reinforcement Learning An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-30T20:55:04.121723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T20:52:06.448213Z digest=sha256:fb88c5bda57e6e56944c04a3c5852ef0fb1f0796a74de9361035f1ebc7104d31

Observation 4bb3849e-ca0e-42b2-9395-bb96262f48b2 · inbound

Beyond Partner Diversity: An Influence-Based Team Steering Framework for Zero-Shot Human-Machine Teaming cites this paper.

Beyond Partner Diversity: An Influence-Based Team Steering Framework for Zero-Shot Human-Machine Teaming An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-19T15:27:38.614899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-19T15:27:22.127896Z digest=sha256:95ed526cf4b89b7f2495bfd256e12c8f1bdd3851282d92615caffe47864f6f6b

Observation 7f4be0b3-7adb-41e8-ab15-2a685a9a60aa · inbound

SwarmHarness: Skill-Based Task Routing via Decentralized Incentive-Aligned AI Agent Networks cites this paper.

SwarmHarness: Skill-Based Task Routing via Decentralized Incentive-Aligned AI Agent Networks An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-06-29T11:43:23.574722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T11:41:30.827427Z digest=sha256:ffb8609baa746e9ae1610916fb25ec2714d1accae351bc662b0fefefd375787e

Observation e61e9a24-c40b-4c47-b376-7d1b108638ee · inbound

CHORUS: Decentralized Multi-Embodiment Collaboration with One VLA Policy cites this paper.

CHORUS: Decentralized Multi-Embodiment Collaboration with One VLA Policy An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-27T10:00:49.182519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T09:55:03.745087Z digest=sha256:8541aa5e77cc9a3ccccb0e5284f1526e99fd7ac4ffd6267a686c339363984fd9

Observation 5c75ce09-454a-4ecf-8e76-0fb44705bfdc · inbound

Self-CTRL: Self-Consistency Training with Reinforcement Learning cites this paper.

Self-CTRL: Self-Consistency Training with Reinforcement Learning An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:08:56.006777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T01:38:48.296421Z digest=sha256:1ab286969bff686428b3e0e7f632a410aed86f98e914f47474ae89db77bf3b7b

Observation 2fa175ba-26b1-47bc-8294-906142bf628d · inbound

Embodied Human-Robot Interaction via Acoustics: A MARL Approach with AcoustoBots for Spatial Data Physicalization cites this paper.

Embodied Human-Robot Interaction via Acoustics: A MARL Approach with AcoustoBots for Spatial Data Physicalization An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-07-08T01:44:26.037734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-08T01:42:38.113751Z digest=sha256:d7dc3ab32a083a3f052e14a79aeb12720b1b33edc63f2923b21c324efc454728

Observation b6c0562f-ae7e-4f06-b25f-d52cc58b1241 · inbound

Social-spatial dependencies for learning visual navigation cites this paper.

Social-spatial dependencies for learning visual navigation An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-07-09T10:26:10.921508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-07-09T10:23:00.725686Z digest=sha256:343bafa791ba4a570927da50db6910dfe457ec3ee9adeb65de77964a0822f99e

Observation 02cff74d-13ab-47c2-a7df-cd902bf4b2d1 · inbound

Reinforcement Learning for Delivery Drone-Based Participatory Sensing in Dynamic Environments cites this paper.

Reinforcement Learning for Delivery Drone-Based Participatory Sensing in Dynamic Environments An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T14:08:04.235006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:08:04.235006Z digest=sha256:7f877d19cbeaaf4c065b62c5059e2f8b44371bf1eed49a699c2f8eb7aa13d9b0

Observation c811c081-173a-4935-8d7f-0fccb6dc2993 · inbound

FedCritic-MIMO: Communication-Efficient Serverless Federated Critic Learning for Massive-MIMO Resource Control in Open and Disaggregated 6G RANs cites this paper.

FedCritic-MIMO: Communication-Efficient Serverless Federated Critic Learning for Massive-MIMO Resource Control in Open and Disaggregated 6G RANs An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T14:52:00.844770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:52:00.844770Z digest=sha256:03c58babfa88894bbf7432956aef2feddb11acb3bd1cbbe227924e3a6638e9c1

Observation 252e686e-8d5a-4464-be9a-213793c6c3fa · inbound

Latent Semantic State Estimation for Reliable Swarming of UAVs under Intermittent Connectivity cites this paper.

Latent Semantic State Estimation for Reliable Swarming of UAVs under Intermittent Connectivity An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-14T04:25:48.181291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:25:48.181291Z digest=sha256:aecc6401e49c692e12435d445dcee8f9a15c2d2f3326a3d646477f1ab0f2f229

Observation 5074cb71-2bf8-4341-9c92-132d71c9fe2e · inbound

Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition cites this paper.

Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T11:20:20.055396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:20:20.055396Z digest=sha256:3fd54db570ec56aeef012eb781514abd5e9dcdb434c3647b7625483d6186c4d5

Observation afa5d363-c047-481b-9121-1601d3146f91 · inbound

Decentralized Multi-Player Q-Learning in Episodic Markov Decision Processes with Information Asymmetry cites this paper.

Decentralized Multi-Player Q-Learning in Episodic Markov Decision Processes with Information Asymmetry An Introduction to Centralized Training for Decentralized Execution in Cooperative Multi-Agent Reinforcement Learning

Reference 149

Resolution
unresolved
no resolver link, observed 2026-08-16T00:06:35.898809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T00:06:35.898809Z digest=sha256:cd90e51d0bab86c064cd309704ac4528b8c0d039615b0559ad04dba61b96ed2f