Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T16:19:57.385629Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 1 inbound Pith citation observation for arXiv:2502.06261.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T16:19:57.385629Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-21T20:40:13.721163Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-21T20:40:35.732465Z
42 of 42 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2502d351-1eec-4dd7-ad7d-9c2cb6102b59 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c1d2711-1abc-47b8-b82c-8db6fcc4c381 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Andrew Bagnell, and Jan Peters
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e9906b90-5667-40d2-9ccc-d79ac42f6c43 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Lillicrap, Fan Hui, Laurent Sifre, George van den Driessche, Thore Graepel, and Demis Hassabis
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 260c663c-9d33-4732-b418-d6c540395658 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Superhuman ai for multiplayer poker
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ed4485b-0a8b-4973-904a-848343e7039f · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Multi-agent deep reinforcement learning: a survey
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 44ff44a0-b69d-46c1-8593-53e56d53de4d · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning A review of cooperative multi-agent deep reinforcement learning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c621892d-1268-4738-8b97-d31465cd2da1 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Game-Theoretic Multiagent Reinforcement Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53d58494-54f7-40e7-b050-f2caec36681b · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning A survey of multi-agent deep reinforcement learning with communication
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 63863f1c-7dbe-4ae1-80ed-7c8c1efad291 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Learning to Communicate in Multi-Agent Reinforcement Learning : A Review
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1fa22990-4611-4c4e-93ae-e3fc907cf14a · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Synchronizing UA V teams for timely data collection and energy transfer by deep reinforcement learning.IEEE Trans
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e2af9672-e080-4e8c-a6d4-cf4f554dec2e · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Gupta, Maxim Egorov, and Mykel J
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 76cdcfb3-9d57-40fe-9989-19df94926e04 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Actor-attention-critic for multi-agent reinforcement learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d735ec4f-2ab2-4cac-b3e2-fd3f511285c5 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Multi-agent game abstraction via graph attention neural network
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 50ad4426-d549-4411-95f0-7ab54952edad · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Foerster, Yannis M
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4d8280e2-0c10-4867-9da6-8745ea8aaa57 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Learning attentional communication for multi-agent cooperation
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4cd0f890-75f4-47d0-807d-7ef2abc8efde · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning On centralized critics in multi-agent reinforcement learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ab5ac30a-7032-499c-a7fa-4c129a7727a2 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Contrasting centralized and decentralized critics in multi-agent reinforcement learning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e2acd4fc-4eb2-4ef4-beba-78cad4a5ecc8 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 29b1d46d-b70b-4eda-8a75-ac8e11450dd9 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Learning when to communicate at scale in multiagent cooperative and competitive tasks
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b1cd5816-7b13-4b7c-9501-07fd00af2ed0 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Learning nearly decomposable value functions via communication minimization
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5399a4ea-c558-4750-9256-e3608063321b · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Tarmac: Targeted multi-agent communication
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f486a015-d6a9-4425-941b-47cc06b575ac · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Succinct and robust multi-agent communication with temporal message control
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 47371f0b-3f35-4051-bba6-955baeb63263 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Multi-agent incentive communication via decentralized teammate modeling
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dc851daa-5e93-4f43-b34f-898263b0d19a · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Efficient Communication via Self-supervised Information Aggregation for Online and Offline Multi-agent Reinforcement Learning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d663d7fd-eb8a-462a-ae88-75fa11f7a47a · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning T2MAC: targeted and trusted multi-agent communication through selective engagement and evidence-driven integration
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f8ef8820-5f49-4e12-8642-0d8229b43d46 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Learning agent communication under limited bandwidth by message pruning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e9d26ae2-e5db-4c43-8407-b819ba851ea4 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Learning individually inferred communication for multi-agent cooperation
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1daecf9d-87cd-4d6c-8e4b-26ab0c57e9cb · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Scalable communication for multi-agent reinforcement learning via transformer-based email mechanism
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fdd81f46-af2f-4fc1-b09d-151fe47407e2 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Paleja, and Matthew C
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 28dcfc46-fc94-4beb-8b44-c963f933ecc4 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Rgmcomm: Return gap minimization via discrete communications in multi-agent reinforcement learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ccd2a045-12b2-4663-ac85-ec7f84772e72 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Learning efficient multi-agent communication: An information bottleneck approach
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 36c00417-400e-47e1-8152-c691e2fc92a4 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Turner, Zoubin Ghahramani, and Sergey Levine
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 750508a3-0530-4fab-b6e2-b517d7b473bf · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Settling the variance of multi-agent policy gradients
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c6b2283d-fff9-476f-993c-66a066145f08 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Bayen, Sham M
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8ce56dc4-6fcd-4ac0-93cd-dcaff29cba82 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Foerster, Gregory Farquhar, Triantafyllos Afouras, Nantas Nardelli, and Shimon Whiteson
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0a93dc96-7135-4326-9758-12a1ae8adbb9 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Oliehoek and Christopher Amato
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7d2499bd-7f49-4c4f-9b0a-e539f1638e5e · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning How, and John Vian
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 096520c6-2008-4126-8b19-96cc2861ea95 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Bayen, and Yi Wu
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3b6c0b15-b62e-4869-bb13-1a6c84b38299 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Reinforcement learning with perturbed rewards
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation efeccdbb-b43b-48f1-8ee1-e0d8329ab0a9 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Boltzmann exploration done right
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8429745e-9246-47e2-a989-00b14c6caff1 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Multi-agent reinforcement learning is a sequence modeling problem
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e113d20f-d310-4260-8ff7-5ad4152560c3 · outbound
Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 43643e14-e495-41d7-b4f5-2f0fd331414b · inbound
Dynamic Generation of Multi-LLM Agents Communication Topologies with Graph Diffusion Models Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.