Pith. sign in

Paper Citation Record · LEDGER

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning

As of 9 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 1 inbound Pith citation observation for arXiv:2502.06261.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.06261 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T16:19:57.385629Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-21T20:40:13.721163Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T20:40:35.732465Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact2
  • verified fuzzy35
  • unresolved4
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2502d351-1eec-4dd7-ad7d-9c2cb6102b59 · outbound

This paper cites Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T16:19:57.231120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:19:57.231120Z digest=sha256:6d42be67c0f813e2b5f6f701a6850605c5e012d81cb0aea1ca0f99ae733b7313

Observation 2c1d2711-1abc-47b8-b82c-8db6fcc4c381 · outbound

This paper cites Andrew Bagnell, and Jan Peters.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Andrew Bagnell, and Jan Peters

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.942225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.235827Z digest=sha256:5f3c2cb72efd81c9d70feed1d928677bb36a475d91aa44a14a9b8339f2725676

Observation e9906b90-5667-40d2-9ccc-d79ac42f6c43 · outbound

This paper cites Lillicrap, Fan Hui, Laurent Sifre, George van den Driessche, Thore Graepel, and Demis Hassabis.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Lillicrap, Fan Hui, Laurent Sifre, George van den Driessche, Thore Graepel, and Demis Hassabis

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.930455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.239709Z digest=sha256:d76ec0d9ce5f62f83c4801260560352ac72c2a073c5d8b613db6650025ba0c75

Observation 260c663c-9d33-4732-b418-d6c540395658 · outbound

This paper cites Superhuman ai for multiplayer poker.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Superhuman ai for multiplayer poker

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T16:19:57.243916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:19:57.243916Z digest=sha256:3453564a23e07e3de73c956707203c59c102c9e9632faec6f7ed640d0a82d281

Observation 0ed4485b-0a8b-4973-904a-848343e7039f · outbound

This paper cites Multi-agent deep reinforcement learning: a survey.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Multi-agent deep reinforcement learning: a survey

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.912218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.247910Z digest=sha256:f79cb0696a69289f671c72bf6a3196ba739de68120d0a7bed1ee4fba5a200428

Observation 44ff44a0-b69d-46c1-8593-53e56d53de4d · outbound

This paper cites A review of cooperative multi-agent deep reinforcement learning.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning A review of cooperative multi-agent deep reinforcement learning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.900717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.252110Z digest=sha256:8c2dab6c9d191cd91c4fbe6a30bc6df5f90c41a98b1b658be3c238785dcec463

Observation c621892d-1268-4738-8b97-d31465cd2da1 · outbound

This paper cites Game-Theoretic Multiagent Reinforcement Learning.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Game-Theoretic Multiagent Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T16:19:57.256137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:19:57.256137Z digest=sha256:c9b2b5dc63a5e1c4da1cc47937e4441863d7f385205ddd63b689877ceac4510c

Observation 53d58494-54f7-40e7-b050-f2caec36681b · outbound

This paper cites A survey of multi-agent deep reinforcement learning with communication.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning A survey of multi-agent deep reinforcement learning with communication

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.890024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.260307Z digest=sha256:9f3e3ab4e63b15647209403aec3ae011f77876623044ea71cfdd2724a555fe0f

Observation 63863f1c-7dbe-4ae1-80ed-7c8c1efad291 · outbound

This paper cites Learning to Communicate in Multi-Agent Reinforcement Learning : A Review.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Learning to Communicate in Multi-Agent Reinforcement Learning : A Review

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-08T16:19:57.447875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.263808Z digest=sha256:2f3d36cc086e601d2bbb4c9aeca4c1c486cd0c362657066a9ba8c9c294f28b33

Observation 1fa22990-4611-4c4e-93ae-e3fc907cf14a · outbound

This paper cites Synchronizing UA V teams for timely data collection and energy transfer by deep reinforcement learning.IEEE Trans.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Synchronizing UA V teams for timely data collection and energy transfer by deep reinforcement learning.IEEE Trans

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.877265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.268322Z digest=sha256:5c6a5838d5587995de0e2e71d37dc046429883d469bdcfa9816b95284b0ce691

Observation e2af9672-e080-4e8c-a6d4-cf4f554dec2e · outbound

This paper cites Gupta, Maxim Egorov, and Mykel J.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Gupta, Maxim Egorov, and Mykel J

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.865314Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.272017Z digest=sha256:6534f3c6ced5d9e73efc0d6d411622fcfd4634c124b83256533c18a31b2a5083

Observation 76cdcfb3-9d57-40fe-9989-19df94926e04 · outbound

This paper cites Actor-attention-critic for multi-agent reinforcement learning.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Actor-attention-critic for multi-agent reinforcement learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.852836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.275766Z digest=sha256:be2a77267466d0281e9b981cf63a3d7a98355a9d7ecf3b1e64184b6454ac3dcd

Observation d735ec4f-2ab2-4cac-b3e2-fd3f511285c5 · outbound

This paper cites Multi-agent game abstraction via graph attention neural network.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Multi-agent game abstraction via graph attention neural network

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.841082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.279261Z digest=sha256:c996790b3d6988636c218e8466d8c43f35b68d41f889612cbdaa123a656a9935

Observation 50ad4426-d549-4411-95f0-7ab54952edad · outbound

This paper cites Foerster, Yannis M.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Foerster, Yannis M

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.828059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.282882Z digest=sha256:c8dd8ef909252c9e5021bd270eb4ba4cc47273414c615efad9659c5cc8c071d7

Observation 4d8280e2-0c10-4867-9da6-8745ea8aaa57 · outbound

This paper cites Learning attentional communication for multi-agent cooperation.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Learning attentional communication for multi-agent cooperation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.815218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.286367Z digest=sha256:4277959b7133b4b01e11244d16fe3b8c81528f802ce6b1de79dbee2aa457dc60

Observation 4cd0f890-75f4-47d0-807d-7ef2abc8efde · outbound

This paper cites On centralized critics in multi-agent reinforcement learning.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning On centralized critics in multi-agent reinforcement learning

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.803138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.290029Z digest=sha256:f3468e956b636f0cb46244448b6dcf3d34cc4a3e3d1e37ea1d2dc9591227cafb

Observation ab5ac30a-7032-499c-a7fa-4c129a7727a2 · outbound

This paper cites Contrasting centralized and decentralized critics in multi-agent reinforcement learning.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Contrasting centralized and decentralized critics in multi-agent reinforcement learning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.790276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.293752Z digest=sha256:25243cfc5aa1fb418b364b3412e7710402c1ee3fe7369484e9bb46dc657516c5

Observation e2acd4fc-4eb2-4ef4-beba-78cad4a5ecc8 · outbound

This paper cites an unresolved cited work.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-08T16:19:57.777375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.297561Z digest=sha256:dfda17f14e47e5c234e7e97e1bdc2470915f16323d85bc2dca8ef78b1fd34318

Observation 29b1d46d-b70b-4eda-8a75-ac8e11450dd9 · outbound

This paper cites Learning when to communicate at scale in multiagent cooperative and competitive tasks.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Learning when to communicate at scale in multiagent cooperative and competitive tasks

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.764455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.301076Z digest=sha256:786e9c5da87561c78acd91bfb69654dbfb0c102f92f389816dd7d32ce5eef50f

Observation b1cd5816-7b13-4b7c-9501-07fd00af2ed0 · outbound

This paper cites Learning nearly decomposable value functions via communication minimization.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Learning nearly decomposable value functions via communication minimization

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.752061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.304345Z digest=sha256:48d3dae852e369c36c39769d59994c16df8ed67ba9a3c1694055a9285be848e0

Observation 5399a4ea-c558-4750-9256-e3608063321b · outbound

This paper cites Tarmac: Targeted multi-agent communication.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Tarmac: Targeted multi-agent communication

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.739756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.308063Z digest=sha256:373170a7cc0e78d2fbb4326d12dc8a22b7d86d5d752b9c97e140bcb15d3f4e07

Observation f486a015-d6a9-4425-941b-47cc06b575ac · outbound

This paper cites Succinct and robust multi-agent communication with temporal message control.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Succinct and robust multi-agent communication with temporal message control

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.727281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.311448Z digest=sha256:682f0a79dc9e96fce62d4c342437d10e002c3911e898b02d35b09dd53dde31ee

Observation 47371f0b-3f35-4051-bba6-955baeb63263 · outbound

This paper cites Multi-agent incentive communication via decentralized teammate modeling.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Multi-agent incentive communication via decentralized teammate modeling

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.714440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.315048Z digest=sha256:cfc756faa23062502c8b0ad4daeeb26e5b849cb77445ad0054a2671422dfa17a

Observation dc851daa-5e93-4f43-b34f-898263b0d19a · outbound

This paper cites Efficient Communication via Self-supervised Information Aggregation for Online and Offline Multi-agent Reinforcement Learning.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Efficient Communication via Self-supervised Information Aggregation for Online and Offline Multi-agent Reinforcement Learning

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-08T16:19:57.429389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.318308Z digest=sha256:fa132830e7c3eed91f6509e6036300122db0faef184e0ca09ca6f05c16cd49af

Observation d663d7fd-eb8a-462a-ae88-75fa11f7a47a · outbound

This paper cites T2MAC: targeted and trusted multi-agent communication through selective engagement and evidence-driven integration.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning T2MAC: targeted and trusted multi-agent communication through selective engagement and evidence-driven integration

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.702005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.321950Z digest=sha256:2c3dcba9b8345bd6743aa40900d47dc2c6029a1a68613205a42f96b2af1aa310

Observation f8ef8820-5f49-4e12-8642-0d8229b43d46 · outbound

This paper cites Learning agent communication under limited bandwidth by message pruning.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Learning agent communication under limited bandwidth by message pruning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.689369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.325582Z digest=sha256:14eabd402e2362c1115c3f7835ce2868b6a61afb15f40312b5a5d0c080113320

Observation e9d26ae2-e5db-4c43-8407-b819ba851ea4 · outbound

This paper cites Learning individually inferred communication for multi-agent cooperation.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Learning individually inferred communication for multi-agent cooperation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.676710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.329291Z digest=sha256:5fbe797bec97e50f36cf34879d268f4512a9d635e806ec453c00ad3b628ebb0e

Observation 1daecf9d-87cd-4d6c-8e4b-26ab0c57e9cb · outbound

This paper cites Scalable communication for multi-agent reinforcement learning via transformer-based email mechanism.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Scalable communication for multi-agent reinforcement learning via transformer-based email mechanism

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.663455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.333133Z digest=sha256:7d46ae29de885258eef4e6a3b5038090fa5d36baca0acbeb8fd6a97ee65d6da4

Observation fdd81f46-af2f-4fc1-b09d-151fe47407e2 · outbound

This paper cites Paleja, and Matthew C.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Paleja, and Matthew C

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.651318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.336752Z digest=sha256:7a1a193ace484858d5f378dd6c9dde5b0300761645d63c32d9a5fb959c26f3b4

Observation 28dcfc46-fc94-4beb-8b44-c963f933ecc4 · outbound

This paper cites Rgmcomm: Return gap minimization via discrete communications in multi-agent reinforcement learning.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Rgmcomm: Return gap minimization via discrete communications in multi-agent reinforcement learning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.639216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.340540Z digest=sha256:4acccb92983dcb8fb19663fcb377a9bb970e4cbeb96d7f14a8adbefd39ac497b

Observation ccd2a045-12b2-4663-ac85-ec7f84772e72 · outbound

This paper cites Learning efficient multi-agent communication: An information bottleneck approach.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Learning efficient multi-agent communication: An information bottleneck approach

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.626274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.344297Z digest=sha256:57d9cded0c619600590752a7d74832956997a5ec9dfc6cd7ef9003b14a407ef6

Observation 36c00417-400e-47e1-8152-c691e2fc92a4 · outbound

This paper cites Turner, Zoubin Ghahramani, and Sergey Levine.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Turner, Zoubin Ghahramani, and Sergey Levine

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.613038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.348020Z digest=sha256:cb32c0293824f8f63238e661d25e858d5e32ca645f30bb9d0ac028b5ee440496

Observation 750508a3-0530-4fab-b6e2-b517d7b473bf · outbound

This paper cites Settling the variance of multi-agent policy gradients.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Settling the variance of multi-agent policy gradients

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.600249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.351746Z digest=sha256:557bbffbea8713d02e6f3a258380f2dc7d21df6e52e2d3a43f27b7082eaa65ca

Observation c6b2283d-fff9-476f-993c-66a066145f08 · outbound

This paper cites Bayen, Sham M.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Bayen, Sham M

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.588140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.355468Z digest=sha256:facc7bad09f8d68a22abd622aa0bc91a5d4396a1e538095a7e0f14059fd72b82

Observation 8ce56dc4-6fcd-4ac0-93cd-dcaff29cba82 · outbound

This paper cites Foerster, Gregory Farquhar, Triantafyllos Afouras, Nantas Nardelli, and Shimon Whiteson.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Foerster, Gregory Farquhar, Triantafyllos Afouras, Nantas Nardelli, and Shimon Whiteson

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.575982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.358946Z digest=sha256:1559a034dccc8651405ec7416ea73ec620a94a1180918775fa5044546e8dd792

Observation 0a93dc96-7135-4326-9758-12a1ae8adbb9 · outbound

This paper cites Oliehoek and Christopher Amato.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Oliehoek and Christopher Amato

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.563650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.362927Z digest=sha256:38ce6232ee3b9197716caf80663c79bd61ae5e22ab199ea80fc1a04dbfb88850

Observation 7d2499bd-7f49-4c4f-9b0a-e539f1638e5e · outbound

This paper cites How, and John Vian.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning How, and John Vian

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.551241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.366371Z digest=sha256:095c25bcaa3ffebfc47f57894a0ea49f7228ff845fbc73936b48187e70e69459

Observation 096520c6-2008-4126-8b19-96cc2861ea95 · outbound

This paper cites Bayen, and Yi Wu.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Bayen, and Yi Wu

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.539002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.370035Z digest=sha256:37a561aeb319fbd5a600056a14d741346860ed58ca5bbc85a2ab63648882aa57

Observation 3b6c0b15-b62e-4869-bb13-1a6c84b38299 · outbound

This paper cites Reinforcement learning with perturbed rewards.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Reinforcement learning with perturbed rewards

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.525368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.373921Z digest=sha256:8eff1a35a03f6844f852d16551eb94a1a93ee7abb17e9c7a9b090cb123bc50e8

Observation efeccdbb-b43b-48f1-8ee1-e0d8329ab0a9 · outbound

This paper cites Boltzmann exploration done right.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Boltzmann exploration done right

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.512319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.378211Z digest=sha256:35fe7badc420a96c8163e1be692057e50d6e44ea99640bf0bf368e6537a6da41

Observation 8429745e-9246-47e2-a989-00b14c6caff1 · outbound

This paper cites Multi-agent reinforcement learning is a sequence modeling problem.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Multi-agent reinforcement learning is a sequence modeling problem

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T16:19:57.499808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.381817Z digest=sha256:af8ac20574088c9f248335da235c4be220a746f6d466dd6781a8e97c457caac8

Observation e113d20f-d310-4260-8ff7-5ad4152560c3 · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 42

Resolution
malformed identifier
raw_fallback, observed 2026-08-08T16:19:57.487352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:19:57.385629Z digest=sha256:e13aa97345955153ec65b8b0d392641ffa4b1fa36462cc899a34639505d4e11d

Pith citing papers

Observation 43643e14-e495-41d7-b4f5-2f0fd331414b · inbound

Dynamic Generation of Multi-LLM Agents Communication Topologies with Graph Diffusion Models cites this paper.

Dynamic Generation of Multi-LLM Agents Communication Topologies with Graph Diffusion Models Reducing Variance Caused by Communication in Decentralized Multi-agent Deep Reinforcement Learning

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-21T20:40:35.734281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T20:40:13.721163Z digest=sha256:ef17a2fca346e3d0af55d740a2b740e95222d67d983d559ed0279334a0cb6d14