Pith. sign in

Paper Citation Record · LEDGER

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning

As of 15 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 0 inbound Pith citation observations for arXiv:1908.02269.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.02269 v4

Coverage vector

measured 33 of 33 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T14:54:44.217833Z

measured 33 of 33 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

33 of 33 outbound references displayed

  • verified exact0
  • verified fuzzy20
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 31f3c23a-40dd-458b-ad90-a2b0870a94d5 · outbound

This paper cites Feudal Multi-Agent Hierarchies for Cooperative Reinforcement Learning.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Feudal Multi-Agent Hierarchies for Cooperative Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T14:54:43.953588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:54:43.953588Z digest=sha256:842d84d93cf4992b8c424d4573d819f7c91f4d4f2444682538a107f68196f082

Observation 84970a71-c76f-4102-8a03-f77eca30e11e · outbound

This paper cites Layer Normalization.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Layer Normalization

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-14T14:54:43.962010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:54:43.962010Z digest=sha256:238bfd9c55b1dd1194399e9232fa0039830df0116af40edd8ba74ff632145af2

Observation 01a9811d-c5d6-4227-aa08-c87f81b65cb1 · outbound

This paper cites The option-critic architecture.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning The option-critic architecture

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:54:45.063890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:43.969778Z digest=sha256:3402e02c3961b99b91ca0f77a6bfb8d5a4dfd9c88562a3e7103fbff791b76def

Observation 52f83007-9283-4099-932a-3b46882bad39 · outbound

This paper cites Measuring collaborative emergent behavior in multi-agent reinforcement learning.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Measuring collaborative emergent behavior in multi-agent reinforcement learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:54:45.041024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:43.982105Z digest=sha256:f49da9bcf7316b7466bde3adddab2e2190db600207d46dfbaeed6086d0608eaa

Observation f7470be0-522b-41e2-8c3d-701365afbee3 · outbound

This paper cites Intrinsically motivated reinforcement learning.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Intrinsically motivated reinforcement learning

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:54:45.003405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:43.990917Z digest=sha256:9fca2d99b18a9c3a659ba6f022c5967edb4a96d9da4fee66111fb730864c80f5

Observation d39358ca-3cd2-4be1-bddb-c0069c2195fb · outbound

This paper cites Learning to communicate with deep multi-agent reinforcement learning.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Learning to communicate with deep multi-agent reinforcement learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-14T14:54:43.998984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:54:43.998984Z digest=sha256:46fd235fb2262b4fc2842d46abc8d12d08b196c69cdbc99cd6eaf0dc1bda0e90

Observation 25bc1adb-21a0-4e6a-83a1-ab4f9f7070b4 · outbound

This paper cites Bayesian action decoder for deep multi-agent reinforcement learning.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Bayesian action decoder for deep multi-agent reinforcement learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:54:44.967010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:44.006749Z digest=sha256:483c7b52bfdf523465a2385da08494590bd74bf7b936953d4a54c3b803a46df0

Observation 1c6afe98-c908-4f7a-b777-9f1ce5a37824 · outbound

This paper cites Counterfactual multi-agent policy gradients.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Counterfactual multi-agent policy gradients

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:54:44.942622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:44.021077Z digest=sha256:11f06f884d08314b6a1f1eeeb9fed89b8d1faf3e355b8d969dbfea8e580b858a

Observation a130423e-82d1-4f7a-b308-117e4156787b · outbound

This paper cites Gupta, Maxim Egorov, and Mykel J.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Gupta, Maxim Egorov, and Mykel J

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:54:44.921896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:44.028507Z digest=sha256:e829500a72007662af2c1735770f0be983a22bba06649a3ee2506c4730017e60

Observation 22556c48-a8d6-4b99-95ff-eafb6a13f8ba · outbound

This paper cites Opponent modeling in deep reinforcement learning.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Opponent modeling in deep reinforcement learning

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:54:44.900238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:44.041530Z digest=sha256:067e3148a0c465d663bbcb5ab3167c0e921d6e01db0c56d8bba8abb818746885

Observation 129480d0-e810-4143-adda-f77976f36055 · outbound

This paper cites A Survey and Critique of Multiagent Deep Reinforcement Learning.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning A Survey and Critique of Multiagent Deep Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-14T14:54:44.049340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:54:44.049340Z digest=sha256:15b4c2eea33e3f515363399d0b01c3f8ae0478c6a0172112944da1e2fb931446

Observation abd0ce00-f827-4af1-bc49-f704efc88603 · outbound

This paper cites an unresolved cited work.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-14T14:54:44.877925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:44.059575Z digest=sha256:c4ae94fa3ff622d332d49826114f351ed1328ec86b35ba4e8c3bb42796effbef

Observation 04cbe614-582f-4f98-bc8a-fa3eb88473ee · outbound

This paper cites A Deep Policy Inference Q-Network for Multi-Agent Systems.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning A Deep Policy Inference Q-Network for Multi-Agent Systems

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-14T14:54:44.066304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:54:44.066304Z digest=sha256:29ed625154a4cbb14167abbfa39bfe7b5c2eea9afb8e28e4f2b1f016cd8e8911

Observation a10cb208-fce8-49c0-831a-ae2a06cf8f8d · outbound

This paper cites Actor-attention-critic for multi-agent reinforcement learning.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Actor-attention-critic for multi-agent reinforcement learning

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:54:44.857945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:44.076673Z digest=sha256:6924c36897eea0868a3b676e96e3eb09f757787d129be0af4b670d68928ae022

Observation 2b989619-573b-485b-a6a0-5c5e368f7a9e · outbound

This paper cites Categorical reparametrization with gumble-softmax.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Categorical reparametrization with gumble-softmax

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-14T14:54:44.083614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:54:44.083614Z digest=sha256:0a375d8f4ef9f7d2a21970ae553112b2d07100bf0168ef7405c3939c8dd15101

Observation da6d8119-f219-4ed5-9504-639efdc9fbbe · outbound

This paper cites Social influence as intrinsic motivation for multi-agent deep reinforcement learning.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Social influence as intrinsic motivation for multi-agent deep reinforcement learning

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:54:44.823208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:44.089869Z digest=sha256:66d43f02613d7afeade173e4c39a4c8ab15b5a6ba1a449dfe0f5cc9f146268c2

Observation 1fc2de41-4159-4a4a-a534-53c3c608b89b · outbound

This paper cites Learning attentional communication for multi-agent cooperation.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Learning attentional communication for multi-agent cooperation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:54:44.801844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:44.096591Z digest=sha256:41863ca2b0831378b5d1fa7c83933f730c6bb5d56488f0afed460276447c9785

Observation 42ebaf7e-bf16-4354-93a7-4a234ef18050 · outbound

This paper cites Reinforcement learning: A survey.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Reinforcement learning: A survey

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:54:44.777650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:44.105639Z digest=sha256:eca0194e543d10422f9b8624558b6a2762347140a073c9c00c093f7d0d7828f0

Observation c93dc2f8-bc49-4efe-afde-4e7b6ccce3df · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Adam: A Method for Stochastic Optimization

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-14T14:54:44.111235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:54:44.111235Z digest=sha256:7e260dce0cf5775830f62b7e7574682fd16c302a619cb2658e77057bbf6a8103

Observation 16709973-0c26-4dc5-80e1-a0517f3bb0ad · outbound

This paper cites Google Research Football: A Novel Reinforcement Learning Environment.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Google Research Football: A Novel Reinforcement Learning Environment

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-14T14:54:44.120179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:54:44.120179Z digest=sha256:58d0f5bb970d5d0ff3309dfcbd788e46e866ace80f1c051fe604e17e0a2b06da

Observation 62416596-1eb8-4359-8097-1aa993049d47 · outbound

This paper cites Multi-Agent Cooperation and the Emergence of (Natural) Language.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Multi-Agent Cooperation and the Emergence of (Natural) Language

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-14T14:54:44.128708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:54:44.128708Z digest=sha256:31fd88d099a510a528312e382dd5d8d4f79a76d99f37df08818d102667cd6ccd

Observation 3e95b0af-57b3-4e7d-be15-dd3959467b37 · outbound

This paper cites Continuous control with deep reinforcement learning.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Continuous control with deep reinforcement learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-14T14:54:44.135861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:54:44.135861Z digest=sha256:03d0d9c4462a6994d5a1e37fce01cb63df0cf6c15f849fa3e770ba6bad54a37a

Observation befe7fff-7d03-473c-9326-e6b6a947ec6e · outbound

This paper cites Markov games as a framework for multi-agent reinforcement learning.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Markov games as a framework for multi-agent reinforcement learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-14T14:54:44.144177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:54:44.144177Z digest=sha256:9da4b0f8a88d6e7cd4a0e4a3239b5f0c731f31877f79be1a2796652ba465329c

Observation 5f1afdca-1cc3-471a-93f9-dd20ffda956b · outbound

This paper cites Multi- agent actor-critic for mixed cooperative-competitive environments.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Multi- agent actor-critic for mixed cooperative-competitive environments

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:54:44.740842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:44.152051Z digest=sha256:e026039149adfd8618ea2b4c6da6a48551b6b2e69e1b82ef161ede05e3d874fb

Observation 6e379780-130d-461f-84e0-42e43404be60 · outbound

This paper cites Maven: Multi- agent variational exploration.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Maven: Multi- agent variational exploration

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:54:44.716233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:44.158250Z digest=sha256:206fb40239d8a9c5af17e09a5ac2c10250cc0f3a9eed489178504782fc95ec9f

Observation 1096e428-e54c-48df-b0e0-ae19377565ef · outbound

This paper cites Emergence of grounded compositional language in multi- agent populations.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Emergence of grounded compositional language in multi- agent populations

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:54:44.697519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:44.163546Z digest=sha256:f9695e2b3c724267a0d7030a4365beebc0de69a44243844d949c4f480700868b

Observation ed03b82e-b625-43bb-91d0-25e61f5229d0 · outbound

This paper cites Rectified linear units improve restricted boltzmann machines.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Rectified linear units improve restricted boltzmann machines

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:54:44.672750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:44.169332Z digest=sha256:7b909563926e90ac7bbfcde8c903d9e93303055ebd6e87c1235988a9c608c235

Observation 8e72d5c7-6e71-4109-a574-adae660e8898 · outbound

This paper cites Qmix: Monotonic value function factorisation for deep multi-agent reinforcement learning.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Qmix: Monotonic value function factorisation for deep multi-agent reinforcement learning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:54:44.652083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:44.176891Z digest=sha256:0a07c2cf6ff274dad0038130c3442ef8c93248a4bf21268b08fa53ede40d19d9

Observation 0b089523-2274-4f6f-829a-0aaad85d73a8 · outbound

This paper cites Opponent modeling in real-time strategy games.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Opponent modeling in real-time strategy games

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:54:44.629619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:44.185152Z digest=sha256:53fd3960800705ab4b4c3b0bd8791678ff77be4a7d859515e4cf6cf77cabca08

Observation 8fdc0df3-96d1-4a4b-bee7-de972d2b003d · outbound

This paper cites Dropout: a simple way to prevent neural networks from overfitting.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Dropout: a simple way to prevent neural networks from overfitting

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-14T14:54:44.191177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:54:44.191177Z digest=sha256:865cace07a15fbf696fd17c15963cf5ff7f1f52bcec2a36d7dcbc3720cba2546

Observation e52cd374-21f5-4329-bc95-5d70881be3f3 · outbound

This paper cites Learning to share and hide intentions using information regularization.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning Learning to share and hide intentions using information regularization

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:54:44.578301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:44.200017Z digest=sha256:06b6b033e815795ebef4ade3ac68a336e023e8f9a06ede15dfe1c1048e191c61

Observation 0d2c43a1-ebe2-4d31-b065-4e220b25a4e7 · outbound

This paper cites On the theory of the brownian motion.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning On the theory of the brownian motion

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:54:44.555735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:44.206855Z digest=sha256:7740e9bceb850fe526dbea7c38887785997d50bccfe9d94d816afc588b0b9f78

Observation 54851eca-7b0a-4646-859b-6c282148b864 · outbound

This paper cites MADDPG + policy mask.

Promoting Coordination through Policy Regularization in Multi-Agent Deep Reinforcement Learning MADDPG + policy mask

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:54:44.533474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-14T14:54:44.217833Z digest=sha256:7894ade1739d4f380c6bd12ea55865ef5760838cee0f153d75f1080818476d0f

Pith citing papers

No inbound Pith citation observations are available.