Pith. sign in

Paper Citation Record · LEDGER

Remembering the Markov Property in Cooperative MARL

As of 18 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2507.18333.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.18333 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:20:45.426207Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact0
  • verified fuzzy18
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c4fa5e91-f1d6-45c6-86fc-945723ea4e3a · outbound

This paper cites Autonomous agents modelling other agents: A comprehensive survey and open problems.

Remembering the Markov Property in Cooperative MARL Autonomous agents modelling other agents: A comprehensive survey and open problems

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.570620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.070848Z digest=sha256:d100698660514681d33da0912201d5f8dfb567eb035c485593b626d45c7ca4e9

Observation 968dbb79-5f4f-43af-9998-6769f5324cf0 · outbound

This paper cites Albrecht, Filippos Christianos, and Lukas Sch\"afer.

Remembering the Markov Property in Cooperative MARL Albrecht, Filippos Christianos, and Lukas Sch\"afer

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.537761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.083380Z digest=sha256:62905c2b6e2044d2e1361bf530e50e09ec7daf207f9fa18fab9a0abf0bcb3096

Observation 540b63bb-d6bb-4ac2-ac8c-5c86412e4a62 · outbound

This paper cites Optimal control of M arkov processes with incomplete state information.

Remembering the Markov Property in Cooperative MARL Optimal control of M arkov processes with incomplete state information

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.509296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.093823Z digest=sha256:adfeba33760f9c814470a326aeb3fff5a74e7bd8092e10143cf87b712d3193a9

Observation 40d5592a-1926-4508-b5b6-acd85a12b731 · outbound

This paper cites The hanabi challenge: A new frontier for ai research.

Remembering the Markov Property in Cooperative MARL The hanabi challenge: A new frontier for ai research

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.098662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.098662Z digest=sha256:c45822d9316d44cb07768c1d9b006f47ecf00e468adf69c2f22dad27577ca98f

Observation 44b23c26-441a-41ba-bf1c-212456047777 · outbound

This paper cites The complexity of decentralized control of markov decision processes.

Remembering the Markov Property in Cooperative MARL The complexity of decentralized control of markov decision processes

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.444200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.105513Z digest=sha256:a9ac90bd8aae53eee24419c6212630fb4aaaf3bdeaf057b322dc79767b7b73fb

Observation 1bdcf2c9-902d-456a-b0ff-2098cb75461e · outbound

This paper cites Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?.

Remembering the Markov Property in Cooperative MARL Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.124963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.124963Z digest=sha256:d7f0147ca40a3871d18fcb96dd2099387be83902b3a305b618b3573d56c5fd2a

Observation 767b0a5c-831e-4cfd-8332-42004233aeec · outbound

This paper cites Smacv2: An improved benchmark for cooperative multi-agent reinforcement learning.

Remembering the Markov Property in Cooperative MARL Smacv2: An improved benchmark for cooperative multi-agent reinforcement learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.406440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.134830Z digest=sha256:1a4657e70b38137fe6c6c2d3baed089294bdea7f841542c26fa2a3d0f891244f

Observation 9318dc74-91e6-4a2c-989f-9a8ce814a138 · outbound

This paper cites Bayesian action decoder for deep multi-agent reinforcement learning.

Remembering the Markov Property in Cooperative MARL Bayesian action decoder for deep multi-agent reinforcement learning

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.376813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.142457Z digest=sha256:7013bb2998ebdfe7ccde1dde197cae94488980a795f1bf791a9ddd5fd990b577

Observation 9803b937-1ae1-49e1-b051-92f459b39cb4 · outbound

This paper cites Studying the Interplay Between the Actor and Critic Representations in Reinforcement Learning.

Remembering the Markov Property in Cooperative MARL Studying the Interplay Between the Actor and Critic Representations in Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.151102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.151102Z digest=sha256:074f898cf1a258561ff631f1c764f9927bc3c2e6887df5a9f8947361e0a58598

Observation 9e07b5b1-b7f7-478c-b092-f0762c3bc2e0 · outbound

This paper cites Deep recurrent q-learning for partially observable mdps.

Remembering the Markov Property in Cooperative MARL Deep recurrent q-learning for partially observable mdps

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.340119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.158224Z digest=sha256:9730a903553a7d5b171f20faed4ce53e3b55a57630027690e934d612e33acb77

Observation 508e3ad0-4800-4341-9f31-13c702551081 · outbound

This paper cites other-play.

Remembering the Markov Property in Cooperative MARL other-play

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.302156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.169641Z digest=sha256:8ae5f6ccf2c09c83b7cc8d4658cd9daf538035315cf5d77e151fb86b0323c26b

Observation 961d6cde-5353-416b-91b5-f694ebd03589 · outbound

This paper cites Off-belief learning.

Remembering the Markov Property in Cooperative MARL Off-belief learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.257877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.179013Z digest=sha256:8b6e7ff8608cde12b496db770239172ffeaeb80add43ac669236d2274f7c6525

Observation d6b2a152-6b39-48f7-8bab-7b0d194a7090 · outbound

This paper cites Planning and acting in partially observable stochastic domains.

Remembering the Markov Property in Cooperative MARL Planning and acting in partially observable stochastic domains

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.185723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.185723Z digest=sha256:6202c2362146a2caa80d5f0429d7f952dfd7c63150ef47d3467b63dfc5ea4838

Observation b75f4662-c49c-44a7-af88-7a7ded0a3d0f · outbound

This paper cites Multi-agent reinforcement learning as a rehearsal for decentralized planning.

Remembering the Markov Property in Cooperative MARL Multi-agent reinforcement learning as a rehearsal for decentralized planning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.191290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.191290Z digest=sha256:046f2b716c4dad6d2917d6e5f69c79834ee867788f165e6e4b34f65f01dfeafb

Observation 7c522fe2-f1e7-4966-a875-f02cf11c1cc1 · outbound

This paper cites Estimating mutual information.

Remembering the Markov Property in Cooperative MARL Estimating mutual information

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.204100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.204100Z digest=sha256:885327b8b78f8b6ed6682e0f4fb3ad72166891aa9a70c16a300970edd940deff

Observation b35c1ed4-a17e-487d-bda5-bd69207728b4 · outbound

This paper cites Nonapproximability results for partially observable M arkov D ecision P rocesses.

Remembering the Markov Property in Cooperative MARL Nonapproximability results for partially observable M arkov D ecision P rocesses

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.152383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.212871Z digest=sha256:ff229b212cefe03414dc8dd1b3c4c2f151944947662eccef85eef36701627bfa

Observation 65ff5392-b109-466f-9f28-5830e7754f26 · outbound

This paper cites Partner Modelling Emerges in Recurrent Agents (But Only When It Matters).

Remembering the Markov Property in Cooperative MARL Partner Modelling Emerges in Recurrent Agents (But Only When It Matters)

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.219324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.219324Z digest=sha256:eabfb3e6e1b934fd770063e650792cb380d9e1954a03722035fb7c0befd60fb1

Observation 62971bf3-cf82-4957-b1ff-1aa10fd3830f · outbound

This paper cites Towards few-shot coordination: Revisiting ad-hoc teamplay challenge in the game of hanabi.

Remembering the Markov Property in Cooperative MARL Towards few-shot coordination: Revisiting ad-hoc teamplay challenge in the game of hanabi

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.100610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.226642Z digest=sha256:77a87a5006229a228f8d2a925f9122a9944c0dd1a954afdc8161f2fe39aaffab

Observation cdacd85c-232d-42b7-834e-ee46f00fa7ad · outbound

This paper cites Optimal and approximate q-value functions for decentralized pomdps.

Remembering the Markov Property in Cooperative MARL Optimal and approximate q-value functions for decentralized pomdps

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.236391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.236391Z digest=sha256:0ffc997d04432574cb139d16d3ae7ddedfcaf6f79339711f0287c96885875e30

Observation 163e3c6b-7ab6-4758-b4c4-75978526783a · outbound

This paper cites A concise introduction to decentralized POMDPs, volume 1.

Remembering the Markov Property in Cooperative MARL A concise introduction to decentralized POMDPs, volume 1

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.250388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.250388Z digest=sha256:4dfe663cbb4de9f34096b5b488992f5f3cb6c16910a4b53cb9784db9fa1df9d9

Observation 1c593a86-beff-412b-ae42-05b63875bdfb · outbound

This paper cites The complexity of M arkov D ecision P rocesses.

Remembering the Markov Property in Cooperative MARL The complexity of M arkov D ecision P rocesses

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.046766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.261928Z digest=sha256:6b9cc7f3d448ace1b63950ac5c4359d78970637b7c3fa731fd5204e72a0e8611

Observation 4dab9ba7-2994-4682-9406-65442990f983 · outbound

This paper cites Benchmarking Multi-Agent Deep Reinforcement Learning Algorithms in Cooperative Tasks.

Remembering the Markov Property in Cooperative MARL Benchmarking Multi-Agent Deep Reinforcement Learning Algorithms in Cooperative Tasks

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.269096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.269096Z digest=sha256:33992293eb202051ddd1856b645b37567a23661c7571c0efc76ac6fcdf1e0d9b

Observation d19a400a-4fe7-4fe0-b2f5-e0dda961b21d · outbound

This paper cites Agent modelling under partial observability for deep reinforcement learning.

Remembering the Markov Property in Cooperative MARL Agent modelling under partial observability for deep reinforcement learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.006244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.275751Z digest=sha256:da9db6e54e0c8c39ce3214d3746040e8d56dcca9218fdfc75c0de2c74f47135d

Observation 2233985a-2caa-4b17-aec3-7f39164ad8bc · outbound

This paper cites Facmac: Factored multi-agent centralised policy gradients.

Remembering the Markov Property in Cooperative MARL Facmac: Factored multi-agent centralised policy gradients

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.282696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.282696Z digest=sha256:524b0095e5afaa59e5653fc5e9e5ee93a7957b4945125179398ea541977c0fb1

Observation 22717ee6-b952-4100-ae0b-81ab47e12a2e · outbound

This paper cites Machine theory of mind.

Remembering the Markov Property in Cooperative MARL Machine theory of mind

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:45.955186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.289933Z digest=sha256:183b66862b0753cb338b57cd801b24d35cffcba3307c0d940d61056c0f82e137

Observation a98797fd-3d9a-44ff-8d22-6ee717188181 · outbound

This paper cites Mutual information between discrete and continuous data sets.

Remembering the Markov Property in Cooperative MARL Mutual information between discrete and continuous data sets

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.307175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.307175Z digest=sha256:5f0866accddf4e33c15eb5b3bbfee8ec0b8fb1d0c47ccae8b120bee18a790812

Observation 0a8b6e83-a659-49fe-9f1c-cd0166d76ff9 · outbound

This paper cites JaxMARL: Multi-Agent RL Environments and Algorithms in JAX.

Remembering the Markov Property in Cooperative MARL JaxMARL: Multi-Agent RL Environments and Algorithms in JAX

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.321654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.321654Z digest=sha256:f2c0d95ce5195deb4e0c191c83a5d2b9ebcd6c79dcc7d599bec69e0c59d50624

Observation 45ed4b3f-9170-4581-969c-a16ff0bef7a7 · outbound

This paper cites The StarCraft Multi-Agent Challenge.

Remembering the Markov Property in Cooperative MARL The StarCraft Multi-Agent Challenge

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.335420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.335420Z digest=sha256:7c7db89d3c3740cccf9d48691c123e502d49f794e39cd83a8a945cbfccc9f7e1

Observation 4824165f-4a2a-4b82-bedc-ec848fa0da91 · outbound

This paper cites Sutton and Andrew G.

Remembering the Markov Property in Cooperative MARL Sutton and Andrew G

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.344743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.344743Z digest=sha256:2dfbea61a4da4e6cb31959ac70260d258c2d3c939ee93d28ade607b09802a45e

Observation ce931b36-4ec8-4cbe-8a96-ee0ae306df07 · outbound

This paper cites Order matters: Agent-by-agent policy optimization.

Remembering the Markov Property in Cooperative MARL Order matters: Agent-by-agent policy optimization

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:45.853665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.358732Z digest=sha256:04672f1172ecb834d389f70dd961d754adbc7624ab7ab0158563da3597fffc23

Observation 04b3004c-d772-4bcc-a29e-1e54b140abac · outbound

This paper cites Emergence of maps in the memories of blind navigation agents.

Remembering the Markov Property in Cooperative MARL Emergence of maps in the memories of blind navigation agents

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:45.834884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.371678Z digest=sha256:d7735c983da9584ed6748a83b09ad8e76925831cc7930e8fb067210e94947fad

Observation be67f0dc-42f3-489e-86ca-122c136aa495 · outbound

This paper cites Learning latent representations to influence multi-agent interaction.

Remembering the Markov Property in Cooperative MARL Learning latent representations to influence multi-agent interaction

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:45.812784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.377556Z digest=sha256:fdbc177ef3846c5d9d8f49567b1d984d646f8d737ee0d2cf74a5b5f1e67e97b4

Observation 95e00ed8-4c54-4fbc-bc6c-e5f12e2e37ea · outbound

This paper cites The surprising effectiveness of ppo in cooperative multi-agent games.

Remembering the Markov Property in Cooperative MARL The surprising effectiveness of ppo in cooperative multi-agent games

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.387404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.387404Z digest=sha256:28da80943d60d22b790176d666de2f537019eac6f643b6e0557af574c4daa3c8

Observation 8d1b8beb-e096-4cce-947e-48a0ebd00ea3 · outbound

This paper cites Heterogeneous-agent reinforcement learning.

Remembering the Markov Property in Cooperative MARL Heterogeneous-agent reinforcement learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.393733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.393733Z digest=sha256:6332be2f75e2865c37cb6f6620ccbebe503a2d503f3e1d831d7ee97410d57400

Observation da94f39f-e44c-43f5-9bee-21cc1147454d · outbound

This paper cites Deep interactive bayesian reinforcement learning via meta-learning.

Remembering the Markov Property in Cooperative MARL Deep interactive bayesian reinforcement learning via meta-learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:45.757902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.408594Z digest=sha256:76b81538a930f0fd988ba115f81a939b7c1d1e7d05ee22e73b6be0ae92107c4f

Observation 13b01a95-80e2-4ca4-9965-6cdd89b796e4 · outbound

This paper cites write newline.

Remembering the Markov Property in Cooperative MARL write newline

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.426207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.426207Z digest=sha256:a658e9a68f2616be69141a3e2a19d6076aa192ea998fa01ab5d9fa09944e0e94

Pith citing papers

No inbound Pith citation observations are available.