Pith. sign in

Paper Citation Record · LEDGER

Remembering the Markov Property in Cooperative MARL

As of 19 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2507.18333.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.18333 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:20:45.426207Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact0
  • verified fuzzy18
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c4fa5e91-f1d6-45c6-86fc-945723ea4e3a · outbound

This paper cites Autonomous agents modelling other agents: A comprehensive survey and open problems.

Remembering the Markov Property in Cooperative MARL Autonomous agents modelling other agents: A comprehensive survey and open problems

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.570620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.070848Z digest=sha256:e2fca52e57dc30202b246293a8db94dd5d356436a46be8690413af93535c9c9f

Observation 968dbb79-5f4f-43af-9998-6769f5324cf0 · outbound

This paper cites Albrecht, Filippos Christianos, and Lukas Sch\"afer.

Remembering the Markov Property in Cooperative MARL Albrecht, Filippos Christianos, and Lukas Sch\"afer

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.537761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.083380Z digest=sha256:6b4a9d3e3e8d8c5d0eac25929604967b0227d127bf8ce56d6b392adc7e6fc90c

Observation 540b63bb-d6bb-4ac2-ac8c-5c86412e4a62 · outbound

This paper cites Optimal control of M arkov processes with incomplete state information.

Remembering the Markov Property in Cooperative MARL Optimal control of M arkov processes with incomplete state information

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.509296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.093823Z digest=sha256:2a7ff8321b3811a96c41ad26e708f1ba3e5c17f7fcf08e2dbbc2649ea57425f7

Observation 40d5592a-1926-4508-b5b6-acd85a12b731 · outbound

This paper cites The hanabi challenge: A new frontier for ai research.

Remembering the Markov Property in Cooperative MARL The hanabi challenge: A new frontier for ai research

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.098662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.098662Z digest=sha256:c45822d9316d44cb07768c1d9b006f47ecf00e468adf69c2f22dad27577ca98f

Observation 44b23c26-441a-41ba-bf1c-212456047777 · outbound

This paper cites The complexity of decentralized control of markov decision processes.

Remembering the Markov Property in Cooperative MARL The complexity of decentralized control of markov decision processes

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.444200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.105513Z digest=sha256:82798efd1543b912e8f85f590d541f3a3f118979b7562b28a834f9c6fd742d03

Observation 1bdcf2c9-902d-456a-b0ff-2098cb75461e · outbound

This paper cites Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?.

Remembering the Markov Property in Cooperative MARL Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.124963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.124963Z digest=sha256:6dd7801b67b62c3d9c82007ac1e092c7202431a69e695720c7693de56194faa7

Observation 767b0a5c-831e-4cfd-8332-42004233aeec · outbound

This paper cites Smacv2: An improved benchmark for cooperative multi-agent reinforcement learning.

Remembering the Markov Property in Cooperative MARL Smacv2: An improved benchmark for cooperative multi-agent reinforcement learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.406440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.134830Z digest=sha256:b39b110d9ccc6bfdbfbeb5866ec060497e13870e03bd1e016165a49647606639

Observation 9318dc74-91e6-4a2c-989f-9a8ce814a138 · outbound

This paper cites Bayesian action decoder for deep multi-agent reinforcement learning.

Remembering the Markov Property in Cooperative MARL Bayesian action decoder for deep multi-agent reinforcement learning

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.376813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.142457Z digest=sha256:5b0775651be39fc2f7351cbe172f13b71f40a83f2858c37fe67dd583a96d4718

Observation 9803b937-1ae1-49e1-b051-92f459b39cb4 · outbound

This paper cites Studying the Interplay Between the Actor and Critic Representations in Reinforcement Learning.

Remembering the Markov Property in Cooperative MARL Studying the Interplay Between the Actor and Critic Representations in Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.151102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.151102Z digest=sha256:074f898cf1a258561ff631f1c764f9927bc3c2e6887df5a9f8947361e0a58598

Observation 9e07b5b1-b7f7-478c-b092-f0762c3bc2e0 · outbound

This paper cites Deep recurrent q-learning for partially observable mdps.

Remembering the Markov Property in Cooperative MARL Deep recurrent q-learning for partially observable mdps

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.340119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.158224Z digest=sha256:0272780fb0388abaf6710ac18ec2d62e39e416085ca1d4366bd9f6187b7094e2

Observation 508e3ad0-4800-4341-9f31-13c702551081 · outbound

This paper cites other-play.

Remembering the Markov Property in Cooperative MARL other-play

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.302156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.169641Z digest=sha256:b1edbc1196c74cc2fae30ce1a6c9869c06ef6be6ea7a7c872151d987ee4176c8

Observation 961d6cde-5353-416b-91b5-f694ebd03589 · outbound

This paper cites Off-belief learning.

Remembering the Markov Property in Cooperative MARL Off-belief learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.257877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.179013Z digest=sha256:e5a9b0622ff9c04616594f569584fef6721ba39be997988f2bd28b85c6cf250f

Observation d6b2a152-6b39-48f7-8bab-7b0d194a7090 · outbound

This paper cites Planning and acting in partially observable stochastic domains.

Remembering the Markov Property in Cooperative MARL Planning and acting in partially observable stochastic domains

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.185723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.185723Z digest=sha256:6202c2362146a2caa80d5f0429d7f952dfd7c63150ef47d3467b63dfc5ea4838

Observation b75f4662-c49c-44a7-af88-7a7ded0a3d0f · outbound

This paper cites Multi-agent reinforcement learning as a rehearsal for decentralized planning.

Remembering the Markov Property in Cooperative MARL Multi-agent reinforcement learning as a rehearsal for decentralized planning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.191290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.191290Z digest=sha256:046f2b716c4dad6d2917d6e5f69c79834ee867788f165e6e4b34f65f01dfeafb

Observation 7c522fe2-f1e7-4966-a875-f02cf11c1cc1 · outbound

This paper cites Estimating mutual information.

Remembering the Markov Property in Cooperative MARL Estimating mutual information

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.204100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.204100Z digest=sha256:885327b8b78f8b6ed6682e0f4fb3ad72166891aa9a70c16a300970edd940deff

Observation b35c1ed4-a17e-487d-bda5-bd69207728b4 · outbound

This paper cites Nonapproximability results for partially observable M arkov D ecision P rocesses.

Remembering the Markov Property in Cooperative MARL Nonapproximability results for partially observable M arkov D ecision P rocesses

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.152383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.212871Z digest=sha256:8709a0909fbe54d629765ec3750a51c24098655172663ca0e09e28d97706d9f7

Observation 65ff5392-b109-466f-9f28-5830e7754f26 · outbound

This paper cites Partner Modelling Emerges in Recurrent Agents (But Only When It Matters).

Remembering the Markov Property in Cooperative MARL Partner Modelling Emerges in Recurrent Agents (But Only When It Matters)

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.219324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.219324Z digest=sha256:eabfb3e6e1b934fd770063e650792cb380d9e1954a03722035fb7c0befd60fb1

Observation 62971bf3-cf82-4957-b1ff-1aa10fd3830f · outbound

This paper cites Towards few-shot coordination: Revisiting ad-hoc teamplay challenge in the game of hanabi.

Remembering the Markov Property in Cooperative MARL Towards few-shot coordination: Revisiting ad-hoc teamplay challenge in the game of hanabi

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.100610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.226642Z digest=sha256:161c138c86b156e2e26ef88c6f535bab770e00727250ad528998a6ed9c07de0c

Observation cdacd85c-232d-42b7-834e-ee46f00fa7ad · outbound

This paper cites Optimal and approximate q-value functions for decentralized pomdps.

Remembering the Markov Property in Cooperative MARL Optimal and approximate q-value functions for decentralized pomdps

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.236391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.236391Z digest=sha256:0ffc997d04432574cb139d16d3ae7ddedfcaf6f79339711f0287c96885875e30

Observation 163e3c6b-7ab6-4758-b4c4-75978526783a · outbound

This paper cites A concise introduction to decentralized POMDPs, volume 1.

Remembering the Markov Property in Cooperative MARL A concise introduction to decentralized POMDPs, volume 1

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.250388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.250388Z digest=sha256:4dfe663cbb4de9f34096b5b488992f5f3cb6c16910a4b53cb9784db9fa1df9d9

Observation 1c593a86-beff-412b-ae42-05b63875bdfb · outbound

This paper cites The complexity of M arkov D ecision P rocesses.

Remembering the Markov Property in Cooperative MARL The complexity of M arkov D ecision P rocesses

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.046766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.261928Z digest=sha256:d5da54d2c07567ea12e21d2ad62190bb9f3cff8dc86adcd2b7b68ee7b3261e8d

Observation 4dab9ba7-2994-4682-9406-65442990f983 · outbound

This paper cites Benchmarking Multi-Agent Deep Reinforcement Learning Algorithms in Cooperative Tasks.

Remembering the Markov Property in Cooperative MARL Benchmarking Multi-Agent Deep Reinforcement Learning Algorithms in Cooperative Tasks

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.269096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.269096Z digest=sha256:33992293eb202051ddd1856b645b37567a23661c7571c0efc76ac6fcdf1e0d9b

Observation d19a400a-4fe7-4fe0-b2f5-e0dda961b21d · outbound

This paper cites Agent modelling under partial observability for deep reinforcement learning.

Remembering the Markov Property in Cooperative MARL Agent modelling under partial observability for deep reinforcement learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:46.006244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.275751Z digest=sha256:4dea07e7a4c5b71e7db2efb0e7e3aa95e5f28a8bb8dfcc9d03923b04c4d16f47

Observation 2233985a-2caa-4b17-aec3-7f39164ad8bc · outbound

This paper cites Facmac: Factored multi-agent centralised policy gradients.

Remembering the Markov Property in Cooperative MARL Facmac: Factored multi-agent centralised policy gradients

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.282696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.282696Z digest=sha256:524b0095e5afaa59e5653fc5e9e5ee93a7957b4945125179398ea541977c0fb1

Observation 22717ee6-b952-4100-ae0b-81ab47e12a2e · outbound

This paper cites Machine theory of mind.

Remembering the Markov Property in Cooperative MARL Machine theory of mind

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:45.955186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.289933Z digest=sha256:edbd34b973eab262bae1a22c4656a8ecdbc83af0a6be955e6b11118fca1e4c73

Observation a98797fd-3d9a-44ff-8d22-6ee717188181 · outbound

This paper cites Mutual information between discrete and continuous data sets.

Remembering the Markov Property in Cooperative MARL Mutual information between discrete and continuous data sets

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.307175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.307175Z digest=sha256:5f0866accddf4e33c15eb5b3bbfee8ec0b8fb1d0c47ccae8b120bee18a790812

Observation 0a8b6e83-a659-49fe-9f1c-cd0166d76ff9 · outbound

This paper cites JaxMARL: Multi-Agent RL Environments and Algorithms in JAX.

Remembering the Markov Property in Cooperative MARL JaxMARL: Multi-Agent RL Environments and Algorithms in JAX

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.321654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.321654Z digest=sha256:f2c0d95ce5195deb4e0c191c83a5d2b9ebcd6c79dcc7d599bec69e0c59d50624

Observation 45ed4b3f-9170-4581-969c-a16ff0bef7a7 · outbound

This paper cites The StarCraft Multi-Agent Challenge.

Remembering the Markov Property in Cooperative MARL The StarCraft Multi-Agent Challenge

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.335420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.335420Z digest=sha256:7c7db89d3c3740cccf9d48691c123e502d49f794e39cd83a8a945cbfccc9f7e1

Observation 4824165f-4a2a-4b82-bedc-ec848fa0da91 · outbound

This paper cites Sutton and Andrew G.

Remembering the Markov Property in Cooperative MARL Sutton and Andrew G

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.344743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.344743Z digest=sha256:2dfbea61a4da4e6cb31959ac70260d258c2d3c939ee93d28ade607b09802a45e

Observation ce931b36-4ec8-4cbe-8a96-ee0ae306df07 · outbound

This paper cites Order matters: Agent-by-agent policy optimization.

Remembering the Markov Property in Cooperative MARL Order matters: Agent-by-agent policy optimization

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:45.853665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.358732Z digest=sha256:9fa5b6f0a2daef5da665d41453f173391d16aacd0ca9d6239637b20007a266f3

Observation 04b3004c-d772-4bcc-a29e-1e54b140abac · outbound

This paper cites Emergence of maps in the memories of blind navigation agents.

Remembering the Markov Property in Cooperative MARL Emergence of maps in the memories of blind navigation agents

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:45.834884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.371678Z digest=sha256:e3e626a1a808e7723891e08d22b75c3c36c28d909bed6acc87b23481e2194464

Observation be67f0dc-42f3-489e-86ca-122c136aa495 · outbound

This paper cites Learning latent representations to influence multi-agent interaction.

Remembering the Markov Property in Cooperative MARL Learning latent representations to influence multi-agent interaction

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:45.812784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.377556Z digest=sha256:f72a230e6a615eabb7728e9a11668430ae5a18e2d055a4c6ebf7929f070ca2ec

Observation 95e00ed8-4c54-4fbc-bc6c-e5f12e2e37ea · outbound

This paper cites The surprising effectiveness of ppo in cooperative multi-agent games.

Remembering the Markov Property in Cooperative MARL The surprising effectiveness of ppo in cooperative multi-agent games

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.387404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.387404Z digest=sha256:28da80943d60d22b790176d666de2f537019eac6f643b6e0557af574c4daa3c8

Observation 8d1b8beb-e096-4cce-947e-48a0ebd00ea3 · outbound

This paper cites Heterogeneous-agent reinforcement learning.

Remembering the Markov Property in Cooperative MARL Heterogeneous-agent reinforcement learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.393733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.393733Z digest=sha256:6332be2f75e2865c37cb6f6620ccbebe503a2d503f3e1d831d7ee97410d57400

Observation da94f39f-e44c-43f5-9bee-21cc1147454d · outbound

This paper cites Deep interactive bayesian reinforcement learning via meta-learning.

Remembering the Markov Property in Cooperative MARL Deep interactive bayesian reinforcement learning via meta-learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:20:45.757902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-15T18:20:45.408594Z digest=sha256:19133bbd21dad27c3e7f713d07a7361aacadb6743b08d50dae108eead81c6fc6

Observation 13b01a95-80e2-4ca4-9965-6cdd89b796e4 · outbound

This paper cites write newline.

Remembering the Markov Property in Cooperative MARL write newline

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T18:20:45.426207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:20:45.426207Z digest=sha256:a658e9a68f2616be69141a3e2a19d6076aa192ea998fa01ab5d9fa09944e0e94

Pith citing papers

No inbound Pith citation observations are available.