Pith. sign in

Paper Citation Record · LEDGER

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance

As of 21 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 0 inbound Pith citation observations for arXiv:2501.10593.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.10593 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T19:06:27.178808Z

measured 37 of 37 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

37 of 37 outbound references displayed

  • verified exact0
  • verified fuzzy23
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d537b2b4-f8f2-4c17-9e38-338ea14f1ee1 · outbound

This paper cites Basis for intentions: Efficient inverse reinforcement learning using past experience, 2022.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Basis for intentions: Efficient inverse reinforcement learning using past experience, 2022

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:29.064232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:26.736762Z digest=sha256:91f62a9a831ad87200a3f24e60d8c0614fb2e27e396c72ae9550a023897f5fcc

Observation 7a89f7a3-5005-4e9b-91da-30d964261d5b · outbound

This paper cites Albrecht and Subramanian Ramamoorthy.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Albrecht and Subramanian Ramamoorthy

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:29.017281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:26.746894Z digest=sha256:32d4beccf89875819b00b14c3f42e6791fc0da1d01869ed250429bce8c1ddd1e

Observation f3242f02-021c-4f88-961e-d0361f63e8e7 · outbound

This paper cites On the Utility of Learning about Humans for Human-AI Coordination.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance On the Utility of Learning about Humans for Human-AI Coordination

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:26.752581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:26.752581Z digest=sha256:03c6cea6109349546838ea3b358d70d9dbcfce1dace1432d89ee445816bfddc1

Observation 28190281-ef35-429a-b64f-fe07fe9e79e9 · outbound

This paper cites Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:26.778989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:26.778989Z digest=sha256:34c82fac29689c6969331ec73dc06863a62e69973be451ee1a5d5d0608010de3

Observation 4ad79d62-4ff7-4c76-acc4-ab5a830543dd · outbound

This paper cites Emergent complexity and zero-shot transfer via unsupervised environment design, 2021.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Emergent complexity and zero-shot transfer via unsupervised environment design, 2021

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.978784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:26.786278Z digest=sha256:833de4fddcd41550e70b9cf1f35e82939b045c19fc801d3a2fad3dc558d5234f

Observation 696fb80f-8639-428e-8cbd-89b48b18b58c · outbound

This paper cites Counterfactual multi-agent policy gradients, 2017.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Counterfactual multi-agent policy gradients, 2017

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.943528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:26.818117Z digest=sha256:2648de633ea8fe2abfdeed7286b2a60e4f9ce3462ec385d42a4735b297304cd5

Observation 18b4baed-7607-44b5-91c4-13a243e35499 · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor, 2018.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor, 2018

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:26.828973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:26.828973Z digest=sha256:39e8c77535e8e75bf31a07ce055bc1fa062203cfcd936b471c4c4021fa15bc2c

Observation 3694a783-74b9-491e-940f-8dc193a0fc09 · outbound

This paper cites The evolution of cultural evolution.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance The evolution of cultural evolution

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.881357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:26.836328Z digest=sha256:74c2b230e2a50ea4087f1d7ac87d158351cbfaaa419f795f455b4b378a32b300

Observation ba55ccc7-2f1a-4dcd-b82e-cada4a6cc873 · outbound

This paper cites an unresolved cited work.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-10T19:06:28.836315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:26.846414Z digest=sha256:7115ee21553adbc32e33ca1d5c918d028846bd444c4fef791c147c8d1bc3b38f

Observation 861f1a24-9d73-4110-b6b9-ea501e3081fd · outbound

This paper cites Reinforcement learning with unsupervised auxiliary tasks, 2016.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Reinforcement learning with unsupervised auxiliary tasks, 2016

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.781669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:26.857339Z digest=sha256:57d2e8e40e8f16196faec30e9c29e362faa4fb83f39a0189cdb04f3945473578

Observation fb0918b5-983d-4fa2-8d0e-2c38e1180c19 · outbound

This paper cites Czarnecki, Iain Dunning, Luke Marris, Guy Lever, Antonio Garc´ıa Casta˜neda, Charlie Beattie, Neil C.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Czarnecki, Iain Dunning, Luke Marris, Guy Lever, Antonio Garc´ıa Casta˜neda, Charlie Beattie, Neil C

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.721807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:26.868220Z digest=sha256:159d85a226dd4c71d29c2eacfda7d33d1fa606d90b7fc111d5f9f35661c409f9

Observation f7a2753b-44eb-4b42-bcdf-5aa763ea501e · outbound

This paper cites Recursive bayesian human intent recognition in shared-control robotics.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Recursive bayesian human intent recognition in shared-control robotics

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:26.880253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:26.880253Z digest=sha256:c767b2674e37731041dfedc8d8fa3d944852b13bd18ff909ee9737f2a71183a0

Observation c21e5ce3-3bbc-4023-9cfc-b3d1657398eb · outbound

This paper cites Losey, and Dorsa Sadigh.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Losey, and Dorsa Sadigh

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.670905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:26.888108Z digest=sha256:067183f17fe55a86e8eb127df907c4bc64815b9c6270113dba6fda412622108c

Observation d077819f-cd94-43bc-98e4-fa3b906f03d8 · outbound

This paper cites Learning dynamics model in reinforcement learning by incorporating the long term future, 2019.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Learning dynamics model in reinforcement learning by incorporating the long term future, 2019

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.635298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:26.905877Z digest=sha256:1a985ccce04392dbe126c6178c03ff14bea3635796c782d67d376c2ca0a95981

Observation 6d49a3f2-990c-4dd7-84ad-83ba94f7f69a · outbound

This paper cites Multi-agent reinforcement learning with multi- step generative models, 2019.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Multi-agent reinforcement learning with multi- step generative models, 2019

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.598765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:26.911966Z digest=sha256:eed1bac3851e3c31e0df7a568bafb6f8114726d83839af66026630fdb6b00e6c

Observation 6161702c-d913-4fa4-8997-27b3a7f2f943 · outbound

This paper cites Social learning strategies.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Social learning strategies

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:26.917903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:26.917903Z digest=sha256:78d530d38afd91eb7ab292a0005a0daaebdd35698f4151ed1df7b63ccf1b2c29

Observation a56f853e-6062-473d-b088-c8eaa831ca3b · outbound

This paper cites Generalization and network design strategies.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Generalization and network design strategies

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.508780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:26.923958Z digest=sha256:77f735fe43e713adbaf14b5195889cbfb4eed031260f128112990f31775ee31d

Observation 129e91d5-557d-48d8-9146-3c3e908aa58e · outbound

This paper cites Rectifier nonlinearities improve neural network acoustic models.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Rectifier nonlinearities improve neural network acoustic models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.453366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:26.929333Z digest=sha256:1aa57926cf45072bc3acfdc6fdd49b4209996994566f21a67029b393a768c1ae

Observation c26c8b57-50ca-4be1-b317-1a35c8410603 · outbound

This paper cites Emergence of grounded compositional language in multi- agent populations, 2018.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Emergence of grounded compositional language in multi- agent populations, 2018

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.370840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:26.936560Z digest=sha256:b672b84711c8296fafe7ebe6dc649449973f6554aa79249c89d3b52a752bff68

Observation b61c1c8b-7e1b-4ca6-92ba-9a9ae73d3ceb · outbound

This paper cites Emergent social learning via multi-agent reinforcement learning, 2021.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Emergent social learning via multi-agent reinforcement learning, 2021

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.332017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:26.947585Z digest=sha256:2caf569af6e724bb8dddedd788c1b19d910ae69ae960d34ea63190f12b9bca75

Observation eeae4ba5-5d80-493a-a368-b540d19c68b1 · outbound

This paper cites Srinivasa.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Srinivasa

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.296403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:26.962941Z digest=sha256:d214614ea9a902b4a3fbaa40e05afb5868532722e3368ec20573d0f70ed211fc

Observation f00b01da-f6db-4277-9387-0382f7875662 · outbound

This paper cites Albrecht.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Albrecht

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.250550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:26.982192Z digest=sha256:91f59f4a62dd591a4092af2e7b63c720c0fd1f784529ac5a1f1e352e8bd856fc

Observation dc68b850-5a84-4e86-98e7-7f65d2f45bb3 · outbound

This paper cites Tenenbaum, Sanja Fidler, and Antonio Torralba.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Tenenbaum, Sanja Fidler, and Antonio Torralba

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.181958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:27.011753Z digest=sha256:acd2e20a5b5c8370c95b512968284084ce98c368d21b6c407b8ddd3337efd74b

Observation b08e9966-4202-4252-be82-8c096ead850a · outbound

This paper cites Modeling others using oneself in multi-agent reinforcement learning, 2018.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Modeling others using oneself in multi-agent reinforcement learning, 2018

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.125248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:27.026167Z digest=sha256:4f88d2ebbe345179668f023852b63b83fe8d9b94fe01df62116ba2d75120c551

Observation 2673280f-6d84-4a65-a09b-c6cee5f12feb · outbound

This paper cites Proximal policy optimization algorithms, 2017.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Proximal policy optimization algorithms, 2017

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:27.034876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:27.034876Z digest=sha256:e8ece1fd18d48a5df1d1c766e7292e049aa232b36463a26f94298b5c12efbf1d

Observation af535b3f-c668-4f6e-a907-9bba527059b1 · outbound

This paper cites Loss is its own reward: Self-supervision for reinforcement learning, 2017.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Loss is its own reward: Self-supervision for reinforcement learning, 2017

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:27.994120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:27.046728Z digest=sha256:961890e0cc9050d5230bf7cdabd0e25ea1242391f8df7b687b6362bb0c81a8e1

Observation 96d206e5-8946-4903-8469-b644aa0c6ee6 · outbound

This paper cites Zhao, Archit Sharma, Karl Pertsch, Jianlan Luo, Sergey Levine, and Chelsea Finn.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Zhao, Archit Sharma, Karl Pertsch, Jianlan Luo, Sergey Levine, and Chelsea Finn

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:27.058046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:27.058046Z digest=sha256:fab7ea68ac1ce1518c8dd3d4c217acc40647ac8156eeba857f98c8ae013698b8

Observation 1cbbcdb1-cd8c-42ef-b820-3a27a3bb2390 · outbound

This paper cites Message-passing approach for threshold models of behavior in networks.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Message-passing approach for threshold models of behavior in networks

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:27.072016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:27.072016Z digest=sha256:a3b170b16cf95a7ed10234c7057f9c2fe85746603b8d5408d7bdcb7101a31197

Observation 4068f696-d52e-414c-b6f1-74108fd1f5e1 · outbound

This paper cites Mastering chess and shogi by self-play with a general reinforcement learning algorithm, 2017.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Mastering chess and shogi by self-play with a general reinforcement learning algorithm, 2017

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:27.080776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:27.080776Z digest=sha256:49fdfa04a2ac236fc44e0e4cac276d41f382bcf3f8be740ea90e27561a978fee

Observation 0e5e48d1-68a2-4b52-bbe4-7563b449296d · outbound

This paper cites Smallwood and Edward J.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Smallwood and Edward J

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:27.918249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:27.094921Z digest=sha256:7e80bdf9da57fc5ba4dd7ec3f5e03ba7cb2b1719b15aff779819cef25fff42c1

Observation 6d44c7e6-4d02-4634-baa4-49da968a5114 · outbound

This paper cites McKee, Matt Botvinick, Edward Hughes, and Richard Everett.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance McKee, Matt Botvinick, Edward Hughes, and Richard Everett

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:27.878794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:27.103156Z digest=sha256:434e92cf17bfc36d59909d440f4c397aaec20b1df8548713c24609ee99cb9411

Observation be4d5c22-43f9-46aa-8646-0eff70c945bb · outbound

This paper cites Pettingzoo: Gym for multi-agent reinforcement learning.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Pettingzoo: Gym for multi-agent reinforcement learning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:27.109331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:27.109331Z digest=sha256:dbae1339250d89e8aa395743aadda46f23cc7d5cd027f22010258ea62e186637

Observation fe531969-7b1a-44a8-9409-6e6f0018e467 · outbound

This paper cites Paired Open-Ended Trailblazer (POET): Endlessly Generating Increasingly Complex and Diverse Learning Environments and Their Solutions.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Paired Open-Ended Trailblazer (POET): Endlessly Generating Increasingly Complex and Diverse Learning Environments and Their Solutions

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:27.125935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:27.125935Z digest=sha256:fa175a8695b9260362dae2acd4d99e5f2582a79a0ad81a095b831e079cfe9172

Observation dc7e7e4a-e1b2-4294-bcfd-d658448eb67e · outbound

This paper cites an unresolved cited work.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-10T19:06:27.831658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:27.151138Z digest=sha256:5c2c7694367b9005845f43d72a001c37b70a99a1b9564fec3611dec19ad48a2d

Observation 10948f30-246e-4d58-bc53-08f5d4e22f80 · outbound

This paper cites Towards generalizability of multi-agent reinforcement learning in graphs with recurrent message passing, 2024.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Towards generalizability of multi-agent reinforcement learning in graphs with recurrent message passing, 2024

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:27.782180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:27.166033Z digest=sha256:31d5dcfdd0f6a94366458b69252dd5ea51d41bbbab6a8b23a30c5ea75cb3dbbf

Observation 12201486-51f9-4452-a6b7-4276c427015c · outbound

This paper cites The surprising effectiveness of ppo in cooperative, multi-agent games, 2021.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance The surprising effectiveness of ppo in cooperative, multi-agent games, 2021

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:27.717210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:06:27.178808Z digest=sha256:a8a869e7fea13d6bbea6ca08dcd4ccbe387549d3e5b09b71deceff4fe684e1d1

Observation 58a30f00-1af9-4bc3-bc2e-861941e7d3c0 · outbound

This paper cites Benchmarking Multi-Agent Deep Reinforcement Learning Algorithms in Cooperative Tasks.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Benchmarking Multi-Agent Deep Reinforcement Learning Algorithms in Cooperative Tasks

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:26.997151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:26.997151Z digest=sha256:76afde0d87870441457b5bc6c9bf79e329ab98c3a7326b06d532ce8d446f2240

Pith citing papers

No inbound Pith citation observations are available.