Pith. sign in

Paper Citation Record · LEDGER

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance

As of 21 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 0 inbound Pith citation observations for arXiv:2501.10593.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.10593 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T19:06:27.178808Z

measured 37 of 37 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

37 of 37 outbound references displayed

  • verified exact0
  • verified fuzzy23
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d537b2b4-f8f2-4c17-9e38-338ea14f1ee1 · outbound

This paper cites Basis for intentions: Efficient inverse reinforcement learning using past experience, 2022.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Basis for intentions: Efficient inverse reinforcement learning using past experience, 2022

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:29.064232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:26.736762Z digest=sha256:4eb409361ce4580594e92590bb9878d29a647934ddef275b2295a12400f2838c

Observation 7a89f7a3-5005-4e9b-91da-30d964261d5b · outbound

This paper cites Albrecht and Subramanian Ramamoorthy.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Albrecht and Subramanian Ramamoorthy

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:29.017281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:26.746894Z digest=sha256:0d46d3b11fe86faab1d476821d7710a1dcd0b2acc997faf6cbf8e83c3fddced0

Observation f3242f02-021c-4f88-961e-d0361f63e8e7 · outbound

This paper cites On the Utility of Learning about Humans for Human-AI Coordination.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance On the Utility of Learning about Humans for Human-AI Coordination

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:26.752581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:26.752581Z digest=sha256:8854f24655132955458d37f9841a74227b881e5b62dbed09afaf5d132defb7e7

Observation 28190281-ef35-429a-b64f-fe07fe9e79e9 · outbound

This paper cites Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Is Independent Learning All You Need in the StarCraft Multi-Agent Challenge?

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:26.778989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:26.778989Z digest=sha256:34c82fac29689c6969331ec73dc06863a62e69973be451ee1a5d5d0608010de3

Observation 4ad79d62-4ff7-4c76-acc4-ab5a830543dd · outbound

This paper cites Emergent complexity and zero-shot transfer via unsupervised environment design, 2021.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Emergent complexity and zero-shot transfer via unsupervised environment design, 2021

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.978784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:26.786278Z digest=sha256:4a0837dc721251a8618f6a69b7ba523676de379e7813e562b75f72a30d617137

Observation 696fb80f-8639-428e-8cbd-89b48b18b58c · outbound

This paper cites Counterfactual multi-agent policy gradients, 2017.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Counterfactual multi-agent policy gradients, 2017

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.943528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:26.818117Z digest=sha256:f99e34cdf6ef3159fa8ee5932c32b4acae8cb9ea1e35ddbbda413ea0185b3819

Observation 18b4baed-7607-44b5-91c4-13a243e35499 · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor, 2018.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor, 2018

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:26.828973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:26.828973Z digest=sha256:39e8c77535e8e75bf31a07ce055bc1fa062203cfcd936b471c4c4021fa15bc2c

Observation 3694a783-74b9-491e-940f-8dc193a0fc09 · outbound

This paper cites The evolution of cultural evolution.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance The evolution of cultural evolution

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.881357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:26.836328Z digest=sha256:a218a239f42c7f5548299b7707b29f4d0be530e7e28e9c8b83ef52951411b309

Observation ba55ccc7-2f1a-4dcd-b82e-cada4a6cc873 · outbound

This paper cites an unresolved cited work.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-10T19:06:28.836315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:26.846414Z digest=sha256:62497daa73f07be8ab289f19962ea5ddd53cf2795d489c83ff93ff41099e9715

Observation 861f1a24-9d73-4110-b6b9-ea501e3081fd · outbound

This paper cites Reinforcement learning with unsupervised auxiliary tasks, 2016.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Reinforcement learning with unsupervised auxiliary tasks, 2016

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.781669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:26.857339Z digest=sha256:5d48f18f9b0eb51c00bf5db6ee7781a53bfcb5e245dbf1082f15a479f589f921

Observation fb0918b5-983d-4fa2-8d0e-2c38e1180c19 · outbound

This paper cites Czarnecki, Iain Dunning, Luke Marris, Guy Lever, Antonio Garc´ıa Casta˜neda, Charlie Beattie, Neil C.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Czarnecki, Iain Dunning, Luke Marris, Guy Lever, Antonio Garc´ıa Casta˜neda, Charlie Beattie, Neil C

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.721807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:26.868220Z digest=sha256:0740b313449a69b0086592a6e593d6c494805c2bc405c0a2075954b4cc4edf86

Observation f7a2753b-44eb-4b42-bcdf-5aa763ea501e · outbound

This paper cites Recursive bayesian human intent recognition in shared-control robotics.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Recursive bayesian human intent recognition in shared-control robotics

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:26.880253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:26.880253Z digest=sha256:c767b2674e37731041dfedc8d8fa3d944852b13bd18ff909ee9737f2a71183a0

Observation c21e5ce3-3bbc-4023-9cfc-b3d1657398eb · outbound

This paper cites Losey, and Dorsa Sadigh.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Losey, and Dorsa Sadigh

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.670905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:26.888108Z digest=sha256:6a52f26b46d506aa979573457b0c6d92fc875f88d67350b786be3a31555bc81f

Observation d077819f-cd94-43bc-98e4-fa3b906f03d8 · outbound

This paper cites Learning dynamics model in reinforcement learning by incorporating the long term future, 2019.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Learning dynamics model in reinforcement learning by incorporating the long term future, 2019

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.635298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:26.905877Z digest=sha256:3d0e7e0809c61066662d1a8d3b938a7af038a9526cdba33347a9744d9e070940

Observation 6d49a3f2-990c-4dd7-84ad-83ba94f7f69a · outbound

This paper cites Multi-agent reinforcement learning with multi- step generative models, 2019.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Multi-agent reinforcement learning with multi- step generative models, 2019

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.598765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:26.911966Z digest=sha256:05799c1f058078166b03c7c69e8f3d06fd87f42380f2ea1b2cc0977cde991ec7

Observation 6161702c-d913-4fa4-8997-27b3a7f2f943 · outbound

This paper cites Social learning strategies.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Social learning strategies

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:26.917903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:26.917903Z digest=sha256:78d530d38afd91eb7ab292a0005a0daaebdd35698f4151ed1df7b63ccf1b2c29

Observation a56f853e-6062-473d-b088-c8eaa831ca3b · outbound

This paper cites Generalization and network design strategies.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Generalization and network design strategies

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.508780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:26.923958Z digest=sha256:08ef6e7f0031f3d0bbaeb937fab08feb930dac2bdeec2203fcf4cb2c1d59014f

Observation 129e91d5-557d-48d8-9146-3c3e908aa58e · outbound

This paper cites Rectifier nonlinearities improve neural network acoustic models.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Rectifier nonlinearities improve neural network acoustic models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.453366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:26.929333Z digest=sha256:d4fb985fd72723f3e7aae6f9ed6ab147c0da027037083b72c008170806979025

Observation c26c8b57-50ca-4be1-b317-1a35c8410603 · outbound

This paper cites Emergence of grounded compositional language in multi- agent populations, 2018.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Emergence of grounded compositional language in multi- agent populations, 2018

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.370840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:26.936560Z digest=sha256:c57b0451809f8958b037f922f3631df6bc3f93aa9acb06c1e5f34103669c319a

Observation b61c1c8b-7e1b-4ca6-92ba-9a9ae73d3ceb · outbound

This paper cites Emergent social learning via multi-agent reinforcement learning, 2021.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Emergent social learning via multi-agent reinforcement learning, 2021

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.332017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:26.947585Z digest=sha256:a534db23786d2bfa6f556e21b78c0be38f890b6a9373cd8a6cf50d6189942ce5

Observation eeae4ba5-5d80-493a-a368-b540d19c68b1 · outbound

This paper cites Srinivasa.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Srinivasa

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.296403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:26.962941Z digest=sha256:1b658b0025dbc03302391045360b13ba8b0c1c156663aeb52d16ca60a866bdf0

Observation f00b01da-f6db-4277-9387-0382f7875662 · outbound

This paper cites Albrecht.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Albrecht

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.250550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:26.982192Z digest=sha256:56d12d80c76af34031a00611de4cc344332fad1fe1ad429097f63a07fdcbe620

Observation dc68b850-5a84-4e86-98e7-7f65d2f45bb3 · outbound

This paper cites Tenenbaum, Sanja Fidler, and Antonio Torralba.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Tenenbaum, Sanja Fidler, and Antonio Torralba

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.181958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:27.011753Z digest=sha256:faa4084996f631913054374ac3778f3341c2cef7f3e4b3f7efc22160056c5521

Observation b08e9966-4202-4252-be82-8c096ead850a · outbound

This paper cites Modeling others using oneself in multi-agent reinforcement learning, 2018.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Modeling others using oneself in multi-agent reinforcement learning, 2018

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:28.125248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:27.026167Z digest=sha256:8b7f84e05694b084d13d6d81e3145f084bc66ff2c058c6c000bb735078508112

Observation 2673280f-6d84-4a65-a09b-c6cee5f12feb · outbound

This paper cites Proximal policy optimization algorithms, 2017.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Proximal policy optimization algorithms, 2017

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:27.034876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:27.034876Z digest=sha256:e8ece1fd18d48a5df1d1c766e7292e049aa232b36463a26f94298b5c12efbf1d

Observation af535b3f-c668-4f6e-a907-9bba527059b1 · outbound

This paper cites Loss is its own reward: Self-supervision for reinforcement learning, 2017.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Loss is its own reward: Self-supervision for reinforcement learning, 2017

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:27.994120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:27.046728Z digest=sha256:30a267c9eee800d240c21e1e47b667efa09c05a252c7b7a0b281c5fa075dee10

Observation 96d206e5-8946-4903-8469-b644aa0c6ee6 · outbound

This paper cites Zhao, Archit Sharma, Karl Pertsch, Jianlan Luo, Sergey Levine, and Chelsea Finn.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Zhao, Archit Sharma, Karl Pertsch, Jianlan Luo, Sergey Levine, and Chelsea Finn

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:27.058046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:27.058046Z digest=sha256:fab7ea68ac1ce1518c8dd3d4c217acc40647ac8156eeba857f98c8ae013698b8

Observation 1cbbcdb1-cd8c-42ef-b820-3a27a3bb2390 · outbound

This paper cites Message-passing approach for threshold models of behavior in networks.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Message-passing approach for threshold models of behavior in networks

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:27.072016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:27.072016Z digest=sha256:a3b170b16cf95a7ed10234c7057f9c2fe85746603b8d5408d7bdcb7101a31197

Observation 4068f696-d52e-414c-b6f1-74108fd1f5e1 · outbound

This paper cites Mastering chess and shogi by self-play with a general reinforcement learning algorithm, 2017.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Mastering chess and shogi by self-play with a general reinforcement learning algorithm, 2017

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:27.080776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:27.080776Z digest=sha256:49fdfa04a2ac236fc44e0e4cac276d41f382bcf3f8be740ea90e27561a978fee

Observation 0e5e48d1-68a2-4b52-bbe4-7563b449296d · outbound

This paper cites Smallwood and Edward J.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Smallwood and Edward J

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:27.918249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:27.094921Z digest=sha256:cb29479654bbd3960e92d3c79d7f32c484f43ea706d9768361731df280ea577d

Observation 6d44c7e6-4d02-4634-baa4-49da968a5114 · outbound

This paper cites McKee, Matt Botvinick, Edward Hughes, and Richard Everett.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance McKee, Matt Botvinick, Edward Hughes, and Richard Everett

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:27.878794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:27.103156Z digest=sha256:ad130c450d1aae3998f02d5ee54ed1baa9f4fbc6582b7d8dd40a23ca17031e4d

Observation be4d5c22-43f9-46aa-8646-0eff70c945bb · outbound

This paper cites Pettingzoo: Gym for multi-agent reinforcement learning.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Pettingzoo: Gym for multi-agent reinforcement learning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:27.109331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:27.109331Z digest=sha256:dbae1339250d89e8aa395743aadda46f23cc7d5cd027f22010258ea62e186637

Observation fe531969-7b1a-44a8-9409-6e6f0018e467 · outbound

This paper cites Paired Open-Ended Trailblazer (POET): Endlessly Generating Increasingly Complex and Diverse Learning Environments and Their Solutions.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Paired Open-Ended Trailblazer (POET): Endlessly Generating Increasingly Complex and Diverse Learning Environments and Their Solutions

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:27.125935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:27.125935Z digest=sha256:fa175a8695b9260362dae2acd4d99e5f2582a79a0ad81a095b831e079cfe9172

Observation dc7e7e4a-e1b2-4294-bcfd-d658448eb67e · outbound

This paper cites an unresolved cited work.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-10T19:06:27.831658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:27.151138Z digest=sha256:b00564fc91a56ca3d2453d527aaf20785c5918e3b983972003395df3de9c71f1

Observation 10948f30-246e-4d58-bc53-08f5d4e22f80 · outbound

This paper cites Towards generalizability of multi-agent reinforcement learning in graphs with recurrent message passing, 2024.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Towards generalizability of multi-agent reinforcement learning in graphs with recurrent message passing, 2024

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:27.782180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:27.166033Z digest=sha256:14e758192b3e193d20c2acf99ba808f22cf015e4fdb9e1366e2e230c73f9a16a

Observation 12201486-51f9-4452-a6b7-4276c427015c · outbound

This paper cites The surprising effectiveness of ppo in cooperative, multi-agent games, 2021.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance The surprising effectiveness of ppo in cooperative, multi-agent games, 2021

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:06:27.717210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T19:06:27.178808Z digest=sha256:926cf79caa15dfe41fe1b9b0237b15c7f85b7e03433da3e137b573c1f471b541

Observation 58a30f00-1af9-4bc3-bc2e-861941e7d3c0 · outbound

This paper cites Benchmarking Multi-Agent Deep Reinforcement Learning Algorithms in Cooperative Tasks.

ColorGrid: A Multi-Agent Non-Stationary Environment for Goal Inference and Assistance Benchmarking Multi-Agent Deep Reinforcement Learning Algorithms in Cooperative Tasks

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:26.997151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:06:26.997151Z digest=sha256:76afde0d87870441457b5bc6c9bf79e329ab98c3a7326b06d532ce8d446f2240

Pith citing papers

No inbound Pith citation observations are available.