Pith. sign in

Paper Citation Record · LEDGER

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings

As of 22 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 0 inbound Pith citation observations for arXiv:2604.25076.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.25076 v1

Coverage vector

measured 20 of 20 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-07T17:09:37.436426Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

20 of 20 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved3
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ba86492d-ac84-4089-b7fb-f327fc67c991 · outbound

This paper cites Learning to play trajectory games against opponents with unknown objectives.

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings Learning to play trajectory games against opponents with unknown objectives

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:23:21.648082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:09:37.436426Z digest=sha256:327c7a8d170209e8ba121a55251c8bb1e2b2fdb3dfa1d0c28ab4dcf67d39f21f

Observation fb8d3454-5367-4b5b-86fe-a3c0fd45099a · outbound

This paper cites Foerster, Francis Song, Edward Hughes, Neil Burch, Iain Dunning, Shimon Whiteson, Matthew Botvinick, and Michael Bowling.

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings Foerster, Francis Song, Edward Hughes, Neil Burch, Iain Dunning, Shimon Whiteson, Matthew Botvinick, and Michael Bowling

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:23:21.652259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:09:37.436426Z digest=sha256:c7544f02df82fec163cae705ffe3c57a05663e8a1f345597e16ef1f9cead41e3

Observation 1cf38d66-d5f3-40ee-ba85-a75c3fa88b91 · outbound

This paper cites Learning to play sequential games versus unknown opponents.

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings Learning to play sequential games versus unknown opponents

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:23:21.656188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:09:37.436426Z digest=sha256:e4ef63805f44f5e96e8d8c0fc515ed209a6e1452ce7022ac963b75ab6d6714bc

Observation 9f77bfa5-104a-4fee-a445-c7449f4e72eb · outbound

This paper cites At-drone: Benchmarking adaptive teaming in multi-drone pursuit.

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings At-drone: Benchmarking adaptive teaming in multi-drone pursuit

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:23:21.664032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:09:37.436426Z digest=sha256:e47a878e479a73c23add643f07a5cff59c948833a521568cf61a2cee508b977c

Observation 53582493-ba4f-467d-96f3-63a980d7fae0 · outbound

This paper cites other-play.

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings other-play

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:23:21.675583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:09:37.436426Z digest=sha256:ca94842daf862fd9cbb6e01b9ac98cc6548907a4c9b9f93729cdc95f9be967dc

Observation af2eba3d-71ba-46cb-967f-1da53279a836 · outbound

This paper cites Td-gammon, a self-teaching backgammon program, achieves master-level play.Neural Comput., 6(2):215–219, March 1994.

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings Td-gammon, a self-teaching backgammon program, achieves master-level play.Neural Comput., 6(2):215–219, March 1994

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:23:21.660114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:09:37.436426Z digest=sha256:7a73763b2a3a73dbf6538af5ae0a9e355500971dec87522a7612062e63b1a279

Observation aa1e39ef-0050-4b25-8c80-506d6faf1b4e · outbound

This paper cites Trajectory diversity for zero-shot coordination.

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings Trajectory diversity for zero-shot coordination

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:23:21.667966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:09:37.436426Z digest=sha256:3457fc3e20e3179ea51fb5bee4f9ae04d78f70262c52dd6ca1a5994d8b9c996d

Observation 583d1763-c9b0-4748-944e-005e6c9f8d5d · outbound

This paper cites Heterogeneous multi-agent zero-shot coordination by coevolution.

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings Heterogeneous multi-agent zero-shot coordination by coevolution

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:23:21.671797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:09:37.436426Z digest=sha256:3481b2962d60ff237e810fad748fe381ecec3ed6cff78822a73bb7d89bbeb75a

Observation 8112b080-639a-41f4-b15b-05e4751d44b8 · outbound

This paper cites Comprehensive overview of reward engineering and shaping in advancing reinforcement learning applications.

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings Comprehensive overview of reward engineering and shaping in advancing reinforcement learning applications

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:23:21.682649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:09:37.436426Z digest=sha256:6858c1cf10145fb1b4e2af7e78f161428e94723f74276734f6a726413a23a105

Observation 25c82948-7241-4cde-8839-13ffad695ad8 · outbound

This paper cites Learning zero-shot cooperation with humans, assuming humans are biased.

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings Learning zero-shot cooperation with humans, assuming humans are biased

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:23:21.693922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:09:37.436426Z digest=sha256:3b4194cfdb987f545040582d307db64abb100177b234d56a6bf4e922e640f0c8

Observation 3ffcbb42-4eba-47d6-b24c-2e85a2e07c5d · outbound

This paper cites an unresolved cited work.

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-05-27T01:23:21.678763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:09:37.436426Z digest=sha256:81f42138d1be791b9891e6de6bec7f0c997c258dd5a4203bad6bd469b2982b03

Observation 3aa11d4a-850c-4a5e-8ea1-0434a61a0f18 · outbound

This paper cites Theory of mind for deep reinforcement learning in hanabi.

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings Theory of mind for deep reinforcement learning in hanabi

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:23:21.716786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:09:37.436426Z digest=sha256:9dddbc3113264da38da49d4cb928cad498fb63b3ae251abf51c8b07ecc844f2d

Observation 4a39509b-55ac-4f28-942f-80d4baba3dbe · outbound

This paper cites Ho, Thomas L.

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings Ho, Thomas L

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:23:21.708444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:09:37.436426Z digest=sha256:b41d388e53a46660af68f98cbeb01ae827b85751c866ee646b8f6d69c40932a0

Observation cdedd4f5-7700-469a-9d7e-d68e87b66153 · outbound

This paper cites Eureka: Human-level reward design via coding large language models.

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings Eureka: Human-level reward design via coding large language models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:23:21.686290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:09:37.436426Z digest=sha256:968e5a61f4479b4240ae62dbb7c6fb1e6ff3cc3fb16128a304f05ec03e6f0566

Observation 8952262a-baa7-468b-ad5a-a39f158d47e3 · outbound

This paper cites an unresolved cited work.

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-05-27T01:23:21.720430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:09:37.436426Z digest=sha256:f3e969a2dccd1412c8e7e47ed846b84160b554f6d3e123e7b1cd85d25d69f5b1

Observation a881e72a-e3ed-469c-a987-4fc8e17d5308 · outbound

This paper cites an unresolved cited work.

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-05-27T01:23:21.712796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:09:37.436426Z digest=sha256:a378093e5bb55dba62d736a74c48d1d4f3244ec4e1c2e804b551a6e39c7d5a6d

Observation 24d8edeb-a4fe-499a-b1dc-7e89241f8a7e · outbound

This paper cites The surprising effectiveness of ppo in cooperative, multi-agent games.

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings The surprising effectiveness of ppo in cooperative, multi-agent games

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:23:21.704713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:09:37.436426Z digest=sha256:7b54e5a5c4b952795a8c7bbee42a98ea039a0b4f847360e7342bc913f394ca1b

Observation de521ac0-c31b-47f0-9068-92d76a5262a3 · outbound

This paper cites Zsc-eval: An evaluation toolkit and benchmark for multi-agent zero-shot coordination.

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings Zsc-eval: An evaluation toolkit and benchmark for multi-agent zero-shot coordination

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:23:21.689778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:09:37.436426Z digest=sha256:ddb54fa5a696ffec522dfe1952a3516f91a561cd6f06c1b0e8e95c3911892964

Observation db19eda7-a250-4443-9ca6-db1ca9baef20 · outbound

This paper cites An efficient end-to-end training approach for zero-shot human-ai coordination.

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings An efficient end-to-end training approach for zero-shot human-ai coordination

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:23:21.697549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:09:37.436426Z digest=sha256:73a516d4556c510a21201390f0480c4d38fd6bb45102aa1e3acd9cb3365f422d

Observation 13431839-d067-490b-ac2d-c31e98639d76 · outbound

This paper cites folder".

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings folder"

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T01:23:21.701268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:09:37.436426Z digest=sha256:f8815503219065dc051ab46717a3de55ce82d53c5d69711f7fe3507fcd45c2eb

Pith citing papers

No inbound Pith citation observations are available.