Pith. sign in

Paper Citation Record · LEDGER

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork

As of 10 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2605.24423.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.24423 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-30T13:41:33.048760Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact1
  • verified fuzzy8
  • unresolved1
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2994aa2f-ea39-486b-9bbb-c0ae1fbbb9ca · outbound

This paper cites org/CorpusID:258845718.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork org/CorpusID:258845718

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.377952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:1ea048a1432247d2fa5164277e692b354f4470bc0ccd0b803b0aef1eb1606177

Observation 265e1385-03b7-47fa-b88a-ab2b6308995f · outbound

This paper cites Gaussian Error Linear Units (GELUs).

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Gaussian Error Linear Units (GELUs)

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-06-30T13:44:40.646681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:cd5c1c41e050e976b94e54788729578789e67ee11f3be33aa875679d58f3e047

Observation 8f643b94-2f42-4010-9287-9968208ae998 · outbound

This paper cites Population Based Training of Neural Networks.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Population Based Training of Neural Networks

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T13:44:40.655360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:6924cd459ae62bc147131a0c64bbbd5011ffff32e42545d8d29d3eaef0305f10

Observation 16824615-afb9-47c1-9d37-92f47a96f050 · outbound

This paper cites org/CorpusID:235313679.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork org/CorpusID:235313679

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.379622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:d8e980f4ee4f5d8aef6e523d15c7bf9d3978f8a727f5e50cc8abc7a6f75a7bca

Observation a8845882-9b53-45fb-b7da-d223bc92992c · outbound

This paper cites Lee, K.-H., Nachum, O., Yang, M.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Lee, K.-H., Nachum, O., Yang, M

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.376310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:9cd348eaad8571b1692f4df372c9400c12f92a20b2fa836b22efe4175d407f9e

Observation 9d35b37b-59e1-43eb-9bb7-aa9eedff9f76 · outbound

This paper cites A Survey of In-Context Reinforcement Learning.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork A Survey of In-Context Reinforcement Learning

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T13:44:40.652995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:cc023f0f305e41ea4c5ea6a101b49892a68651b090f30eca9e8e017797cc5e7d

Observation fdddb94a-055c-40ba-a3a0-bf0bb16f9ce6 · outbound

This paper cites Nikulin, A., Zisman, I., Zemtsov, A., and Kurenkov, V.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Nikulin, A., Zisman, I., Zemtsov, A., and Kurenkov, V

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.381274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:54beaa500468a4704b9a44215da2fba47b7f306cc5136b2ab9fd0827e1a8b87e

Observation 3cf610d3-6094-4450-b28e-6c46046c07d0 · outbound

This paper cites Papoudakis, G., Christianos, F., Sch¨afer, L., and Albrecht, S.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Papoudakis, G., Christianos, F., Sch¨afer, L., and Albrecht, S

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.384709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:76f9854a7929414836e84d61be7a576d7588cc89400b8b1615160633d7623c69

Observation 094f98ac-c233-4892-81d0-2936901bb134 · outbound

This paper cites Rahman, A., Fosong, E., Carlucho, I., and Albrecht, S.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Rahman, A., Fosong, E., Carlucho, I., and Albrecht, S

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.382977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:294fb68ffa8ac21a2e3aba47195bede38556b2c09c79d046a83b574e837cd716

Observation 8c372c56-e4e3-46d5-9d5f-629fc0451f94 · outbound

This paper cites Rahman, M., Cui, J., and Stone, P.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Rahman, M., Cui, J., and Stone, P

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.386230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:ab188f890ae5574fd1d1c34a72857086285ffe765d90342add30444aa3651580

Observation 68c6b7eb-5f46-4fe6-a84f-fd68489ca1ed · outbound

This paper cites Proximal Policy Optimization Algorithms.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Proximal Policy Optimization Algorithms

Reference 11

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T13:44:40.650034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:701c6a3769bd965cfba3b704c3f8a5ea0c98b84d6a114b63376a3581a4832905

Observation 1622cf25-36bc-4b78-ac98-ee0b0c320d51 · outbound

This paper cites Game-Theoretic Multiagent Reinforcement Learning.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Game-Theoretic Multiagent Reinforcement Learning

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T13:44:40.658304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:bf1574eea9dc981d7c2ee43226971375062780be24f64d67ba994f5c9046cae9

Observation 297e9e37-ef01-4313-b3fc-9ccc8ac17698 · outbound

This paper cites 13 Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork A.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork 13 Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork A

Reference 13

Resolution
malformed identifier
raw_fallback, observed 2026-07-09T01:05:50.369574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:c6c7af091ca7cc5cda5515df52d0259f39ce913de8b367299852bbde902a4458

Observation 257aa6cb-219e-4e42-9005-ebe973523059 · outbound

This paper cites an unresolved cited work.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Unresolved cited work

Reference 14

Resolution
malformed identifier
raw_fallback, observed 2026-07-09T01:05:50.371255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:6d97c639fbfcf92141b2cd020b48a22f4071c70d6f82d55feb0b684f7896b545

Observation b726f50b-16e6-439b-8866-9d280035ccb8 · outbound

This paper cites an unresolved cited work.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-07-09T01:05:50.372914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:126818efc7f1ea0fd917365e0f441e821a79d212d69c06fe7ab957b2e95c3957

Observation 1d53bfb5-76d2-4fab-af37-51755952b90b · outbound

This paper cites These layers perform channel-wise feature transformation without spatial mixing.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork These layers perform channel-wise feature transformation without spatial mixing

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-07-09T01:05:50.374630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:8bec22ef6a539a477cbb944c5292ee5f61583a8ba6e1bef8ac8cceb8c1d0a7b7

Observation c178aa6a-d8fe-46b7-a1db-799a383de70a · outbound

This paper cites with prior.

Benchmarking the Limits of In-Context Reinforcement Learning for Ad-Hoc Teamwork with prior

Reference 17

Resolution
malformed identifier
raw_fallback, observed 2026-07-09T01:05:50.367729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:41:33.048760Z digest=sha256:881c621b055429273ec4e3706f809143bdcdbb1130c5aa4dd51c0191242eb6a5

Pith citing papers

No inbound Pith citation observations are available.