Pith. sign in

Paper Citation Record · LEDGER

Training Small LLMs as Spatial Multi-Agent Policies

As of 10 August 2026, this Paper Citation Record lists 82 of 82 outbound references and 0 inbound Pith citation observations for arXiv:2608.01425.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.01425 v1

Coverage vector

measured 82 of 82 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T00:17:19.162842Z

measured 82 of 82 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

82 of 82 outbound references displayed

  • verified exact1
  • verified fuzzy33
  • unresolved48
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a97d8b63-9663-4fa8-87f7-6a9852f58212 · outbound

This paper cites Voyager: An Open-Ended Embodied Agent with Large Language Models.

Training Small LLMs as Spatial Multi-Agent Policies Voyager: An Open-Ended Embodied Agent with Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.689848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.689848Z digest=sha256:1e046700575063356eab4d5657fc6151b504006c618f57c58aa150b6b8df35a2

Observation aff42b47-9994-4087-9ba5-d6dac4fa6c1e · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.696932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.696932Z digest=sha256:3b43b9f55ade5af9b4adb644feb3ca8824e56321a59fcadd0e23a2d0701d6d55

Observation 2de7eabb-a2ba-4075-943c-a1239c942983 · outbound

This paper cites arXiv preprint arXiv:2601.17152 , year=.

Training Small LLMs as Spatial Multi-Agent Policies arXiv preprint arXiv:2601.17152 , year=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.703643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.703643Z digest=sha256:7bf7b506e020bf6f16ecdca5fb4820cf15c29f33935dc9fae02cd5b224e4fa02

Observation b12b963c-b795-4405-aed9-6c0bba901d64 · outbound

This paper cites OpenReview , year=.

Training Small LLMs as Spatial Multi-Agent Policies OpenReview , year=

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.709484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.709484Z digest=sha256:b291b69121162cd4849a4d42ee85b5dde8d6352eb2b794e379d4ea14b970d266

Observation 6c284024-588c-4f79-8f89-4e5bd08b8576 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Training Small LLMs as Spatial Multi-Agent Policies Advances in Neural Information Processing Systems , volume=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.714665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.714665Z digest=sha256:45b4cd84c5d64bc4993d3a0465c4b01df1365d60e1a3b6fcbca6cb8f5073d17d

Observation 4c865463-b048-4ba8-bae9-50add8251865 · outbound

This paper cites arXiv preprint arXiv:2510.01586 , year=.

Training Small LLMs as Spatial Multi-Agent Policies arXiv preprint arXiv:2510.01586 , year=

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.720447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.720447Z digest=sha256:87b6266ea53fd335d3bc8332b4e2c16f15f6bf62b375a1f1b8e4f6c90b27171f

Observation 0576f503-603f-4e3f-8182-8163a01bb468 · outbound

This paper cites Closed-Loop Vision-Language Planning for Multi-Agent Coordination.

Training Small LLMs as Spatial Multi-Agent Policies Closed-Loop Vision-Language Planning for Multi-Agent Coordination

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.726467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.726467Z digest=sha256:da87d4854249b587ed153cf932ad5dfa56fb647e1efb82c604d13afa64915385

Observation 6b6ad325-012c-4c22-9c49-247357ae114d · outbound

This paper cites UC Santa Barbara , year=.

Training Small LLMs as Spatial Multi-Agent Policies UC Santa Barbara , year=

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.868559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.732211Z digest=sha256:884e9dad50875bbc3c1330b64e1abe20a5ccd425821e0b4798d60cfff9000210

Observation e94b610e-713a-4c07-825a-0edb073fcbda · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

Training Small LLMs as Spatial Multi-Agent Policies Understanding R1-Zero-Like Training: A Critical Perspective

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.737411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.737411Z digest=sha256:b8be37250f5583d7464583c6c655a3c58b86daa3aeadc9ec880f1a63006e66fa

Observation 2f03ed0d-1683-4da8-af96-da53025443df · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Training Small LLMs as Spatial Multi-Agent Policies Advances in Neural Information Processing Systems , volume=

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.743725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.743725Z digest=sha256:335029439f1aa6f17edb14752256851e1c6b21c87b4c8949c59372926781b5b7

Observation 1a588016-8a39-4dff-bb02-d5a245d662a8 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Training Small LLMs as Spatial Multi-Agent Policies Proximal Policy Optimization Algorithms

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.749113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.749113Z digest=sha256:837605999b25f27c00f004e8154974c4e8bcf35cb825203ce72ea072a6d70754

Observation 8521a6a6-5341-4500-a4d6-f3c3b96eabde · outbound

This paper cites International Conference on Machine Learning , pages=.

Training Small LLMs as Spatial Multi-Agent Policies International Conference on Machine Learning , pages=

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.840127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.758844Z digest=sha256:8c9b02c66217e9c3b0bdfdc8bfba9efa130a0a8311e5e89491094e98534aaf2f

Observation 4840f87f-4e23-461d-8e09-ce62ef5e7172 · outbound

This paper cites NeurIPS Workshop on Foundation Models for Decision Making , year=.

Training Small LLMs as Spatial Multi-Agent Policies NeurIPS Workshop on Foundation Models for Decision Making , year=

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.822807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.769822Z digest=sha256:089e6e82671793cca5ed0fa029ae3d4c6c6ef6cf2d9648493b8f46b393b253fe

Observation 281d48e2-fd5f-4f85-bb87-24c1eb599cd7 · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , year=.

Training Small LLMs as Spatial Multi-Agent Policies Proceedings of the AAAI Conference on Artificial Intelligence , year=

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.804134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.779453Z digest=sha256:ff2a6bb10836fa30234e7bf6568423ea30ef3c5586146899b1a2fad9555fcfcb

Observation 35d66fd2-3521-47fd-abe6-b9fa4f4ff948 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year=.

Training Small LLMs as Spatial Multi-Agent Policies Advances in Neural Information Processing Systems (NeurIPS) , year=

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.787311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.785160Z digest=sha256:f2e83ee6b87f37a452d67e7e8bbeb65023037bb75e6be885d986039c1c10097a

Observation bb4a014f-d93e-4aeb-b37f-2f5256887198 · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

Training Small LLMs as Spatial Multi-Agent Policies Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.769709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.790147Z digest=sha256:c7045f375b074491c778ac474547ad173fdafa12615a27a8e39b9bf1639bc5a8

Observation 7d236f9d-0376-402b-bc1c-dfcefd257f76 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year=.

Training Small LLMs as Spatial Multi-Agent Policies Advances in Neural Information Processing Systems (NeurIPS) , year=

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.751698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.795288Z digest=sha256:1a660faef926074f3496958cf6d0f65d2a095cbd2b91f7cbe8c5aa57a8227c29

Observation b31dce2a-ab47-4d36-8973-bd6747d3c956 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year=.

Training Small LLMs as Spatial Multi-Agent Policies Advances in Neural Information Processing Systems (NeurIPS) , year=

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.733834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.800200Z digest=sha256:03dd193d612211790b35913153aa726836948ccac1163ada1d68de18cbc2fa4f

Observation d2e3b472-7bea-484b-b0fd-562c5707ebb5 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year=.

Training Small LLMs as Spatial Multi-Agent Policies Advances in Neural Information Processing Systems (NeurIPS) , year=

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.717022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.805288Z digest=sha256:373785c99483e4474d265a2ed91a5588b8d484a2b313c8d2f5179ad6678114ea

Observation b3a1fb5d-c5e6-4ec0-b7c9-70c168ad9ac4 · outbound

This paper cites Springer , year=.

Training Small LLMs as Spatial Multi-Agent Policies Springer , year=

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.695508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.810810Z digest=sha256:d58a2582f7e4c58d3459c37809a7a3ef25da991846cf113c1a6ba63648ca05f9

Observation 6e57958b-43fb-4984-a586-4443d56da4d3 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Training Small LLMs as Spatial Multi-Agent Policies LoRA: Low-Rank Adaptation of Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.816275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.816275Z digest=sha256:5bea7d55736487c461ec4a7502985ef51d9c6a33bc46da1228b1112007fad6d6

Observation c1a833ca-92fa-4c3f-b1a5-ad0d3cad9a9d · outbound

This paper cites International Conference on Machine Learning , year=.

Training Small LLMs as Spatial Multi-Agent Policies International Conference on Machine Learning , year=

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.677463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.822297Z digest=sha256:427ef0c2fbef1e636165781aafdff5555705e6021f7fc7496d9d52d04cfcc9b5

Observation 351da279-b29a-4b49-b18b-d5b61739b9eb · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Training Small LLMs as Spatial Multi-Agent Policies Advances in Neural Information Processing Systems , volume=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.828752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.828752Z digest=sha256:f9e910dc800abd3f7e72d180e3964f86ff99bf64ed562877079a970c1cc3017d

Observation 23479c98-ff00-4dfa-992c-b82043a4e229 · outbound

This paper cites MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework.

Training Small LLMs as Spatial Multi-Agent Policies MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.833955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.833955Z digest=sha256:47d362e84e1f4324e8d238fce1f9c610511edb3e537ed5ddb50018dcdbb4817e

Observation 3434c501-f3ac-42d2-b063-930a559fc8ce · outbound

This paper cites ACL , year=.

Training Small LLMs as Spatial Multi-Agent Policies ACL , year=

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.647587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.839710Z digest=sha256:3baa4c8e82e43c6b858c32df26991a7294c29d25eb4ef8f487ed6cadd2d35e3d

Observation 70af42e8-0a44-4878-9741-25ad902fdfc7 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Training Small LLMs as Spatial Multi-Agent Policies Advances in Neural Information Processing Systems , volume=

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.844676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.844676Z digest=sha256:e15ca9b41543c08caf0067defa79092c4b253c23cd76420d2e3a78ee4407b574

Observation d67e673f-7d17-42f8-a067-bb02a803aea5 · outbound

This paper cites and Shen, Yelong and Wallis, Phillip and Allen-Zhu, Zeyuan and Li, Yuanzhi and Wang, Shean and Wang, Lu and Chen, Weizhu , booktitle=.

Training Small LLMs as Spatial Multi-Agent Policies and Shen, Yelong and Wallis, Phillip and Allen-Zhu, Zeyuan and Li, Yuanzhi and Wang, Shean and Wang, Lu and Chen, Weizhu , booktitle=

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.850417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.850417Z digest=sha256:d351b4fb40d821ec58769720cd804ca3a3c1e7f224fc22af488010f7b6d4c0be

Observation 4dfda6af-2954-4ccc-9a9c-78698920f64f · outbound

This paper cites Ghost in the Minecraft: Generally Capable Agents for Open-World Environments via Large Language Models with Text-based Knowledge and Memory.

Training Small LLMs as Spatial Multi-Agent Policies Ghost in the Minecraft: Generally Capable Agents for Open-World Environments via Large Language Models with Text-based Knowledge and Memory

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.854928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.854928Z digest=sha256:b2d42364e017f4b50975de7c2e630eb7ff5902f30ffdaaa834c5bbf9c3b48189

Observation a1e8b022-ed96-4af7-9f0e-770c5e5c6d99 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.607884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.859747Z digest=sha256:5b20700ea3f44b8e4034fc60cf8d9ab81729963959ca865a6a07440cfcf06071

Observation a286020b-7751-4734-ac4b-c6e17a7d02e5 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.590631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.864434Z digest=sha256:f1cbb6c6c5b08434efd1eca5c85687bff808c50da382d0b213e66dff46710f53

Observation 6c585948-2051-4a80-a9e5-1eaa5c4b91e6 · outbound

This paper cites International Conference on Learning Representations (ICLR) , year=.

Training Small LLMs as Spatial Multi-Agent Policies International Conference on Learning Representations (ICLR) , year=

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.572774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.869706Z digest=sha256:186c3a42713235b96488f46181b0f7bbb661c0115d5611fd5e7f7507e0b1a9d9

Observation 80db720c-85ad-4f7f-a485-d8af810fac3f · outbound

This paper cites Conference on Robot Learning (CoRL) , year=.

Training Small LLMs as Spatial Multi-Agent Policies Conference on Robot Learning (CoRL) , year=

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.875905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.875905Z digest=sha256:963f0dfb87a6ea786c2b941d89bb74ab4adceacbed5e25affcb8d297d94c21e5

Observation 0f972238-1c77-4327-b00b-f931b383e25b · outbound

This paper cites Conference on Robot Learning (CoRL) , year=.

Training Small LLMs as Spatial Multi-Agent Policies Conference on Robot Learning (CoRL) , year=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.880997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.880997Z digest=sha256:8a5ccb8862e3eb7c3f7c20067045cff70f6a457c3d44026e00c849661bfa520f

Observation 3eefc640-8de1-4b56-8967-db89ca05c890 · outbound

This paper cites IEEE International Conference on Robotics and Automation (ICRA) , year=.

Training Small LLMs as Spatial Multi-Agent Policies IEEE International Conference on Robotics and Automation (ICRA) , year=

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.529731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.886656Z digest=sha256:8a196d66eddb975d5c6a298b698184ee87e71918cfd5c41c66bdd37d0a6027b9

Observation 6ee843cd-c5da-4d20-8dae-407dbf1297c5 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.891696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.891696Z digest=sha256:4d28de857f2873acd13ed28a11a43703971d8199340febc174b37e3a4881f2f3

Observation 50abe3fe-4824-40bc-8362-7c5e3f8c01c0 · outbound

This paper cites Efficient Guided Generation for Large Language Models.

Training Small LLMs as Spatial Multi-Agent Policies Efficient Guided Generation for Large Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.899014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.899014Z digest=sha256:9dbc75fdc4868cdbb2c07a5adb91720eef9a29e5e9d94db330e8b95cd72d37ca

Observation 3b668109-28ad-42f5-9ff9-993a5fa4ad31 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year=.

Training Small LLMs as Spatial Multi-Agent Policies Advances in Neural Information Processing Systems (NeurIPS) , year=

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.501153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.904272Z digest=sha256:ecf27f16d8615f74f8a215a6e861df189b2b3a3c105432adcfc979886b54d662

Observation b8ed0b0c-df6a-4c95-8538-5e662421c4d8 · outbound

This paper cites 2025 , note=.

Training Small LLMs as Spatial Multi-Agent Policies 2025 , note=

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.480863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.910855Z digest=sha256:04ce35d099ebebf20db36bda6a7c2924aa9f13582c768f6846ae15abe75689c5

Observation cbe7a7b1-ded3-432c-998d-8f591ead5018 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.462829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.916313Z digest=sha256:a0a3c94fad30ce23ac2b3bec4c5794efa34402ac8c60c6b7147d929d84951985

Observation ea4d3b2b-2173-45cd-8185-ea230b9d4968 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.446121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.921996Z digest=sha256:b46152aba42298bf1b255b37808df52bd8b2c43ec2901b530c3fbbf8e8316de5

Observation 099afb6d-1268-4920-b234-5036b0fe6c98 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.430371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.927241Z digest=sha256:7ea808ef0b0a0e5703e38211a51f4ee36bd2833ea21e5d64e262c80d4da70fbe

Observation b12cd2aa-80fc-4ab2-9d5f-1df7132057b6 · outbound

This paper cites Advances in Neural Information Processing Systems (NeurIPS) , year=.

Training Small LLMs as Spatial Multi-Agent Policies Advances in Neural Information Processing Systems (NeurIPS) , year=

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.413440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.933257Z digest=sha256:d60fa3e04f4f51886c09c105b796b60e9800ae1925cc51ea217c874a3f3cf365

Observation 892fa440-2f8a-4f92-ba0f-2d8734c00959 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.394338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.938232Z digest=sha256:1c1ef7424f1bd88492410f2cc8875c4a54b7af6b79845aeb078db657ad40a95a

Observation 18e5a08a-154f-424b-b87e-54c41f148893 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.943167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.943167Z digest=sha256:dcd98d1f329de4ea21d7671be6b26683ce3b6afd2a937707210ad4820bb486e0

Observation 7f693115-a410-42fa-b2f0-661c294ac217 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.948504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.948504Z digest=sha256:2104bf6d93e9ec2f6c07c68efc6afb5fb158cd18fd5433c5195e6da564010afb

Observation c093510d-5867-4870-bafd-7ca458ccfe5b · outbound

This paper cites and Chandar, Sarath and Burch, Neil and Lanctot, Marc and Song, H.

Training Small LLMs as Spatial Multi-Agent Policies and Chandar, Sarath and Burch, Neil and Lanctot, Marc and Song, H

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.355411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.953219Z digest=sha256:5d3d8424831544bf021466db68501882450d6a54070b61d168132fcf8215d1db

Observation d7dbceaf-98ba-4b84-9951-587a4e13783d · outbound

This paper cites Proceedings of the 16th International Conference on Autonomous Agents and Multiagent Systems (AAMAS) , year=.

Training Small LLMs as Spatial Multi-Agent Policies Proceedings of the 16th International Conference on Autonomous Agents and Multiagent Systems (AAMAS) , year=

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.338289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.958388Z digest=sha256:c88ab91d59c169f9ecde81e11c79388a2331a76c75abc7ccf1e43647421bb3be

Observation 90bb8644-b115-4a79-8631-2a3c3bf2a99c · outbound

This paper cites and Griffiths, Thomas L.

Training Small LLMs as Spatial Multi-Agent Policies and Griffiths, Thomas L

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.321830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.963478Z digest=sha256:45806265f6b0a7360bcd963d6bb3a258f31f38cfa276ce1793030812805e0c47

Observation a0eccb0c-3296-4040-a020-e778a140476b · outbound

This paper cites Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology (UIST) , year=.

Training Small LLMs as Spatial Multi-Agent Policies Proceedings of the 36th Annual ACM Symposium on User Interface Software and Technology (UIST) , year=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:18.968298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:18.968298Z digest=sha256:9211b0f7738c74665df57f1303976771920328ec11b1c64f00222e52cf69fd50

Observation 7908b1a1-b13f-4add-aed4-8534b8cded8e · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.296262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.973189Z digest=sha256:e873224dbc41857842b16d593fc744cb23411724134e0ad2fee5bc4d576d2025

Observation 96fa44f0-7e4e-40a4-b57e-ca9d028302f4 · outbound

This paper cites and Precup, Doina and Singh, Satinder , journal=.

Training Small LLMs as Spatial Multi-Agent Policies and Precup, Doina and Singh, Satinder , journal=

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.280455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.977900Z digest=sha256:47f252a74fa68be97ab6ccecb10ef09937caa5cae85e4d9eecdcd6bc99ede520

Observation 0525e42e-14f9-42fb-904c-03edd38e1c91 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.264604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.983484Z digest=sha256:afdefd4b9a4f62e99e5d6afbcd4f548a00861b9c7923472b41b46a0348ff14c1

Observation 5eda816b-5dbe-453d-9aa5-5a7dfa8698d1 · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

Training Small LLMs as Spatial Multi-Agent Policies Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.245258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.988541Z digest=sha256:eec484674161cf94f1390440469cafd491016e5c1fbeba39924ef84bfa1e79b3

Observation 3b71affc-c92d-4308-96a8-428307d4bdca · outbound

This paper cites International Conference on Learning Representations (ICLR) , year=.

Training Small LLMs as Spatial Multi-Agent Policies International Conference on Learning Representations (ICLR) , year=

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.227093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.993534Z digest=sha256:3a385f00c87e2ec91a9d623826107440f1ceb7b888b707a1f328fa3058dc52f5

Observation ed4071c0-a341-418b-95a3-59bb87720136 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.209083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:18.999934Z digest=sha256:16d179ffddeb8f45b8ea03cd98f2961fd431b01c60555d679e6c2cc03aeb8259

Observation e0b86ce6-fe95-4fe9-a2fb-5ff0681e48a5 · outbound

This paper cites Findings of the Association for Computational Linguistics: EACL , year=.

Training Small LLMs as Spatial Multi-Agent Policies Findings of the Association for Computational Linguistics: EACL , year=

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.192653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.020776Z digest=sha256:d8930cb904961f88f0d9513c1e9cd8aa52aab9c68c0bb7cd8a21b6693bc09c36

Observation 045eb470-05de-40f5-82db-752649c50f95 · outbound

This paper cites Planning with Macro-Actions in Decentralized.

Training Small LLMs as Spatial Multi-Agent Policies Planning with Macro-Actions in Decentralized

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.177195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.026518Z digest=sha256:ab590a5d5a003c61537461f2afca0207b1052cd7af51aa065374f1631244d44d

Observation 208abd0b-4dc1-49f1-9c62-0b32969abecb · outbound

This paper cites Proceedings of the 3rd Conference on Robot Learning (CoRL) , series =.

Training Small LLMs as Spatial Multi-Agent Policies Proceedings of the 3rd Conference on Robot Learning (CoRL) , series =

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.161210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.031859Z digest=sha256:eb7f9452ea10624c244a15f90e7a30bbbf78bddd42f413162b9abd7b8976ad35

Observation 6fa2ef9f-167e-4585-883b-af4bfbf96553 · outbound

This paper cites Melting Pot 2.0.

Training Small LLMs as Spatial Multi-Agent Policies Melting Pot 2.0

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:19.036828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:19.036828Z digest=sha256:023e5603b9b844da3e8447d0ad452f733988a1fb1f59375cdb3a671dd70a41ac

Observation 4446c246-3a87-4cab-8a68-bba161cd121a · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.144954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.042331Z digest=sha256:1daba5de363051d164564779b5ab5700e0b080639620c43c9388b45786bd10ef

Observation ca5ef68b-b704-4e09-b427-a00fb561974a · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.129244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.047580Z digest=sha256:0c75a706d83e8c781a507ef4ca5401dc7e603b71022911b09c437d20284bcc57

Observation 37a23397-6dbb-4921-98ba-c35f6f86013b · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 68

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.112323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.052882Z digest=sha256:bb5ae7b1375a7594c26e49114b5b94954134f52d969d58bea5baaa0fa631dc55

Observation be1ee015-6b03-40fa-8258-0f241e4d48e2 · outbound

This paper cites K.; Griffiths, T.

Training Small LLMs as Spatial Multi-Agent Policies K.; Griffiths, T

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.095697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.058640Z digest=sha256:adec03f924110cf74a9da36f68b64c38ef797c01ded904891543272434a60823

Observation f51176cd-a51d-4833-92ac-0f9649304b43 · outbound

This paper cites Beyond Scalar Rewards: Dense Feedback for LLM Policy Synthesis in Sequential Social Dilemmas.

Training Small LLMs as Spatial Multi-Agent Policies Beyond Scalar Rewards: Dense Feedback for LLM Policy Synthesis in Sequential Social Dilemmas

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:19.064587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:19.064587Z digest=sha256:efc4293d7c869c06fee1afa8a3063e196624413c3d71cfebdc8ae374e49d88ca

Observation 074dcf2b-4495-4183-9722-abe3ed2c6eb1 · outbound

This paper cites J.; Shen, Y.; Wallis, P.; Allen-Zhu, Z.; Li, Y.; Wang, S.; Wang, L.; and Chen, W.

Training Small LLMs as Spatial Multi-Agent Policies J.; Shen, Y.; Wallis, P.; Allen-Zhu, Z.; Li, Y.; Wang, S.; Wang, L.; and Chen, W

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.079512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.069699Z digest=sha256:1ef65f692dd94d2a404fc927f836b15617abac6bf73787a7ecddd82692bfaac8

Observation 6e6c1218-4dcb-43f3-96ef-eef2c6215eec · outbound

This paper cites Shapley-Coop: Credit Assignment for Emergent Cooperation in Self-Interested LLM Agents.

Training Small LLMs as Spatial Multi-Agent Policies Shapley-Coop: Credit Assignment for Emergent Cooperation in Self-Interested LLM Agents

Reference 72

Resolution
verified exact
local_arxiv, observed 2026-08-06T00:17:19.440446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.076571Z digest=sha256:073d971928d17a20bb5245fdc005e48d2662cb0319093d58247358bc8f48df48

Observation c3d2e64b-13ae-4f2b-9695-637cb1ae01da · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 73

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.063475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.082763Z digest=sha256:711f1f5d8fc4f21518ed107696fda9c9460a1e664b1b008cffbbe8b1e2117cf9

Observation 5f08f481-6680-4341-8cc4-ff7a8559f6f3 · outbound

This paper cites Z.; Phillips, M.; Tuyls, K.; Du \'e \ n ez-Guzm \'a n, E.

Training Small LLMs as Spatial Multi-Agent Policies Z.; Phillips, M.; Tuyls, K.; Du \'e \ n ez-Guzm \'a n, E

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.047296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.088260Z digest=sha256:71a103040acb02ddcaa6e55bd9c3b380e661bd19616b0388ed3a5698a4b7b1ad

Observation 4e82f720-32cc-46fc-bd4a-9162abb3a677 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:19.093339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:19.093339Z digest=sha256:9d6c1aaf50568adb82fa8bb38872fa4c3711c1385cd7c5aed870c077825beda0

Observation 7b0df47a-78fc-41e5-9a9d-ef91e1a0e6ac · outbound

This paper cites Z.; Zambaldi, V.; Lanctot, M.; Marecki, J.; and Graepel, T.

Training Small LLMs as Spatial Multi-Agent Policies Z.; Zambaldi, V.; Lanctot, M.; Marecki, J.; and Graepel, T

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:20.030609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.098155Z digest=sha256:6df0500a1c33477cf6116855c416b374d25598d9331770fdbcc66885a8e7a26d

Observation 4ef82487-1800-488d-8341-243b57cc5ee5 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 77

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:20.013919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.103067Z digest=sha256:ed33460888a5e769d590e1a53095ec5ea12b557e174a343728106f45f270250b

Observation 9fb159ba-b3df-4cd0-b1b3-dfc0ae6029be · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:19.109264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:19.109264Z digest=sha256:04c770644f38c0d2760a22609b71a65f0c3c34911a548847fbbc75c4a0afc19c

Observation 3b9ef1fc-705d-4fcb-b573-359e092ebda6 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 79

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:19.996489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.115120Z digest=sha256:8a14f02bb8eb49d8d14ffa8a649bdcedf298d570bb9ceee10f931878e6e6898b

Observation 393301f8-4665-4d83-9de7-38482675db31 · outbound

This paper cites S.; Rios, M.; Fonseca, Y.; Giraldo, L.

Training Small LLMs as Spatial Multi-Agent Policies S.; Rios, M.; Fonseca, Y.; Giraldo, L

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:19.975414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.120501Z digest=sha256:961d4022c6561ba41af7414b486bdd3be5fb71c0e717565d1ceca0d657f0cfe9

Observation c5f27482-1d88-405c-9225-5ccec37cbaf9 · outbound

This paper cites Z.; Zambaldi, V.; Beattie, C.; Tuyls, K.; and Graepel, T.

Training Small LLMs as Spatial Multi-Agent Policies Z.; Zambaldi, V.; Beattie, C.; Tuyls, K.; and Graepel, T

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:19.958505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.126320Z digest=sha256:094c99ba64cd7c398f05e8905ba8c04e3f2d7b324effaa083865e9e07f41a215

Observation a032285e-e12f-460b-8ab8-5d5fd80d6392 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:19.131642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:19.131642Z digest=sha256:ceeb1542c2469e004a8ad7f6174535734c5520db62b5b016faaf86e4b4cb5c7c

Observation 23aee3d0-7425-4fee-b340-f63ba4738826 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Training Small LLMs as Spatial Multi-Agent Policies DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:19.136940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:19.136940Z digest=sha256:3a8210a84f7d70ce10e8fc6a2a2b444f921304f1c035a122f9c3294d9c4f4df1

Observation bf7d81a3-36db-414c-ad10-6dcefc65d953 · outbound

This paper cites S.; Precup, D.; and Singh, S.

Training Small LLMs as Spatial Multi-Agent Policies S.; Precup, D.; and Singh, S

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T00:17:19.941061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.141764Z digest=sha256:9430631eec3f8b14da67c14292aee43d341e57a149f3025a3979c773cb6e51ac

Observation a35f7338-3833-459b-9157-b94ed8552733 · outbound

This paper cites On the Planning Abilities of Large Language Models : A Critical Investigation.

Training Small LLMs as Spatial Multi-Agent Policies On the Planning Abilities of Large Language Models : A Critical Investigation

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:19.146958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:19.146958Z digest=sha256:b9ab822408be6362d6e07bb36df514c256e749d65c5ec69a5502103956956c55

Observation 177a0a35-6a89-47ef-9c81-af255d2df0de · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 86

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:19.924618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.152230Z digest=sha256:e2e2f0f05d6d9809549db89459f641a588cf4a8fb08f53f6dc94793434eefc1f

Observation 34d0d686-e323-4f85-ad92-f58c18a118ed · outbound

This paper cites Low-Rank Agent-Specific Adaptation (LoRASA) for Multi-Agent Policy Learning.

Training Small LLMs as Spatial Multi-Agent Policies Low-Rank Agent-Specific Adaptation (LoRASA) for Multi-Agent Policy Learning

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-06T00:17:19.157771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T00:17:19.157771Z digest=sha256:581ae249429ac529c34a548f41800b934740b8c53c907e9b68f20077f1bc5908

Observation 747efd9e-4ff6-476a-82c9-b01992e752e3 · outbound

This paper cites an unresolved cited work.

Training Small LLMs as Spatial Multi-Agent Policies Unresolved cited work

Reference 88

Resolution
unresolved
raw_fallback, observed 2026-08-06T00:17:19.906827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T00:17:19.162842Z digest=sha256:2bbd8376c59d4a2fa48c7e032d675769499706e4a71a38386b6c384a2cd716f3

Pith citing papers

No inbound Pith citation observations are available.