Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T16:57:01.718096Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 5 inbound Pith citation observations for arXiv:2502.06060.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T16:57:01.718096Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:00:53.877395Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-19T07:32:09.425973Z
43 of 43 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c3a65100-2e0f-424c-889b-d3855b56bad2 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Do As I Can, Not As I Say: Grounding Language in Robotic Affordances
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82269cb0-6c43-4206-b60e-bb0349a54d42 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d1411885-52ba-484b-9954-f2d1d1bc9662 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Sparks of Artificial General Intelligence: Early experiments with GPT-4
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 803c69ef-361d-4977-b95c-e6f653a82a6c · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bb82cba9-8dc6-4512-afd3-3031fbd2d34a · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Ho, Thomas L
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 94c093dc-24a3-4329-b144-fe9175949061 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a4a05442-3376-414f-8d4f-72ea7717feac · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Optimizing Robotic Manipulation with Decision-RWKV: A Recurrent Sequence Modeling Approach for Lifelong Learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e7b51152-717d-48c8-ab6f-73cf92841f4d · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Miller, Sasha Mitts, Adithya Renduchintala, Stephen Roller, Dirk Rowe, Weiyan Shi, Joe Spisak, Alexander Wei, David Wu, Hugh Zhang, and Markus Zijlstra
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ed4ef97-699d-4132-8d05-361ebe8bf570 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Frank and Noah D
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c59bec0d-5c7d-4c24-bccc-d42737c7cbe7 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e5ebcad-1ce4-46a6-8f6a-8bd5b423d22a · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a762d21c-7fc4-418c-99c3-c0772f6901bd · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc7e3208-0a00-472b-894f-b98b816d5e20 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Other- Play
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2f7442a0-b8fe-4c37-bea2-a9a0ce36215e · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d7d2303e-97d9-4ec2-946f-28fed295ec62 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning How Well Can a Long Sequence Model Model Long Sequences? Comparing Architechtural Inductive Biases on Long-Context Abilities
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c50d80ce-5814-4147-a54b-62d404ea2b00 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bdcd4efd-7c2d-42bf-aa95-eef8ab8a9cf4 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8d1c974c-6301-415b-bae0-1b9c816eb016 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d78b4802-27a3-4b01-8284-ce04436d6d68 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Hidden Agenda: a Social Deduction Game with Diverse Learned Equilibria
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b9fab916-f919-4763-9ae0-5c910b4b35ac · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3635e34f-e469-49e2-8704-34332322c219 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f49ff4b5-ac22-446d-8833-4c97c287a483 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5f7ae03b-e11f-4eee-93b8-a7beb62a043e · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2b224264-2c23-4853-b5f8-c2f8c8b868a4 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3995a771-aa8a-4de3-81db-cc6c6b24386f · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a42f230c-1963-4977-9d4d-f9e96780ac98 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Learning to communicate about shared procedural abstractions
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2661ece3-9c85-4199-9481-737b0e06edeb · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2fa00c6e-339e-46ec-8bd5-6bd3832d717e · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Training language models to follow instructions with human feedback
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b0df138-354e-4b88-984c-e67ba61d28ff · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning O’Brien, Carrie J
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6edb8572-601d-4e33-a15a-ab35cc3a17dd · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e912628-25be-47fa-a17c-6dcebc2e7e29 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 493521fc-d916-4604-ba17-648c2501faa0 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d97236c-f11e-449b-a795-6bae400c00d6 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Clever Hans or Neural Theory of Mind? Stress Testing Social Reasoning in Large Language Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d064cd7-dda9-4876-b173-ef880458fcdb · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning GPT-4 Doesn't Know It's Wrong: An Analysis of Iterative Prompting for Reasoning Problems
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19451257-c0e9-4641-92d3-31179d4a4e02 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Attention Is All You Need
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a32063cf-a45b-4d49-b935-53e29b1ad566 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a8d5c4cb-100f-4c6b-be19-372bff2584bb · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6a85906f-9f91-4430-a8e7-fd65d033e350 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Voyager: An Open-Ended Embodied Agent with Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3543e8b-a77b-4390-bea8-73c7e09aeff4 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Self-Rewarding Language Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c0458b4-6fb2-4a7d-b06f-00ede8fca473 · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning wait” in a room until something changes in the environment, or “go
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7f886fcd-e415-4c2d-8fdb-38584b05a71a · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Unresolved cited work
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation baddf2b3-91c9-4f54-8480-39f1a8382b7b · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d16c14f6-979f-4cee-8499-8609add9293b · outbound
Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning In Advances in Neural Information Processing Systems (NeurIPS)
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6f4b43c0-30b2-4d6a-aa41-a68e9ba6cd59 · inbound
AI Agent Behavioral Science Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning
Reference 128
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ed57130-317f-43e1-add1-ca4c61bd1a67 · inbound
Chasing Moving Targets with Online Self-Play Reinforcement Learning for Safer Language Models Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation deb9a075-edd8-439d-a3ea-2abfb0e03a88 · inbound
Bayesian Social Deduction with Graph-Informed Language Models Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fb4baca6-46fd-4e13-ba7d-a435926ffbd3 · inbound
Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5d7dd24d-003d-4756-b280-7027d65ebcb3 · inbound
Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.