Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T15:22:45.112343Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 3 inbound Pith citation observations for arXiv:2502.06470.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T15:22:45.112343Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:02:22.194525Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T22:37:25.772308Z
64 of 64 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6cd02e11-537a-4cc7-98e6-8f2b1124796a · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks , " * write output.state after.block = add.period write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 206b6ca1-d8df-48a8-8663-cab022113c61 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a877984e-bb44-4910-8304-89fc8beed9f0 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Understanding intermediate layers using linear classifier probes
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc8883a8-ff25-4446-a322-8650f92e5d4f · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks When Benchmarks are Targets: Revealing the Sensitivity of Large Language Model Leaderboards
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b887c14-4be0-4c01-838d-5625f8d50d92 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks S.; Jenner, E.; Casper, S.; Sourbut, O.; Edelman, B
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0746af7b-cddc-4883-ad73-40ed2558eedd · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks theory of mind
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation de61f0dc-12f4-466a-97a9-87f51cb6a01f · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 29c2bc8d-cc02-46d0-95ea-e758b4a3f63c · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Defending Against Unforeseen Failure Modes with Latent Adversarial Training
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcf5387b-fd12-4489-aeac-4e22d6570019 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Designing a Dashboard for Transparency and Control of Conversational AI
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2efbbbad-a974-46bf-96ad-17a5026d30f3 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 31dc4491-51cd-4743-b9b5-81d2460fdd41 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks AI capabilities can be significantly improved without expensive retraining
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 573e6e4f-86e2-48e0-b037-5afba48391e0 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9b9d80f4-7702-4146-9f78-c8b104a6801c · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7ed83bd0-703b-4e47-aaf9-67b0cddebfc8 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30945236-e2b6-4f7c-b67e-affae368ed34 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Intelligent Virtual Assistants with LLM-based Process Automation
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f902a8a-de8b-4d30-a549-24b9c9e8a546 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 10da55c9-07aa-4dec-9455-7a4d5f6f9e2c · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Risks from Learned Optimization in Advanced Machine Learning Systems
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b8e680f-b342-4078-9380-d66d987ad91d · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks AI safety via debate
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 933ebd2d-c29a-4a63-af67-3f2f0d5b513c · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unveiling Theory of Mind in Large Language Models: A Parallel to Single Neurons in the Human Brain
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7702c32-0ac5-45b9-8bd8-3d9076b32dac · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Uncovering Deceptive Tendencies in Language Models: A Simulated Company AI Assistant
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16b7ce84-6f18-43fd-9187-040d03b6a75d · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Y.; Kramar, J.; Brown-Cohen, J.; Albanie, S.; Bulian, J.; Agarwal, R.; Lindner, D.; Tang, Y.; Goodman, N.; and Shah, R
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 80c60722-ab96-48ed-864c-c318fbbdf98a · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3d3ad469-529f-4724-9e18-127e5e95c28b · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks M.; Kundu, A.; Jawhar, S.; Park, J.; and Jurewicz, M
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fdf842a5-b312-4806-a4d4-f53831836d5e · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8f2bd0b3-d6aa-4505-9e26-d1f0ae003aec · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Q.; Stepputtis, S.; Campbell, J.; Hughes, D.; Lewis, C
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 89d5da1c-959a-405f-bed3-532545066a16 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks D.; Dombrowski, A.-K.; Goel, S.; Mukobi, G.; Helm-Burger, N.; Lababidi, R.; Justen, L.; Liu, A
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9774ac63-f2b6-4993-8d80-81d6ad9d400a · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4b8c5de1-8909-4e67-8d58-6f9a875dd5e6 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Rethinking Machine Unlearning for Large Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57490763-0d01-4943-8e3b-de84b2446e4c · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks S.; Cope, D.; and Schoots, N
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 893bfeee-1300-4124-874d-4ebec08404da · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks R.; Baranchuk, M.; Strohmeier, M.; Bolina, V.; Torr, P.; Hammond, L.; and de Witt, C
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4e1c8841-c02f-4d93-8722-3e329ad24936 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Welfare Diplomacy: Benchmarking Language Model Cooperation
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c97cda4-e145-48cc-a58b-ba06e4c45106 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c3de1ec6-b36a-475a-b61e-2c634e000f78 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks GPT-4 Technical Report
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d6a2552-dec3-405b-a988-6d2b7e4b2826 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 36d6f1b0-29c1-450d-baf5-214319ce0742 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Generative Agents: Interactive Simulacra of Human Behavior
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9484155d-7282-4476-bcba-696e0bc11699 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks S.; Goldstein, S.; O’Gara, A.; Chen, M.; and Hendrycks, D
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d071f3cc-b5e0-4a43-b38e-4e0cb6c9ffa1 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 21f19ec5-6798-4ca7-b008-49f6f07a591b · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation de7d4bf4-d26a-4ecd-84c7-e22222b2a28e · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e7ac3f25-1cb4-40d0-a48f-ab403d9f84fe · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5f9dde6a-183b-4b27-85fd-b14beac53a30 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0ee4f544-be7d-4c64-8f5d-0b8459439d51 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Transformers represent belief state geometry in their residual stream
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25876f44-6dc4-4f82-8e67-aa834d41b1c2 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks H.; Zhou, X.; Choi, Y.; Goldberg, Y.; Sap, M.; and Shwartz, V
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 24f2bfbf-c2ba-4a96-8ef6-bba8fa773dfd · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb44456a-1b3b-4e55-8dd2-7a80243ecfe0 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 50395c04-7f3c-4874-a8e1-c12984b57b65 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3c8fa8f0-da92-4b99-8b1a-dc740b0cf524 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks LLM Theory of Mind and Alignment: Opportunities and Risks
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dba2078f-4897-407b-926d-62dfa5b805c7 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks LLMs achieve adult human performance on higher-order theory of mind tasks
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61779350-4ab4-4f45-aaff-8a8e8e3a7d48 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1ed7a4d9-058e-4b99-94e5-7960f1058790 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7ceda1a8-ad88-497e-b6b1-651955a74066 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks GenSim: A General Social Simulation Platform with Large Language Model based Agents
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6855c7d3-7d6d-4126-952b-77ceacf3eeb9 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Steering Language Models With Activation Engineering
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1649c9ec-82ab-4eb1-a22c-e240d09e33e8 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Large Language Models Fail on Trivial Alterations to Theory-of-Mind Tasks
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 715b3a84-e7c1-4a58-8a95-fe7eb84c2449 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks AI Sandbagging: Language Models can Strategically Underperform on Evaluations
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66fc2730-c3a7-4b5c-8c4c-5e0cbf9f2459 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation da5ba737-5f0e-48bc-95d7-8df83b7aaaca · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks P.; and Morency, L.-P
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 62454549-30fa-4401-b93a-ebd7638e1291 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a8c3addc-f454-4b1d-981b-340b4fffbb23 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a0d907c1-d251-469c-b824-fdd4f6acdd2d · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 98d547f2-6be6-438e-84a3-e244994410cb · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0b58a3d6-b9b1-42c3-896e-a3f41e1f58e0 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks A Careful Examination of Large Language Model Performance on Grade School Arithmetic
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 658e47e3-2348-49f2-95c5-1b1e74bab22a · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0bd53ef2-e478-4077-8ee6-18c833023895 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fea45c52-340f-4f29-92e7-1f2acde4c355 · outbound
A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Representation Engineering: A Top-Down Approach to AI Transparency
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 822582f7-0c6d-429d-9f80-a40485da0b2b · inbound
Agents Require Metacognitive and Strategic Reasoning to Succeed in the Coming Labor Markets A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 841517d8-08d0-4bd8-9042-0a5d2ad0f541 · inbound
Does Theory of Mind Improvement Really Benefit Human-AI Interactions? Empirical Findings from Interactive Evaluations A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 50bbfd55-d1c8-4153-99a4-f826df319961 · inbound
Toward Human-Centered Multi-Agent Systems: Integrating Cognition, Culture, Values, and Cooperation in AI Agents A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.