Pith. sign in

Paper Citation Record · LEDGER

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks

As of 10 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 3 inbound Pith citation observations for arXiv:2502.06470.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.06470 v1

Coverage vector

measured 64 of 64 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T15:22:45.112343Z

measured 67 of 67 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:02:22.194525Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T22:37:25.772308Z

Reference resolution

64 of 64 outbound references displayed

  • verified exact0
  • verified fuzzy11
  • unresolved53
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6cd02e11-537a-4cc7-98e6-8f2b1124796a · outbound

This paper cites , " * write output.state after.block = add.period write newline.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.769888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.769888Z digest=sha256:fb46362cd2fb939d9d512cc710b7b8b501166fede06f062de329244b3c929242

Observation 206b6ca1-d8df-48a8-8663-cab022113c61 · outbound

This paper cites write newline.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.776121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.776121Z digest=sha256:bfa73253acebed418300cc78b30fa51a6c4657f437b8ab726bd74d26b57e282f

Observation a877984e-bb44-4910-8304-89fc8beed9f0 · outbound

This paper cites Understanding intermediate layers using linear classifier probes.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Understanding intermediate layers using linear classifier probes

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.782384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.782384Z digest=sha256:6772ed3dff19f9dc24181ff52f37edbc63451ba993a2923a32c025fa1fe7e1ae

Observation cc8883a8-ff25-4446-a322-8650f92e5d4f · outbound

This paper cites When Benchmarks are Targets: Revealing the Sensitivity of Large Language Model Leaderboards.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks When Benchmarks are Targets: Revealing the Sensitivity of Large Language Model Leaderboards

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.788604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.788604Z digest=sha256:02c6ca8ebf3b8fa9b2045d010e823bb0a8534ee698ed81456dc9b7a2d5140e72

Observation 0b887c14-4be0-4c01-838d-5625f8d50d92 · outbound

This paper cites S.; Jenner, E.; Casper, S.; Sourbut, O.; Edelman, B.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks S.; Jenner, E.; Casper, S.; Sourbut, O.; Edelman, B

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:46.210568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.794498Z digest=sha256:0198f29d4f2386fbb4644281cf71c7de7fbbce633301ae84d4d5486153032db9

Observation 0746af7b-cddc-4883-ad73-40ed2558eedd · outbound

This paper cites theory of mind.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks theory of mind

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:46.191345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.800330Z digest=sha256:7e5923315bfdd8225f3ef6a4c4284c5fc7f0e8b15ded4af2cce5eb4cfba8fda9

Observation de61f0dc-12f4-466a-97a9-87f51cb6a01f · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:46.173063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.805659Z digest=sha256:068576290a47dde6d9c633258b904b0ca6a3feeec7845c4d5167ce19ecad6430

Observation 29c2bc8d-cc02-46d0-95ea-e758b4a3f63c · outbound

This paper cites Defending Against Unforeseen Failure Modes with Latent Adversarial Training.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Defending Against Unforeseen Failure Modes with Latent Adversarial Training

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.811321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.811321Z digest=sha256:82764626ad8f7961553cb227c0bb54307690ac01048a69460c7fba229f67b3f9

Observation dcf5387b-fd12-4489-aeac-4e22d6570019 · outbound

This paper cites Designing a Dashboard for Transparency and Control of Conversational AI.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Designing a Dashboard for Transparency and Control of Conversational AI

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.817749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.817749Z digest=sha256:3d4cf5a3e36bddc5d676e44f359eda129d0d1fec800e7b618ec5904687a7b0a1

Observation 2efbbbad-a974-46bf-96ad-17a5026d30f3 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:46.154807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.824077Z digest=sha256:7b10cec615a6b712ed69faa62a1a991386b391cf501385a4b3de367c67bbdd83

Observation 31dc4491-51cd-4743-b9b5-81d2460fdd41 · outbound

This paper cites AI capabilities can be significantly improved without expensive retraining.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks AI capabilities can be significantly improved without expensive retraining

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.829284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.829284Z digest=sha256:60275e054e51e50a8b32bc99971ff04d9d7e5c9843f3068dfe9df2e646720670

Observation 573e6e4f-86e2-48e0-b037-5afba48391e0 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:46.137282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.834477Z digest=sha256:9f98d381d33fe32810f90120fde25ee03c722b1f13546070bb80f4613e0b5e90

Observation 9b9d80f4-7702-4146-9f78-c8b104a6801c · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:46.120361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.839605Z digest=sha256:fd861265457cf6778d4115801a1573521d40b1867ecb172a73a4ab544d637e39

Observation 7ed83bd0-703b-4e47-aaf9-67b0cddebfc8 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.845044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.845044Z digest=sha256:30d914b2b6e1e0556e9d60b9d719771b7114dd63e3d5120429f79455f72b8349

Observation 30945236-e2b6-4f7c-b67e-affae368ed34 · outbound

This paper cites Intelligent Virtual Assistants with LLM-based Process Automation.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Intelligent Virtual Assistants with LLM-based Process Automation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.850403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.850403Z digest=sha256:5ae2c78648e8be0b2af6f64181f7b988cb1e84986730c20494f474ff876d0bdb

Observation 8f902a8a-de8b-4d30-a549-24b9c9e8a546 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:46.092685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.855893Z digest=sha256:ff70a5e64033c9e87914fa5141d73377c4c31b6bccc7004999140dd7ac42084c

Observation 10da55c9-07aa-4dec-9455-7a4d5f6f9e2c · outbound

This paper cites Risks from Learned Optimization in Advanced Machine Learning Systems.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Risks from Learned Optimization in Advanced Machine Learning Systems

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.860935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.860935Z digest=sha256:d29d83c4d757ea4a9923e8f26031453d1baeae1a364c6bee1d02b8ceb4c0b79a

Observation 2b8e680f-b342-4078-9380-d66d987ad91d · outbound

This paper cites AI safety via debate.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks AI safety via debate

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.871075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.871075Z digest=sha256:b88071fcb01e48796d5cc4decda5eb406ffbfdbcc51c0b929fb1fd85e19e22e9

Observation 933ebd2d-c29a-4a63-af67-3f2f0d5b513c · outbound

This paper cites Unveiling Theory of Mind in Large Language Models: A Parallel to Single Neurons in the Human Brain.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unveiling Theory of Mind in Large Language Models: A Parallel to Single Neurons in the Human Brain

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.876548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.876548Z digest=sha256:e8351ff81f02e1d09f4d4f9853767922e06cde28776262c2c2a5126c7ce960c3

Observation f7702c32-0ac5-45b9-8bd8-3d9076b32dac · outbound

This paper cites Uncovering Deceptive Tendencies in Language Models: A Simulated Company AI Assistant.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Uncovering Deceptive Tendencies in Language Models: A Simulated Company AI Assistant

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.881881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.881881Z digest=sha256:ad6a211d4d7f6ca4a4ec5cadf5dc341847d57ada55aac760e3e983408ed9538e

Observation 16b7ce84-6f18-43fd-9187-040d03b6a75d · outbound

This paper cites Y.; Kramar, J.; Brown-Cohen, J.; Albanie, S.; Bulian, J.; Agarwal, R.; Lindner, D.; Tang, Y.; Goodman, N.; and Shah, R.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Y.; Kramar, J.; Brown-Cohen, J.; Albanie, S.; Bulian, J.; Agarwal, R.; Lindner, D.; Tang, Y.; Goodman, N.; and Shah, R

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:46.075378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.887246Z digest=sha256:44882a75a66a5f2d867a95e16e82c3e66d5fd08e14577e1108a2128a5c8dedec

Observation 80c60722-ab96-48ed-864c-c318fbbdf98a · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:46.057915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.892203Z digest=sha256:3d3c2fb80634d01730d1e50b015c479f01b73acc2a9bdb27b4d4f8086c7f6f76

Observation 3d3ad469-529f-4724-9e18-127e5e95c28b · outbound

This paper cites M.; Kundu, A.; Jawhar, S.; Park, J.; and Jurewicz, M.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks M.; Kundu, A.; Jawhar, S.; Park, J.; and Jurewicz, M

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:46.040914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.897457Z digest=sha256:f7a9b583761508c59f49ee6a18e5a6b72db6180dbabaed5e669f336dbeec850d

Observation fdf842a5-b312-4806-a4d4-f53831836d5e · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:46.024838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.902260Z digest=sha256:a9947b9bcc75e79b31267dea9eb63c12e99df19199349e446c51b378d66f9b8c

Observation 8f2bd0b3-d6aa-4505-9e26-d1f0ae003aec · outbound

This paper cites Q.; Stepputtis, S.; Campbell, J.; Hughes, D.; Lewis, C.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Q.; Stepputtis, S.; Campbell, J.; Hughes, D.; Lewis, C

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:46.009315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.907051Z digest=sha256:596942d19815004375678f81d7797e6a03f21e29dbc2641ce890688a4cecf6a0

Observation 89d5da1c-959a-405f-bed3-532545066a16 · outbound

This paper cites D.; Dombrowski, A.-K.; Goel, S.; Mukobi, G.; Helm-Burger, N.; Lababidi, R.; Justen, L.; Liu, A.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks D.; Dombrowski, A.-K.; Goel, S.; Mukobi, G.; Helm-Burger, N.; Lababidi, R.; Justen, L.; Liu, A

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:45.992958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.911934Z digest=sha256:8380b77c27e523b3c5a4a4ceb9be8629dc7e2fda3385b0240b15cd731211fbba

Observation 9774ac63-f2b6-4993-8d80-81d6ad9d400a · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.976152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.917280Z digest=sha256:6959f8c727ce40756f0469288f828cfb3ce90328ec6d373f065c04fa473808ec

Observation 4b8c5de1-8909-4e67-8d58-6f9a875dd5e6 · outbound

This paper cites Rethinking Machine Unlearning for Large Language Models.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Rethinking Machine Unlearning for Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.922410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.922410Z digest=sha256:8358290a5d7ed91ffb9db10b7b564c577af28329fb712f21a819d48b996495a4

Observation 57490763-0d01-4943-8e3b-de84b2446e4c · outbound

This paper cites S.; Cope, D.; and Schoots, N.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks S.; Cope, D.; and Schoots, N

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:45.959915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.927713Z digest=sha256:3851505af0a2c805103844c3323adf29805cb4cc9c023be0c4903e36c65bf011

Observation 893bfeee-1300-4124-874d-4ebec08404da · outbound

This paper cites R.; Baranchuk, M.; Strohmeier, M.; Bolina, V.; Torr, P.; Hammond, L.; and de Witt, C.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks R.; Baranchuk, M.; Strohmeier, M.; Bolina, V.; Torr, P.; Hammond, L.; and de Witt, C

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:45.943348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.932771Z digest=sha256:1eb870c938b1908a2983f202a113e87e717d06539ecdecf7f5b1da263e419cdd

Observation 4e1c8841-c02f-4d93-8722-3e329ad24936 · outbound

This paper cites Welfare Diplomacy: Benchmarking Language Model Cooperation.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Welfare Diplomacy: Benchmarking Language Model Cooperation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.937791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.937791Z digest=sha256:a6199646aeb323ec51b42b62312d271aa600ef05f8d5d3f8da43aec4c94aa4b6

Observation 3c97cda4-e145-48cc-a58b-ba06e4c45106 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.926410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.943052Z digest=sha256:7cfdbde636489c8f49c5d01ecab5c571a76ac2d77bf2c231f6d7a00ca3fc4d9b

Observation c3de1ec6-b36a-475a-b61e-2c634e000f78 · outbound

This paper cites GPT-4 Technical Report.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks GPT-4 Technical Report

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.948112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.948112Z digest=sha256:65df2701bc2800a6e8b6121cf1ae6ea0e7ea6d51ccd08b869eceac9809e10102

Observation 1d6a2552-dec3-405b-a988-6d2b7e4b2826 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.910445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.953589Z digest=sha256:468e3b91e5732023970078f00befe0d2716cea0cae0b580fdfe15c89fad0df62

Observation 36d6f1b0-29c1-450d-baf5-214319ce0742 · outbound

This paper cites Generative Agents: Interactive Simulacra of Human Behavior.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Generative Agents: Interactive Simulacra of Human Behavior

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.958547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.958547Z digest=sha256:c8669986832c11fc6075a1fc3a128430f3652b503f0d7d247fa1a55cc55e6333

Observation 9484155d-7282-4476-bcba-696e0bc11699 · outbound

This paper cites S.; Goldstein, S.; O’Gara, A.; Chen, M.; and Hendrycks, D.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks S.; Goldstein, S.; O’Gara, A.; Chen, M.; and Hendrycks, D

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:45.892387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.964464Z digest=sha256:4399c79df8c83cfc7fdf9eaf25960c154d8952c86f90c23994b90263c46ea04d

Observation d071f3cc-b5e0-4a43-b38e-4e0cb6c9ffa1 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.876076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.969455Z digest=sha256:77879dcb36f97dcf37ceb19c2a77aa35c32562df396b4ecbfdfe60271fcb7004

Observation 21f19ec5-6798-4ca7-b008-49f6f07a591b · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.858963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.974926Z digest=sha256:4d4740a82680f6a222016ccc97821f745b69227749fdff86ea48f50a533a2131

Observation de7d4bf4-d26a-4ecd-84c7-e22222b2a28e · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.841531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.980417Z digest=sha256:85c7157ba49e6ec82bb75defa4e740910c2129adfda770ce56708915ab902855

Observation e7ac3f25-1cb4-40d0-a48f-ab403d9f84fe · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.821466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.985768Z digest=sha256:2e02594b42c4f793d618111179a75a0bc1ff10cf9a67e98e7f9a8aa0e7b2f9dd

Observation 5f9dde6a-183b-4b27-85fd-b14beac53a30 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.802716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.991158Z digest=sha256:a125f0d33d4db570de8f72dd457458113d713b97bbbbcc32cb1a1c48310a8a6b

Observation 0ee4f544-be7d-4c64-8f5d-0b8459439d51 · outbound

This paper cites Transformers represent belief state geometry in their residual stream.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Transformers represent belief state geometry in their residual stream

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.996501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.996501Z digest=sha256:84a89742287643203d498ee861f8aa151893e28afbbd873dbfa3114a8443502d

Observation 25876f44-6dc4-4f82-8e67-aa834d41b1c2 · outbound

This paper cites H.; Zhou, X.; Choi, Y.; Goldberg, Y.; Sap, M.; and Shwartz, V.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks H.; Zhou, X.; Choi, Y.; Goldberg, Y.; Sap, M.; and Shwartz, V

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:45.785287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.001999Z digest=sha256:b95370f2886890faad5deeb5aef8786b2b30e0ba1de50c7649c11bca84d2972c

Observation 24f2bfbf-c2ba-4a96-8ef6-bba8fa773dfd · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:45.007633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:45.007633Z digest=sha256:dfdd875bea7704144eea840584c437ac65e55067de60b713a8eeb2ef8cf03461

Observation eb44456a-1b3b-4e55-8dd2-7a80243ecfe0 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.768040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.014937Z digest=sha256:80126eb04c62b9b79c9d390e67fb76aaf8e872e00002d4e449b12705cda37411

Observation 50395c04-7f3c-4874-a8e1-c12984b57b65 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.751506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.019830Z digest=sha256:6601cb446d0f0b548a025a4da38fc95625b31c0e48fef4be74ff49e0e799bd9a

Observation 3c8fa8f0-da92-4b99-8b1a-dc740b0cf524 · outbound

This paper cites LLM Theory of Mind and Alignment: Opportunities and Risks.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks LLM Theory of Mind and Alignment: Opportunities and Risks

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:45.024982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:45.024982Z digest=sha256:7b8e37371cdeee84b6a0e79fe70173b2f3c87ae195be1222de33f64017e5d412

Observation dba2078f-4897-407b-926d-62dfa5b805c7 · outbound

This paper cites LLMs achieve adult human performance on higher-order theory of mind tasks.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks LLMs achieve adult human performance on higher-order theory of mind tasks

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:45.030149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:45.030149Z digest=sha256:37b61c5688768b04210d5d3efd03572d43e0823c35c4d5a5c018c3e5f5cc2e6c

Observation 61779350-4ab4-4f45-aaff-8a8e8e3a7d48 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.735490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.035504Z digest=sha256:ef10fa1029abfa42dd15192ec92960c150d29c5b16f9b5e4b766a0c6633c405c

Observation 1ed7a4d9-058e-4b99-94e5-7960f1058790 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.719878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.040599Z digest=sha256:06946cb272955cc683fef5f28ebadf20dd7a081a91921ecd271d97b73701f65e

Observation 7ceda1a8-ad88-497e-b6b1-651955a74066 · outbound

This paper cites GenSim: A General Social Simulation Platform with Large Language Model based Agents.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks GenSim: A General Social Simulation Platform with Large Language Model based Agents

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:45.045448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:45.045448Z digest=sha256:3f22fb3087f66f1f1908f5287bf290cb2c2411d226555e54f86f99640874f3a5

Observation 6855c7d3-7d6d-4126-952b-77ceacf3eeb9 · outbound

This paper cites Steering Language Models With Activation Engineering.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Steering Language Models With Activation Engineering

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:45.050879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:45.050879Z digest=sha256:b3feaac3d129267e854d9a5d971e1e07d76d9b81121c2c44dcf274eda15052a2

Observation 1649c9ec-82ab-4eb1-a22c-e240d09e33e8 · outbound

This paper cites Large Language Models Fail on Trivial Alterations to Theory-of-Mind Tasks.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Large Language Models Fail on Trivial Alterations to Theory-of-Mind Tasks

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:45.056255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:45.056255Z digest=sha256:90c7069adc364100d6e4e0b0f952864bee84c031103055b7383de50af54b79b9

Observation 715b3a84-e7c1-4a58-8a95-fe7eb84c2449 · outbound

This paper cites AI Sandbagging: Language Models can Strategically Underperform on Evaluations.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks AI Sandbagging: Language Models can Strategically Underperform on Evaluations

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:45.061524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:45.061524Z digest=sha256:fb92124f53cf8b1a8e9e70b196fa11e5590bc749a7a7d381ba6310d7686c57cd

Observation 66fc2730-c3a7-4b5c-8c4c-5e0cbf9f2459 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.703545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.067050Z digest=sha256:3a167ea21aa35051d7be37e04a55b039efc6d85a8a7e6f98bbb81c3b6d635f7b

Observation da5ba737-5f0e-48bc-95d7-8df83b7aaaca · outbound

This paper cites P.; and Morency, L.-P.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks P.; and Morency, L.-P

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:45.686701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.071851Z digest=sha256:5e2c206265bf8940606f164f1955260547c776b0f79ff0f9426fd7b0eddac192

Observation 62454549-30fa-4401-b93a-ebd7638e1291 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.669506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.077014Z digest=sha256:f817536e3b843696041274171d20eacab96fa542da45bd190339f2536794af20

Observation a8c3addc-f454-4b1d-981b-340b4fffbb23 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.649870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.082337Z digest=sha256:df93f0da1767963f5de906aaaf2d9d6742db8a07e69d589ed173e23e42a6fd79

Observation a0d907c1-d251-469c-b824-fdd4f6acdd2d · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.630885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.087249Z digest=sha256:d7c94e87301056690a4a71d9510fc094a41c4cadca2522b15684b7b49447bfb0

Observation 98d547f2-6be6-438e-84a3-e244994410cb · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.613924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.092277Z digest=sha256:d58dc8911a8e8888fdbcd8bc042d74202986b7fa01825454cffcd0b0a385aa77

Observation 0b58a3d6-b9b1-42c3-896e-a3f41e1f58e0 · outbound

This paper cites A Careful Examination of Large Language Model Performance on Grade School Arithmetic.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks A Careful Examination of Large Language Model Performance on Grade School Arithmetic

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:45.097354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:45.097354Z digest=sha256:b97a5bd6ca0c58b286d3bea43cfb9515a63d1d7aa6bdedc10ff1184386ea32bb

Observation 658e47e3-2348-49f2-95c5-1b1e74bab22a · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.597345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.102646Z digest=sha256:c8b8c4666336986e9a2561f58f314430f9d8efc1fda396d98c672c6545bc69db

Observation 0bd53ef2-e478-4077-8ee6-18c833023895 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.580230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.107446Z digest=sha256:49aa9e8c0ea123c2155cac1fc2ed69bb41700f489ef3cff91d169bb54fc9a16a

Observation fea45c52-340f-4f29-92e7-1f2acde4c355 · outbound

This paper cites Representation Engineering: A Top-Down Approach to AI Transparency.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Representation Engineering: A Top-Down Approach to AI Transparency

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:45.112343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:45.112343Z digest=sha256:86790d8a20770a3e627fd58d151a924a2a765b53d24cdbda0b18c53784d9aa01

Pith citing papers

Observation 822582f7-0c6d-429d-9f80-a40485da0b2b · inbound

Agents Require Metacognitive and Strategic Reasoning to Succeed in the Coming Labor Markets cites this paper.

Agents Require Metacognitive and Strategic Reasoning to Succeed in the Coming Labor Markets A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:02:22.194525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:02:22.194525Z digest=sha256:170fed800ebb60f6749925a74c5051dc811111cb2ef44c4112203dfcf2fb06ba

Observation 841517d8-08d0-4bd8-9042-0a5d2ad0f541 · inbound

Does Theory of Mind Improvement Really Benefit Human-AI Interactions? Empirical Findings from Interactive Evaluations cites this paper.

Does Theory of Mind Improvement Really Benefit Human-AI Interactions? Empirical Findings from Interactive Evaluations A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-19T17:57:42.627226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-19T17:55:04.458343Z digest=sha256:445f938f6edcfbf87393936969156728fea2cfbb466277e2e19b89e2ab5e8837

Observation 50bbfd55-d1c8-4153-99a4-f826df319961 · inbound

Toward Human-Centered Multi-Agent Systems: Integrating Cognition, Culture, Values, and Cooperation in AI Agents cites this paper.

Toward Human-Centered Multi-Agent Systems: Integrating Cognition, Culture, Values, and Cooperation in AI Agents A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-02T22:37:25.774226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T18:47:36.189582Z digest=sha256:c27850ad1312f4b072eea3dc84427463c81f73d773fe3c2ac0003bc0708ae9c1