Pith. sign in

Paper Citation Record · LEDGER

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks

As of 9 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 3 inbound Pith citation observations for arXiv:2502.06470.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.06470 v1

Coverage vector

measured 64 of 64 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T15:22:45.112343Z

measured 67 of 67 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:02:22.194525Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T22:37:25.772308Z

Reference resolution

64 of 64 outbound references displayed

  • verified exact0
  • verified fuzzy11
  • unresolved53
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6cd02e11-537a-4cc7-98e6-8f2b1124796a · outbound

This paper cites , " * write output.state after.block = add.period write newline.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.769888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.769888Z digest=sha256:9bd2cbfa903bdd8090a710665b10d0b45af11afa10d1526c4520748129dd0fb6

Observation 206b6ca1-d8df-48a8-8663-cab022113c61 · outbound

This paper cites write newline.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.776121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.776121Z digest=sha256:342df44293337ca048d2fd285b8d439cb66be5df5858a4d5b1beede647478a4f

Observation a877984e-bb44-4910-8304-89fc8beed9f0 · outbound

This paper cites Understanding intermediate layers using linear classifier probes.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Understanding intermediate layers using linear classifier probes

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.782384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.782384Z digest=sha256:777833449e4152128f65207066aa0a64be792feceb17e09f1bd5c399d8698b0c

Observation cc8883a8-ff25-4446-a322-8650f92e5d4f · outbound

This paper cites When Benchmarks are Targets: Revealing the Sensitivity of Large Language Model Leaderboards.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks When Benchmarks are Targets: Revealing the Sensitivity of Large Language Model Leaderboards

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.788604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.788604Z digest=sha256:692aa7e60ca52ebf1087605cf5d5fa1711145fb562d337988b43af73876ea18f

Observation 0b887c14-4be0-4c01-838d-5625f8d50d92 · outbound

This paper cites S.; Jenner, E.; Casper, S.; Sourbut, O.; Edelman, B.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks S.; Jenner, E.; Casper, S.; Sourbut, O.; Edelman, B

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:46.210568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.794498Z digest=sha256:62b0dcf942ebb2cc0d03f4ae213f21b9f1bf253973fb75bccf0e5d09ba9edd10

Observation 0746af7b-cddc-4883-ad73-40ed2558eedd · outbound

This paper cites theory of mind.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks theory of mind

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:46.191345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.800330Z digest=sha256:e31fafd064821816a45d0266112fafeabf561f1e28bcb5db7c5a1be4eee58f31

Observation de61f0dc-12f4-466a-97a9-87f51cb6a01f · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:46.173063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.805659Z digest=sha256:082152f0951f64f4350608f44122a2622c8f2574e3889da13dfe603252e864a0

Observation 29c2bc8d-cc02-46d0-95ea-e758b4a3f63c · outbound

This paper cites Defending Against Unforeseen Failure Modes with Latent Adversarial Training.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Defending Against Unforeseen Failure Modes with Latent Adversarial Training

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.811321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.811321Z digest=sha256:52ce3089747027b75c0a93e7879968b78755d27137461bf14c99605be2c32c25

Observation dcf5387b-fd12-4489-aeac-4e22d6570019 · outbound

This paper cites Designing a Dashboard for Transparency and Control of Conversational AI.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Designing a Dashboard for Transparency and Control of Conversational AI

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.817749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.817749Z digest=sha256:75eaa09e8fde3ed76b876c2ff547bc323c8d8b6ae155c3c59d396c9b04d996a9

Observation 2efbbbad-a974-46bf-96ad-17a5026d30f3 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:46.154807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.824077Z digest=sha256:74cd150368ca2cbf2cccb61df058dcb47c8120f93117c2af27f99e80fc43263d

Observation 31dc4491-51cd-4743-b9b5-81d2460fdd41 · outbound

This paper cites AI capabilities can be significantly improved without expensive retraining.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks AI capabilities can be significantly improved without expensive retraining

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.829284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.829284Z digest=sha256:e65fa18e3460957ba7d0c1d87a72d978d9d9c4d0eeae30918d50c33c9f6f83bf

Observation 573e6e4f-86e2-48e0-b037-5afba48391e0 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:46.137282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.834477Z digest=sha256:cedd64b192bf95510edc77d2e3c2ffbc70a87de7e6e08b05f5d9f5d76860603a

Observation 9b9d80f4-7702-4146-9f78-c8b104a6801c · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:46.120361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.839605Z digest=sha256:c0511d6819e89e58aa184c04a17d44a537e214c17f2d8cee41e030cbbebd66db

Observation 7ed83bd0-703b-4e47-aaf9-67b0cddebfc8 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.845044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.845044Z digest=sha256:791527f452002ea489eb2d7ed4fb90261789fb5d0e8abbeb5d1b7b2fcbaf25e6

Observation 30945236-e2b6-4f7c-b67e-affae368ed34 · outbound

This paper cites Intelligent Virtual Assistants with LLM-based Process Automation.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Intelligent Virtual Assistants with LLM-based Process Automation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.850403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.850403Z digest=sha256:758d9db28a2df2303e87ec04c5ef78c5af65ae9ce8e866a8e9ffed1e2e9746f7

Observation 8f902a8a-de8b-4d30-a549-24b9c9e8a546 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:46.092685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.855893Z digest=sha256:26bce526d61a825cb7d0a09888234fb804dccc5affcea056bc5b4eea7625f972

Observation 10da55c9-07aa-4dec-9455-7a4d5f6f9e2c · outbound

This paper cites Risks from Learned Optimization in Advanced Machine Learning Systems.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Risks from Learned Optimization in Advanced Machine Learning Systems

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.860935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.860935Z digest=sha256:d08e0e434c33a720f62dbcd11e116e34995a003c036792e809103284ab100406

Observation 2b8e680f-b342-4078-9380-d66d987ad91d · outbound

This paper cites AI safety via debate.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks AI safety via debate

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.871075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.871075Z digest=sha256:c528d663b728818e336a71494bb30ca16daf009226248131613290c507c1b995

Observation 933ebd2d-c29a-4a63-af67-3f2f0d5b513c · outbound

This paper cites Unveiling Theory of Mind in Large Language Models: A Parallel to Single Neurons in the Human Brain.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unveiling Theory of Mind in Large Language Models: A Parallel to Single Neurons in the Human Brain

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.876548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.876548Z digest=sha256:6debcc271c4d20666571b87fe1aaec20e435edba27f2fdafca3d7abeec2b4357

Observation f7702c32-0ac5-45b9-8bd8-3d9076b32dac · outbound

This paper cites Uncovering Deceptive Tendencies in Language Models: A Simulated Company AI Assistant.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Uncovering Deceptive Tendencies in Language Models: A Simulated Company AI Assistant

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.881881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.881881Z digest=sha256:4b64761fcb84e8d955d6404cceced86379e83e78ed2004cce31987c601f1089c

Observation 16b7ce84-6f18-43fd-9187-040d03b6a75d · outbound

This paper cites Y.; Kramar, J.; Brown-Cohen, J.; Albanie, S.; Bulian, J.; Agarwal, R.; Lindner, D.; Tang, Y.; Goodman, N.; and Shah, R.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Y.; Kramar, J.; Brown-Cohen, J.; Albanie, S.; Bulian, J.; Agarwal, R.; Lindner, D.; Tang, Y.; Goodman, N.; and Shah, R

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:46.075378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.887246Z digest=sha256:c2dfe4c7cac3dc7bc20aa8d3eefa62b791c5ae9d6ca8721dec240adea73a038f

Observation 80c60722-ab96-48ed-864c-c318fbbdf98a · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:46.057915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.892203Z digest=sha256:2647e8f5713e7aa6889d7748371f2be9ef8cff4f6a4219fe0bf806374452d7ae

Observation 3d3ad469-529f-4724-9e18-127e5e95c28b · outbound

This paper cites M.; Kundu, A.; Jawhar, S.; Park, J.; and Jurewicz, M.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks M.; Kundu, A.; Jawhar, S.; Park, J.; and Jurewicz, M

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:46.040914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.897457Z digest=sha256:c7775f4288824ed2c0b432740f7017d52af3eff16966ee1ca7ab1bb265313fe4

Observation fdf842a5-b312-4806-a4d4-f53831836d5e · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:46.024838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.902260Z digest=sha256:44ab27b4638fad14f58e666d3d7cdb9901271c5240f8d4292ae52fd5f7e5e479

Observation 8f2bd0b3-d6aa-4505-9e26-d1f0ae003aec · outbound

This paper cites Q.; Stepputtis, S.; Campbell, J.; Hughes, D.; Lewis, C.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Q.; Stepputtis, S.; Campbell, J.; Hughes, D.; Lewis, C

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:46.009315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.907051Z digest=sha256:a9d332fb73b37afeca8846b516bfd211508a317e29531ae503ae4719be42c335

Observation 89d5da1c-959a-405f-bed3-532545066a16 · outbound

This paper cites D.; Dombrowski, A.-K.; Goel, S.; Mukobi, G.; Helm-Burger, N.; Lababidi, R.; Justen, L.; Liu, A.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks D.; Dombrowski, A.-K.; Goel, S.; Mukobi, G.; Helm-Burger, N.; Lababidi, R.; Justen, L.; Liu, A

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:45.992958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.911934Z digest=sha256:4d8ae3b0e499d40e356e78e119f38234465afa6d231497b5050116dd784a1c79

Observation 9774ac63-f2b6-4993-8d80-81d6ad9d400a · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.976152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.917280Z digest=sha256:dcf78f908b24a6601188e56daff0b01575cbbd484a4772be8cda0db2959d9744

Observation 4b8c5de1-8909-4e67-8d58-6f9a875dd5e6 · outbound

This paper cites Rethinking Machine Unlearning for Large Language Models.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Rethinking Machine Unlearning for Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.922410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.922410Z digest=sha256:59b107dfcb219b346b1a9cde124d69d834df16f981fe728fae233e696a09e70b

Observation 57490763-0d01-4943-8e3b-de84b2446e4c · outbound

This paper cites S.; Cope, D.; and Schoots, N.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks S.; Cope, D.; and Schoots, N

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:45.959915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.927713Z digest=sha256:2c7b08cbaadf4e5f093571e8809f55357192d6cc45548af506331ef7be7862da

Observation 893bfeee-1300-4124-874d-4ebec08404da · outbound

This paper cites R.; Baranchuk, M.; Strohmeier, M.; Bolina, V.; Torr, P.; Hammond, L.; and de Witt, C.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks R.; Baranchuk, M.; Strohmeier, M.; Bolina, V.; Torr, P.; Hammond, L.; and de Witt, C

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:45.943348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.932771Z digest=sha256:337b9129a0b39dc639bad18fd511276b085a891d23f45fe7e02036778deb4679

Observation 4e1c8841-c02f-4d93-8722-3e329ad24936 · outbound

This paper cites Welfare Diplomacy: Benchmarking Language Model Cooperation.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Welfare Diplomacy: Benchmarking Language Model Cooperation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.937791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.937791Z digest=sha256:e05a0522228f88ed852201f117755cd399728420d06ce9d7a540cc5330d1d338

Observation 3c97cda4-e145-48cc-a58b-ba06e4c45106 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.926410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.943052Z digest=sha256:c8ee09f14a429397133963d6dc8e3ad7884d53a83d4cd72536faafdf2eea2808

Observation c3de1ec6-b36a-475a-b61e-2c634e000f78 · outbound

This paper cites GPT-4 Technical Report.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks GPT-4 Technical Report

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.948112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.948112Z digest=sha256:173bf2ebf77b9463411d3548f91f8804856ed224e1cda44ceebb38f301754054

Observation 1d6a2552-dec3-405b-a988-6d2b7e4b2826 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.910445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.953589Z digest=sha256:c85a558823e59cb5eb5ac9a9ac589db9de24da20392efa8d259ee2f8563b6dc6

Observation 36d6f1b0-29c1-450d-baf5-214319ce0742 · outbound

This paper cites Generative Agents: Interactive Simulacra of Human Behavior.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Generative Agents: Interactive Simulacra of Human Behavior

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.958547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.958547Z digest=sha256:ad36ea32b69a7836e43e06c9a48236e997d6da5705fed68e1eeba4588a663265

Observation 9484155d-7282-4476-bcba-696e0bc11699 · outbound

This paper cites S.; Goldstein, S.; O’Gara, A.; Chen, M.; and Hendrycks, D.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks S.; Goldstein, S.; O’Gara, A.; Chen, M.; and Hendrycks, D

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:45.892387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.964464Z digest=sha256:77938f710a56213cbcf419a04d9640792d57393ccb7598fc1af7de815de585b8

Observation d071f3cc-b5e0-4a43-b38e-4e0cb6c9ffa1 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.876076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.969455Z digest=sha256:ba7d7b938d66a857d88562e1a95cfd7084b22e4eaa31a3f1c41c66dc3478bb4a

Observation 21f19ec5-6798-4ca7-b008-49f6f07a591b · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.858963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.974926Z digest=sha256:4445a5b6681d2e913d7f36317d3ef68896523e9c446e793880b93f791c06087e

Observation de7d4bf4-d26a-4ecd-84c7-e22222b2a28e · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.841531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.980417Z digest=sha256:6de942adbb56da638284a2e81f11073de68a05c6481f4ed23ba7b812213fee7b

Observation e7ac3f25-1cb4-40d0-a48f-ab403d9f84fe · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.821466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.985768Z digest=sha256:9b7e15fe235b4593c7d342ff989f793b73a5be2ddeee0e2a1deaceb00cc7a482

Observation 5f9dde6a-183b-4b27-85fd-b14beac53a30 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.802716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:44.991158Z digest=sha256:0216143464b5c14151a4eb3cc2b93ee26bbf1f728408ab3fd1f8da3e5af0ac56

Observation 0ee4f544-be7d-4c64-8f5d-0b8459439d51 · outbound

This paper cites Transformers represent belief state geometry in their residual stream.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Transformers represent belief state geometry in their residual stream

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:44.996501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:44.996501Z digest=sha256:128fc0152acbf0b5087150a9dfab422b770964515352b2ff022a36e55d8bd498

Observation 25876f44-6dc4-4f82-8e67-aa834d41b1c2 · outbound

This paper cites H.; Zhou, X.; Choi, Y.; Goldberg, Y.; Sap, M.; and Shwartz, V.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks H.; Zhou, X.; Choi, Y.; Goldberg, Y.; Sap, M.; and Shwartz, V

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:45.785287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.001999Z digest=sha256:2b846e9fef820c12242a05124e3cd5f374ed6410f06202ff8c38360caaa7925b

Observation 24f2bfbf-c2ba-4a96-8ef6-bba8fa773dfd · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:45.007633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:45.007633Z digest=sha256:901b5e31a9cb097bd6a074703c510502f7c7d2472cca48de437a621d7c52ccfb

Observation eb44456a-1b3b-4e55-8dd2-7a80243ecfe0 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.768040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.014937Z digest=sha256:71d92c201e9d20529b2f8f51ad2f5b36810762f022bf4e2a791361938fc808ac

Observation 50395c04-7f3c-4874-a8e1-c12984b57b65 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.751506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.019830Z digest=sha256:80e019e5785640955207ec66a4896d1aac2dcb1944086e586c87300ac4c145ac

Observation 3c8fa8f0-da92-4b99-8b1a-dc740b0cf524 · outbound

This paper cites LLM Theory of Mind and Alignment: Opportunities and Risks.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks LLM Theory of Mind and Alignment: Opportunities and Risks

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:45.024982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:45.024982Z digest=sha256:b8ac884c4a0fdad8b6ea80a0318c1a06628312374a5191fdf2b785fdadfd6164

Observation dba2078f-4897-407b-926d-62dfa5b805c7 · outbound

This paper cites LLMs achieve adult human performance on higher-order theory of mind tasks.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks LLMs achieve adult human performance on higher-order theory of mind tasks

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:45.030149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:45.030149Z digest=sha256:8c43e1ee36303d1a749228c70df853607697895f0ebc5a4fc930f8008dcb72b0

Observation 61779350-4ab4-4f45-aaff-8a8e8e3a7d48 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.735490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.035504Z digest=sha256:aabd5f861d13919ca1072fc9805119d33f455d07ab57a3399ed253c1d07e4054

Observation 1ed7a4d9-058e-4b99-94e5-7960f1058790 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.719878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.040599Z digest=sha256:6de31173d989f04e8a75fddfe5d736e092c3f617cae9327ae65b87168b76057f

Observation 7ceda1a8-ad88-497e-b6b1-651955a74066 · outbound

This paper cites GenSim: A General Social Simulation Platform with Large Language Model based Agents.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks GenSim: A General Social Simulation Platform with Large Language Model based Agents

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:45.045448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:45.045448Z digest=sha256:95908c287270e1f0a7b5ac33ec10a51db18ba50bb75738c4aa91495283a69ec1

Observation 6855c7d3-7d6d-4126-952b-77ceacf3eeb9 · outbound

This paper cites Steering Language Models With Activation Engineering.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Steering Language Models With Activation Engineering

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:45.050879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:45.050879Z digest=sha256:ecf560eba37caf2842c9723ef3e6108d866e2cf1a5a55d089395c4f793972460

Observation 1649c9ec-82ab-4eb1-a22c-e240d09e33e8 · outbound

This paper cites Large Language Models Fail on Trivial Alterations to Theory-of-Mind Tasks.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Large Language Models Fail on Trivial Alterations to Theory-of-Mind Tasks

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:45.056255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:45.056255Z digest=sha256:68c24766e375d2f5a2d2637c2f215c4c69b8a59b9e12901ed7e1bb55b9463ee4

Observation 715b3a84-e7c1-4a58-8a95-fe7eb84c2449 · outbound

This paper cites AI Sandbagging: Language Models can Strategically Underperform on Evaluations.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks AI Sandbagging: Language Models can Strategically Underperform on Evaluations

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:45.061524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:45.061524Z digest=sha256:55ef58f99837480dd10c2fa14a7898c9b859a1edae8a55d8191d7416e42dc0f1

Observation 66fc2730-c3a7-4b5c-8c4c-5e0cbf9f2459 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.703545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.067050Z digest=sha256:a44a345873d40cca5c95089be39d03cd0665c1b84288f0cd630fbc300f7b197b

Observation da5ba737-5f0e-48bc-95d7-8df83b7aaaca · outbound

This paper cites P.; and Morency, L.-P.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks P.; and Morency, L.-P

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T15:22:45.686701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.071851Z digest=sha256:875a6c060c4ddee3c0439363317c3f38b6fc513391ae9d998aa01ca177c24452

Observation 62454549-30fa-4401-b93a-ebd7638e1291 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.669506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.077014Z digest=sha256:bc55d8a58f9f1435beeff7267ae978706f709d1137ca939431969deff44cea9b

Observation a8c3addc-f454-4b1d-981b-340b4fffbb23 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.649870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.082337Z digest=sha256:26c914bed008643c4817735c3d0da7479bb2bf3bb7a9d0548c926ed59729388e

Observation a0d907c1-d251-469c-b824-fdd4f6acdd2d · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.630885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.087249Z digest=sha256:7e6250c209b49db7b89ebbbd6f39f6a9d1d107ee516f33534267c064952c6452

Observation 98d547f2-6be6-438e-84a3-e244994410cb · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.613924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.092277Z digest=sha256:1a40af27e93576a5c99f33115f528f04ed48085ba0c87d9cec6fd6a47ba2d6b7

Observation 0b58a3d6-b9b1-42c3-896e-a3f41e1f58e0 · outbound

This paper cites A Careful Examination of Large Language Model Performance on Grade School Arithmetic.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks A Careful Examination of Large Language Model Performance on Grade School Arithmetic

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:45.097354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:45.097354Z digest=sha256:1f48053c9b65b7ae8d19580dcdab165c9fa2ca96e63fdf30407276c05191bfad

Observation 658e47e3-2348-49f2-95c5-1b1e74bab22a · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.597345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.102646Z digest=sha256:dd6fe2092e0373934b6b4bc18f650e96b1e9c46fbaeeb547abe97cbe1ddb7a52

Observation 0bd53ef2-e478-4077-8ee6-18c833023895 · outbound

This paper cites an unresolved cited work.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-08T15:22:45.580230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T15:22:45.107446Z digest=sha256:ceea2fcc582a0b12ac2e82bfe86a983ae7a72332f6152242acf4a0b577ebfc2c

Observation fea45c52-340f-4f29-92e7-1f2acde4c355 · outbound

This paper cites Representation Engineering: A Top-Down Approach to AI Transparency.

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks Representation Engineering: A Top-Down Approach to AI Transparency

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-08T15:22:45.112343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T15:22:45.112343Z digest=sha256:0877db7f64dc0aa8ff3737604b7845306106e6a6e054d33eb41208f096675435

Pith citing papers

Observation 822582f7-0c6d-429d-9f80-a40485da0b2b · inbound

Agents Require Metacognitive and Strategic Reasoning to Succeed in the Coming Labor Markets cites this paper.

Agents Require Metacognitive and Strategic Reasoning to Succeed in the Coming Labor Markets A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:02:22.194525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:02:22.194525Z digest=sha256:4b464dfd9a496836369970a252e1f548b7a5714235f4d73128dd86eeda5871b6

Observation 841517d8-08d0-4bd8-9042-0a5d2ad0f541 · inbound

Does Theory of Mind Improvement Really Benefit Human-AI Interactions? Empirical Findings from Interactive Evaluations cites this paper.

Does Theory of Mind Improvement Really Benefit Human-AI Interactions? Empirical Findings from Interactive Evaluations A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-19T17:57:42.627226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-19T17:55:04.458343Z digest=sha256:795a77734e5a4d8354c70fb0f550aec6a7d9bb475ce91ea087ee4740626da75b

Observation 50bbfd55-d1c8-4153-99a4-f826df319961 · inbound

Toward Human-Centered Multi-Agent Systems: Integrating Cognition, Culture, Values, and Cooperation in AI Agents cites this paper.

Toward Human-Centered Multi-Agent Systems: Integrating Cognition, Culture, Values, and Cooperation in AI Agents A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-02T22:37:25.774226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T18:47:36.189582Z digest=sha256:bfbfdf706be0de5fe797f4983aee1d5fd4285d9928e1e6ddb97d5ff6fe3e8da0