Pith. sign in

Paper Citation Record · LEDGER

FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2310.15421.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.15421 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T04:32:42.318584Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T23:17:29.774947Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation dca4600e-da31-434b-8300-929d10bb0b04 · inbound

Code Simulation as a Proxy for High-order Tasks in Large Language Models cites this paper.

Code Simulation as a Proxy for High-order Tasks in Large Language Models FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-09T04:32:42.318584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:32:42.318584Z digest=sha256:6fcbe92a68507d73c731e59e475e520507b3034d3aa90dffd5b54a2541e040aa

Observation 376a9404-d10c-4302-84ff-918afe3a131b · inbound

SocialMaze: A Benchmark for Evaluating Social Reasoning in Large Language Models cites this paper.

SocialMaze: A Benchmark for Evaluating Social Reasoning in Large Language Models FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:18.777761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:18.777761Z digest=sha256:89fe9830a4ab6cb07baa73460a184db3f3d3e6ecc46becaf26390373b2b1cf3b

Observation f6ea748b-9302-42d0-b5ba-23bdd750f2f9 · inbound

Beyond Context to Cognitive Appraisal: Emotion Reasoning as a Theory of Mind Benchmark for Large Language Models cites this paper.

Beyond Context to Cognitive Appraisal: Emotion Reasoning as a Theory of Mind Benchmark for Large Language Models FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:45.594867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:10:45.594867Z digest=sha256:ba8670fb8cbb572dd9f3f1a6705e7e0d535c52cbb55baeb484afed254ac3fe30

Observation e5f44678-fdfe-4b2f-8e5a-a448a1081d57 · inbound

Does It Make Sense to Speak of Introspection in Large Language Models? cites this paper.

Does It Make Sense to Speak of Introspection in Large Language Models? FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T10:30:37.634383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:30:37.634383Z digest=sha256:6ad28c99ef4d8d8a3ab4d9546c293f81bb70d7b0316fe2c18a02b6f6514aa00f

Observation 4e089954-2179-40a6-b911-2530b8d02f70 · inbound

Intentionally Unintentional: GenAI Exceptionalism and the First Amendment cites this paper.

Intentionally Unintentional: GenAI Exceptionalism and the First Amendment FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:15.706574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:15.706574Z digest=sha256:58b08b6c6e43658a29a6d37755299e48241f13c9d751a685a7825cb853d6a402

Observation 82b6199b-6fc3-4418-85b2-4319bb337cba · inbound

The Decrypto Benchmark for Multi-Agent Reasoning and Theory of Mind cites this paper.

The Decrypto Benchmark for Multi-Agent Reasoning and Theory of Mind FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-06T22:49:38.805574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:49:38.805574Z digest=sha256:23ccfc367c8b017e17e34ff60b58f6b3e0192093b3989583f841f541a7500387

Observation 5e06b11d-0b66-4ad5-951d-9eef5f8694cc · inbound

Mind the Motions: Benchmarking Theory-of-Mind in Everyday Body Language cites this paper.

Mind the Motions: Benchmarking Theory-of-Mind in Everyday Body Language FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:04:17.452832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T18:04:07.256885Z digest=sha256:946ed87190fbc1b35b1a31b76fa2bc0439124f3b8be6cfde2f16ed113cc6a486

Observation ee95f333-29cd-410b-827a-547663070bde · inbound

Shadow-Loom: Causal Reasoning over Graphical World Models of Narratives cites this paper.

Shadow-Loom: Causal Reasoning over Graphical World Models of Narratives FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T06:15:38.644850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T18:41:38.814086Z digest=sha256:0925818f9f628320bcb4b42dd0c34bf5f96dda894bf1e29efd27391ba6e77516

Observation b8612cc8-e1c1-4027-866c-0dcae289a8ec · inbound

Reinforcing Human Behavior Simulation via Verbal Feedback cites this paper.

Reinforcing Human Behavior Simulation via Verbal Feedback FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T07:24:02.439889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-21T07:21:48.649289Z digest=sha256:ee70356329462db2a9339289dd86725d42b8a8c1c3e6e2b166fafe33ed10412a

Observation 3eec17da-3b16-4d4e-8181-5fb696eb3057 · inbound

PerspectiveGap: A Benchmark for Multi-Agent Orchestration Prompting cites this paper.

PerspectiveGap: A Benchmark for Multi-Agent Orchestration Prompting FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-02T23:17:29.776514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T18:19:06.340769Z digest=sha256:e42b3e14c238c7bc57260afa1acd0224ac42923039c74dd7033f6280656cdd20

Observation ea89d470-3d85-432e-9ded-37d101a518d0 · inbound

PerspectiveGap: A Benchmark for Multi-Agent Orchestration Prompting cites this paper.

PerspectiveGap: A Benchmark for Multi-Agent Orchestration Prompting FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-14T18:11:45.966902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T18:11:45.966902Z digest=sha256:c0726cf245b3ba8738b41abc8bef2996641ef515ee499d633d9a5b054832dda5

Observation 088b21c5-b056-4d5f-bb8b-1219dc7f646d · inbound

Triadic Werewolf: A Jester Role for Multi-Hop Theory of Mind in LLMs cites this paper.

Triadic Werewolf: A Jester Role for Multi-Hop Theory of Mind in LLMs FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T19:13:53.397462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T04:48:58.181882Z digest=sha256:59f50004443cf033c32ac2349560d71a6f2ffa85ddd9184c2248b6198cc682c8

Observation c8fcd3cf-db62-4f3a-a3ed-eeb48e03be77 · inbound

Training with (Swap) Regret Loss in a Single-Layer Self-Attention Model: A Case Study on the Probability Simplex cites this paper.

Training with (Swap) Regret Loss in a Single-Layer Self-Attention Model: A Case Study on the Probability Simplex FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 71

Resolution
unresolved
no resolver link, observed 2026-07-31T23:51:55.026505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T23:51:55.026505Z digest=sha256:8cbfc18569f469845f502e6432801ab680faa80d51be61e77ea0f15a4bd82a50

Observation 8db2ee0f-190a-4eb2-85eb-4c69f7fed3cc · inbound

Evaluating Theory of Mind in Reasoning Models: Robustness over Reasoning cites this paper.

Evaluating Theory of Mind in Reasoning Models: Robustness over Reasoning FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T19:57:56.664963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:57:56.664963Z digest=sha256:9dd525d06112a0ab1c261842f1391cec10834e7b88ac0de33e36cfbe8431bcf2