Pith. sign in

Paper Citation Record · LEDGER

HI-TOM: A Benchmark for Evaluating Higher-Order Theory of Mind Reasoning in Large Language Models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2310.16755.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.16755 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:04:34.719074Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T06:29:37.131424Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 93bfd8cb-0986-4884-9228-76953960d739 · inbound

THiNK: Can Large Language Models Think-aloud? cites this paper.

THiNK: Can Large Language Models Think-aloud? HI-TOM: A Benchmark for Evaluating Higher-Order Theory of Mind Reasoning in Large Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:34.719074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:34.719074Z digest=sha256:41558339b954b416c8bacbd29f558c824b05d59eed27e3dc1cedc477a5f8939c

Observation d99fafda-3e55-4188-8842-b0537ab993a0 · inbound

The Decrypto Benchmark for Multi-Agent Reasoning and Theory of Mind cites this paper.

The Decrypto Benchmark for Multi-Agent Reasoning and Theory of Mind HI-TOM: A Benchmark for Evaluating Higher-Order Theory of Mind Reasoning in Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T22:49:38.482038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:49:38.482038Z digest=sha256:e7c81416ec15d60a7d2735706c910999b32c97278132bf4d1fcc57bd85d2edbd

Observation ae2ae7dc-ff6e-4cbb-b296-a770c15b2b4e · inbound

Chinese Short-Form Creative Content Generation via Explanation-Oriented Multi-Objective Optimization cites this paper.

Chinese Short-Form Creative Content Generation via Explanation-Oriented Multi-Objective Optimization HI-TOM: A Benchmark for Evaluating Higher-Order Theory of Mind Reasoning in Large Language Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:52:06.073964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T20:50:53.673037Z digest=sha256:a55605e9072a0d498890d7a371629ce1024d852da4810c56d34c4352687fa2f0

Observation 482e5646-d75c-4dac-ae15-fdb5ac14ba9d · inbound

Reinforcing Human Behavior Simulation via Verbal Feedback cites this paper.

Reinforcing Human Behavior Simulation via Verbal Feedback HI-TOM: A Benchmark for Evaluating Higher-Order Theory of Mind Reasoning in Large Language Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:24:02.476776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-21T07:21:48.649289Z digest=sha256:7bcfcd42baf8daed455d8dbcb5f4647116e9bc8cfdc8644e602249d41d5124c0

Observation 79ef1173-ad4c-4cfe-8b61-6bb88139e717 · inbound

Social World Model for Lifelong Social Intelligence cites this paper.

Social World Model for Lifelong Social Intelligence HI-TOM: A Benchmark for Evaluating Higher-Order Theory of Mind Reasoning in Large Language Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:29:37.133283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T14:29:11.743627Z digest=sha256:a2b6d9499f3a1b32ab23feea5168d91287f7fc98d90841d9dc217038cafe9d35