Pith. sign in

Paper Citation Record · LEDGER

HI-TOM: A Benchmark for Evaluating Higher-Order Theory of Mind Reasoning in Large Language Models

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2310.16755.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.16755 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:04:34.719074Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T06:29:37.131424Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 93bfd8cb-0986-4884-9228-76953960d739 · inbound

THiNK: Can Large Language Models Think-aloud? cites this paper.

THiNK: Can Large Language Models Think-aloud? HI-TOM: A Benchmark for Evaluating Higher-Order Theory of Mind Reasoning in Large Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:34.719074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:34.719074Z digest=sha256:41558339b954b416c8bacbd29f558c824b05d59eed27e3dc1cedc477a5f8939c

Observation d99fafda-3e55-4188-8842-b0537ab993a0 · inbound

The Decrypto Benchmark for Multi-Agent Reasoning and Theory of Mind cites this paper.

The Decrypto Benchmark for Multi-Agent Reasoning and Theory of Mind HI-TOM: A Benchmark for Evaluating Higher-Order Theory of Mind Reasoning in Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T22:49:38.482038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:49:38.482038Z digest=sha256:ef8b4ecb821509ab4e5ed72b11017132f07212594bdbbab2b206c70a2011fbea

Observation ae2ae7dc-ff6e-4cbb-b296-a770c15b2b4e · inbound

Chinese Short-Form Creative Content Generation via Explanation-Oriented Multi-Objective Optimization cites this paper.

Chinese Short-Form Creative Content Generation via Explanation-Oriented Multi-Objective Optimization HI-TOM: A Benchmark for Evaluating Higher-Order Theory of Mind Reasoning in Large Language Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:52:06.073964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T20:50:53.673037Z digest=sha256:012fcfdda4146adbbbf42a7e9ff0d7e9b7887b6e583d51894d35befb6fb561bf

Observation 482e5646-d75c-4dac-ae15-fdb5ac14ba9d · inbound

Reinforcing Human Behavior Simulation via Verbal Feedback cites this paper.

Reinforcing Human Behavior Simulation via Verbal Feedback HI-TOM: A Benchmark for Evaluating Higher-Order Theory of Mind Reasoning in Large Language Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:24:02.476776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-21T07:21:48.649289Z digest=sha256:4272cb77bd8ae2a95f9e3ba77a00d181ddf4425b09ad0bc3b68ee624a5a71372

Observation 79ef1173-ad4c-4cfe-8b61-6bb88139e717 · inbound

Social World Model for Lifelong Social Intelligence cites this paper.

Social World Model for Lifelong Social Intelligence HI-TOM: A Benchmark for Evaluating Higher-Order Theory of Mind Reasoning in Large Language Models

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:29:37.133283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T14:29:11.743627Z digest=sha256:c92c7de4028160a5dcee5ae631754669a4db003ca68ee2f6dc88837d117b8d10