Pith. sign in

Paper Citation Record · LEDGER

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models

As of 8 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2506.14224.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.14224 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:28:26.619534Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact2
  • verified fuzzy16
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fd2d8bbe-6533-46ae-956f-67c9aadf7717 · outbound

This paper cites GPT-4 Technical Report.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:22.948708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:22.948708Z digest=sha256:c07d29e177344c91a35f51d7d3bd00fe528b06c7d86f377167d67acc213c1007

Observation 87f72791-366a-4e56-a4cb-f8866ea1f799 · outbound

This paper cites Do llms exhibit human-like reasoning? evaluating theory of mind in llms for open-ended responses.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Do llms exhibit human-like reasoning? evaluating theory of mind in llms for open-ended responses

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:37.218093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:23.009505Z digest=sha256:63febb4a56b58fa75a698550a9e263be5f23fe6875540c3447c0112e01c06591

Observation aa45be4f-e851-4a68-9d40-da8bbce87f85 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:23.130188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:23.130188Z digest=sha256:e23f4296c3146282bf4473a9ad0076c1d966db622005f28159b796aa39bfe376

Observation bde74608-22b4-4b74-ada7-4cd6a1224c81 · outbound

This paper cites Hallucination of Multimodal Large Language Models: A Survey.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Hallucination of Multimodal Large Language Models: A Survey

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:23.244204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:23.244204Z digest=sha256:ab2344e91667df53592500782e3b087662cb1de3436eeac643f5584d7d0609fe

Observation 55c22fba-95c5-4f34-b47d-e88240dd0711 · outbound

This paper cites theory of mind.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models theory of mind

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:36.966034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:23.327105Z digest=sha256:83587688478f952509395f46cbec16c5e6c067519abc4c91f9d77dd72a12e47e

Observation 31da6107-d971-465d-a0f9-041b363b03ec · outbound

This paper cites an unresolved cited work.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:28:36.762234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:23.437074Z digest=sha256:adde8fd8bbcf7682f0284a337824ec9a8e109526c6230adae295e2b8cb731a5b

Observation 33c20f2f-5e9c-4280-8c2c-499f650af28e · outbound

This paper cites Sparks of Artificial General Intelligence: Early experiments with GPT-4.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Sparks of Artificial General Intelligence: Early experiments with GPT-4

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:23.499151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:23.499151Z digest=sha256:5b75a42704c15e3dbd17e65b048637df82244aa68f23b4e0c9cca4b1b03beb44

Observation c2345e0f-fd69-4acf-b503-ca5bf8d94200 · outbound

This paper cites Through the theory of mind's eye: Reading minds with multimodal video large language models.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Through the theory of mind's eye: Reading minds with multimodal video large language models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:23.564854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:23.564854Z digest=sha256:32a2b7878e152dae1b654b582907590035edad6d2ca1db9094255f1d09ae63f2

Observation 806965e1-0ce4-4b89-b66d-9da4b55bbcb6 · outbound

This paper cites Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:23.677281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:23.677281Z digest=sha256:440e4da4bfe3797e59f60a9409802d87dec4f0534fd6cd3b975b0112609485c6

Observation 35991cdc-34e0-4217-9f59-6aa764bda43f · outbound

This paper cites The Llama 3 Herd of Models.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models The Llama 3 Herd of Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:23.776668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:23.776668Z digest=sha256:26c1e7bc6e8659c2667ab2099503189ada97cb5f089b7513600a863806756e81

Observation 3cc93695-7564-430b-b4bd-0945222aaadf · outbound

This paper cites Who is Mistaken?.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Who is Mistaken?

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:28:27.010971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:23.869613Z digest=sha256:059a44267fbe8d05e87356c08321db3a576bd554bb2fd801fd6d9aeaad4ebbd5

Observation 93e44000-eab0-43db-8a37-182334a3cdc5 · outbound

This paper cites M., and Dillon, M.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models M., and Dillon, M

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:36.582680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:23.936271Z digest=sha256:f943a6f470fadb4d2b796c9e20320ec0240832e3d87f720e67a1f37db2864354

Observation 5ca30607-29df-4e9c-8e4e-cc2deb0c23c0 · outbound

This paper cites an unresolved cited work.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:28:36.322279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:24.030959Z digest=sha256:1073107ab5f8f1a4401f07f11b44a42113a47807b60ad92b38780a6b34cb744e

Observation 54bd068a-a2ce-434b-bfda-c4b6c17ca3da · outbound

This paper cites Distributional vectors encode referential attributes.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Distributional vectors encode referential attributes

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:36.092670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:24.105076Z digest=sha256:c8e91b14f8e733428998a27899ccf2a1d6aca5aa013e3ab106d23184f858b175

Observation 19e4d656-eba0-4d4d-a1b5-e51cbc3c0fbc · outbound

This paper cites Mistral 7B.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Mistral 7B

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:24.202247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:24.202247Z digest=sha256:0468f80fd7db999188fa02478d209691f7dc84261e82877038ad3c40f03dcc45

Observation 207098c1-491d-41e7-a4c4-791f40507776 · outbound

This paper cites MMT o M - QA : Multimodal theory of mind question answering.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models MMT o M - QA : Multimodal theory of mind question answering

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:24.323659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:24.323659Z digest=sha256:9f46012092bc68e85ee66d14c9ebbf81d24bb64e5d52aeec3e0008caf7e468c7

Observation 818afb4d-91ae-48c4-a0b6-0061a4073722 · outbound

This paper cites Evaluating Large Language Models in Theory of Mind Tasks.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Evaluating Large Language Models in Theory of Mind Tasks

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:24.444328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:24.444328Z digest=sha256:f8b1a7d1bbed46931c596d73ec011f5de812bd6c7612deef1e691723379020e7

Observation 9f5bbe94-3938-4c56-af15-3a1a2aa71bfc · outbound

This paper cites Evaluating large language models in theory of mind tasks.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Evaluating large language models in theory of mind tasks

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:35.865227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:24.488322Z digest=sha256:a9e761b961c3d8c2ed220f0a623be919f4029d33bf9d95c83d257d54ab600780

Observation c790b181-2db8-43eb-8b26-9c26ddab3fa7 · outbound

This paper cites What’s in an embedding? Analyzing word embeddings through multilingual evaluation.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models What’s in an embedding? Analyzing word embeddings through multilingual evaluation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:35.638751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:24.610920Z digest=sha256:d3ca9129d1c40beef9dca0591ab65173787f6c7ef8b615a907090736bdc6c16e

Observation 0582280e-b90f-4db8-99eb-cc394f11f674 · outbound

This paper cites Revisiting the evaluation of theory of mind through question answering.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Revisiting the evaluation of theory of mind through question answering

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:35.459458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:24.709661Z digest=sha256:8f84b67f5bf6a7248b21cbb7b3ee85d9adcf1eb1e11612789bf2707f2d72cc6c

Observation e4ead157-f5aa-474c-9393-83bf28cea292 · outbound

This paper cites Inference-time intervention: Eliciting truthful answers from a language model.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Inference-time intervention: Eliciting truthful answers from a language model

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:35.222704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:24.801191Z digest=sha256:9f51e326ff4aa39573beb8af4044c4cdcdcfb9cd080f96db76751db74a9057e2

Observation 235e191c-b256-4637-8238-077aaf6fce10 · outbound

This paper cites DeepSeek-V3 Technical Report.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models DeepSeek-V3 Technical Report

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:24.852810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:24.852810Z digest=sha256:724cd9893548886f1fbd2ea12a3bed48e633272dbe463047e48c37a6462cd54d

Observation 33ebb050-394e-4dec-92a1-c6fb30297b48 · outbound

This paper cites Large Model Strategic Thinking, Small Model Efficiency: Transferring Theory of Mind in Large Language Models.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Large Model Strategic Thinking, Small Model Efficiency: Transferring Theory of Mind in Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:24.942418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:24.942418Z digest=sha256:8903494563f854e386d02a0162bd2e746b72067e1e4e1d4d99912f2c154a3942

Observation 0167ac2f-a22a-4dec-86e7-475019775c64 · outbound

This paper cites Towards a holistic landscape of situated theory of mind in large language models.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Towards a holistic landscape of situated theory of mind in large language models

Reference 24

Resolution
verified exact
doi, observed 2026-08-07T00:28:26.780658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:25.051440Z digest=sha256:d43b700d78db1107217e13aee9f788744a9229e65acbf280975ba63516898436

Observation 22cdedd6-ae45-4b79-9524-65c4b6c5edd5 · outbound

This paper cites A review on machine theory of mind.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models A review on machine theory of mind

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:34.996724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:25.122747Z digest=sha256:632734883d1b0cae502fb47d221fc96fdc1948ad5c794d5a98e5604c4dd68b4c

Observation 3aca20ec-dbf5-4269-bc68-00a205d13a24 · outbound

This paper cites Evaluating theory of mind in question answering.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Evaluating theory of mind in question answering

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:25.167889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:25.167889Z digest=sha256:315be25134976056ba9df543b3eccbb6687b5e6dddd7cbfab29a57907b037ef0

Observation b432b1bf-fac7-47ec-a019-7ed73ce1fd7b · outbound

This paper cites Theory of Mind as Intrinsic Motivation for Multi-Agent Reinforcement Learning.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Theory of Mind as Intrinsic Motivation for Multi-Agent Reinforcement Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:25.231349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:25.231349Z digest=sha256:ec5581788dfe94eca6544bc5867bf2b14a9ef17d0ef3605ceca11f6fac8b2a20

Observation c2e44e34-44e8-44a5-ae70-64ebd83a6716 · outbound

This paper cites Neural theory-of-mind? on the limits of social intelligence in large LM s.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Neural theory-of-mind? on the limits of social intelligence in large LM s

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:25.311237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:25.311237Z digest=sha256:e08e6401b71c796a4ce8516aa31d60f38f9f76cb37b8d0f92de4c767a7dfb31b

Observation 13646d49-ec21-4f07-aff8-1b8d7399473e · outbound

This paper cites Symmetric machine theory of mind.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Symmetric machine theory of mind

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:34.793911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:25.364111Z digest=sha256:083fe1fc8f4742d23a7cf37a7f4cd690303751912e2196cd3f58e4fa7c316b7b

Observation 0555ea13-19fe-4d7e-957e-33de2bce04a4 · outbound

This paper cites H., Zhou, X., Choi, Y., Goldberg, Y., Sap, M., and Shwartz, V.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models H., Zhou, X., Choi, Y., Goldberg, Y., Sap, M., and Shwartz, V

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:34.560220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:25.434622Z digest=sha256:7c0d3601d901d76896fadf5027132538927709612e9194e59be5b4eeeedad448

Observation ee61ebd7-42ca-409b-83f3-bf72f2888641 · outbound

This paper cites Muma-tom: Multi-modal multi-agent theory of mind.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Muma-tom: Multi-modal multi-agent theory of mind

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:34.338427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:25.540625Z digest=sha256:c37d0efe8ac89392d517c99d5fe465da6f17bae645e82783f541f28dd04827c1

Observation a596cc83-ea41-43cc-be45-acf1aa7b9b22 · outbound

This paper cites Agent: A benchmark for core psychological reasoning.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Agent: A benchmark for core psychological reasoning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:34.180366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:25.668627Z digest=sha256:5b75519c0e56bf510f00202a0c5bb099958a0ad7059d0bddc374395f0853950c

Observation 5f10b3df-74e7-47a0-9dc2-c389bcb32a5f · outbound

This paper cites and Lernould, A.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models and Lernould, A

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:25.799002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:25.799002Z digest=sha256:a2e0be88c516d082fe2d91c9793f24c85cea9b28139a6b1fadec12793ae64f46

Observation 8025005c-1164-430c-99ad-003536a40d14 · outbound

This paper cites W., Albergo, D., Borghini, G., Pansardi, O., Scaliti, E., Gupta, S., Saxena, K., Rufo, A., Panzeri, S., Manzi, G., et al.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models W., Albergo, D., Borghini, G., Pansardi, O., Scaliti, E., Gupta, S., Saxena, K., Rufo, A., Panzeri, S., Manzi, G., et al

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:33.968330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:25.931531Z digest=sha256:5102a3aa28e3fd19cd92d2e27b95c541c8b72b4817560a5cfb11710ae6278e4e

Observation 7dc1d12d-5d3f-4278-8758-95563df25a10 · outbound

This paper cites Doubao-1.5-pro: Exploring the ultimate balance between model performance and inference efficiency, 2025.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Doubao-1.5-pro: Exploring the ultimate balance between model performance and inference efficiency, 2025

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:33.812277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:26.016543Z digest=sha256:d7b9eb12706997bbcaffd2c368c86382e45ea665aa56a7b9ff86ece74b8b73f8

Observation dd9926bf-df0e-4a7b-953c-8bb85c6796e3 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:26.094789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:26.094789Z digest=sha256:10108539e2ec20f4c43c0ba590ccd5a80d0bc24e5dd27058dcad92925d88c36c

Observation 4f45d0ac-3793-4ce4-b70e-57931e169a76 · outbound

This paper cites Large Language Models Fail on Trivial Alterations to Theory-of-Mind Tasks.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Large Language Models Fail on Trivial Alterations to Theory-of-Mind Tasks

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:26.225894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:26.225894Z digest=sha256:4a24fd17dc503749be06b4d4f9f9ac7ef345ebe9916df180712bd12e778098e6

Observation bfe4c5de-14c3-45f0-a1d0-80c40e5a4c8f · outbound

This paper cites Theory of Mind in Large Language Models: Examining Performance of 11 State-of-the-Art models vs. Children Aged 7-10 on Advanced Tests.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Theory of Mind in Large Language Models: Examining Performance of 11 State-of-the-Art models vs. Children Aged 7-10 on Advanced Tests

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:26.300936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:26.300936Z digest=sha256:cfb0ca5a652ae97b7b408ebaedd12cf7a3519cbe7e0fe6d89fbbf93e32a06273

Observation 99e5bde0-cc33-4141-b767-e98ad2aaf563 · outbound

This paper cites an unresolved cited work.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:28:32.973373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:26.352822Z digest=sha256:d5dadaab4914473acdb66127daecf3ba417348554fb61ad01679077c23d8302d

Observation 86ec5611-5393-48fc-9d03-d633a903a312 · outbound

This paper cites Hi- T o M : A benchmark for evaluating higher-order theory of mind reasoning in large language models.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Hi- T o M : A benchmark for evaluating higher-order theory of mind reasoning in large language models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:26.405960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:26.405960Z digest=sha256:ab329fa7d2314d57aae6b8458c5e1b3a5bf2c9e2eff160092a1e0983acaebcbb

Observation 6888fed0-8533-4557-9b08-c80522c948e6 · outbound

This paper cites Tomvalley: Evaluating the theory of mind reasoning of llms in realistic social context.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Tomvalley: Evaluating the theory of mind reasoning of llms in realistic social context

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:28.624645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T00:28:26.458512Z digest=sha256:c5472870e9e5330c3a412c757cd281440ed046a1922f7dff21db4cc546cb7e56

Observation e0acd5bb-3522-4a1c-a658-4ccc7d573a77 · outbound

This paper cites OpenToM: A Comprehensive Benchmark for Evaluating Theory-of-Mind Reasoning Capabilities of Large Language Models.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models OpenToM: A Comprehensive Benchmark for Evaluating Theory-of-Mind Reasoning Capabilities of Large Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:26.516974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:26.516974Z digest=sha256:5b636f9200da2e157254b22dc89d088e9fec15f226999e93123c0f51af9e1116

Observation 19cc8dad-7779-48b0-9afa-8716ffd6d1b2 · outbound

This paper cites Evaluating and Enhancing LLMs Agent based on Theory of Mind in Guandan: A Multi-Player Cooperative Game under Imperfect Information.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Evaluating and Enhancing LLMs Agent based on Theory of Mind in Guandan: A Multi-Player Cooperative Game under Imperfect Information

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:26.571658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:26.571658Z digest=sha256:df00cbb89f2d5c5be76aaa39d162fb03af138e677e7c1749cc1a168b6b187b60

Observation 0f8c7b11-1f9e-4f65-a046-f54dbed01f70 · outbound

This paper cites How FaR Are Large Language Models From Agents with Theory-of-Mind?.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models How FaR Are Large Language Models From Agents with Theory-of-Mind?

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:26.619534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:26.619534Z digest=sha256:1f92813bcca8bffc9ef5559603431d9971f0b00b2f12cc22bbd4f91bd482e895

Pith citing papers

No inbound Pith citation observations are available.