Pith. sign in

Paper Citation Record · LEDGER

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models

As of 14 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2506.14224.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.14224 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:28:26.619534Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact2
  • verified fuzzy16
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fd2d8bbe-6533-46ae-956f-67c9aadf7717 · outbound

This paper cites GPT-4 Technical Report.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:22.948708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:22.948708Z digest=sha256:81a930d3537fdd00236622088c40ae85b4baee2db165104fd9a08fac28e36f5f

Observation 87f72791-366a-4e56-a4cb-f8866ea1f799 · outbound

This paper cites Do llms exhibit human-like reasoning? evaluating theory of mind in llms for open-ended responses.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Do llms exhibit human-like reasoning? evaluating theory of mind in llms for open-ended responses

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:37.218093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:23.009505Z digest=sha256:f0018657cca28aebf155bf965e37b14fa46b0a623f39198701c8d9479227983f

Observation aa45be4f-e851-4a68-9d40-da8bbce87f85 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:23.130188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:23.130188Z digest=sha256:2fca4268af568d2ddfbec2e6013365e28a00a676b818d09aa71dd2e93e3aa275

Observation bde74608-22b4-4b74-ada7-4cd6a1224c81 · outbound

This paper cites Hallucination of Multimodal Large Language Models: A Survey.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Hallucination of Multimodal Large Language Models: A Survey

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:23.244204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:23.244204Z digest=sha256:bb603c54b1a29de2435a2edd7fac75dfbb834301112d79b4476d0a83bc264328

Observation 55c22fba-95c5-4f34-b47d-e88240dd0711 · outbound

This paper cites theory of mind.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models theory of mind

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:36.966034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:23.327105Z digest=sha256:3d31c81d45513a530512606ad89e949ec5344946dd208317d601936ea6b68bde

Observation 31da6107-d971-465d-a0f9-041b363b03ec · outbound

This paper cites an unresolved cited work.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:28:36.762234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:23.437074Z digest=sha256:77f039e661a7c7dc00f48677441bdb10cc7c17b45e31fae49a09ae32ca1cf6a3

Observation 33c20f2f-5e9c-4280-8c2c-499f650af28e · outbound

This paper cites Sparks of Artificial General Intelligence: Early experiments with GPT-4.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Sparks of Artificial General Intelligence: Early experiments with GPT-4

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:23.499151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:23.499151Z digest=sha256:2f98edae6865a58c28d959bb5f603e75de8c6d7603b0d624ec2b6cc3b1ae4f68

Observation c2345e0f-fd69-4acf-b503-ca5bf8d94200 · outbound

This paper cites Through the theory of mind's eye: Reading minds with multimodal video large language models.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Through the theory of mind's eye: Reading minds with multimodal video large language models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:23.564854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:23.564854Z digest=sha256:b5b50e5d9dc329cab7d88b7890440c72891f365c7e180b433075987326ae76f2

Observation 806965e1-0ce4-4b89-b66d-9da4b55bbcb6 · outbound

This paper cites Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Minigrid & Miniworld: Modular & Customizable Reinforcement Learning Environments for Goal-Oriented Tasks

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:23.677281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:23.677281Z digest=sha256:c469047f17f74d8d407d38285b671513963c358cea444ea550bd13af184956c7

Observation 35991cdc-34e0-4217-9f59-6aa764bda43f · outbound

This paper cites The Llama 3 Herd of Models.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models The Llama 3 Herd of Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:23.776668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:23.776668Z digest=sha256:4cb90bd7cd233ae4758be705951e34376bcf4a686a4fcca3201e0067fc51cb0c

Observation 3cc93695-7564-430b-b4bd-0945222aaadf · outbound

This paper cites Who is Mistaken?.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Who is Mistaken?

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:28:27.010971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:23.869613Z digest=sha256:a55d2180bdafee0b18e6c1edb558dd398f8479bd3973e6d6e0ccc2c6a0674f34

Observation 93e44000-eab0-43db-8a37-182334a3cdc5 · outbound

This paper cites M., and Dillon, M.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models M., and Dillon, M

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:36.582680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:23.936271Z digest=sha256:ae9767b6492b8609aef0ad4f8a8267bf4c982ad13b3b30f1c7fc8ec45d4c7960

Observation 5ca30607-29df-4e9c-8e4e-cc2deb0c23c0 · outbound

This paper cites an unresolved cited work.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:28:36.322279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:24.030959Z digest=sha256:9f16fe3883a34f897c74592b69c57b252d2066e40e60cf2867f2c753861d1469

Observation 54bd068a-a2ce-434b-bfda-c4b6c17ca3da · outbound

This paper cites Distributional vectors encode referential attributes.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Distributional vectors encode referential attributes

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:36.092670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:24.105076Z digest=sha256:3f6babb3fbdf4da53519fe18a03cc2c89a609c6e536effdbb7248734f0628218

Observation 19e4d656-eba0-4d4d-a1b5-e51cbc3c0fbc · outbound

This paper cites Mistral 7B.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Mistral 7B

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:24.202247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:24.202247Z digest=sha256:a636a1a541f4d88d43309083f9deb70f0a9cef5d7a1fa1278b697f6b8fca41cd

Observation 207098c1-491d-41e7-a4c4-791f40507776 · outbound

This paper cites MMT o M - QA : Multimodal theory of mind question answering.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models MMT o M - QA : Multimodal theory of mind question answering

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:24.323659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:24.323659Z digest=sha256:2d8d2bb97fe49ebf5309ff3a08cb568da72ea6d4abd17b5c847386385793144a

Observation 818afb4d-91ae-48c4-a0b6-0061a4073722 · outbound

This paper cites Evaluating Large Language Models in Theory of Mind Tasks.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Evaluating Large Language Models in Theory of Mind Tasks

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:24.444328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:24.444328Z digest=sha256:317654412dec1e626dae74da001ee2ed2ae0216be6d3b6711da22c1be458d3ef

Observation 9f5bbe94-3938-4c56-af15-3a1a2aa71bfc · outbound

This paper cites Evaluating large language models in theory of mind tasks.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Evaluating large language models in theory of mind tasks

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:35.865227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:24.488322Z digest=sha256:75f208f3896905f45f02c2f7fbe4431d1bcb431a5bdb1a027e06352f9af68029

Observation c790b181-2db8-43eb-8b26-9c26ddab3fa7 · outbound

This paper cites What’s in an embedding? Analyzing word embeddings through multilingual evaluation.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models What’s in an embedding? Analyzing word embeddings through multilingual evaluation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:35.638751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:24.610920Z digest=sha256:62a4d515cf009c70ebced0377cccdf3019986a3d7361a50af0e670597581c28d

Observation 0582280e-b90f-4db8-99eb-cc394f11f674 · outbound

This paper cites Revisiting the evaluation of theory of mind through question answering.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Revisiting the evaluation of theory of mind through question answering

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:35.459458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:24.709661Z digest=sha256:fb8d82bbf96cbffa86ac26ba1f38e4c60b20141d0648a56f5073d50b487ea777

Observation e4ead157-f5aa-474c-9393-83bf28cea292 · outbound

This paper cites Inference-time intervention: Eliciting truthful answers from a language model.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Inference-time intervention: Eliciting truthful answers from a language model

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:35.222704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:24.801191Z digest=sha256:cddbb94f78b11e928e45460f06be65008458720249ea12d03f51b367f48b4754

Observation 235e191c-b256-4637-8238-077aaf6fce10 · outbound

This paper cites DeepSeek-V3 Technical Report.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models DeepSeek-V3 Technical Report

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:24.852810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:24.852810Z digest=sha256:bf8d0de9865f51e6862b4421e11721fbe4360fe54d390fd08e6327d23af08228

Observation 33ebb050-394e-4dec-92a1-c6fb30297b48 · outbound

This paper cites Large Model Strategic Thinking, Small Model Efficiency: Transferring Theory of Mind in Large Language Models.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Large Model Strategic Thinking, Small Model Efficiency: Transferring Theory of Mind in Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:24.942418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:24.942418Z digest=sha256:e7bbdde2563b6ae54ba91fd95bfa6bd9b08a6e8c2ba1e919ae108f14ce15e211

Observation 0167ac2f-a22a-4dec-86e7-475019775c64 · outbound

This paper cites Towards a holistic landscape of situated theory of mind in large language models.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Towards a holistic landscape of situated theory of mind in large language models

Reference 24

Resolution
verified exact
doi, observed 2026-08-07T00:28:26.780658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:25.051440Z digest=sha256:68378a4bcd7fdbf86436769a197ded22ee03769f6b1dae0e6e3df54c528599be

Observation 22cdedd6-ae45-4b79-9524-65c4b6c5edd5 · outbound

This paper cites A review on machine theory of mind.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models A review on machine theory of mind

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:34.996724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:25.122747Z digest=sha256:b10810cbbc49d4c742e40af5fd00814d78b9c3af41be3fc5352d99cdda722d36

Observation 3aca20ec-dbf5-4269-bc68-00a205d13a24 · outbound

This paper cites Evaluating theory of mind in question answering.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Evaluating theory of mind in question answering

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:25.167889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:25.167889Z digest=sha256:1013f5a77aa54b3290920142e14a1fc378e2948cf46bd791359e5b194af2ef18

Observation b432b1bf-fac7-47ec-a019-7ed73ce1fd7b · outbound

This paper cites Theory of Mind as Intrinsic Motivation for Multi-Agent Reinforcement Learning.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Theory of Mind as Intrinsic Motivation for Multi-Agent Reinforcement Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:25.231349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:25.231349Z digest=sha256:84e8b14ff66f4d626b212fa851b92ab9c9462ddce5b05147630f535ab849c4be

Observation c2e44e34-44e8-44a5-ae70-64ebd83a6716 · outbound

This paper cites Neural theory-of-mind? on the limits of social intelligence in large LM s.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Neural theory-of-mind? on the limits of social intelligence in large LM s

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:25.311237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:25.311237Z digest=sha256:cd6d6fb07ff51ea86bdb58bc526e94dbd5b619324968db0f2a9d329912e48f26

Observation 13646d49-ec21-4f07-aff8-1b8d7399473e · outbound

This paper cites Symmetric machine theory of mind.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Symmetric machine theory of mind

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:34.793911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:25.364111Z digest=sha256:6e4d0966e295b8290fb0c0b4a01ae9200445772b3a905eecddf7083313c2add0

Observation 0555ea13-19fe-4d7e-957e-33de2bce04a4 · outbound

This paper cites H., Zhou, X., Choi, Y., Goldberg, Y., Sap, M., and Shwartz, V.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models H., Zhou, X., Choi, Y., Goldberg, Y., Sap, M., and Shwartz, V

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:34.560220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:25.434622Z digest=sha256:804d1c730e73513239b24595fbc6f5ccc7efe60cc42ae922726ed881fdf1b1ac

Observation ee61ebd7-42ca-409b-83f3-bf72f2888641 · outbound

This paper cites Muma-tom: Multi-modal multi-agent theory of mind.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Muma-tom: Multi-modal multi-agent theory of mind

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:34.338427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:25.540625Z digest=sha256:90866932781b7e295467c9092fa660818baf5ff1d3e3d2c953b99f9a3472ad23

Observation a596cc83-ea41-43cc-be45-acf1aa7b9b22 · outbound

This paper cites Agent: A benchmark for core psychological reasoning.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Agent: A benchmark for core psychological reasoning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:34.180366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:25.668627Z digest=sha256:cdbc9a6deb50ab55ed723352628e4caef9df2bf04559fb3852ed0b74228448a0

Observation 5f10b3df-74e7-47a0-9dc2-c389bcb32a5f · outbound

This paper cites and Lernould, A.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models and Lernould, A

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:25.799002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:25.799002Z digest=sha256:c538e3d72dd6ea2e75596f33b38ace08a7f959fe3672256c3845692c0040d881

Observation 8025005c-1164-430c-99ad-003536a40d14 · outbound

This paper cites W., Albergo, D., Borghini, G., Pansardi, O., Scaliti, E., Gupta, S., Saxena, K., Rufo, A., Panzeri, S., Manzi, G., et al.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models W., Albergo, D., Borghini, G., Pansardi, O., Scaliti, E., Gupta, S., Saxena, K., Rufo, A., Panzeri, S., Manzi, G., et al

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:33.968330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:25.931531Z digest=sha256:0ddd9223eeb96907087864ce01b8d5eeb47cb279a568272e8322047ccd7143ac

Observation 7dc1d12d-5d3f-4278-8758-95563df25a10 · outbound

This paper cites Doubao-1.5-pro: Exploring the ultimate balance between model performance and inference efficiency, 2025.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Doubao-1.5-pro: Exploring the ultimate balance between model performance and inference efficiency, 2025

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:33.812277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:26.016543Z digest=sha256:0caafc6fd666f6aa22ae43152cc1088d3fa052c19cfb160c97789844e04fb582

Observation dd9926bf-df0e-4a7b-953c-8bb85c6796e3 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:26.094789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:26.094789Z digest=sha256:10b84ab33dcc8cd8f25dfff43dd733635a1a0498a3e7c4afaf565438d6c15129

Observation 4f45d0ac-3793-4ce4-b70e-57931e169a76 · outbound

This paper cites Large Language Models Fail on Trivial Alterations to Theory-of-Mind Tasks.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Large Language Models Fail on Trivial Alterations to Theory-of-Mind Tasks

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:26.225894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:26.225894Z digest=sha256:c8ab0db07bf292c7b9bab429346c4366f7d4742c8ffa3da96fb52dfaf0149816

Observation bfe4c5de-14c3-45f0-a1d0-80c40e5a4c8f · outbound

This paper cites Theory of Mind in Large Language Models: Examining Performance of 11 State-of-the-Art models vs. Children Aged 7-10 on Advanced Tests.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Theory of Mind in Large Language Models: Examining Performance of 11 State-of-the-Art models vs. Children Aged 7-10 on Advanced Tests

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:26.300936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:26.300936Z digest=sha256:990d51a1f37bb9fd0a779adde925ab1bc3b6568ef8f4d5d117799e78e455ce3b

Observation 99e5bde0-cc33-4141-b767-e98ad2aaf563 · outbound

This paper cites an unresolved cited work.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:28:32.973373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:26.352822Z digest=sha256:110c4dbf50278ebac9ec8bbee001927971d533622a7b615668d383bb3f97bb74

Observation 86ec5611-5393-48fc-9d03-d633a903a312 · outbound

This paper cites Hi- T o M : A benchmark for evaluating higher-order theory of mind reasoning in large language models.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Hi- T o M : A benchmark for evaluating higher-order theory of mind reasoning in large language models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:26.405960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:26.405960Z digest=sha256:c391646a16447d773028dfad860fe74bcaf923dc70c4293cc072db88a96fac4d

Observation 6888fed0-8533-4557-9b08-c80522c948e6 · outbound

This paper cites Tomvalley: Evaluating the theory of mind reasoning of llms in realistic social context.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Tomvalley: Evaluating the theory of mind reasoning of llms in realistic social context

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:28:28.624645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T00:28:26.458512Z digest=sha256:1e88f2dd52cbbfa5644d0fa1e477aa3cebb299f6c77f4ae22efec0af2293b50e

Observation e0acd5bb-3522-4a1c-a658-4ccc7d573a77 · outbound

This paper cites OpenToM: A Comprehensive Benchmark for Evaluating Theory-of-Mind Reasoning Capabilities of Large Language Models.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models OpenToM: A Comprehensive Benchmark for Evaluating Theory-of-Mind Reasoning Capabilities of Large Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:26.516974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:26.516974Z digest=sha256:c2877e2a618cf4c6b2bf2b5d67b8eb87cda34abad3352c840b6fdbaf4b3c7606

Observation 19cc8dad-7779-48b0-9afa-8716ffd6d1b2 · outbound

This paper cites Evaluating and Enhancing LLMs Agent based on Theory of Mind in Guandan: A Multi-Player Cooperative Game under Imperfect Information.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models Evaluating and Enhancing LLMs Agent based on Theory of Mind in Guandan: A Multi-Player Cooperative Game under Imperfect Information

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:26.571658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:26.571658Z digest=sha256:a20c2bf5ab4bcc91c6b02bffdf3b028d070f22b3acea47ab170c86050a758149

Observation 0f8c7b11-1f9e-4f65-a046-f54dbed01f70 · outbound

This paper cites How FaR Are Large Language Models From Agents with Theory-of-Mind?.

From Black Boxes to Transparent Minds: Evaluating and Enhancing the Theory of Mind in Multimodal Large Language Models How FaR Are Large Language Models From Agents with Theory-of-Mind?

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T00:28:26.619534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:28:26.619534Z digest=sha256:ee5365e0c059596378c13e53f663ae296a752948539d0784f6c07b8bf083f903

Pith citing papers

No inbound Pith citation observations are available.