Pith. sign in

Paper Citation Record · LEDGER

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM

As of 8 August 2026, this Paper Citation Record lists 87 of 87 outbound references and 4 inbound Pith citation observations for arXiv:2505.24238.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24238 v2

Coverage vector

measured 87 of 87 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:33:13.820971Z

measured 91 of 91 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:39:03.028373Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

87 of 87 outbound references displayed

  • verified exact1
  • verified fuzzy18
  • unresolved67
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation cd373f87-c04c-4530-bc32-e4af6da0f6fd · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.080923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.080923Z digest=sha256:c74cf743962a312c55e498ded18dca1bf47c2e7b931a5d45673a4c0f0ba691fb

Observation 1f5eb507-8925-4c9a-b2f3-080a49e013e3 · outbound

This paper cites Qwen2.5-VL Technical Report.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Qwen2.5-VL Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.122153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.122153Z digest=sha256:8dfd0fe05d264b102e02bc8d60646d10073c21d8b4299cadb72c3e6b8172670a

Observation e235ea63-d9dd-4ac5-bd0b-e70324949e95 · outbound

This paper cites Mitigating Open-Vocabulary Caption Hallucinations.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mitigating Open-Vocabulary Caption Hallucinations

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.187286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.187286Z digest=sha256:67e7d4cdf953469f2fb46f31c4a269741b615c94e36c3d79c7b8e585aeaf829c

Observation 0ac1d8d1-47ac-46b8-a4fa-ba1fb0af3714 · outbound

This paper cites MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.268894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.268894Z digest=sha256:d90ce5388158b484dc19a1af18f56386be207ad66af35c4849a508ad03bca287

Observation a02ce98e-e834-4a5d-8358-f84a06c5ca25 · outbound

This paper cites Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.382734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.382734Z digest=sha256:79a32f166ff072b99ca2642c8757dd43aa980f5497e3d25e111175264599fe6b

Observation 6b2a2fc9-cbcb-4800-9df7-e62cbf88b301 · outbound

This paper cites Mitigating Hallucination in Visual Language Models with Visual Supervision.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mitigating Hallucination in Visual Language Models with Visual Supervision

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.475621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.475621Z digest=sha256:46e43b8a971833411aaf8ca5eeb2b0e578ad3d8aaf8c1debbccbe248b6b1509f

Observation 60e44d62-1f95-4794-bbb8-ed07a8806100 · outbound

This paper cites Puzzlevqa: Diagnosing multimodal reasoning challenges of language models with abstract visual patterns.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Puzzlevqa: Diagnosing multimodal reasoning challenges of language models with abstract visual patterns

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.558504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.558504Z digest=sha256:381a3b9b476b33e5f7c38926ce008c8de3f103e4abecd53f548e6d435cefeda8

Observation e6c82916-4191-463e-9612-10cbbf6c873d · outbound

This paper cites Zero-shot generalizable incremental learning for vision-language object detection.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Zero-shot generalizable incremental learning for vision-language object detection

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.634033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.634033Z digest=sha256:98ac84cc84da61bf1a9fb2baa03d3a48d194a4a8fc8da30ca8f357d7208587e2

Observation d23fb554-b6b2-4c64-80be-3802ef89af64 · outbound

This paper cites MR-GDINO: Efficient Open-World Continual Object Detection.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM MR-GDINO: Efficient Open-World Continual Object Detection

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:33:14.676390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:33:06.701427Z digest=sha256:a9921f06bcf5cc486c3c29d89be0d29fb58dad9ecf8ddcf15d56690b407100b8

Observation f50f399f-d566-4461-af0e-1f313c9b77a1 · outbound

This paper cites Lpt: Long-tailed prompt tuning for image classification.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Lpt: Long-tailed prompt tuning for image classification

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.750153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.750153Z digest=sha256:e7550c2fb072f72636701ff317e538890a8c55cad543e90cdac1a333d35c5d54

Observation 40226a05-9e99-486a-8a25-e68cf003dc0d · outbound

This paper cites A Survey on In-context Learning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM A Survey on In-context Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.816254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.816254Z digest=sha256:11311bbffa69f651e08df39b2279565ee365eb8d9e25ea13c459e6f2bda891ea

Observation ea35cb9c-7fd7-4ce0-bc0d-e7d524fc5b9f · outbound

This paper cites Virgo: A Preliminary Exploration on Reproducing o1-like MLLM.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Virgo: A Preliminary Exploration on Reproducing o1-like MLLM

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.915732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.915732Z digest=sha256:889bdef067998622057e793486c11a3863c764898a6183f4b14646db26abfdb1

Observation 49e1bd80-b7be-43fc-863d-fb2bf75fdb27 · outbound

This paper cites Detecting hallucinations in large language models using semantic entropy.Nature, 630(8017):625–630, 2024.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Detecting hallucinations in large language models using semantic entropy.Nature, 630(8017):625–630, 2024

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.022394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.022394Z digest=sha256:f26dee1badbdc23855ace8e6ae99b88e1f660589737255bdacec94d4299cb363

Observation b3dc9fb6-d758-4d73-af6e-464ab485d758 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.090750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.090750Z digest=sha256:45f21b64be1ed504525586a49f505b97b738f021e783c3abd1309d98e173c900

Observation 6cd15c42-a470-4dbe-9f0d-e99b8db63c67 · outbound

This paper cites GPTScore: Evaluate as you desire.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM GPTScore: Evaluate as you desire

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.151002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.151002Z digest=sha256:34cad54162951558833a072f7dee903fb30c7941056a8096af9ec35fb3e27a9a

Observation 85588250-a0aa-44c6-a5fe-3bd9c163fbff · outbound

This paper cites The capacity for moral self-correction in large language models.Parameters, 109(1010):1011.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM The capacity for moral self-correction in large language models.Parameters, 109(1010):1011

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.271503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.271503Z digest=sha256:7fd72525483f266a8513c8bba7880af35352e47c9e4b8f078b14536c29752af5

Observation 131f8f24-faff-4e6e-9cc9-ecdc554449ba · outbound

This paper cites Are Language Models Puzzle Prodigies? Algorithmic Puzzles Unveil Serious Challenges in Multimodal Reasoning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Are Language Models Puzzle Prodigies? Algorithmic Puzzles Unveil Serious Challenges in Multimodal Reasoning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.360301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.360301Z digest=sha256:5a5bb8fbc960a0c2e560c43c7f9e62a64a7b08b4ebdb35d0646cd786bdb8066f

Observation d385e19b-f4e6-4955-89e5-25a3370bbfea · outbound

This paper cites The Llama 3 Herd of Models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM The Llama 3 Herd of Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.475863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.475863Z digest=sha256:2cb6690d971643bc93dcee785b2f3cbc3853068c3c086e0c6631c7e60d2f52bd

Observation 82fb3068-1a9f-4e79-a8f4-edb963f00afd · outbound

This paper cites A Survey on LLM-as-a-Judge.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM A Survey on LLM-as-a-Judge

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.571919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.571919Z digest=sha256:970fe27cb504546cbb5618a009b39a35ad795ef3f712b7ef2cf7543dff026586

Observation a4bcd11f-f250-480f-bf44-546db8c7a328 · outbound

This paper cites Hallusionbench: an advanced diagnostic suite for entangled language hallucination and visual illusion in large vision-language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Hallusionbench: an advanced diagnostic suite for entangled language hallucination and visual illusion in large vision-language models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.664840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.664840Z digest=sha256:0e8bfbd0eec0f1205d37057957f1ee643399dcb44447f4174d8a9e78f08c745a

Observation aaf1ac63-d878-4cb2-807b-c8f4228efe91 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.772797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.772797Z digest=sha256:b5e14e5a3a1ae50c77dfdc77a110a99f968d1a93f23845a1a2295ea4e4f32e7b

Observation e4ea4dc8-3afa-4b0e-a85f-889900e9f05d · outbound

This paper cites When Continue Learning Meets Multimodal Large Language Model: A Survey.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM When Continue Learning Meets Multimodal Large Language Model: A Survey

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.869192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.869192Z digest=sha256:16c8ed606f424170e821a3763bb59eaf572e984e73d98744afafe363267508d0

Observation 3c11f0aa-8b18-422f-9198-fd055ab6103c · outbound

This paper cites GPT-4o System Card.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM GPT-4o System Card

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.959565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.959565Z digest=sha256:ba5a23f1fa2f65211460d14015c4c47e93a2ab98d66f7bc7464689471c44c268

Observation efeed5f0-0c9e-4944-a8b8-e7aa228c618e · outbound

This paper cites OpenAI o1 System Card.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM OpenAI o1 System Card

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.049492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.049492Z digest=sha256:971953c2bc5c5ed8a4f0c55f6b221b05a25bef7a02638515638410bfa8d7981b

Observation ffa0a01f-ae22-4626-935d-4c149f169e29 · outbound

This paper cites Towards mitigating LLM hallucination via self reflection.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Towards mitigating LLM hallucination via self reflection

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:18.503341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:33:08.183950Z digest=sha256:527aa8e2723a9b2490302cec567e48d2cdf1b07e09eb092355963c9e629582ad

Observation 62c94a89-a059-4823-bb1d-acd48b4486d5 · outbound

This paper cites Hallucination augmented contrastive learning for multimodal large language model.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Hallucination augmented contrastive learning for multimodal large language model

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.290558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.290558Z digest=sha256:d0126477c50909c2afeab8441e041168294d4a876e22b817b9f1b65810817b89

Observation a72d196d-62cc-4209-83be-37aec9055dc7 · outbound

This paper cites MME-CoT: Benchmarking Chain-of-Thought in Large Multimodal Models for Reasoning Quality, Robustness, and Efficiency.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM MME-CoT: Benchmarking Chain-of-Thought in Large Multimodal Models for Reasoning Quality, Robustness, and Efficiency

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.379110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.379110Z digest=sha256:7846dbe2f5b27491d86629474988c5f99bb85cde2bf06afe592bcdccdc240097

Observation d3eabd59-e4da-4a04-a2a9-c88aed94f37e · outbound

This paper cites Decoupling representation and classifier for long-tailed recognition.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Decoupling representation and classifier for long-tailed recognition

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:18.166965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:33:08.480767Z digest=sha256:7747d56e0514d44fad791986e55018977cb1ea12fc6a4877ba6de8c697ef64ee

Observation 2e8ba4f8-f945-4fb9-9e87-57bb85616c2f · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Gonzalez, Hao Zhang, and Ion Stoica

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.559186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.559186Z digest=sha256:cab89c7db9e718bb2ada8886baeaada8dcaaa02705227afecb2eb706d2f56319

Observation c3e2ac17-0813-463f-9f74-946b2dfbbd0c · outbound

This paper cites SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.631430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.631430Z digest=sha256:803b06700774b42a97693c6f861506d680bd06fa335432bb9f918fc023704879

Observation 5dc004e3-889e-47bb-aa86-7f110d4a3be6 · outbound

This paper cites LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.729239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.729239Z digest=sha256:ffa45a663f95feda77757a75bb90314429baf8b9e19082c4833dbd95566d7282

Observation 3042c996-c970-4899-9c34-7b2afc9e70d9 · outbound

This paper cites Salad-bench: A hierarchical and comprehensive safety benchmark for large language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Salad-bench: A hierarchical and comprehensive safety benchmark for large language models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.797773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.797773Z digest=sha256:da207ccd38061869a40d6835f99e111c86aaae19fb47d60c9640b00e8e8d028d

Observation ebe72b6c-d1fb-434d-b5a8-7dd39b936bf4 · outbound

This paper cites Long-tailed visual recognition via gaussian clouded logit adjustment.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Long-tailed visual recognition via gaussian clouded logit adjustment

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:17.903767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:33:08.934739Z digest=sha256:0374e8d248ac9d770f7c26008237aab3428c9da674313e7addf92953f7971c90

Observation 4b24e925-a8de-49f7-ae0e-ae79231d2290 · outbound

This paper cites Evaluating object hallucination in large vision-language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Evaluating object hallucination in large vision-language models

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:17.659774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:33:09.006807Z digest=sha256:37b83c1c2513f3c12641fcc1967eb9cec519abb66de9b8acb7ce958bd9c88644

Observation a46d00ea-67a9-4727-8ff8-6fe63c8f0dda · outbound

This paper cites Omnibench: Towards the future of universal omni-language models.arXiv preprint arXiv:2409.15272, 2024.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Omnibench: Towards the future of universal omni-language models.arXiv preprint arXiv:2409.15272, 2024

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.096630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.096630Z digest=sha256:35a1af6042c5f97e235917a56c7c732ca9eb6dd3aa73ec39e465bed3c8440283

Observation ef6c1503-b847-4245-861e-fa84899b4c71 · outbound

This paper cites DeepSeek-V3 Technical Report.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM DeepSeek-V3 Technical Report

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.169761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.169761Z digest=sha256:ad5e02813878538f3707dba2d39942154f5babea4c28ec02075d57d7a4f5d8c8

Observation 169dedb4-ccf8-4215-b20e-99bfc5d5e067 · outbound

This paper cites Mitigat- ing hallucination in large multi-modal models via robust instruction tuning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mitigat- ing hallucination in large multi-modal models via robust instruction tuning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:17.441938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:33:09.244473Z digest=sha256:7e4d80a037f78b259d91cdf7e9914ad7c1cd1dd85175957d25f486cdcd9e1234

Observation a549b860-d5f8-4579-88d9-6e0627e2fb8f · outbound

This paper cites Visual instruction tuning, 2023.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Visual instruction tuning, 2023

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:17.186401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:33:09.337797Z digest=sha256:71f18cfb61dcaaa41658b326b93fe7c90258c8b370253f1be61548b15e971fad

Observation 2551a763-37bd-41ce-931b-e4db63c3bf7d · outbound

This paper cites Visual-RFT: Visual Reinforcement Fine-Tuning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Visual-RFT: Visual Reinforcement Fine-Tuning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.432965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.432965Z digest=sha256:1f8930a124e6695ff417d8a86471dc7c07f8573ea605790c903a5cdb8a41d605

Observation 7f6f5d91-2270-4ad3-be3f-e631d7dceb6c · outbound

This paper cites Decoupled weight decay regularization.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Decoupled weight decay regularization

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.543779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.543779Z digest=sha256:48a5dbe4007ea784aff0f37ca40d2090f7fe1bac938bd00842841647781e154c

Observation b7779aaa-102d-49ed-8517-7739ee88a52b · outbound

This paper cites Mathvista: Evaluating mathematical reasoning of foundation models in visual contexts.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mathvista: Evaluating mathematical reasoning of foundation models in visual contexts

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.643603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.643603Z digest=sha256:6077c4f8c7630360682ef943806aa1f6cde0f655a0e4e6c996bb8b1995b37d53

Observation 50b2646c-45b8-47df-92bf-d05cb8e0947f · outbound

This paper cites Learn to explain: Multimodal reasoning via thought chains for science question answering.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Learn to explain: Multimodal reasoning via thought chains for science question answering

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.735339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.735339Z digest=sha256:9ab7c1c48d903b334c238d88475cca70cb539ffaad815580733d1625fda2c717

Observation 0c0d71be-894d-4563-aa5a-2caaf14733de · outbound

This paper cites Ursa: Understanding and verifying chain-of-thought reasoning in multimodal mathematics.arXiv preprint arXiv:2501.04686, 2025.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Ursa: Understanding and verifying chain-of-thought reasoning in multimodal mathematics.arXiv preprint arXiv:2501.04686, 2025

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.808767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.808767Z digest=sha256:5dc508ec78b5541bcb0ef3b177edc88a81d996d13232aa746fc33ac817dd804d

Observation 63567212-d451-45be-8661-10868da2d31b · outbound

This paper cites Ok-vqa: A visual question answering benchmark requiring external knowledge.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Ok-vqa: A visual question answering benchmark requiring external knowledge

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.913245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.913245Z digest=sha256:8566d72d04ba28713389187bd9aee9d3ad06748874b7eb2661f5209fa2b34a6c

Observation d1fc6cd4-025a-416c-a8d9-0a324dbcea08 · outbound

This paper cites Chartqa: A benchmark for question answering about charts with visual and logical reasoning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Chartqa: A benchmark for question answering about charts with visual and logical reasoning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.990090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.990090Z digest=sha256:57c5150fc62abefa2a83e99b628c51e0de3135cf22fa263aaa2704d35bfee7d1

Observation d12ecdd1-663f-4408-a3f2-485ff7baeab6 · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.071390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.071390Z digest=sha256:cad75a83b846ec3d7a36d733d7719d06c2fead485fd07b4c9da94558f972468c

Observation 4d684078-fce3-4628-b438-3871fe71aeab · outbound

This paper cites Compositional chain-of- thought prompting for large multimodal models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Compositional chain-of- thought prompting for large multimodal models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.139052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.139052Z digest=sha256:213f6f3117fd98f124130c6220e6c7e1999e9bdf78f56a8d86a34add4b987769

Observation a53719f6-9c58-428f-b724-d963b96de839 · outbound

This paper cites Pytorch: An imperative style, high-performance deep learning library.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Pytorch: An imperative style, high-performance deep learning library

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.217648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.217648Z digest=sha256:b0007f5d52f3daf70d4c9f81ec2a71500700604001c92a7be02a4c250fdcfc30

Observation 04a56abd-815d-43aa-81bf-37bc82357b95 · outbound

This paper cites LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.310457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.310457Z digest=sha256:5739b77cca61b08165416e69402986401f29f466d7ee17f67fce2e350974e9f6

Observation 9e83b7d7-971b-4220-8f1f-5ce5d603d714 · outbound

This paper cites ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.405095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.405095Z digest=sha256:87fc01374974e5291beec7310d56f783d24c18fab4bc1e6f7b8eb154d9b53903

Observation 698b7e53-57d5-4cb9-a2ee-331fa319c673 · outbound

This paper cites Mitigating object hallucination in mllms via data-augmented phrase-level alignment.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mitigating object hallucination in mllms via data-augmented phrase-level alignment

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:16.787632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:33:10.547565Z digest=sha256:ce48fe9bc4134f9a2ddb7ee48436681380b5df467a7240147a9ec171f0fc5a56

Observation 284f205c-e24d-49d4-98ed-c7382e884264 · outbound

This paper cites Proximal Policy Optimization Algorithms.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Proximal Policy Optimization Algorithms

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.633888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.633888Z digest=sha256:0ecda17c2eed4ca800d6ef35cf7f00ffb5fcf2f18c30c08d7890ecff9e62f5cf

Observation a0bec2fa-ad09-47da-9b3d-dcfc443a20cb · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.745345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.745345Z digest=sha256:d6992bed346aab455e868e89dd261e17d4d9f6ea26f6fe5319ed94293300282b

Observation b2b2aefe-d4c4-4b50-9c4c-dffc83574f91 · outbound

This paper cites Monte carlo tree search: A review of recent modifications and applications.Artificial Intelligence Review, 56(3):2497–2562, 2023.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Monte carlo tree search: A review of recent modifications and applications.Artificial Intelligence Review, 56(3):2497–2562, 2023

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.856167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.856167Z digest=sha256:4cfd2f3c074bc051e6700bd727d83678785d548d5bcae16523ecca0df3a4fee9

Observation dcba0640-8e74-499b-936a-ba7fdccf62b7 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Gemini: A Family of Highly Capable Multimodal Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.924496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.924496Z digest=sha256:0b0f51e81138ce714f85c333b142bb66537cd3a20eefcfcd3d2de33d26111fd8

Observation 443738d3-a1a4-4e01-8fa7-10e641d3eecb · outbound

This paper cites QVQ: To See the World with Wisdom.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM QVQ: To See the World with Wisdom

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:16.606887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:33:11.048068Z digest=sha256:93f37d7884e5cc0aac2bba38f0d12bc16f592921b411bb6bc21ab024671588e5

Observation 61e93b22-71a1-4bf0-957f-2c3634b8c1f2 · outbound

This paper cites Uncertainty-Based Abstention in LLMs Improves Safety and Reduces Hallucinations.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Uncertainty-Based Abstention in LLMs Improves Safety and Reduces Hallucinations

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:11.153366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:11.153366Z digest=sha256:8526795e3bf54c3e311c0689145487a6f159cc80393ca984153dbd3793ada18a

Observation 8084a308-e4ac-4805-9d83-a724b7daae11 · outbound

This paper cites Eyes wide shut? exploring the visual shortcomings of multimodal llms.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Eyes wide shut? exploring the visual shortcomings of multimodal llms

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:11.217618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:11.217618Z digest=sha256:5992c4548034ac22ce26202c13df05cfef0b51750e36c782bb7b5e8ab6c0d2db

Observation e08b3ec0-020d-434d-8a05-7a02209f9913 · outbound

This paper cites Measuring multimodal mathematical reasoning with math-vision dataset.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Measuring multimodal mathematical reasoning with math-vision dataset

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:16.486557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:33:11.281006Z digest=sha256:85ec35074861f0ba67a06a10e7a54ad852a6df477934a664e4782bf3be4dcc69

Observation 22da4bca-b3c8-4658-8779-f635c78267e6 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:11.350592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:11.350592Z digest=sha256:83c0af760b788c390f02e77e1387717cdb9cb4565f48b817d8b528089b39da80

Observation 1772c607-1e3e-4cc8-a011-797bd52d4277 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Chain-of-thought prompting elicits reasoning in large language models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:11.422440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:11.422440Z digest=sha256:9ef288caae126ffe70184d525c0c22d9a7e5fa353e5dc182522ac9156ea7b96b

Observation edbb8b53-7b56-4357-9147-0fd3e8b09e71 · outbound

This paper cites RITUAL: Random Image Transformations as a Universal Anti-hallucination Lever in Large Vision Language Models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM RITUAL: Random Image Transformations as a Universal Anti-hallucination Lever in Large Vision Language Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:11.492919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:11.492919Z digest=sha256:9c96cf98e69e04a891fab21a9d41ffca66e4600cb4cc730becee0ce2fad79c20

Observation 630b0b44-60ff-4031-b5bd-136c5134da95 · outbound

This paper cites Autohallusion: Automatic gen- eration of hallucination benchmarks for vision-language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Autohallusion: Automatic gen- eration of hallucination benchmarks for vision-language models

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:16.317333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:33:11.562584Z digest=sha256:2ac7cb6a6a0d57255728cc965115145abeae1b3ae55d335f97715a24982c0629

Observation 2767d4b3-7a1e-44f6-9233-16d9eaf5c5ab · outbound

This paper cites Grok 3 Beta — The Age of Reasoning Agents.https://x.ai/grok, 2025.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Grok 3 Beta — The Age of Reasoning Agents.https://x.ai/grok, 2025

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:16.123781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:33:11.613404Z digest=sha256:a4d063bcf46054e5881205718568554710d10cf5d10017747ffd7a24b6c4828e

Observation 9b1c93ed-31f5-4c5e-aef0-419a9196b660 · outbound

This paper cites Mitigating object hallucination via concentric causal attention.Advances in Neural Information Processing Systems, 37:92012–92035, 2024.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mitigating object hallucination via concentric causal attention.Advances in Neural Information Processing Systems, 37:92012–92035, 2024

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.957279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:33:11.724385Z digest=sha256:d6b5f39fbeb62e11b385a132e40e1702d2ede5c2a0755add8b0793e7f2ba2591

Observation 6ce3eff8-0eda-42fa-bd40-3d76fa9bab4f · outbound

This paper cites Llava-cot: Let vision language models reason step-by-step, 2024.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Llava-cot: Let vision language models reason step-by-step, 2024

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.820797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:33:11.795732Z digest=sha256:5a2ebb4a63974a50081de7761e123873eff211d01ff280bc9b9666a3055ff0e4

Observation eebbacd3-290d-40b6-9616-cec260b62609 · outbound

This paper cites Qwen2.5 Technical Report.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Qwen2.5 Technical Report

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:11.894050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:11.894050Z digest=sha256:179a6e3de6824e385b24c7345870ebfde521f82744bc2e9f0540b7f00fbba908

Observation 98b48a0d-4c30-470a-b2d7-380e837c3767 · outbound

This paper cites Soft-prompting with graph-of-thought for multi-modal representation learning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Soft-prompting with graph-of-thought for multi-modal representation learning

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.698313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:33:11.952400Z digest=sha256:27c0f0ebe81ef0f2fcf8a80704ea8a73cc8a19f1912b42562e417b08a9377a68

Observation 9304223c-8c7d-4574-b644-7bb3046630c9 · outbound

This paper cites Mitigating hallucination in large vision- language models via modular attribution and intervention.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mitigating hallucination in large vision- language models via modular attribution and intervention

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.543195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:33:12.053954Z digest=sha256:68f7546afa0aabe1b355f760321054b8e15f601a20053fbef849dd9e5a2afbe7

Observation 73dcefd7-a958-4dd6-bded-72c5e627c82b · outbound

This paper cites R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.177383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.177383Z digest=sha256:9032533193c745bb2852bdd1f8ec8ec6d89a0ec2d4650243b73b4450cca4382b

Observation 0b48fbce-4c79-415a-aeac-4e3e872724ae · outbound

This paper cites Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.258432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.258432Z digest=sha256:fe36a592acc6d930c0791a8fd0c95207de40c7dd82dff4746053febf8ad09c96

Observation c618bb1f-9378-4f4e-a8e8-05ca272c661b · outbound

This paper cites React: Synergizing reasoning and acting in language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM React: Synergizing reasoning and acting in language models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.383675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.383675Z digest=sha256:eb7336cca749f8a16f60c50fe81039f2eefc9848f556f5a710e54d5a0e1fe5fa

Observation ed50d571-49f4-4efd-a3c0-5fc29315c7a7 · outbound

This paper cites LIMO: Less is More for Reasoning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM LIMO: Less is More for Reasoning

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.448500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.448500Z digest=sha256:9d19cdf8d803edef9aa74cfbbfb6cbe1c91d5dae286866d4038db1b2769b718b

Observation 3cb4a8e8-4452-4317-ac9e-074e4f911ef1 · outbound

This paper cites Hallucidoctor: Mitigating hallucinatory toxicity in visual instruction data.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Hallucidoctor: Mitigating hallucinatory toxicity in visual instruction data

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.557605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.557605Z digest=sha256:7381d366b14ca27bff422dea5b5a1b226799f5f64cd0193b1365b5908b2fd292

Observation 9d8616b4-cc9f-478c-83fe-fa1433f61ea9 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.653152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.653152Z digest=sha256:4497279d17293cb877644fd06df8420c975dcfeba783fe087c468f2cfc828e62

Observation 0f2fb295-f912-4fbd-8059-15d7e34eb919 · outbound

This paper cites Rlhf-v: Towards trustworthy mllms via behavior alignment from fine-grained correctional human feedback.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Rlhf-v: Towards trustworthy mllms via behavior alignment from fine-grained correctional human feedback

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.750492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.750492Z digest=sha256:8037836870414bb4111d9006233d04b479cc4c328ca27c674ac81ffba3582674

Observation 9ebaadc9-87c8-41a0-ba69-f23f61aa0c81 · outbound

This paper cites LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.857251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.857251Z digest=sha256:b9b8b40146c37d989f47c240fb176b702551ad456b338389f07c2b45de3cb608

Observation 5e0dc656-f057-48a7-a72b-70435f63512d · outbound

This paper cites Reflective instruction tuning: Mitigating hallucinations in large vision-language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Reflective instruction tuning: Mitigating hallucinations in large vision-language models

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.380205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:33:12.980715Z digest=sha256:7a4fa5a4710ac718312a21378dc3bdee338875a9233d8e88e8c1cd85d769392f

Observation 5a298254-9648-4782-abb3-752e69622b8f · outbound

This paper cites Mathverse: Does your multi-modal llm truly see the diagrams in visual math problems? InEuropean Conference on Computer Vision, pages 169–186.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mathverse: Does your multi-modal llm truly see the diagrams in visual math problems? InEuropean Conference on Computer Vision, pages 169–186

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:13.073898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:13.073898Z digest=sha256:f0b332a4ed50666523b0d2cbfbb43a2b36253f8e8be7e71c9a5469ea3f99117f

Observation 140db751-7cb8-4809-9e30-2e832334feb1 · outbound

This paper cites VL-Uncertainty: Detecting Hallucination in Large Vision-Language Model via Uncertainty Estimation.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM VL-Uncertainty: Detecting Hallucination in Large Vision-Language Model via Uncertainty Estimation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:13.190579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:13.190579Z digest=sha256:46b5bc320c968c7f9b3689de2592dda0155ad540492dbfedd84faaaf2d9df155

Observation 334b295f-408d-44bd-8026-aa20d981a015 · outbound

This paper cites The Lessons of Developing Process Reward Models in Mathematical Reasoning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM The Lessons of Developing Process Reward Models in Mathematical Reasoning

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:13.281932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:13.281932Z digest=sha256:d848e17b4fc93c4396557907c8592594d1b35bcda5e3e9f2b4d4f79295fb1cde

Observation 5fd38f61-18d5-478f-b822-1218dce4fd63 · outbound

This paper cites Multimodal chain-of-thought reasoning in language models.Transactions on Machine Learning Research.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Multimodal chain-of-thought reasoning in language models.Transactions on Machine Learning Research

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:13.378434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:13.378434Z digest=sha256:329d2ea42a13623983a582aaf40dadaa965f6c6935ef7cb591f2154076af7da5

Observation daedcae0-b3a7-437e-b406-b2b3c87dc327 · outbound

This paper cites Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:13.470889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:13.470889Z digest=sha256:b83b30a73d046c6e7610f1de798de5ab138ef44c82474b8105d358510cdcc3d4

Observation cfacc5c1-dfc2-4639-8b13-4af198dd8327 · outbound

This paper cites A picture is worth a graph: A blueprint debate paradigm for multimodal reasoning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM A picture is worth a graph: A blueprint debate paradigm for multimodal reasoning

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.196268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:33:13.563966Z digest=sha256:df3777d4acbffdc0dcd44d900ebbc640c09e958bbd537eef99dad43526579f62

Observation 7f35f66e-4cbe-463e-a511-e8f97f085ad1 · outbound

This paper cites Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:13.663322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:13.663322Z digest=sha256:dc6f8aeee95e93b9a578e69af1b87aad44bb8a28c82364e616581c4c4aaae532

Observation 0db258ac-a6f2-4765-b849-f5e6ecfc0fd7 · outbound

This paper cites Relying on the unreliable: The impact of language models’ reluctance to express uncertainty.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Relying on the unreliable: The impact of language models’ reluctance to express uncertainty

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.038180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:33:13.747662Z digest=sha256:c0e5af1be3e29ecb2a60034fa8d6503f95a43f9a61441b59c7c99ebaed878a41

Observation 0d179a1e-f3fd-41b8-a842-23e070afc6e7 · outbound

This paper cites the answer is [answer in the input].

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM the answer is [answer in the input]

Reference 87

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T12:33:14.881658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:33:13.820971Z digest=sha256:8acb6a56349822e8dc2e889a60598cab0dc6798dbf147de3bb8c1243a81ac58f

Pith citing papers

Observation 8f61cce6-5c2d-4d17-8b3f-2c8cc3ad9045 · inbound

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models cites this paper.

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-05T10:39:03.028373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:39:03.028373Z digest=sha256:e4b6a7e7d06c8189ac8c775f7781a51c5c297c92343357fc61d6b20d29e1b69e

Observation 93d009a1-1467-4909-ae1a-aa45d4e09569 · inbound

Understanding the Role of Hallucination in Reinforcement Post-Training of Multimodal Reasoning Models cites this paper.

Understanding the Role of Hallucination in Reinforcement Post-Training of Multimodal Reasoning Models MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:53:16.282414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T20:48:52.130130Z digest=sha256:7c7b7448e349a8d5f9fe53b452ec9dca9d1fc9de23023d5313076869d7d9d82a

Observation 8db78eb4-7cbe-4125-83e5-637815967363 · inbound

HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering cites this paper.

HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM

Reference 230

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T23:54:45.651801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-09T23:51:47.724033Z digest=sha256:0304d1797e493116bb000981919d99d33b0c056990797cc0eaeeedde5165b7f0

Observation 45eb3f94-6839-4143-946f-13d83c6c4310 · inbound

SoccerRef-Agents: Multi-Agent System for Automated Soccer Refereeing cites this paper.

SoccerRef-Agents: Multi-Agent System for Automated Soccer Refereeing MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:46:13.723544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T08:03:57.287297Z digest=sha256:72702f970b9f340d9decd613b9ef96b9f693e5bfca93f3e605aaf486d425235a