Pith. sign in

Paper Citation Record · LEDGER

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM

As of 18 August 2026, this Paper Citation Record lists 87 of 87 outbound references and 4 inbound Pith citation observations for arXiv:2505.24238.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24238 v2

Coverage vector

measured 87 of 87 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:33:13.820971Z

measured 91 of 91 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T10:39:03.028373Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

87 of 87 outbound references displayed

  • verified exact1
  • verified fuzzy18
  • unresolved67
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation cd373f87-c04c-4530-bc32-e4af6da0f6fd · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.080923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.080923Z digest=sha256:4aef3fb29c813af1336ccccc95969119690dda1491b039257edb45d30df7ef24

Observation 1f5eb507-8925-4c9a-b2f3-080a49e013e3 · outbound

This paper cites Qwen2.5-VL Technical Report.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Qwen2.5-VL Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.122153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.122153Z digest=sha256:ced9a2ef6bbd51bf6144210c77f90c4ea6adf1d1174325d769c8d580d529d761

Observation e235ea63-d9dd-4ac5-bd0b-e70324949e95 · outbound

This paper cites Mitigating Open-Vocabulary Caption Hallucinations.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mitigating Open-Vocabulary Caption Hallucinations

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.187286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.187286Z digest=sha256:c96b02dbb9d372e43341d28c4d5da0d460f6983af2bf8150f8f3d424af375de2

Observation 0ac1d8d1-47ac-46b8-a4fa-ba1fb0af3714 · outbound

This paper cites MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM MM-IQ: Benchmarking Human-Like Abstraction and Reasoning in Multimodal Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.268894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.268894Z digest=sha256:fe27c9fec7dc94df20ec817b5e726efc05e724bfcf9d18b26c0e465714a2dca0

Observation a02ce98e-e834-4a5d-8358-f84a06c5ca25 · outbound

This paper cites Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.382734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.382734Z digest=sha256:621656e1e8c5d845e72012c2a60950ffd573dd290091502edc0f32ccd8744f72

Observation 6b2a2fc9-cbcb-4800-9df7-e62cbf88b301 · outbound

This paper cites Mitigating Hallucination in Visual Language Models with Visual Supervision.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mitigating Hallucination in Visual Language Models with Visual Supervision

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.475621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.475621Z digest=sha256:a712fa1b13f3a0462a2d64f2a63ad808269c01671d59c378f7701296ef7d42f6

Observation 60e44d62-1f95-4794-bbb8-ed07a8806100 · outbound

This paper cites Puzzlevqa: Diagnosing multimodal reasoning challenges of language models with abstract visual patterns.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Puzzlevqa: Diagnosing multimodal reasoning challenges of language models with abstract visual patterns

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.558504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.558504Z digest=sha256:fee012930301bccce7b898acfd7f355d0ec51e0a23f147d3e3161c52b64fd36d

Observation e6c82916-4191-463e-9612-10cbbf6c873d · outbound

This paper cites Zero-shot generalizable incremental learning for vision-language object detection.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Zero-shot generalizable incremental learning for vision-language object detection

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.634033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.634033Z digest=sha256:902206ca5ced741faa7d5416995a4cb675f97353b40e4d8dee6fc324c7062e1a

Observation d23fb554-b6b2-4c64-80be-3802ef89af64 · outbound

This paper cites MR-GDINO: Efficient Open-World Continual Object Detection.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM MR-GDINO: Efficient Open-World Continual Object Detection

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:33:14.676390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:33:06.701427Z digest=sha256:9c47f4125e5a5519c0c2fb480987aaefceb12fe36f9650ce453a58f9184c8dee

Observation f50f399f-d566-4461-af0e-1f313c9b77a1 · outbound

This paper cites Lpt: Long-tailed prompt tuning for image classification.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Lpt: Long-tailed prompt tuning for image classification

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.750153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.750153Z digest=sha256:21ccb63d259831fab1491195ba966aa4b3c1fb2ea5bbdb8d7f65a647a0294631

Observation 40226a05-9e99-486a-8a25-e68cf003dc0d · outbound

This paper cites A Survey on In-context Learning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM A Survey on In-context Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.816254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.816254Z digest=sha256:4344698351a4491aaefeefcc8480aa4b61060b434f3dcefcf4539e36c24a9be7

Observation ea35cb9c-7fd7-4ce0-bc0d-e7d524fc5b9f · outbound

This paper cites Virgo: A Preliminary Exploration on Reproducing o1-like MLLM.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Virgo: A Preliminary Exploration on Reproducing o1-like MLLM

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:06.915732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:06.915732Z digest=sha256:0bf4500cb74e9ce991cc06b844f98377d234104c37a007544b7718bbf591d039

Observation 49e1bd80-b7be-43fc-863d-fb2bf75fdb27 · outbound

This paper cites Detecting hallucinations in large language models using semantic entropy.Nature, 630(8017):625–630, 2024.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Detecting hallucinations in large language models using semantic entropy.Nature, 630(8017):625–630, 2024

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.022394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.022394Z digest=sha256:e61d37992ce7cdb181543dbbc8038be3782b0340eb9b35f742301979d74404a9

Observation b3dc9fb6-d758-4d73-af6e-464ab485d758 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.090750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.090750Z digest=sha256:f4929ea40bdf65b40271e1ab8367f3c129c84d6cda29730910d269e52bbd7613

Observation 6cd15c42-a470-4dbe-9f0d-e99b8db63c67 · outbound

This paper cites GPTScore: Evaluate as you desire.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM GPTScore: Evaluate as you desire

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.151002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.151002Z digest=sha256:0e02a83ed2b4fdd414ade7dd02024343a8034982dc55553157847af49f441090

Observation 85588250-a0aa-44c6-a5fe-3bd9c163fbff · outbound

This paper cites The capacity for moral self-correction in large language models.Parameters, 109(1010):1011.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM The capacity for moral self-correction in large language models.Parameters, 109(1010):1011

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.271503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.271503Z digest=sha256:961cebc20e8ef4a2985ceaf4203bb3bd7bf24431873c58a189f54a0f8b383ac7

Observation 131f8f24-faff-4e6e-9cc9-ecdc554449ba · outbound

This paper cites Are Language Models Puzzle Prodigies? Algorithmic Puzzles Unveil Serious Challenges in Multimodal Reasoning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Are Language Models Puzzle Prodigies? Algorithmic Puzzles Unveil Serious Challenges in Multimodal Reasoning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.360301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.360301Z digest=sha256:75371ab7807eda99bac512ada8d48cb541c6fa47ff4d57d7c90039dc6cbf99db

Observation d385e19b-f4e6-4955-89e5-25a3370bbfea · outbound

This paper cites The Llama 3 Herd of Models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM The Llama 3 Herd of Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.475863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.475863Z digest=sha256:0b38f9eb23021ce7593d29ca965b308d5b1fc30f7a64a4ccc0cdb3bdde08f22c

Observation 82fb3068-1a9f-4e79-a8f4-edb963f00afd · outbound

This paper cites A Survey on LLM-as-a-Judge.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM A Survey on LLM-as-a-Judge

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.571919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.571919Z digest=sha256:8c7692e253cae18118c5bfcce7c4896dc326d61b507c0cb57b59f6cb271f2461

Observation a4bcd11f-f250-480f-bf44-546db8c7a328 · outbound

This paper cites Hallusionbench: an advanced diagnostic suite for entangled language hallucination and visual illusion in large vision-language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Hallusionbench: an advanced diagnostic suite for entangled language hallucination and visual illusion in large vision-language models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.664840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.664840Z digest=sha256:5a08193bafdbd3d61f581a34ace1bbc52130e492d722b76c9f4f01a480e327f7

Observation aaf1ac63-d878-4cb2-807b-c8f4228efe91 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.772797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.772797Z digest=sha256:3e8025ce61aeeec30b7fadd8055770654150a2e2556d5a0faf4c87e00a569863

Observation e4ea4dc8-3afa-4b0e-a85f-889900e9f05d · outbound

This paper cites When Continue Learning Meets Multimodal Large Language Model: A Survey.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM When Continue Learning Meets Multimodal Large Language Model: A Survey

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.869192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.869192Z digest=sha256:edfab16a6291efb4a90bb9f662fdf0c9277ac89067b38e24759141a28bbf14b5

Observation 3c11f0aa-8b18-422f-9198-fd055ab6103c · outbound

This paper cites GPT-4o System Card.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM GPT-4o System Card

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:07.959565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:07.959565Z digest=sha256:ade471e250ccdfd629bc4c9825f24640e9a500930148fd4083608d308b4f52b5

Observation efeed5f0-0c9e-4944-a8b8-e7aa228c618e · outbound

This paper cites OpenAI o1 System Card.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM OpenAI o1 System Card

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.049492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.049492Z digest=sha256:e1dfcf90abec6cc6867e57a8f041dfce2493953be4265849171a282045dc1404

Observation ffa0a01f-ae22-4626-935d-4c149f169e29 · outbound

This paper cites Towards mitigating LLM hallucination via self reflection.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Towards mitigating LLM hallucination via self reflection

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:18.503341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:33:08.183950Z digest=sha256:6ba37cb618627b0ddc6637e9d9ecff7e27da2faee88bd9814c80c473729d2a85

Observation 62c94a89-a059-4823-bb1d-acd48b4486d5 · outbound

This paper cites Hallucination augmented contrastive learning for multimodal large language model.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Hallucination augmented contrastive learning for multimodal large language model

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.290558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.290558Z digest=sha256:6285110f9d43c25f92900b9326f656abde20a486f7ff53c7c253a24cfbf7af61

Observation a72d196d-62cc-4209-83be-37aec9055dc7 · outbound

This paper cites MME-CoT: Benchmarking Chain-of-Thought in Large Multimodal Models for Reasoning Quality, Robustness, and Efficiency.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM MME-CoT: Benchmarking Chain-of-Thought in Large Multimodal Models for Reasoning Quality, Robustness, and Efficiency

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.379110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.379110Z digest=sha256:521b0dc69ddcfb88c2aa4e3cade026a84f47456ecaa89b612c763145fa830a48

Observation d3eabd59-e4da-4a04-a2a9-c88aed94f37e · outbound

This paper cites Decoupling representation and classifier for long-tailed recognition.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Decoupling representation and classifier for long-tailed recognition

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:18.166965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:33:08.480767Z digest=sha256:a935396e582c4d99ddfc6b57a8cee425a8795957d56c79311c2be36dd735fe94

Observation 2e8ba4f8-f945-4fb9-9e87-57bb85616c2f · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Gonzalez, Hao Zhang, and Ion Stoica

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.559186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.559186Z digest=sha256:2515a1a716fa3ea490da509d7b22207e73466433ae7900ecff022d53b72bbb71

Observation c3e2ac17-0813-463f-9f74-946b2dfbbd0c · outbound

This paper cites SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.631430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.631430Z digest=sha256:633941a13df67587a43a8123bebbd6fb2cb58531582940ed5a1ef299fa37a9dd

Observation 5dc004e3-889e-47bb-aa86-7f110d4a3be6 · outbound

This paper cites LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.729239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.729239Z digest=sha256:df2e6d7ff9e223921caa30bd1c9a19cf4e6dd09ce753fd1b2bcd2c40d7d6e743

Observation 3042c996-c970-4899-9c34-7b2afc9e70d9 · outbound

This paper cites Salad-bench: A hierarchical and comprehensive safety benchmark for large language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Salad-bench: A hierarchical and comprehensive safety benchmark for large language models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:08.797773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:08.797773Z digest=sha256:95dfc0e293e95f72a8099b6e1a14c4c90a7c06677ad68e491a457a22d440db50

Observation ebe72b6c-d1fb-434d-b5a8-7dd39b936bf4 · outbound

This paper cites Long-tailed visual recognition via gaussian clouded logit adjustment.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Long-tailed visual recognition via gaussian clouded logit adjustment

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:17.903767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:33:08.934739Z digest=sha256:3bd5988414e7ef2fa6337b6bc4c3b3f9a22fc9e4e4b7f21701f7aeb406c7beea

Observation 4b24e925-a8de-49f7-ae0e-ae79231d2290 · outbound

This paper cites Evaluating object hallucination in large vision-language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Evaluating object hallucination in large vision-language models

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:17.659774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:33:09.006807Z digest=sha256:7549e4da1d178ebc381fb209417465394b7a824d1ce7eef131f8c632e5f91082

Observation a46d00ea-67a9-4727-8ff8-6fe63c8f0dda · outbound

This paper cites Omnibench: Towards the future of universal omni-language models.arXiv preprint arXiv:2409.15272, 2024.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Omnibench: Towards the future of universal omni-language models.arXiv preprint arXiv:2409.15272, 2024

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.096630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.096630Z digest=sha256:a5187cadc5ce39373d79b4e930244f153f0cc3b37e41ee6a1fef8454d52a90f1

Observation ef6c1503-b847-4245-861e-fa84899b4c71 · outbound

This paper cites DeepSeek-V3 Technical Report.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM DeepSeek-V3 Technical Report

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.169761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.169761Z digest=sha256:3c0a6260ef4ef70290e1acfa594b82edc5f1c8982e4718ee7fb1c1fdd01ac966

Observation 169dedb4-ccf8-4215-b20e-99bfc5d5e067 · outbound

This paper cites Mitigat- ing hallucination in large multi-modal models via robust instruction tuning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mitigat- ing hallucination in large multi-modal models via robust instruction tuning

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:17.441938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:33:09.244473Z digest=sha256:f1281444f15808d7427f694afccc1ce349fb492c6ff8e856a8c80a79236e336f

Observation a549b860-d5f8-4579-88d9-6e0627e2fb8f · outbound

This paper cites Visual instruction tuning, 2023.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Visual instruction tuning, 2023

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:17.186401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:33:09.337797Z digest=sha256:2e6ca575b82e3346c83556cef8111de0cd29086a4c68c6ee46b32916bd8486e1

Observation 2551a763-37bd-41ce-931b-e4db63c3bf7d · outbound

This paper cites Visual-RFT: Visual Reinforcement Fine-Tuning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Visual-RFT: Visual Reinforcement Fine-Tuning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.432965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.432965Z digest=sha256:189dbb0685fedfa0777fef405b1b44464b8e956c16f6d5d25ca4c599dd1e78e9

Observation 7f6f5d91-2270-4ad3-be3f-e631d7dceb6c · outbound

This paper cites Decoupled weight decay regularization.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Decoupled weight decay regularization

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.543779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.543779Z digest=sha256:1733e887935f8f8bfdfbbb0bbbde5037b2e1caa76506d36f55c8b3f154051a08

Observation b7779aaa-102d-49ed-8517-7739ee88a52b · outbound

This paper cites Mathvista: Evaluating mathematical reasoning of foundation models in visual contexts.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mathvista: Evaluating mathematical reasoning of foundation models in visual contexts

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.643603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.643603Z digest=sha256:dbce9a47a28c0cfbf72cc8083f7069496c0f4178dde32450a2965320869175a0

Observation 50b2646c-45b8-47df-92bf-d05cb8e0947f · outbound

This paper cites Learn to explain: Multimodal reasoning via thought chains for science question answering.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Learn to explain: Multimodal reasoning via thought chains for science question answering

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.735339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.735339Z digest=sha256:8424f96f8455c0f6cb6676c65336f310f877309d5f0db339c8a1d455bb381472

Observation 0c0d71be-894d-4563-aa5a-2caaf14733de · outbound

This paper cites Ursa: Understanding and verifying chain-of-thought reasoning in multimodal mathematics.arXiv preprint arXiv:2501.04686, 2025.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Ursa: Understanding and verifying chain-of-thought reasoning in multimodal mathematics.arXiv preprint arXiv:2501.04686, 2025

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.808767Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.808767Z digest=sha256:e8621bcf6c052926ef3578950a05e74c602ac7dc7fb7833bc8e9be94d0d4601f

Observation 63567212-d451-45be-8661-10868da2d31b · outbound

This paper cites Ok-vqa: A visual question answering benchmark requiring external knowledge.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Ok-vqa: A visual question answering benchmark requiring external knowledge

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.913245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.913245Z digest=sha256:dbc05b186dc0fa381342f4671cad91469267218b2f38b48a344b870a7467bb78

Observation d1fc6cd4-025a-416c-a8d9-0a324dbcea08 · outbound

This paper cites Chartqa: A benchmark for question answering about charts with visual and logical reasoning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Chartqa: A benchmark for question answering about charts with visual and logical reasoning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:09.990090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:09.990090Z digest=sha256:33f8525b8b2de5e1da82da7fbaf70f11b33fdb4f522a61b7fd7c51ee9eee9be4

Observation d12ecdd1-663f-4408-a3f2-485ff7baeab6 · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.071390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.071390Z digest=sha256:9624e3b466b26c3e8911c617e867420b07017e2e0e00b86f5ec16ec9bbb9ccbe

Observation 4d684078-fce3-4628-b438-3871fe71aeab · outbound

This paper cites Compositional chain-of- thought prompting for large multimodal models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Compositional chain-of- thought prompting for large multimodal models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.139052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.139052Z digest=sha256:be304f316ec2893310e231d3e0e540a7221c2a716d1bba961a0bd714b3a8c132

Observation a53719f6-9c58-428f-b724-d963b96de839 · outbound

This paper cites Pytorch: An imperative style, high-performance deep learning library.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Pytorch: An imperative style, high-performance deep learning library

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.217648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.217648Z digest=sha256:7f8101c2410e615b806e76f4b1272d7077abf27c8503e4e838e3a467bd21b742

Observation 04a56abd-815d-43aa-81bf-37bc82357b95 · outbound

This paper cites LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.310457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.310457Z digest=sha256:9a885ddf5fc229f1070517765f00bb4d30aeb2248b617b52daf4a52908d08336

Observation 9e83b7d7-971b-4220-8f1f-5ce5d603d714 · outbound

This paper cites ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM ReCEval: Evaluating Reasoning Chains via Correctness and Informativeness

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.405095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.405095Z digest=sha256:c606f0bb87850d09e1241f3fb4ec0695eead68dbced9161fc1edb086431a08fa

Observation 698b7e53-57d5-4cb9-a2ee-331fa319c673 · outbound

This paper cites Mitigating object hallucination in mllms via data-augmented phrase-level alignment.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mitigating object hallucination in mllms via data-augmented phrase-level alignment

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:16.787632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:33:10.547565Z digest=sha256:028ccd790d96f159f48ecfc8ab046f72101883b5d1e7538acfb2070958c78807

Observation 284f205c-e24d-49d4-98ed-c7382e884264 · outbound

This paper cites Proximal Policy Optimization Algorithms.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Proximal Policy Optimization Algorithms

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.633888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.633888Z digest=sha256:09741645a986af7f3d61710d5578c46ecc7affaa51cee9ca210dfddf4ef1f9d4

Observation a0bec2fa-ad09-47da-9b3d-dcfc443a20cb · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.745345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.745345Z digest=sha256:af966525defd19b7b537234f2dde4854ee1a4bd16a2fb89405f3c86588cc151e

Observation b2b2aefe-d4c4-4b50-9c4c-dffc83574f91 · outbound

This paper cites Monte carlo tree search: A review of recent modifications and applications.Artificial Intelligence Review, 56(3):2497–2562, 2023.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Monte carlo tree search: A review of recent modifications and applications.Artificial Intelligence Review, 56(3):2497–2562, 2023

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.856167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.856167Z digest=sha256:bf0dd75466ff1b9b921571b222b59cc9cd1e77640fd2123bd96bb0027cc59329

Observation dcba0640-8e74-499b-936a-ba7fdccf62b7 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Gemini: A Family of Highly Capable Multimodal Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:10.924496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:10.924496Z digest=sha256:80bc45f880db75f2cdb36cfdc2db83968574a6c48079ec0db69dcb7f378a269a

Observation 443738d3-a1a4-4e01-8fa7-10e641d3eecb · outbound

This paper cites QVQ: To See the World with Wisdom.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM QVQ: To See the World with Wisdom

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:16.606887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:33:11.048068Z digest=sha256:25b0af3efea32746eded6ef48deb15ea7aee784ea08e9677f006e0c8bd12e11d

Observation 61e93b22-71a1-4bf0-957f-2c3634b8c1f2 · outbound

This paper cites Uncertainty-Based Abstention in LLMs Improves Safety and Reduces Hallucinations.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Uncertainty-Based Abstention in LLMs Improves Safety and Reduces Hallucinations

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:11.153366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:11.153366Z digest=sha256:db3cf2c38d1c81ed14b0bc704ad045c097b28c3d2ac3be83a8301bcfb76f94a9

Observation 8084a308-e4ac-4805-9d83-a724b7daae11 · outbound

This paper cites Eyes wide shut? exploring the visual shortcomings of multimodal llms.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Eyes wide shut? exploring the visual shortcomings of multimodal llms

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:11.217618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:11.217618Z digest=sha256:54ff2fe14a9614aba40b4914bddcff7aaf4aef935756da7282d9cf512d15bd7b

Observation e08b3ec0-020d-434d-8a05-7a02209f9913 · outbound

This paper cites Measuring multimodal mathematical reasoning with math-vision dataset.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Measuring multimodal mathematical reasoning with math-vision dataset

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:16.486557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:33:11.281006Z digest=sha256:8e0118c7596bba8aa329d9cbe0e1dececd275b9c338f1e743604655a28ea8553

Observation 22da4bca-b3c8-4658-8779-f635c78267e6 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:11.350592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:11.350592Z digest=sha256:e4d7d0609dd7200c375e5e5ba3c7245a5f0685036452aeb391e0136f2cfac020

Observation 1772c607-1e3e-4cc8-a011-797bd52d4277 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Chain-of-thought prompting elicits reasoning in large language models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:11.422440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:11.422440Z digest=sha256:ab5ec78890e566e1393e4d80af0a072213bcc2e323beeb5873447953e146953c

Observation edbb8b53-7b56-4357-9147-0fd3e8b09e71 · outbound

This paper cites RITUAL: Random Image Transformations as a Universal Anti-hallucination Lever in Large Vision Language Models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM RITUAL: Random Image Transformations as a Universal Anti-hallucination Lever in Large Vision Language Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:11.492919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:11.492919Z digest=sha256:e75396c6d8ba3e73dec7f4d674a7d4ef0bfca8bc518e177410a408b218c27f34

Observation 630b0b44-60ff-4031-b5bd-136c5134da95 · outbound

This paper cites Autohallusion: Automatic gen- eration of hallucination benchmarks for vision-language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Autohallusion: Automatic gen- eration of hallucination benchmarks for vision-language models

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:16.317333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:33:11.562584Z digest=sha256:039eafa28ab437687d805cd3ec199933f3cc673554bc28bdab877108c8037800

Observation 2767d4b3-7a1e-44f6-9233-16d9eaf5c5ab · outbound

This paper cites Grok 3 Beta — The Age of Reasoning Agents.https://x.ai/grok, 2025.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Grok 3 Beta — The Age of Reasoning Agents.https://x.ai/grok, 2025

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:16.123781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:33:11.613404Z digest=sha256:cb874bd5bc38a2a07acda4c783649eff0bd59f8ad93a4fced7e4a3211e90d984

Observation 9b1c93ed-31f5-4c5e-aef0-419a9196b660 · outbound

This paper cites Mitigating object hallucination via concentric causal attention.Advances in Neural Information Processing Systems, 37:92012–92035, 2024.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mitigating object hallucination via concentric causal attention.Advances in Neural Information Processing Systems, 37:92012–92035, 2024

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.957279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:33:11.724385Z digest=sha256:7e65de7a2b1eeac2432df45d3e15369cba4df48195eb2c1b70c078b75279a560

Observation 6ce3eff8-0eda-42fa-bd40-3d76fa9bab4f · outbound

This paper cites Llava-cot: Let vision language models reason step-by-step, 2024.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Llava-cot: Let vision language models reason step-by-step, 2024

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.820797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:33:11.795732Z digest=sha256:24471d2ffe311d7e61f1a1f9920be9538c52e86202aa7329c9ccb9a81aec7c6a

Observation eebbacd3-290d-40b6-9616-cec260b62609 · outbound

This paper cites Qwen2.5 Technical Report.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Qwen2.5 Technical Report

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:11.894050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:11.894050Z digest=sha256:e14a71cb01b0382f50b03176c763fe74d762fa7303acf60e3b40304de86e8181

Observation 98b48a0d-4c30-470a-b2d7-380e837c3767 · outbound

This paper cites Soft-prompting with graph-of-thought for multi-modal representation learning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Soft-prompting with graph-of-thought for multi-modal representation learning

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.698313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:33:11.952400Z digest=sha256:8ffae45b191f694377674f9f5d58ac3c13199bcce7cd0e12c3ef337af339dba4

Observation 9304223c-8c7d-4574-b644-7bb3046630c9 · outbound

This paper cites Mitigating hallucination in large vision- language models via modular attribution and intervention.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mitigating hallucination in large vision- language models via modular attribution and intervention

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.543195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:33:12.053954Z digest=sha256:34b9fe84fad977ffe4bb898869af63c53d3e6a4c4123582a93bfd2297453a6eb

Observation 73dcefd7-a958-4dd6-bded-72c5e627c82b · outbound

This paper cites R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.177383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.177383Z digest=sha256:f16ac2b36cdbeaaa75afcd11d2db81d77857e1cf786ad2d9d145896aca980042

Observation 0b48fbce-4c79-415a-aeac-4e3e872724ae · outbound

This paper cites Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.258432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.258432Z digest=sha256:87d21ab91e049a8ffff67abea1b6fab2e6b7d8bb0ebe72b16549d3b43bc238bb

Observation c618bb1f-9378-4f4e-a8e8-05ca272c661b · outbound

This paper cites React: Synergizing reasoning and acting in language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM React: Synergizing reasoning and acting in language models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.383675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.383675Z digest=sha256:1d1ebf0291948e7af62c6161786a3a96a788aaa33da5c121d01fcb2cc49e0958

Observation ed50d571-49f4-4efd-a3c0-5fc29315c7a7 · outbound

This paper cites LIMO: Less is More for Reasoning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM LIMO: Less is More for Reasoning

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.448500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.448500Z digest=sha256:73a210c3c240146391e9355292fa2e22be5af3ea5746862c775574227b6db2a2

Observation 3cb4a8e8-4452-4317-ac9e-074e4f911ef1 · outbound

This paper cites Hallucidoctor: Mitigating hallucinatory toxicity in visual instruction data.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Hallucidoctor: Mitigating hallucinatory toxicity in visual instruction data

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.557605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.557605Z digest=sha256:3cc8dcfaa4f97209b8266136fe7160c3b702031c0f4e6fefae7c63236fe2f3fe

Observation 9d8616b4-cc9f-478c-83fe-fa1433f61ea9 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.653152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.653152Z digest=sha256:a5df23b8566feb7002339529ae574724287f6f058846bf1c79473c19ca01e457

Observation 0f2fb295-f912-4fbd-8059-15d7e34eb919 · outbound

This paper cites Rlhf-v: Towards trustworthy mllms via behavior alignment from fine-grained correctional human feedback.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Rlhf-v: Towards trustworthy mllms via behavior alignment from fine-grained correctional human feedback

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.750492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.750492Z digest=sha256:67c6d09fa4defb69e4040b905776ddb182c28fad5d1b470fb75fcfec6aefa80c

Observation 9ebaadc9-87c8-41a0-ba69-f23f61aa0c81 · outbound

This paper cites LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:12.857251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:12.857251Z digest=sha256:bae8a6195d151e2b2a2c60c4e6c8ce256440e5260b1c2c5bedd2cd14f99b9964

Observation 5e0dc656-f057-48a7-a72b-70435f63512d · outbound

This paper cites Reflective instruction tuning: Mitigating hallucinations in large vision-language models.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Reflective instruction tuning: Mitigating hallucinations in large vision-language models

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.380205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:33:12.980715Z digest=sha256:6fdba33938752354b521e6a40e5fc2aac5210ba78781dfd4dd170e7037413b35

Observation 5a298254-9648-4782-abb3-752e69622b8f · outbound

This paper cites Mathverse: Does your multi-modal llm truly see the diagrams in visual math problems? InEuropean Conference on Computer Vision, pages 169–186.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Mathverse: Does your multi-modal llm truly see the diagrams in visual math problems? InEuropean Conference on Computer Vision, pages 169–186

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:13.073898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:13.073898Z digest=sha256:853d565c9dd632239e0a66116cd6ea813d54e09a3bac5c9843d9f30be565f7d0

Observation 140db751-7cb8-4809-9e30-2e832334feb1 · outbound

This paper cites VL-Uncertainty: Detecting Hallucination in Large Vision-Language Model via Uncertainty Estimation.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM VL-Uncertainty: Detecting Hallucination in Large Vision-Language Model via Uncertainty Estimation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:13.190579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:13.190579Z digest=sha256:19b776b3d1470e98ab5c00f40c6c6cd6b272d244afffe0e0c0de1c1681aa4f86

Observation 334b295f-408d-44bd-8026-aa20d981a015 · outbound

This paper cites The Lessons of Developing Process Reward Models in Mathematical Reasoning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM The Lessons of Developing Process Reward Models in Mathematical Reasoning

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:13.281932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:13.281932Z digest=sha256:a32eb5bd7c064c651b66313f26aaba89c2265d3870fe5697cc891dbaa121029c

Observation 5fd38f61-18d5-478f-b822-1218dce4fd63 · outbound

This paper cites Multimodal chain-of-thought reasoning in language models.Transactions on Machine Learning Research.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Multimodal chain-of-thought reasoning in language models.Transactions on Machine Learning Research

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:13.378434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:13.378434Z digest=sha256:293ddc55f7e68f29d3065a1f192936cc54b0fb22db0529e224b85b736d41eb9a

Observation daedcae0-b3a7-437e-b406-b2b3c87dc327 · outbound

This paper cites Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:13.470889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:13.470889Z digest=sha256:9630b844ae074e106874940ad8ed806abd8b9cfcb1726987eac631d4c4031f51

Observation cfacc5c1-dfc2-4639-8b13-4af198dd8327 · outbound

This paper cites A picture is worth a graph: A blueprint debate paradigm for multimodal reasoning.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM A picture is worth a graph: A blueprint debate paradigm for multimodal reasoning

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.196268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:33:13.563966Z digest=sha256:edba4c917c21cce002c44aff10efe8d83f3e079435f7ced942927f7fcaf01296

Observation 7f35f66e-4cbe-463e-a511-e8f97f085ad1 · outbound

This paper cites Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Thinking Before Looking: Improving Multimodal LLM Reasoning via Mitigating Visual Hallucination

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T12:33:13.663322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:33:13.663322Z digest=sha256:599689b82b5a019d7c6feb9efd2683457b2d2d9712b43f575914e7edbf7dae27

Observation 0db258ac-a6f2-4765-b849-f5e6ecfc0fd7 · outbound

This paper cites Relying on the unreliable: The impact of language models’ reluctance to express uncertainty.

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM Relying on the unreliable: The impact of language models’ reluctance to express uncertainty

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:33:15.038180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:33:13.747662Z digest=sha256:3dbd0ad3d21fab62fa99c895d8d7ce59ab19853d0c0c454582f4ee3a0b0c17f3

Observation 0d179a1e-f3fd-41b8-a842-23e070afc6e7 · outbound

This paper cites the answer is [answer in the input].

MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM the answer is [answer in the input]

Reference 87

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T12:33:14.881658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T12:33:13.820971Z digest=sha256:ed13b8091e555d1d94538d3900cddba5cf70887b70ce2bd1068a6be511ee8f26

Pith citing papers

Observation 8f61cce6-5c2d-4d17-8b3f-2c8cc3ad9045 · inbound

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models cites this paper.

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-05T10:39:03.028373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:39:03.028373Z digest=sha256:bcb0a91ce799688cca185a46286e0a1c83d35baa864249ec028ab49ca9bb50f6

Observation 93d009a1-1467-4909-ae1a-aa45d4e09569 · inbound

Understanding the Role of Hallucination in Reinforcement Post-Training of Multimodal Reasoning Models cites this paper.

Understanding the Role of Hallucination in Reinforcement Post-Training of Multimodal Reasoning Models MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:53:16.282414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T20:48:52.130130Z digest=sha256:411a3b70aa0a24e289a38e42746552bd3028f64353f5793a95852e1a180b15be

Observation 8db78eb4-7cbe-4125-83e5-637815967363 · inbound

HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering cites this paper.

HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM

Reference 230

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T23:54:45.651801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-09T23:51:47.724033Z digest=sha256:f6886e102b224498f11b0073556e9e187f1659dcd20b16abd1a9fc6f7fb57bee

Observation 45eb3f94-6839-4143-946f-13d83c6c4310 · inbound

SoccerRef-Agents: Multi-Agent System for Automated Soccer Refereeing cites this paper.

SoccerRef-Agents: Multi-Agent System for Automated Soccer Refereeing MIRAGE: Assessing Hallucination in Multimodal Reasoning Chains of MLLM

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:46:13.723544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-08T08:03:57.287297Z digest=sha256:8bf67ba25ba8fb49f84bceced40779796745c7c4e7bbe15dc052fb0400fd8b8e