Pith. sign in

Paper Citation Record · LEDGER

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models

As of 8 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:2507.15652.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.15652 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:32:34.571356Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact0
  • verified fuzzy12
  • unresolved30
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2130b2a7-bad8-4f6e-87ed-024feb8423c5 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.328327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.328327Z digest=sha256:5cdd1b9e6588658fd8faf1946204c246fff5daee1d7ecabbf9ef6eaf66256467

Observation 17c919b7-2fdc-491c-b2cc-9c284d6233e7 · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.332331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.332331Z digest=sha256:31c2e78085f5efaab77219c61007f940e0495d8edef0e4c933434f316d8f7a5e

Observation a0314bc4-10ea-407c-9e23-8526f66cd422 · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.441636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.441636Z digest=sha256:33e1260b04e678a38540e715c2142d419095387adf34b7ba93a5f3f18ea499be

Observation d560e042-2d66-46c3-9ac1-0b5f3b2520d9 · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.445684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.445684Z digest=sha256:b7c847f14e14f2d4bcf0f82cea186a7ed7337a57cfb9f20205221e689032201a

Observation fcf4d8bd-c3ec-4d25-8c61-ea90b40d5844 · outbound

This paper cites GPT-4 Technical Report.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models GPT-4 Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.449497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.449497Z digest=sha256:33198120c3d810512dce7034a30056edea235a1f322be08eb643ab76e3096de5

Observation 593148cc-a9ac-4956-8084-0a63a7873b55 · outbound

This paper cites Improved baselines with visual instruction tuning,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Improved baselines with visual instruction tuning,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:35.067278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:32:34.452527Z digest=sha256:30aa54203cafd09f47be6667b816ed13c6247488d08cb481c1a8ba5b9c555e61

Observation d9b48a2b-da9c-441b-806f-bbcc140f388d · outbound

This paper cites ANOLE: An Open, Autoregressive, Native Large Multimodal Models for Interleaved Image-Text Generation.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models ANOLE: An Open, Autoregressive, Native Large Multimodal Models for Interleaved Image-Text Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.455945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.455945Z digest=sha256:a0c58067c6942b9e48d68dbaabeaa28a29422c421e74e4208fb37ae12c5662c5

Observation 277e86f3-a42b-477d-b914-a29cc941b790 · outbound

This paper cites Hallucination of Multimodal Large Language Models: A Survey.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Hallucination of Multimodal Large Language Models: A Survey

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.459305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.459305Z digest=sha256:ee7cd17159b7249fe009f2cc06692f7b97f5512282d023effd45016e512918ab

Observation 782f701a-cbda-4b39-9432-d2fffed004b8 · outbound

This paper cites A Survey on Hallucination in Large Vision-Language Models.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models A Survey on Hallucination in Large Vision-Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.463265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.463265Z digest=sha256:31f960694a787809299d8f18a9f533f22e374aedd45517042f212f9272cd6b0e

Observation 6b3533b2-517e-4638-adca-81adca8f7420 · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Evaluating Object Hallucination in Large Vision-Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.466909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.466909Z digest=sha256:b4e35a5f2a72d22c5e784c79040e6a1d0e4efd142e13b4c63470aba5024f0e9e

Observation 0c1bf0fe-28d4-4ef2-a113-4ddc74b94663 · outbound

This paper cites Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.470098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.470098Z digest=sha256:cb7a46854a8fc52d8f1f2a19eeca38a44a9a1b30bd2a2e6d5da1026b499754be

Observation 6a2627c0-c68d-4aff-87a2-6975aa66c9b1 · outbound

This paper cites A Survey of Hallucination in Large Foundation Models.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models A Survey of Hallucination in Large Foundation Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.473083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.473083Z digest=sha256:34518b3fc39416dbd017deaefd180e26c46396a39af6bbac80b07babdea2871d

Observation 4180b87b-fc99-4342-8b17-1fff9bf367cf · outbound

This paper cites HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.476129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.476129Z digest=sha256:47708c57404cc20654e863b785afa7129fa02db53ce57f8c90a159c076ccdd81

Observation 67f260e9-e7ee-46aa-a772-65389cf85913 · outbound

This paper cites Advancing Medical Imaging with Language Models: A Journey from N-grams to ChatGPT.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Advancing Medical Imaging with Language Models: A Journey from N-grams to ChatGPT

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.479419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.479419Z digest=sha256:74f1f2a5b2e822f1a1b75791440640933dffc69c916faf803e540012cc5af9b2

Observation c0ece340-1017-479e-b169-fb1deadea48f · outbound

This paper cites ChatCAD: Interactive Computer-Aided Diagnosis on Medical Image using Large Language Models.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models ChatCAD: Interactive Computer-Aided Diagnosis on Medical Image using Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.483196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.483196Z digest=sha256:0b19f1637b6c64ee650341b9972b17452e6c5ec6d1d04355c6bcbdd030a4bfb2

Observation e8257ff2-a84e-465c-81a6-ac3e9c0bef4c · outbound

This paper cites A survey on multimodal large language models for autonomous driving,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models A survey on multimodal large language models for autonomous driving,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:35.056723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:32:34.486759Z digest=sha256:abd0fc10ee6592a79d21171581a6c1bc929beebc6cf13dcd08f908c008f17ad8

Observation 348d23ca-efb2-4462-a8be-a161e57965cc · outbound

This paper cites Drivemlm: Aligning multi-modal large language models with behavioral planning states for autonomous driving,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Drivemlm: Aligning multi-modal large language models with behavioral planning states for autonomous driving,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.490330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.490330Z digest=sha256:91ad88576b6d908b61661a63f5f4e27f781b6c0e1c69cae3004c032b72f3ec21

Observation e82b4d40-579c-4b8a-ac39-81300039953e · outbound

This paper cites Evaluating a large language model on searching for gui layouts,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Evaluating a large language model on searching for gui layouts,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:35.046906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:32:34.493938Z digest=sha256:0e5d9ab874d841eb3e2552a3e3c1b68b8ae93b45ec90aa6a56296a4aaf145c29

Observation d0c1b8c4-852a-40da-9bdd-f068354066ca · outbound

This paper cites DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.496883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.496883Z digest=sha256:f215e3dc36ebca77eec72cb99e317bf90ecce2aa4b80ddef3be3384c82ecd797

Observation d5142ab2-76c8-4829-b805-da10cfb555fe · outbound

This paper cites In-Context Sharpness as Alerts: An Inner Representation Perspective for Hallucination Mitigation.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models In-Context Sharpness as Alerts: An Inner Representation Perspective for Hallucination Mitigation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.499975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.499975Z digest=sha256:d05f2145710261c2cfed4feac40194737c8ffb255d452c5ac6ca2b7d67d8c52a

Observation cabb3c46-a82a-4a36-9996-9b49f7605afe · outbound

This paper cites LLMs Know More Than They Show: On the Intrinsic Representation of LLM Hallucinations.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models LLMs Know More Than They Show: On the Intrinsic Representation of LLM Hallucinations

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.503372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.503372Z digest=sha256:8981d576e051eb9a32cccad77fb13d9b9a570521867dac305e4363aa9d191f41

Observation 7c6b1805-2852-4866-9e1b-31ee16c1f0bf · outbound

This paper cites Llama SLayer 8B: Shallow Layers Hold the Key to Knowledge Injection.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Llama SLayer 8B: Shallow Layers Hold the Key to Knowledge Injection

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.506797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.506797Z digest=sha256:daacc5f053708ebb6cc0b37670ee1cef930c263adb90e72027ec85fc2c059b28

Observation 4684c0d9-e17a-4473-94a5-90294e5ce1e2 · outbound

This paper cites Insights into LLM Long-Context Failures: When Transformers Know but Don't Tell.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Insights into LLM Long-Context Failures: When Transformers Know but Don't Tell

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.510625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.510625Z digest=sha256:ecc74c17f89673497531ab763b61204913b554571639791a9c737ab1b22782a5

Observation 6ef8ed9b-9686-4ae0-af3d-8573f8159d15 · outbound

This paper cites Knowledge Mechanisms in Large Language Models: A Survey and Perspective.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Knowledge Mechanisms in Large Language Models: A Survey and Perspective

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.514275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.514275Z digest=sha256:6bcc50ee19da6135bf0f3920d5799e236e20619e2a7875aec13920f93f96bfe6

Observation 3d1fcd93-d40d-4938-9cb3-0713876ef7b7 · outbound

This paper cites MLLM can see? Dynamic Correction Decoding for Hallucination Mitigation.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models MLLM can see? Dynamic Correction Decoding for Hallucination Mitigation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.517821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.517821Z digest=sha256:1afe9ec668e927af6cf4b75392c2e88c775d2d7d0a096c760ad9e22195cf7164

Observation 08b89adc-3f7e-4935-9b94-7cb210beee41 · outbound

This paper cites Unleashing region understanding in intermediate layers for mllm-based referring expression generation,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Unleashing region understanding in intermediate layers for mllm-based referring expression generation,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:35.036052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:32:34.521189Z digest=sha256:b2f17e51505a5a2d6e31507be112778f3f4048feb2c187b80e23c76467a30eda

Observation d6b394d2-9faa-4b38-a9e2-9b4122b41ac9 · outbound

This paper cites Branchynet: Fast inference via early exiting from deep neural networks,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Branchynet: Fast inference via early exiting from deep neural networks,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:35.026046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:32:34.524262Z digest=sha256:817cd2d838d20c41d2cab705bcd4e489416b60465025d185875c427ceb1c87cc

Observation 1c694b9b-4ce1-4710-9d8f-3a0225167cdd · outbound

This paper cites Depth-Adaptive Transformer.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Depth-Adaptive Transformer

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.526903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.526903Z digest=sha256:4e7135dbce5f68c00267ba66a741d8f75afc1d3383cae281510b11b294468802

Observation 9e416501-821c-4235-b285-89d855e11dae · outbound

This paper cites Confident adaptive language modeling,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Confident adaptive language modeling,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:35.016928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:32:34.529784Z digest=sha256:3b3dc389e3877f188acd30ece990a08ad8eb0f0f583a0eb9bffc6623d853991d

Observation 122b16fc-ee3d-4561-8175-2e63efa27f97 · outbound

This paper cites Hallucination augmented contrastive learning for multimodal large language model,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Hallucination augmented contrastive learning for multimodal large language model,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:35.007082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:32:34.532510Z digest=sha256:78aab6df5fc140be255b07abf9a1da198693be2a941012edd5b8a3eee1ec9da8

Observation 18cf72b5-94b0-4ae0-a7ec-10d30ecc1f81 · outbound

This paper cites Delve into Visual Contrastive Decoding for Hallucination Mitigation of Large Vision-Language Models.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Delve into Visual Contrastive Decoding for Hallucination Mitigation of Large Vision-Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.535229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.535229Z digest=sha256:d24e0b52167998debb4899778d736a2351e3aa5099915f83eed4feee97a5e33c

Observation 797cb5d4-3cbb-48dd-b67b-cfb374d74172 · outbound

This paper cites Mitigating object hallucina- tions in large vision-language models through visual contrastive decoding,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Mitigating object hallucina- tions in large vision-language models through visual contrastive decoding,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:34.995099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:32:34.538089Z digest=sha256:978b5beb7e4a80edbed259c94289d5c96d3687d384c098d3aeae207991c7040e

Observation d6d1e12e-8049-433f-a92f-0c05434cee76 · outbound

This paper cites Mirage in the Eyes: Hallucination Attack on Multi-modal Large Language Models with Only Attention Sink.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Mirage in the Eyes: Hallucination Attack on Multi-modal Large Language Models with Only Attention Sink

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.540866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.540866Z digest=sha256:2615b11b0120e5ef03a2e5a28e5d0175cec45d546a734554a87ae2b5aa5c8bd0

Observation 95d0c6dc-2568-4125-b628-eef990370a11 · outbound

This paper cites Hallucidoctor: Mitigating hallucinatory toxicity in visual instruction data,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Hallucidoctor: Mitigating hallucinatory toxicity in visual instruction data,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:34.984548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:32:34.544752Z digest=sha256:638647abb6384083abfdaf1a29b1057e2234e6ef28f5dbef515557c7e9626944

Observation 3fb021ba-c070-4165-94c6-e41693bede4e · outbound

This paper cites Look Twice Before You Answer: Memory-Space Visual Retracing for Hallucination Mitigation in Multimodal Large Language Models.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Look Twice Before You Answer: Memory-Space Visual Retracing for Hallucination Mitigation in Multimodal Large Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.547689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.547689Z digest=sha256:33f9cdbc444562955a33d031d21a7486bf41d23d54e36aff96e65cb79aeca87f

Observation 5831311d-7ac5-4f57-a000-47a7816de68c · outbound

This paper cites HELPD: Mitigating Hallucination of LVLMs by Hierarchical Feedback Learning with Vision-enhanced Penalty Decoding.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models HELPD: Mitigating Hallucination of LVLMs by Hierarchical Feedback Learning with Vision-enhanced Penalty Decoding

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.551576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.551576Z digest=sha256:20967215001f6a1084646a5cff7265b786a0b2559459f4779978787a1e267dba

Observation f4ff1d17-dd9f-4aeb-a2ab-d1ad4bcb5b74 · outbound

This paper cites Opera: Alleviating hallucination in multi-modal large language models via over-trust penalty and retrospection-allocation,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Opera: Alleviating hallucination in multi-modal large language models via over-trust penalty and retrospection-allocation,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:34.974822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:32:34.555323Z digest=sha256:9412fd9fe837d685d7f353c73a226bb9ada67baa148e78cda7a925cee948f3ba

Observation 73cd3b5b-2e83-4bba-ab18-0a85af58df7e · outbound

This paper cites Instructblip: Towards general-purpose vision-language models with instruction tuning,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Instructblip: Towards general-purpose vision-language models with instruction tuning,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:34.965168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:32:34.558864Z digest=sha256:4cc3fcc71de4a128a7b6f130232808582e11c6d00374f1566fcb1a303e5eca3e

Observation 04aee0e3-03be-4a8f-95bc-bf62b28f4756 · outbound

This paper cites MiniGPT-4: Enhancing vision-language understanding with advanced large language models,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models MiniGPT-4: Enhancing vision-language understanding with advanced large language models,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:34.954851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T15:32:34.561703Z digest=sha256:957a9d7a5d1c4856e9aa45fc7b3ee2cae2b929be6b3a87948502d386ac8cde57

Observation 7bb93f49-e7d5-4f2c-946b-0155508fedfc · outbound

This paper cites Qwen Technical Report.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Qwen Technical Report

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.564854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.564854Z digest=sha256:a88d1692508ae7693965e127e2050ec6737479ef584108d52b97b705a16e0b37

Observation 60c6ffac-10c6-440e-ac75-fe2fbdae59e4 · outbound

This paper cites MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.567976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.567976Z digest=sha256:80fcbcf1db524a8da5bf89b19c87e90bed64bca42d07314c8300cb61347cd190

Observation ec82dec8-964a-44bd-a05c-eef91f3066bf · outbound

This paper cites Object Hallucination in Image Captioning.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Object Hallucination in Image Captioning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.571356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.571356Z digest=sha256:e608b7889126d7efe06aa59dc89327cf68850614d4ba2ca3a1435d136939ba03

Pith citing papers

No inbound Pith citation observations are available.