Pith. sign in

Paper Citation Record · LEDGER

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models

As of 10 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:2507.15652.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.15652 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:32:34.571356Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact0
  • verified fuzzy12
  • unresolved30
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2130b2a7-bad8-4f6e-87ed-024feb8423c5 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.328327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.328327Z digest=sha256:9ac2ebf4ebf9620280a5abc2c2011e7142ba2edf792e5975dc5d0403abeba57d

Observation 17c919b7-2fdc-491c-b2cc-9c284d6233e7 · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.332331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.332331Z digest=sha256:3d88c4e521a9246a00c1f4000e896dc3f6267224bbac35200e1b37f00b6fe8d5

Observation a0314bc4-10ea-407c-9e23-8526f66cd422 · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.441636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.441636Z digest=sha256:abec9100a7056e967cda36dc4168d56e1ceb11a3b0faec04eec625d24fadbc5e

Observation d560e042-2d66-46c3-9ac1-0b5f3b2520d9 · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.445684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.445684Z digest=sha256:9412dc3d143426c0eec5e4451334a9d998f471dee3677bb2d591302d1587715a

Observation fcf4d8bd-c3ec-4d25-8c61-ea90b40d5844 · outbound

This paper cites GPT-4 Technical Report.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models GPT-4 Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.449497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.449497Z digest=sha256:fff9f80526dbc8c3b71d8effa096dc4378ee39517cfbf30d3b101445943e3605

Observation 593148cc-a9ac-4956-8084-0a63a7873b55 · outbound

This paper cites Improved baselines with visual instruction tuning,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Improved baselines with visual instruction tuning,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:35.067278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:32:34.452527Z digest=sha256:f87323460580802d145b3c05254761c54b0c56ccc0b57f0ee76a5c3aabab6099

Observation d9b48a2b-da9c-441b-806f-bbcc140f388d · outbound

This paper cites ANOLE: An Open, Autoregressive, Native Large Multimodal Models for Interleaved Image-Text Generation.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models ANOLE: An Open, Autoregressive, Native Large Multimodal Models for Interleaved Image-Text Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.455945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.455945Z digest=sha256:9ae699df5f2a831035ce657f7f6bb809309ce444d706ed91319fe5afb77e1b34

Observation 277e86f3-a42b-477d-b914-a29cc941b790 · outbound

This paper cites Hallucination of Multimodal Large Language Models: A Survey.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Hallucination of Multimodal Large Language Models: A Survey

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.459305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.459305Z digest=sha256:04b1c2dc3a95c8ae5bb230088a1a2f9cca3a829bb742fe1d85dfb3eddc85d29d

Observation 782f701a-cbda-4b39-9432-d2fffed004b8 · outbound

This paper cites A Survey on Hallucination in Large Vision-Language Models.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models A Survey on Hallucination in Large Vision-Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.463265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.463265Z digest=sha256:772e878d1cb6193e210e7e9e14052727dc6e6ce69be42a6faa5f22e21f6f8199

Observation 6b3533b2-517e-4638-adca-81adca8f7420 · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Evaluating Object Hallucination in Large Vision-Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.466909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.466909Z digest=sha256:0a4a3fd0be5d7521a94819f68b8d74388659026971d6101c471c5bdf07cff450

Observation 0c1bf0fe-28d4-4ef2-a113-4ddc74b94663 · outbound

This paper cites Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.470098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.470098Z digest=sha256:603a1dae8ea2454929d72d0e0cb857712104f775e519000571fc6840aa9b3989

Observation 6a2627c0-c68d-4aff-87a2-6975aa66c9b1 · outbound

This paper cites A Survey of Hallucination in Large Foundation Models.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models A Survey of Hallucination in Large Foundation Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.473083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.473083Z digest=sha256:a274273630a4a4e0b7dc71f98a03014d828b0cb7580866f0ab3b4f19c6b4f4bb

Observation 4180b87b-fc99-4342-8b17-1fff9bf367cf · outbound

This paper cites HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.476129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.476129Z digest=sha256:57adef258a6dd4bd8d532466fdfe4b26efcfab2318bf819e3913d827c008c473

Observation 67f260e9-e7ee-46aa-a772-65389cf85913 · outbound

This paper cites Advancing Medical Imaging with Language Models: A Journey from N-grams to ChatGPT.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Advancing Medical Imaging with Language Models: A Journey from N-grams to ChatGPT

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.479419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.479419Z digest=sha256:1ebeffd8074bb6d712ba9eea42f30119e4ca925f844b7cac37e06ca8c9589376

Observation c0ece340-1017-479e-b169-fb1deadea48f · outbound

This paper cites ChatCAD: Interactive Computer-Aided Diagnosis on Medical Image using Large Language Models.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models ChatCAD: Interactive Computer-Aided Diagnosis on Medical Image using Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.483196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.483196Z digest=sha256:fa67f8747c95b4d404cb46d5426f0c751293016b0833385be3ae565e7db27b80

Observation e8257ff2-a84e-465c-81a6-ac3e9c0bef4c · outbound

This paper cites A survey on multimodal large language models for autonomous driving,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models A survey on multimodal large language models for autonomous driving,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:35.056723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:32:34.486759Z digest=sha256:391e091646f319ca52e52ee11f717d54a673c584ef08e7fea2f1616073971aab

Observation 348d23ca-efb2-4462-a8be-a161e57965cc · outbound

This paper cites Drivemlm: Aligning multi-modal large language models with behavioral planning states for autonomous driving,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Drivemlm: Aligning multi-modal large language models with behavioral planning states for autonomous driving,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.490330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.490330Z digest=sha256:ded5082be534510e2cdf18a7ea7682b1dc44d97b85639d353ec70aaf5d33aaaf

Observation e82b4d40-579c-4b8a-ac39-81300039953e · outbound

This paper cites Evaluating a large language model on searching for gui layouts,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Evaluating a large language model on searching for gui layouts,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:35.046906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:32:34.493938Z digest=sha256:36ff7136080929df953b8816984c0d5c0ecc93df286ac14145e8869369bc30c7

Observation d0c1b8c4-852a-40da-9bdd-f068354066ca · outbound

This paper cites DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.496883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.496883Z digest=sha256:f94e1d4e3d4853283407eb59b60ee77030765b73f26da1b3c6878b84578045be

Observation d5142ab2-76c8-4829-b805-da10cfb555fe · outbound

This paper cites In-Context Sharpness as Alerts: An Inner Representation Perspective for Hallucination Mitigation.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models In-Context Sharpness as Alerts: An Inner Representation Perspective for Hallucination Mitigation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.499975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.499975Z digest=sha256:8f33705a9afc164ce82b172c66bc202abef26c93be1408d9baad8c6c302100c4

Observation cabb3c46-a82a-4a36-9996-9b49f7605afe · outbound

This paper cites LLMs Know More Than They Show: On the Intrinsic Representation of LLM Hallucinations.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models LLMs Know More Than They Show: On the Intrinsic Representation of LLM Hallucinations

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.503372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.503372Z digest=sha256:69a2c38ed54d308348417a3aeae3d3273b745f256c51d34f8bd35cb46fcb61ed

Observation 7c6b1805-2852-4866-9e1b-31ee16c1f0bf · outbound

This paper cites Llama SLayer 8B: Shallow Layers Hold the Key to Knowledge Injection.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Llama SLayer 8B: Shallow Layers Hold the Key to Knowledge Injection

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.506797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.506797Z digest=sha256:49fa074cca4555e5bcca7ff3520a09b707c89b8b9b59675d8332baaa5bd67e27

Observation 4684c0d9-e17a-4473-94a5-90294e5ce1e2 · outbound

This paper cites Insights into LLM Long-Context Failures: When Transformers Know but Don't Tell.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Insights into LLM Long-Context Failures: When Transformers Know but Don't Tell

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.510625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.510625Z digest=sha256:47342be24dedd721e832305dea8f3b8909671b768ec7666fcf87219692a2079d

Observation 6ef8ed9b-9686-4ae0-af3d-8573f8159d15 · outbound

This paper cites Knowledge Mechanisms in Large Language Models: A Survey and Perspective.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Knowledge Mechanisms in Large Language Models: A Survey and Perspective

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.514275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.514275Z digest=sha256:0ab52c48aeba4d1f0019737ca871aae1b00f9bf03e16a846909c550a427d1472

Observation 3d1fcd93-d40d-4938-9cb3-0713876ef7b7 · outbound

This paper cites MLLM can see? Dynamic Correction Decoding for Hallucination Mitigation.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models MLLM can see? Dynamic Correction Decoding for Hallucination Mitigation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.517821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.517821Z digest=sha256:76157789086ac9bc8153f6928636af50f52e48f0a50d2bb1945198e8ed476060

Observation 08b89adc-3f7e-4935-9b94-7cb210beee41 · outbound

This paper cites Unleashing region understanding in intermediate layers for mllm-based referring expression generation,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Unleashing region understanding in intermediate layers for mllm-based referring expression generation,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:35.036052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:32:34.521189Z digest=sha256:08bfd28d699af83ad0366c9f6cc7711d0f8fbbad97a1961aa64fb0f80226771e

Observation d6b394d2-9faa-4b38-a9e2-9b4122b41ac9 · outbound

This paper cites Branchynet: Fast inference via early exiting from deep neural networks,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Branchynet: Fast inference via early exiting from deep neural networks,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:35.026046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:32:34.524262Z digest=sha256:b23626eaab9e3ab23967e5fa1f444579e27e43de0c25d1539265b9b2b8e87a9e

Observation 1c694b9b-4ce1-4710-9d8f-3a0225167cdd · outbound

This paper cites Depth-Adaptive Transformer.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Depth-Adaptive Transformer

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.526903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.526903Z digest=sha256:a3a9e4aed83f7a3adb69325d62806113cb8da789a928a26405a5083fdae83d16

Observation 9e416501-821c-4235-b285-89d855e11dae · outbound

This paper cites Confident adaptive language modeling,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Confident adaptive language modeling,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:35.016928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:32:34.529784Z digest=sha256:641a6966c47ac958da950771f14685bb1c4e90f86a4b5a03595ffab4cf36477c

Observation 122b16fc-ee3d-4561-8175-2e63efa27f97 · outbound

This paper cites Hallucination augmented contrastive learning for multimodal large language model,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Hallucination augmented contrastive learning for multimodal large language model,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:35.007082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:32:34.532510Z digest=sha256:363977a2b5173c85c3ca2bb473c401ae2aad8d00ecebdca075d7a7fce1232a3e

Observation 18cf72b5-94b0-4ae0-a7ec-10d30ecc1f81 · outbound

This paper cites Delve into Visual Contrastive Decoding for Hallucination Mitigation of Large Vision-Language Models.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Delve into Visual Contrastive Decoding for Hallucination Mitigation of Large Vision-Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.535229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.535229Z digest=sha256:9fd6387f3ce26b98a4f3f0e461655bad6bbe4ba247896844f679f6eb1da0a6c7

Observation 797cb5d4-3cbb-48dd-b67b-cfb374d74172 · outbound

This paper cites Mitigating object hallucina- tions in large vision-language models through visual contrastive decoding,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Mitigating object hallucina- tions in large vision-language models through visual contrastive decoding,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:34.995099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:32:34.538089Z digest=sha256:7b27a47af56c75e7846d323a74a6e88e808605d664b92d176c7fe1678d2c3e18

Observation d6d1e12e-8049-433f-a92f-0c05434cee76 · outbound

This paper cites Mirage in the Eyes: Hallucination Attack on Multi-modal Large Language Models with Only Attention Sink.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Mirage in the Eyes: Hallucination Attack on Multi-modal Large Language Models with Only Attention Sink

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.540866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.540866Z digest=sha256:829d6ec4d32c5fd3e71568e6b8b0a74c43040c62b18aa93f74bf537ed539dfb2

Observation 95d0c6dc-2568-4125-b628-eef990370a11 · outbound

This paper cites Hallucidoctor: Mitigating hallucinatory toxicity in visual instruction data,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Hallucidoctor: Mitigating hallucinatory toxicity in visual instruction data,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:34.984548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:32:34.544752Z digest=sha256:e0715bfa34366087950cb2a4ebe505021c0a4cd8ab67683ba66b2ae75e7c70fd

Observation 3fb021ba-c070-4165-94c6-e41693bede4e · outbound

This paper cites Look Twice Before You Answer: Memory-Space Visual Retracing for Hallucination Mitigation in Multimodal Large Language Models.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Look Twice Before You Answer: Memory-Space Visual Retracing for Hallucination Mitigation in Multimodal Large Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.547689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.547689Z digest=sha256:80e8c9090cc7ba89e5216e484badeac53c9cc60d95b8b02976049cbb6d451656

Observation 5831311d-7ac5-4f57-a000-47a7816de68c · outbound

This paper cites HELPD: Mitigating Hallucination of LVLMs by Hierarchical Feedback Learning with Vision-enhanced Penalty Decoding.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models HELPD: Mitigating Hallucination of LVLMs by Hierarchical Feedback Learning with Vision-enhanced Penalty Decoding

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.551576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.551576Z digest=sha256:ab32170ce6159aa0cea2ff7adc2b7dc39e93922e89f9a67c6798e1938ec0d7ad

Observation f4ff1d17-dd9f-4aeb-a2ab-d1ad4bcb5b74 · outbound

This paper cites Opera: Alleviating hallucination in multi-modal large language models via over-trust penalty and retrospection-allocation,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Opera: Alleviating hallucination in multi-modal large language models via over-trust penalty and retrospection-allocation,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:34.974822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:32:34.555323Z digest=sha256:30fd228d0079d18b7baf461919bca9c0c94c6b09f253eb807767ea445a418b9f

Observation 73cd3b5b-2e83-4bba-ab18-0a85af58df7e · outbound

This paper cites Instructblip: Towards general-purpose vision-language models with instruction tuning,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Instructblip: Towards general-purpose vision-language models with instruction tuning,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:34.965168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:32:34.558864Z digest=sha256:73c44c151177738b9a856f3aaf73630f18562632ee9592a69f4af6d3da8dab64

Observation 04aee0e3-03be-4a8f-95bc-bf62b28f4756 · outbound

This paper cites MiniGPT-4: Enhancing vision-language understanding with advanced large language models,.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models MiniGPT-4: Enhancing vision-language understanding with advanced large language models,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:32:34.954851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T15:32:34.561703Z digest=sha256:1132d60aadf1836a749e1657b1730181c8df3a210a8c4aa03b3e4c78cfe1e1cc

Observation 7bb93f49-e7d5-4f2c-946b-0155508fedfc · outbound

This paper cites Qwen Technical Report.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Qwen Technical Report

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.564854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.564854Z digest=sha256:c7e12bd0a95d4f921ab3e074503db22f6793d5c0db42ff8315c034704766a602

Observation 60c6ffac-10c6-440e-ac75-fe2fbdae59e4 · outbound

This paper cites MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.567976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.567976Z digest=sha256:fbc99e0c48d3362c2ab3514eaf15370ff171777e9a84dfa35e886f7ca0197cf3

Observation ec82dec8-964a-44bd-a05c-eef91f3066bf · outbound

This paper cites Object Hallucination in Image Captioning.

Extracting Visual Facts from Intermediate Layers for Mitigating Hallucinations in Multimodal Large Language Models Object Hallucination in Image Captioning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T15:32:34.571356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:32:34.571356Z digest=sha256:5288605afcaff971fc9f2231e75d5fcc4b0b4ea7c3454e26a3e1f495a749349a

Pith citing papers

No inbound Pith citation observations are available.