Pith. sign in

Paper Citation Record · LEDGER

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy

As of 11 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 1 inbound Pith citation observation for arXiv:2605.20965.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.20965 v1

Coverage vector

measured 78 of 78 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-21T05:03:24.950836Z

measured 79 of 79 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T04:32:13.571729Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

78 of 78 outbound references displayed

  • verified exact6
  • verified fuzzy65
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5addce12-de5d-4551-b6fd-3a5fb7e36ef2 · outbound

This paper cites 2025 , pages =.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy 2025 , pages =

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.318495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:bb7d14f95f5c65ee0a0c0b0fb4737faa82e97953ae97f80ea2f9195a03445907

Observation 50bad945-8ea6-4d02-887f-3e3c34d27c35 · outbound

This paper cites Advances in Neural Information Processing Systems , volume =.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Advances in Neural Information Processing Systems , volume =

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.398625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:14257af5bfd1098c3fbdac60513d88cc2d603201d46dbd06e0a58e967b556390

Observation 39d2979b-1e1d-4f66-903e-b7a90ba0f2e6 · outbound

This paper cites The Twelfth International Conference on Learning Representations , year =.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy The Twelfth International Conference on Learning Representations , year =

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.457717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:063c9da7f30e5ea1f14c5ec09c6494b61818b069759a19cc1c16d9135340a8c7

Observation 4cffa676-7f1c-4001-bb1c-7761a4dbf32c · outbound

This paper cites LLaVA-NeXT: Improved reasoning, OCR, and world knowledge , url=.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy LLaVA-NeXT: Improved reasoning, OCR, and world knowledge , url=

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.403057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:877324220b81166dd351fd5605a702de36d7a9f62ce2d0f048f6aa20b0ddf9ed

Observation 8d9df3d1-6fc3-4a3d-86b2-a613c44266b5 · outbound

This paper cites The Thirteenth International Conference on Learning Representations , year =.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy The Thirteenth International Conference on Learning Representations , year =

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.332235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:fd8e80686beceef94949756a03d455b1ef72e0b6b783c360468e01584c87ada2

Observation 91bfb8b9-9a69-4a26-ada7-207a33415392 · outbound

This paper cites The Twelfth International Conference on Learning Representations , year =.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy The Twelfth International Conference on Learning Representations , year =

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.334062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:8f7b3f4dfd4f146063de8bee76338d8b4c19e8c7a5b0b62a65da9d1c106bba84

Observation 2e3bf15f-ca86-4d88-a4c2-bb5e822fdfba · outbound

This paper cites Proceedings of the.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Proceedings of the

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.339443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:f61d119a01498a7f0709df274ec222b0ffb7509b3d056c4536494215a23655f1

Observation 935e1f9a-8d0d-40c9-ba8b-74159cd77a5d · outbound

This paper cites Advances in Neural Information Processing Systems , volume =.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Advances in Neural Information Processing Systems , volume =

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.345504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:8522504daf49f89afbb43de3354f12fc647d12b33bb6ff31bfbd3c68768caff6

Observation ed004046-de83-44bd-8694-ee70e2e8e32c · outbound

This paper cites Proceedings of the.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Proceedings of the

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.380981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:cc0c18bc3e023cbd7f0fce3da64f99e5d4fb3d1fe59d7e914544053231a5401d

Observation 8287beee-582c-4b5c-9ca9-61b5e622433e · outbound

This paper cites and Stepputtis, Simon and Morency, Louis-Philippe and Ramanan, Deva and Sycara, Katia and Xie, Yaqi , booktitle=.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy and Stepputtis, Simon and Morency, Louis-Philippe and Ramanan, Deva and Sycara, Katia and Xie, Yaqi , booktitle=

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.335677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:651d7331b1b4ae3025fe0644c744edde15ee05aac143e1ed02b65af0221d79c0

Observation 78a339ed-f7ec-4aed-8638-f92a3b06067b · outbound

This paper cites Proceedings of the.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Proceedings of the

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.322218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:19b8f7044e865031cdf7b44368149d29013879f9c89334a57429c4f3d5af89ff

Observation 504e346c-dade-463c-935e-148dc1f7469b · outbound

This paper cites Proceedings of the 42nd International Conference on Machine Learning , pages =.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Proceedings of the 42nd International Conference on Machine Learning , pages =

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.434620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:2c2a5e44e34b744bc8528ab4427d833fb5bde09a9225210bdd0b5f0912590d3e

Observation 1980b34f-26d1-4a36-9d6a-59643052af80 · outbound

This paper cites Cracking the Code of Hallucination in LVLMs with Vision-aware Head Divergence , booktitle =.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Cracking the Code of Hallucination in LVLMs with Vision-aware Head Divergence , booktitle =

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.418525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:95473f68b1c3a08b24829985f3d63b5f7cfa9e1beab7df22860f040c1b4e129c

Observation 66b26d5f-def2-4633-bf10-01d09e6063d6 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.395109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:794c2eb94f22c8bb5b662b2f37b0acf0cd46347bd6b8cc45c52c9b507a254fdb

Observation cb297a53-8c2d-453f-8378-d3aea84aa0c4 · outbound

This paper cites Steering LVLM s via Sparse Autoencoder for Hallucination Mitigation.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Steering LVLM s via Sparse Autoencoder for Hallucination Mitigation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.391041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:100e95749300adf1babd8b54718c862fde5ad63e907fee0ca653434f04607583

Observation ab399c63-0588-4300-8910-ca7ee4fd7b33 · outbound

This paper cites The Thirteenth International Conference on Learning Representations , year=.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy The Thirteenth International Conference on Learning Representations , year=

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.388972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:de7f98da9aa858c41e77b7b4d87e3de2a5592242ecd735c0513158f17ce5ef6e

Observation 5211e989-f6b6-42b7-a49b-befae56d8a76 · outbound

This paper cites Science China Information Sciences , volume =.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Science China Information Sciences , volume =

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.400649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:9540cb372189a4b1046e897035ad478a6910e25a93bf845ce05ad988034018f2

Observation e5797cd2-9d7e-4ffa-a075-05597709da91 · outbound

This paper cites Shallow Focus, Deep Fixes: Enhancing Shallow Layers Vision Attention Sinks to Alleviate Hallucination in LVLM s.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Shallow Focus, Deep Fixes: Enhancing Shallow Layers Vision Attention Sinks to Alleviate Hallucination in LVLM s

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.451785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:5a0e309a0c826917a09a6310faedfdc814b927b4262826a70b9a0a7bcee2ab56

Observation 72fbf803-3074-475f-9908-80cc27541cc0 · outbound

This paper cites Self-Introspective Decoding: Alleviating Hallucinations for Large Vision-Language Models , year =.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Self-Introspective Decoding: Alleviating Hallucinations for Large Vision-Language Models , year =

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.371506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:f39ce3c117c2ad48b3901c85f4e397484d227019704bb1efa899e0400a26507d

Observation ffaedc10-4729-4c95-8696-da771b018882 · outbound

This paper cites Seeing Far and Clearly: Mitigating Hallucinations in MLLMs with Attention Causal Decoding , pages=.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Seeing Far and Clearly: Mitigating Hallucinations in MLLMs with Attention Causal Decoding , pages=

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.373490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:a2fa903d58b8b6cbb621ffe94c021a15e15defe5f6ee7207472be48b8934b2c9

Observation 6dddc3dc-745e-436d-b0d0-78e81dc29974 · outbound

This paper cites VQAG uider: Guiding Multimodal Large Language Models to Answer Complex Video Questions.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy VQAG uider: Guiding Multimodal Large Language Models to Answer Complex Video Questions

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.396981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:aac9bcf383a52601303e26f5c63405e22c18e1ce0e8329ce259d7b237251082a

Observation 1391780a-ca82-4ed9-b8c5-6f2414782cd9 · outbound

This paper cites Grounding Multimodal Large Language Models to the World , year =.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Grounding Multimodal Large Language Models to the World , year =

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.442741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:874cb1177a58483bac69a994ff9b02443300a3a6c78ebb81f01272ef9979c8dc

Observation 53b95153-fce4-4fc0-8feb-4d1ea4970e49 · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.367467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:1f6525c58bf120853996292e72ba46d14488904d88a4696bd169cce9c5cb3ad1

Observation 790534c1-189e-431f-b124-2e97265eddd2 · outbound

This paper cites Object Hallucination in Image Captioning.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Object Hallucination in Image Captioning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.429067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:eb03245c1230d585f650ffe14bf717a4c6699538f1280519b2f5a76fdd27f853

Observation 0b4e52e8-0972-47b0-9d19-40025a99f77b · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Evaluating Object Hallucination in Large Vision-Language Models

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.426580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:9e29b8032abc7f010ff501b249764e8cb6863a0a9e394711ed3d4d6e01a0e554

Observation 6e538687-242c-4fc7-9f21-6ea76a28a7dd · outbound

This paper cites Aligning Large Multimodal Models with Factually Augmented RLHF.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Aligning Large Multimodal Models with Factually Augmented RLHF

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.365702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:09801bee6a57f331aca17dd699946ebcdff6d792495400b70aafe4d1de4560e6

Observation 81e10910-ea46-47d1-b38c-de0cca28abc8 · outbound

This paper cites Improved Baselines with Visual Instruction Tuning , pages=.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Improved Baselines with Visual Instruction Tuning , pages=

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.369661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:4348b3ea8a7692e1fc5bdc09eaa67d367f77a901aa71afdf34492b7e935c8d47

Observation b70591cb-b77f-46c8-9ab1-d9298cf6aa43 · outbound

This paper cites Lawrence.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Lawrence

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.377395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:21c62373b0791432254a913929dd20d270b964ba5f600b9ba2255106d0dea519

Observation 9397b85a-edb2-4fea-9b14-331e482f5fb7 · outbound

This paper cites 2022 , booktitle =.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy 2022 , booktitle =

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.337843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:ab2fad8a444a524673d743d3e79d5629b7bbc37401df4facbb49c73dbf0c7aed

Observation 4471b5f8-30fb-4bb3-b656-fd2bbdad13f7 · outbound

This paper cites and Manning, Christopher D.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy and Manning, Christopher D

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.359523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:fb6301e82339cd8e556975d8f83e78310a4b018b242721bc4b20ce3369ee6866

Observation bd4d8531-5ef8-43d2-ae14-176f7aafa71f · outbound

This paper cites Proceedings of the Thirty-Ninth.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Proceedings of the Thirty-Ninth

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.361171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:3cda75d3f7d15a66ff889e68040825968d594424cae481805749ecd9ffe27068

Observation 3092540f-5dfb-48bc-ad2f-acea838f65c4 · outbound

This paper cites DAMRO : Dive into the Attention Mechanism of LVLM to Reduce Object Hallucination.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy DAMRO : Dive into the Attention Mechanism of LVLM to Reduce Object Hallucination

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.363807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:014502ed827e740024b741132b7236ec91141a97f871a7b19d60a46ce0dc721f

Observation c305203f-7bbf-4694-98ca-c8e000649096 · outbound

This paper cites Proceedings of the 42nd International Conference on Machine Learning , volume =.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Proceedings of the 42nd International Conference on Machine Learning , volume =

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.379271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:97d2086b8ecb110a4136624e4986f95209d5610002d70aaedc9ca19422daa82d

Observation e3ae1760-548c-4bfc-9cf1-3d6836fea397 · outbound

This paper cites I mage I n W ords: Unlocking Hyper-Detailed Image Descriptions.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy I mage I n W ords: Unlocking Hyper-Detailed Image Descriptions

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.440783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:9a96ccf36d9dfd4c603bd5d8923dd2a76f926cbc89057b57ede5c50baf4ad949

Observation faad4691-b282-4b5a-8146-cb5f2cd13068 · outbound

This paper cites and Petryk, Suzanne and Gonzalez, Joseph E.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy and Petryk, Suzanne and Gonzalez, Joseph E

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.383378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:6a64d55296a51bc1139fb052671bc9df63fe7a37e1e36c041c05c02749766f89

Observation e903a601-5fca-40c4-8797-0e5b086497e3 · outbound

This paper cites and Kachinthaya, Anish and Zou, Haodi and Canny, John and Gonzalez, Joseph E.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy and Kachinthaya, Anish and Zou, Haodi and Canny, John and Gonzalez, Joseph E

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.341665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:50245c6db439475f7c311b86239f0efe946fe83ff78e4b0a0e14c083c0eeb3b8

Observation d2de85e2-1b1a-4a2d-b1d1-9b204fbcd16f · outbound

This paper cites Proceedings of the Fortieth.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Proceedings of the Fortieth

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.353491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:1bbdbc3ecb117b4fc1e2170da3c7b212e624b3421934888378c7647a19a8a693

Observation cdda3410-c3f5-4d45-b264-f44f57783833 · outbound

This paper cites Mitigating object hallucinations in large vision-language models with assembly of global and local attention.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Mitigating object hallucinations in large vision-language models with assembly of global and local attention

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.355134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:02cce594aa6a3120be88679aa06a76189e3789d229961d398fe664ecc633352a

Observation 89fb5d30-e96e-41f6-a8d6-113fa2008dfb · outbound

This paper cites Qwen3-VL Technical Report.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Qwen3-VL Technical Report

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:03:57.748585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:43c2d57abca14c44ba30302686cd5c9f62dc72162c2cab5fe859f4b6b1e27a5b

Observation b0cc244c-5a35-4b54-b35e-a4de574eeab6 · outbound

This paper cites M., Petryk, S., Gonzalez, J.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy M., Petryk, S., Gonzalez, J

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.349350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:425e59ca88a04658986f84548ac3defee51048d5fad75dc3d392acb61b502761

Observation b2157b1c-65b9-4d55-9171-ecf14829c648 · outbound

This paper cites VQAG uider: Guiding multimodal large language models to answer complex video questions.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy VQAG uider: Guiding multimodal large language models to answer complex video questions

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.351014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:bfc19cdac72c831ac6295cd5f0235aa38b2aa7a6c8df071e5ad13219f1be0972

Observation c398ac7f-a14b-4893-b0fc-dea7ba50ad6e · outbound

This paper cites Alphaedit: Null-space constrained knowledge editing for language models.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Alphaedit: Null-space constrained knowledge editing for language models

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.347169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:1b569d442e13e27c9b6b007697c867bb6210723c70c29b6434134dda356004fd

Observation 4126e6e4-216b-4dff-8482-0cfaa9c1e668 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:03:57.738260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:38410fca540457486538f20b31c3af410562931e5596d05188459d0df1559151

Observation 70f5342e-0676-404a-a140-4610b69f09dc · outbound

This paper cites M., and Soricut, R.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy M., and Soricut, R

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.357503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:2d38ea865b2f4a4065a1b7894b14beec63f0c9fd0ebf9d3f12b6c325388cbb0a

Observation 28f55735-5886-422e-b163-3ce57e23dba3 · outbound

This paper cites DAMRO : Dive into the attention mechanism of LVLM to reduce object hallucination.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy DAMRO : Dive into the attention mechanism of LVLM to reduce object hallucination

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.385178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:9ffe42386fb384f10d14be2b7809157e770e62f1138bed080891468e809bb081

Observation e1f86604-f1f5-4c7e-bc5f-d6422986d9d5 · outbound

This paper cites Cracking the code of hallucination in lvlms with vision-aware head divergence.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Cracking the code of hallucination in lvlms with vision-aware head divergence

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.343307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:075bba46f1d3af9144d721656a24dc8d624d8d7077db7090325209808d75235f

Observation d23ccd14-82a8-45b2-be05-8a760429e054 · outbound

This paper cites Steering LVLM s via sparse autoencoder for hallucination mitigation.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Steering LVLM s via sparse autoencoder for hallucination mitigation

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.387306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:d931acaca3600d8a535eea3f1d36b7e36b5445756010f080adbdaf1e0d6010ea

Observation a95b11c6-4557-4ec7-a21f-4e5805f217c1 · outbound

This paper cites Medical mllm is vulnerable: cross-modality jailbreak and mismatched attacks on medical multimodal large language models.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Medical mllm is vulnerable: cross-modality jailbreak and mismatched attacks on medical multimodal large language models

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.324582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:c9b062ca95dae8899d5f8f9837eaac1f526083933bb0f8b68f738f38c3c32cea

Observation e6a8e62e-319c-4340-aa9e-363ae9b81953 · outbound

This paper cites Self-introspective decoding: Alleviating hallucinations for large vision-language models.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Self-introspective decoding: Alleviating hallucinations for large vision-language models

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.328425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:729d1f18333b402f59cea5c2ff19db98e61d337d3effc882686d8fb88de78b9d

Observation d33e5ccc-2979-4a59-a55b-20b74ef2d5e7 · outbound

This paper cites Visual attention never fades: Selective progressive attention recalibration for detailed image captioning in multimodal large language models.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Visual attention never fades: Selective progressive attention recalibration for detailed image captioning in multimodal large language models

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.326282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:d146ffc0eb901ca2d36e3537728dfa778bd34bc68a6d051738fd6aa68c957fa7

Observation 21db0c55-76dc-4b0e-b73d-406381d4ca73 · outbound

This paper cites an unresolved cited work.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-05-21T05:03:58.329914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:d6b14174f34aed9f5fc9199261b9602f9eeff92660b81d5e677b5c8a860c83e5

Observation ccd35b3d-dd12-47bf-958a-3aa4ca362f3d · outbound

This paper cites an unresolved cited work.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-05-21T05:03:58.316984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:95ca0e5d4e07d4e9f1fd5f0d321e33123765a5f690adcf8c4a760e802151a272

Observation d9c06841-2fa7-4a9f-a9fa-e3a3ae41fa2f · outbound

This paper cites Toward robust hyper-detailed image captioning: A multiagent approach and dual evaluation metrics for factuality and coverage.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Toward robust hyper-detailed image captioning: A multiagent approach and dual evaluation metrics for factuality and coverage

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.453880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:7cebff3593477cbcde5c6ff486a6a4b02fd09eaec869cb4a887f9c63e024e370

Observation 65f76eba-441a-43be-9a2a-dee7dbfb3447 · outbound

This paper cites Mitigating object hallucinations in large vision-language models through visual contrastive decoding.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Mitigating object hallucinations in large vision-language models through visual contrastive decoding

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.448049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:5ce967986afd293c4ae1305b88a02ffd9a1e15b46e7dbd2d41843d70e67d3696

Observation 5611ff2c-db40-4053-9037-4c2a449532fa · outbound

This paper cites Evaluating object hallucination in large vision-language models.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Evaluating object hallucination in large vision-language models

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.446225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:da01b2c4af2906d350f8542538fae0f74eb02969b6095d5eab63477db401e39b

Observation 514cee3d-1ec6-41d6-8a00-4635478dd8ae · outbound

This paper cites an unresolved cited work.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-05-21T05:03:58.444294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:4fe448ce0f1c819ca186bf4b01d69864bfccd819627ef3c3e3f5bbc1e026fcde

Observation 2232bbb1-d5bc-45de-9302-1baac210546b · outbound

This paper cites an unresolved cited work.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-05-21T05:03:58.320541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:942fc7da7cf67fbd275d1b0ec6292a48f306ba781e3cf80eebe79218411fac02

Observation b5ec508f-dde1-4cac-88a1-aa4d80316f42 · outbound

This paper cites an unresolved cited work.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-05-21T05:03:58.392435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:223f3174d066b89dcf0c0e9c7232e74c0776e24efef85c578f59bc62c1c134f8

Observation f3f30ad8-af11-4550-b6ce-c53cbbb6a042 · outbound

This paper cites an unresolved cited work.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-05-21T05:03:58.450098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:3f20e9e2ad3ba4e4e3101fddc015dc33bc7bd299d773b604ccee3116b4405c1a

Observation a119bb20-2032-438a-90eb-c3df165bebe5 · outbound

This paper cites A Survey on Hallucination in Large Vision-Language Models.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy A Survey on Hallucination in Large Vision-Language Models

Reference 67

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:03:57.742123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:79985e999325db8b7a09a627defb66e7121269655d7085f8bc56f619d4ebf8a1

Observation a4f459f3-71a8-4f03-8744-af9db3515c86 · outbound

This paper cites Grounding multimodal large language models to the world.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Grounding multimodal large language models to the world

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.438467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:3a1e91962911e1d8a8627ee4272b56fb85faef836c65723f492f80529844eb05

Observation f493d763-1250-49de-b949-7206cd367d85 · outbound

This paper cites M., Kachinthaya, A., Zou, H., Canny, J., Gonzalez, J.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy M., Kachinthaya, A., Zou, H., Canny, J., Gonzalez, J

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.432476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:5f35c39e04abd1f50326af7dba47f336c0c075b12ed8e9230b2f3b0f3ee85ab9

Observation 4cd863a7-e8d8-4514-a740-1977e656f6b8 · outbound

This paper cites A., Burns, K., Darrell, T., and Saenko, K.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy A., Burns, K., Darrell, T., and Saenko, K

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.430537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:4704ca633ba9664060ebb8f2639f2d7282cea8834064ae5d8f1df6e0e4486ffb

Observation d32622f5-ce42-4a56-a452-ac5d5edcf060 · outbound

This paper cites Aligning large multimodal models with factually augmented RLHF.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Aligning large multimodal models with factually augmented RLHF

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.436788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:e8dce18bd4d44f74552845978b9ab7703f5ae25f0d68640ec0ed50e533b7b854

Observation fcaa1e07-ac44-4d02-96ed-b11f23810162 · outbound

This paper cites Seeing far and clearly: Mitigating hallucinations in mllms with attention causal decoding.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Seeing far and clearly: Mitigating hallucinations in mllms with attention causal decoding

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.421893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:7e91cca89c3fe2e48b8ae9b55cf9fa41e451ff6a52007b2770e96f46141bfb27

Observation 95f92f49-2ca6-4105-808f-b2684e93a568 · outbound

This paper cites Q., Stepputtis, S., Morency, L.-P., Ramanan, D., Sycara, K., and Xie, Y.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Q., Stepputtis, S., Morency, L.-P., Ramanan, D., Sycara, K., and Xie, Y

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.425527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:1e5fcd21f17a53ab5b7dbdfe801052337dfa19af2a3c8fd8cc20883ed747ea9b

Observation 92aa281a-8177-48dc-a630-c45c51970597 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 74

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:03:57.751902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:5b6d92e1f10d0a5c4b8cac01cc7f0da8ea6ec7742389dc63cc3710c34b8bb3f8

Observation 16014604-876e-4b5f-80b8-b8a1dd9c0bd2 · outbound

This paper cites Caption Anything: Interactive Image Description with Diverse Multimodal Controls.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Caption Anything: Interactive Image Description with Diverse Multimodal Controls

Reference 75

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T05:03:57.755094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:2e0013db385b6cf7d0a2329c95ebf6d092e40456c4972fca25d9202bc583b907

Observation d92bad9c-2be4-4b9a-9b6f-5ec1ec8371b9 · outbound

This paper cites ESMC: mllm-based embedding selection for explainable multiple clustering.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy ESMC: mllm-based embedding selection for explainable multiple clustering

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.414695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:f0a989d7e6ea4aed7b3f0e6098f5e07885ae401e7cf0baba6683464c7ae54313

Observation 1f187497-832e-40ef-a976-584a7f91a835 · outbound

This paper cites Clearsight: Visual signal enhancement for object hallucination mitigation in multimodal large language models.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Clearsight: Visual signal enhancement for object hallucination mitigation in multimodal large language models

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.408584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:cb61c3093af27657df66fdc19c90b28c67361f6a8b5296669e5af58336546079

Observation 402abc82-d652-49db-9f8d-9969e5da4902 · outbound

This paper cites Woodpecker: Hallucination correction for multimodal large language models.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Woodpecker: Hallucination correction for multimodal large language models

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.406837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:885f0115528a95378f68bc7e771b38559cfd37bf2d86ecb5b8b8483e7eb555c5

Observation 2492f909-13b7-4f33-801f-a7cc989d3a0f · outbound

This paper cites Rlhf-v: Towards trustworthy mllms via behavior alignment from fine-grained correctional human feedback.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Rlhf-v: Towards trustworthy mllms via behavior alignment from fine-grained correctional human feedback

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.410824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:a0160a5886888d343d006fd1d6f8bda2fc9208d9dcb822dfaf34197d7283d116

Observation fae9d251-beab-4c99-b0e6-3c941a492d8b · outbound

This paper cites Tell your model where to attend: Post-hoc attention steering for llms.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Tell your model where to attend: Post-hoc attention steering for llms

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.412509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:2924159780ad773da2836b9b6ce6eaa5c007457e4541e79aa42bd83a92a0e9db

Observation 0834bdec-3428-4757-9062-a80d52918982 · outbound

This paper cites Vldrive: Vision-augmented lightweight mllms for efficient language-grounded autonomous driving.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Vldrive: Vision-augmented lightweight mllms for efficient language-grounded autonomous driving

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.404765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:8bdbede97c0fefbc06a9da6460143f3ec5179d9a347df447436172114a427bfc

Observation e11059bf-75a1-4f27-815d-1a16835d44ed · outbound

This paper cites Shallow focus, deep fixes: Enhancing shallow layers vision attention sinks to alleviate hallucination in LVLM s.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Shallow focus, deep fixes: Enhancing shallow layers vision attention sinks to alleviate hallucination in LVLM s

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.416394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:1f923b21283ab30e0c0d441d1b47ed5af363314002c90078367b3b149cc62183

Observation 6cb9b9ad-0d4b-4048-ae98-d22806a91254 · outbound

This paper cites Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization

Reference 83

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:03:57.745343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:4bf6db9f43ebf5f14ca6aaeedc27a5ffca39d8e4eb5de3b5b6420f9ea703efb9

Observation 7236b3c1-60a3-4aa0-9d12-d3ca1e1cd61e · outbound

This paper cites Minigpt-4: Enhancing vision-language understanding with advanced large language models.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy Minigpt-4: Enhancing vision-language understanding with advanced large language models

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T05:03:58.455987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:29fa23294efe0da579e2e810151b7a5e491d25f12d8b2ac110a1397a781c8e05

Observation ccc84fe3-98ea-44e5-ba62-fdd1d7bdd8e5 · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 85

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:03:57.759989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-21T05:03:24.950836Z digest=sha256:a087e7452a943aee70fcc9a49c6f4be5696cc91b7d96b54178fcf2b43e372dc2

Pith citing papers

Observation 3f711ddb-7c26-45cb-82b4-96b8566223d3 · inbound

Failing to See or Failing to Know? Attributing Errors in Vision-Language Models cites this paper.

Failing to See or Failing to Know? Attributing Errors in Vision-Language Models Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-04T04:32:13.571729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:32:13.571729Z digest=sha256:3126415807683ccf4eb02bcfcf913f804e9f85a297f9f6c45277442bffc5e0a7