Pith. sign in

Paper Citation Record · LEDGER

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

As of 13 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 26 inbound Pith citation observations for arXiv:2411.09968.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.09968 v1

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T20:12:25.404728Z

measured 72 of 72 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 26 of 26 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T16:45:44.844437Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T19:20:06.531800Z

Reference resolution

46 of 46 outbound references displayed

  • verified exact0
  • verified fuzzy11
  • unresolved34
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bdf3ce59-f4fb-443c-b551-e1cdc020408b · outbound

This paper cites GPT-4 Technical Report.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.203681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.203681Z digest=sha256:7e773413feab1b85ee1d7bd2e0c0bc1261f2236500f9e53f324288c644f7414f

Observation 7edc2242-ccb9-4131-bdc0-7323c428ea06 · outbound

This paper cites Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.208627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.208627Z digest=sha256:8c0b9a3df75700a70b2270763a63349473c91341c5e8a6bcd37e45918a74f588

Observation b6ec5f85-d198-4d90-9c18-217dfa8d8ad4 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.214251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.214251Z digest=sha256:75f5bf67a7d866fe8b434fd915203bff56c23d4996c78e4b35325fa4d8f1249f

Observation fdebd2eb-b4a3-4829-94c9-6d9ea2d514b6 · outbound

This paper cites Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.218918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.218918Z digest=sha256:33eaa79bb57131fdbb65433649c8e8799820f524bebe6544e97ef09cd9da1860

Observation 17949207-2723-4a66-818d-56cc299b48a8 · outbound

This paper cites An image is worth 1/2 tokens after layer 2: Plug-and-play inference acceleration for large vision-language models.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs An image is worth 1/2 tokens after layer 2: Plug-and-play inference acceleration for large vision-language models

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:12:26.042804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T20:12:25.223467Z digest=sha256:db38f4cf189bf83bb3c07fe09ecc2dae8525042bb6f165e2f6525716b3b76d26

Observation e08f02a7-5035-45cd-af36-091819707549 · outbound

This paper cites Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Internvl: Scaling up vision foundation mod- els and aligning for generic visual-linguistic tasks

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:12:26.026650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T20:12:25.227708Z digest=sha256:ec637fd2cf7f9e864b0d6548e1718be72ccefb6563a30f82eefb75a9282761d2

Observation 50dc7d26-2604-4912-b175-5a42100c1be3 · outbound

This paper cites HALC: Object Hallucination Reduction via Adaptive Focal-Contrast Decoding.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs HALC: Object Hallucination Reduction via Adaptive Focal-Contrast Decoding

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.232011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.232011Z digest=sha256:477c38a25d48ea0bf8b3e5ffce7b819ba1247c9fad844fd345fb0e3c5404f466

Observation 90dd223b-9406-44c2-91c4-9c6ec5305db1 · outbound

This paper cites DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.236127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.236127Z digest=sha256:c98399d1d0796901fd2a29245be20f836f9ffc9fcdb566c4f92ef8eb602c7ae9

Observation 28930993-de74-4e22-ac99-749a59bcc78e · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Training Verifiers to Solve Math Word Problems

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.241066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.241066Z digest=sha256:f0c913f1a405cbd9b14e65e480790bd4fde47bbacb1c225362c0c8c1ff6ab535

Observation 0fee5eec-0fbb-46b3-969d-889ba0becc2b · outbound

This paper cites Instructblip: Towards general-purpose vision-language models with instruction tuning.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Instructblip: Towards general-purpose vision-language models with instruction tuning

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:12:26.012141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T20:12:25.245069Z digest=sha256:23fd38a6daf31e381975ea5d2e816be1d36568bcbc49d8c55bf23134458e575a

Observation e4944e9e-bd67-4a88-92c1-f0a5547785b1 · outbound

This paper cites The Llama 3 Herd of Models.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs The Llama 3 Herd of Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.248563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.248563Z digest=sha256:bb03ca272225c465dcecab548e22e02a2409d83c9d400698ffbd9267d3ec27ea

Observation 28a73a8d-5055-4bab-a60d-3fca6d663680 · outbound

This paper cites BLINK: Multimodal Large Language Models Can See but Not Perceive.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs BLINK: Multimodal Large Language Models Can See but Not Perceive

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.253133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.253133Z digest=sha256:56a6907c6bf64dc0c7bda902dffc0d3ed4d33f9fa8cc2a699db5d414b9894c26

Observation 67f5bda6-596a-4d9a-a239-64c14c72b8df · outbound

This paper cites MultiModal-GPT: A Vision and Language Model for Dialogue with Humans.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs MultiModal-GPT: A Vision and Language Model for Dialogue with Humans

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.257623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.257623Z digest=sha256:354fb297e0b0b7af6400e473cf721692e453576b125322adb823add846d9f504

Observation 249facda-1336-49e5-8d2b-9fedfb069ed3 · outbound

This paper cites Making the v in vqa matter: Elevating the role of image understanding in visual question answer- ing.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Making the v in vqa matter: Elevating the role of image understanding in visual question answer- ing

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.263326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.263326Z digest=sha256:da37abbca9ac42c369e9d7d45fd16bfda1d0b6eac71fab74183ccfa698d99664

Observation 08f94172-720d-4d48-b4e7-c60aa6411e96 · outbound

This paper cites HallusionBench: An Advanced Diagnostic Suite for Entangled Language Hallucination and Visual Illusion in Large Vision-Language Models.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs HallusionBench: An Advanced Diagnostic Suite for Entangled Language Hallucination and Visual Illusion in Large Vision-Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.268178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.268178Z digest=sha256:0c853c7df28446fa15023ee2c2da67c8a25a0809bb918df7e6557ab0affc0b57

Observation 04fd8043-e9b7-43e6-8cd3-5d6d94a60980 · outbound

This paper cites Vizwiz grand challenge: Answering visual questions from blind people.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Vizwiz grand challenge: Answering visual questions from blind people

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.273662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.273662Z digest=sha256:168f68c807e07ca47f01ad292b59679e94c9b8a9780b2524b18d8bdf6df60ea9

Observation 1dd6f320-4415-40c5-ae84-4ffd40b13db7 · outbound

This paper cites Opera: Alleviating hallucination in multi-modal large language models via over-trust penalty and retrospection- allocation.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Opera: Alleviating hallucination in multi-modal large language models via over-trust penalty and retrospection- allocation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:12:25.980565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T20:12:25.278308Z digest=sha256:4cdf6c21ae48a22402feab8776bf1a645e18ca4322aabe5c084fdb1425e8a576

Observation 71b4606c-a735-4a59-bdc2-9cb188612e7e · outbound

This paper cites Gqa: A new dataset for real-world visual reasoning and compositional question answering.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Gqa: A new dataset for real-world visual reasoning and compositional question answering

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.282846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.282846Z digest=sha256:b5cfed7804b45df8f0f675688b444eba638c25a59e0823ef8959ef2f207cd458

Observation 0b3c6c01-cae0-42e2-a720-6c6dbbdb855a · outbound

This paper cites Mistral 7B.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Mistral 7B

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.287281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.287281Z digest=sha256:817aaead05b4532c7aed0ff6554718f13d0defb9b4a31acf455425f8c5d4f63e

Observation ba3f93f4-6b9b-4d6a-ae35-89e96d5df01a · outbound

This paper cites Mitigating object hal- lucinations in large vision-language models through visual contrastive decoding.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Mitigating object hal- lucinations in large vision-language models through visual contrastive decoding

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:12:25.958089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T20:12:25.292043Z digest=sha256:8bdd71a40805021cb5dcfec57031c018876ca75c53d2d3c182a8bd58263b114a

Observation 12e930f2-abdb-4369-a725-0fdd98b63cdb · outbound

This paper cites SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.296636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.296636Z digest=sha256:8669fcc19321874b6cede01dc34b27811bcb28c491cc593386c96ae5e93cc2d9

Observation 2fac54ec-00b3-4245-bab3-54e5af2ebd1b · outbound

This paper cites Fine-tuning multimodal llms to follow zero-shot demonstrative instructions.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Fine-tuning multimodal llms to follow zero-shot demonstrative instructions

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:12:25.944435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T20:12:25.300590Z digest=sha256:615db3a6ab1d534a84c10d0732064c5ba953efedf4e348a822e6cfb5d95cd9c3

Observation 4b66e17d-6d70-4075-8d5b-d02e123e53d6 · outbound

This paper cites Inference-time intervention: Elicit- ing truthful answers from a language model.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Inference-time intervention: Elicit- ing truthful answers from a language model

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:12:25.930773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T20:12:25.304976Z digest=sha256:23345b73c0d4d455ba1bb6877b9f886eed8d7be575e4608be4b6a44bb271888e

Observation 41c52b43-b604-4860-be98-45ad83773317 · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Evaluating Object Hallucination in Large Vision-Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.309482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.309482Z digest=sha256:88f4bdf50663407fa11bbcc9c26075eebb0e39dd7a6a67d47782d7c26c691a4d

Observation a3a96c06-e4e9-428f-82e6-e6caed302a18 · outbound

This paper cites LLaMA-VID: An Image is Worth 2 Tokens in Large Language Models.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs LLaMA-VID: An Image is Worth 2 Tokens in Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.313654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.313654Z digest=sha256:508cd57535aff6ace2e77066699dafe48eaa7a499615537656f6abb8a48511de

Observation faf51b02-6611-4def-a105-ed2e629d979c · outbound

This paper cites Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.317748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.317748Z digest=sha256:863d5f0d5c14e42be0ae9379cfb3b5cfd8984bbaa5c8922b1d7bd4ad1373feb9

Observation 101b9bf6-7512-4ad7-83b4-f8a685d48526 · outbound

This paper cites TruthfulQA: Measuring How Models Mimic Human Falsehoods.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs TruthfulQA: Measuring How Models Mimic Human Falsehoods

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.323100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.323100Z digest=sha256:50ed39afcc91a25c6772c0ec7673bc7d325f1b4ce7c87fd1278368ee6353769a

Observation 3627312f-4097-4036-88b1-15e0e8897df0 · outbound

This paper cites Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.327970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.327970Z digest=sha256:e39db7043335915ee11b947c931d8a8e90b17b4989b26c01c51b1fe2f3b78ddf

Observation 0cfc0dfa-21d9-4b2e-8541-baf5e2644961 · outbound

This paper cites Visual instruction tuning.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Visual instruction tuning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:12:25.916731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T20:12:25.332405Z digest=sha256:1f9b5b0466e7a737cafc7ffa50ae9d0378112f96c2ea14627b18071bf2df86f7

Observation 6090bd95-5c06-4e1a-abdd-f86a34d4c252 · outbound

This paper cites Mitigating Dialogue Hallucination for Large Vision Language Models via Adversarial Instruction Tuning.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Mitigating Dialogue Hallucination for Large Vision Language Models via Adversarial Instruction Tuning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.335894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.335894Z digest=sha256:1bf329eb1f340a5461aeb7ce71ceb00af6e00eeb81697d40fe410e6980357ea1

Observation 60a55ec2-3aac-4279-a420-b4cac8b0513b · outbound

This paper cites Object Hallucination in Image Captioning.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Object Hallucination in Image Captioning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.339770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.339770Z digest=sha256:86cb727471265f3dfe4d05f7aa009db8d29556b9510b2bef1cf27a3e9af6e111

Observation 512cb94c-b825-4d7c-9595-460fdb5b935b · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Gemini: A Family of Highly Capable Multimodal Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.343655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.343655Z digest=sha256:49745bd800596eedc5cf2528e023bf438a132463d4cfd3161d694399c18fe3dc

Observation 96e4df4f-10ad-412a-940b-42dded73a0ad · outbound

This paper cites Qwen2.5: A party of foundation models, 2024.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Qwen2.5: A party of foundation models, 2024

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:12:25.903605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T20:12:25.347772Z digest=sha256:4a3f642ff384e7759b5e52c8bd405ced68fcfd6df168250a4bd32da7b71618e2

Observation 37a4881e-785b-4bf9-b3e8-5a552002ccb0 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs LLaMA: Open and Efficient Foundation Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.352607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.352607Z digest=sha256:7b8d220afaaccf2cb1c02e2009ccd7513053a521b0542b0aad2fdcccb93ca7f8

Observation 2d09e604-e5bb-45b0-bb55-56cd6ae902f4 · outbound

This paper cites Label Words are Anchors: An Information Flow Perspective for Understanding In-Context Learning.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Label Words are Anchors: An Information Flow Perspective for Understanding In-Context Learning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.357101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.357101Z digest=sha256:54e8ee29a1cae1fec9cb00adafb66d35ca8ddb1feed048d0b7d9f23fcb23b466

Observation 6d34a6f8-646b-4d3b-961e-020806eddbfb · outbound

This paper cites Dopra: Decoding over-accumulation penalization and re-allocation in specific weighting layer.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Dopra: Decoding over-accumulation penalization and re-allocation in specific weighting layer

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:12:25.891262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T20:12:25.361722Z digest=sha256:5d682e954e9c3d0d4f24812fe03ab8bc8e0aa28d1667ec735e3908d89cc03b05

Observation ad6b1db3-5500-43e3-a82f-2eb847c59763 · outbound

This paper cites Efficient Streaming Language Models with Attention Sinks.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Efficient Streaming Language Models with Attention Sinks

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.365865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.365865Z digest=sha256:cec2d73a701536de78c89c27f513a224af18b4de053f3ddbdbcd04bb47a6c9e8

Observation bef5eb58-35af-4d93-bf56-ec2992193af6 · outbound

This paper cites Mitigating Object Hallucination via Concentric Causal Attention.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Mitigating Object Hallucination via Concentric Causal Attention

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.370246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.370246Z digest=sha256:a2c11744fe3da6b879d5396ebf020f38b1c957451ee15471d813be3ddaaa6d0a

Observation 50911a35-d0c8-4b9a-b77b-c795299b24d3 · outbound

This paper cites Qwen2 Technical Report.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Qwen2 Technical Report

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.374701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.374701Z digest=sha256:a2f5366d6c18d8d6a114b3420d9daf17a70fb785e7b5cc8a9c84b9b44293e0d7

Observation 20c63d24-b751-40d1-8413-7eb3c0896b70 · outbound

This paper cites A Survey on Multimodal Large Language Models.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs A Survey on Multimodal Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.378642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.378642Z digest=sha256:02549a7383632477f41c96feb725f37e15c074c3d9c97aa4bcc493f4371058e5

Observation e2ae72f2-a74e-4f25-b16c-9c8de09397e2 · outbound

This paper cites MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.382754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.382754Z digest=sha256:5f6d640baffaaf68816d7632a6e525b31002ac2f97feb6ad4473c5fc92414785

Observation 8dae5ba5-8f92-49ef-b141-71c705dc3706 · outbound

This paper cites Unveiling and Harnessing Hidden Attention Sinks: Enhancing Large Language Models without Training through Attention Calibration.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Unveiling and Harnessing Hidden Attention Sinks: Enhancing Large Language Models without Training through Attention Calibration

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.386894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.386894Z digest=sha256:cf40cf48ab70433422ecaf28b6339f62f6196a55cf2e328eca0ef4152a6acbf4

Observation 986840d2-860e-4ea8-aadf-d8a3198c3ec6 · outbound

This paper cites Less is more: Mit- igating multimodal hallucination from an eos decision per- spective.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Less is more: Mit- igating multimodal hallucination from an eos decision per- spective

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T20:12:25.877905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-12T20:12:25.392161Z digest=sha256:0e6f12e7e30551f973fbf65386097deafd3c561606418aa192cf49197b4769ff

Observation a42786b8-adcd-447c-9b03-8663bfeb7d2c · outbound

This paper cites Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.396200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.396200Z digest=sha256:40be806899ce8394795308bb52df3662594afc62a962148a87e3141635f4e504

Observation 0e289667-89a8-4c1a-bbaa-5a491e60b570 · outbound

This paper cites From Redundancy to Relevance: Information Flow in LVLMs Across Reasoning Tasks.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs From Redundancy to Relevance: Information Flow in LVLMs Across Reasoning Tasks

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T20:12:25.400685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.400685Z digest=sha256:c62b576347d398c6bf3f1c00aae4aa904b08c9a678a6191c900a785af904ac72

Observation ac1c0026-83ec-4945-988b-9196c20bd4c7 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 46

Resolution
malformed identifier
no resolver link, observed 2026-08-12T20:12:25.404728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:12:25.404728Z digest=sha256:aac1f54e1b97239b0e8dc84c328c66e9eb5964a6e722dd3c23cec3df9c502a64

Pith citing papers

Observation 83c98cf1-a1eb-4065-94db-c6cafa868ab6 · inbound

Hallucination of Multimodal Large Language Models: A Survey cites this paper.

Hallucination of Multimodal Large Language Models: A Survey Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 213

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:33:33.392872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-11T12:33:32.631346Z digest=sha256:864c80b193076b1312c0fbcdc0059ade0fbf3e9dfc2348b2f33c4c7549afd920

Observation a995798a-d4f3-4859-90c3-06a9585b31c1 · inbound

Enhancing Multimodal Large Language Models Complex Reason via Similarity Computation cites this paper.

Enhancing Multimodal Large Language Models Complex Reason via Similarity Computation Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T16:45:44.844437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T16:45:44.844437Z digest=sha256:ba6bef6bacbf30ef222bb7abd958bca898cb097eb5debdfab2cc7181cf6877e4

Observation 28872082-b61b-4bc2-a25c-581e6cef690d · inbound

Cracking the Code of Hallucination in LVLMs with Vision-aware Head Divergence cites this paper.

Cracking the Code of Hallucination in LVLMs with Vision-aware Head Divergence Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T12:43:20.946064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T12:43:20.946064Z digest=sha256:6fef52272d6bfd177d87651ba63c014d7b6d6a9e8d813e56ba372b8a40140a88

Observation bc7c7a60-b5cd-455f-9cbd-42add66843ff · inbound

Toward Modality Gap: Vision Prototype Learning for Weakly-supervised Semantic Segmentation with CLIP cites this paper.

Toward Modality Gap: Vision Prototype Learning for Weakly-supervised Semantic Segmentation with CLIP Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-11T00:13:13.296393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:13:13.296393Z digest=sha256:cae96624fe03de9d7b100cb891d64bdd05345af8800ce204e486c004223b0838

Observation 7e15d85e-44ab-453c-a203-02c5a20a7bdb · inbound

First-place Solution for Streetscape Shop Sign Recognition Competition cites this paper.

First-place Solution for Streetscape Shop Sign Recognition Competition Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-10T22:06:47.769539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:06:47.769539Z digest=sha256:df783e9b6ca0ff0f2a809ac7f7b5d515a24eb2f53d1e8678c74b7a82cd99dd4e

Observation ec6a6d8e-3533-4188-af1d-8de9b73f68c0 · inbound

Visual Attention Never Fades: Selective Progressive Attention ReCalibration for Detailed Image Captioning in Multimodal Large Language Models cites this paper.

Visual Attention Never Fades: Selective Progressive Attention ReCalibration for Detailed Image Captioning in Multimodal Large Language Models Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-09T15:26:18.988905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T15:26:18.988905Z digest=sha256:99dae0fc08b849756719b19818b04b66a624d1fd83f5d5e7bf11536909c334a2

Observation 3ebc0f3a-e102-4dca-b569-c3eb4c548913 · inbound

Mitigating Hallucination in Large Vision-Language Models via Adaptive Attention Calibration cites this paper.

Mitigating Hallucination in Large Vision-Language Models via Adaptive Attention Calibration Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-19T12:52:18.038564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-19T12:48:44.324236Z digest=sha256:0adb213f58566a511170ad8618624c8530f580dc134958040b9d8ee6096a5690

Observation bd025866-4c0b-4594-a251-6f09b8e08d4a · inbound

Mitigating Object Hallucination via Robust Local Perception Search cites this paper.

Mitigating Object Hallucination via Robust Local Perception Search Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T05:54:03.925130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:54:03.925130Z digest=sha256:61ea24beda116912b57ad1fce6ca245dd6dd12b0616597a50b7f5e4efae1e4d6

Observation 40febfed-cfd1-45df-b202-9a1c71f8ff8a · inbound

Dense360: Dense Understanding from Omnidirectional Panoramas cites this paper.

Dense360: Dense Understanding from Omnidirectional Panoramas Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T00:23:55.899256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:23:55.899256Z digest=sha256:92dd293d4987c59a1f5bb26e310eb89fd04c7d2b08e69bfa0dc727d8d4c1858a

Observation 85e6f998-a196-4e31-9939-9d1029a12a5b · inbound

DenseWorld-1M: Towards Detailed Dense Grounded Caption in the Real World cites this paper.

DenseWorld-1M: Towards Detailed Dense Grounded Caption in the Real World Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-06T21:27:40.755138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:27:40.755138Z digest=sha256:3c233c289c9b402496cceab375498b8fe1b03655e390a1b3bc91e95bdcd080eb

Observation 39613d2d-a4c9-482a-b3a3-f5002c3b70f0 · inbound

Kwai Keye-VL Technical Report cites this paper.

Kwai Keye-VL Technical Report Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T20:45:10.173386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:45:10.173386Z digest=sha256:cbb3c5ff7bd78bd04d2a40362f688895c3d0827a0e70dda5290028f7ea70ef4d

Observation d01f3afc-539a-41da-ae2a-b9f69b3bb80b · inbound

MCA-LLaVA: Manhattan Causal Attention for Reducing Hallucination in Large Vision-Language Models cites this paper.

MCA-LLaVA: Manhattan Causal Attention for Reducing Hallucination in Large Vision-Language Models Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-06T18:09:11.917632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:09:11.917632Z digest=sha256:7451e30b009b1dcc1156f2db2f49fe4120fc2b8e5f4176fbb7b8fb9a6931091a

Observation bacc4ae4-e3ee-494c-9aba-71a237fc9cf9 · inbound

METEOR: Multi-Encoder Collaborative Token Pruning for Efficient Vision Language Models cites this paper.

METEOR: Multi-Encoder Collaborative Token Pruning for Efficient Vision Language Models Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-06T13:18:05.927482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:18:05.927482Z digest=sha256:d086e5b18a7cea97a3b785e6ff7f26edd1e9326fe683e9c5749bd0bee6db038f

Observation 62c4938f-dd71-462f-864a-de109ec7fbf6 · inbound

ART: Attention Replacement Technique to Improve Factuality in LLMs cites this paper.

ART: Attention Replacement Technique to Improve Factuality in LLMs Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:30:50.979622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T19:48:22.430128Z digest=sha256:a7dbabbcfe7b08b56472c6da678d93ea96d6e61a671d490b555d5cddf96770c7

Observation 37287928-a3b5-44a7-bc42-95ff7d4b799c · inbound

Attention Sink in Transformers: A Survey on Utilization, Interpretation, and Mitigation cites this paper.

Attention Sink in Transformers: A Survey on Utilization, Interpretation, and Mitigation Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 115

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:05:58.215429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-10T16:17:09.834609Z digest=sha256:64a0402d95165c139eccf58f75600550a21e9ba031b0e4ad4f0d7ce0ca91a1af

Observation 5747b385-5d87-4b9b-aba8-13ecf1503094 · inbound

Uncertainty-Aware Exploratory Direct Preference Optimization for Multimodal Large Language Models cites this paper.

Uncertainty-Aware Exploratory Direct Preference Optimization for Multimodal Large Language Models Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:51:09.369022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T17:00:36.574362Z digest=sha256:0f4783ab86fa2cd1553a3ffd871193d01d5a2eb9d6f318150cdf7baa4682b8a4

Observation e982e3b9-b9e9-42e6-8717-eae43d9fc9fa · inbound

When Looking Is Not Enough: Visual Attention Structure Reveals Hallucination in MLLMs cites this paper.

When Looking Is Not Enough: Visual Attention Structure Reveals Hallucination in MLLMs Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T01:57:05.213662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-13T01:55:07.172228Z digest=sha256:b6f2e5fa313951c968df558ecc473fe77ec06f5464a33dced3005d220ec242ad

Observation 5db316ef-8f41-418d-aa82-0024bc6a2b87 · inbound

Addressing Exacerbated Attention Sink for Source-Free Cross-Domain Few-Shot Learning cites this paper.

Addressing Exacerbated Attention Sink for Source-Free Cross-Domain Few-Shot Learning Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 37

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T22:34:01.812681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-29T22:28:35.155099Z digest=sha256:b847a65498748c30ad53859d64d1b47479c6011234dee8c9aa2d15370f888e5d

Observation 92e378c1-1714-4936-9518-e44b121c1eea · inbound

Kwai Keye-VL-2.0 Technical Report cites this paper.

Kwai Keye-VL-2.0 Technical Report Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-07-03T04:27:37.089106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-27T13:53:10.352603Z digest=sha256:4e43a94e158340c3efa48778b9e8f0ea8c793bd753708ee63962c3c0b8936c70

Observation 84a780bb-3bc0-4210-9694-dd90f5009521 · inbound

Omni-Perception Policy Optimization for Multimodal Emotion Reasoning cites this paper.

Omni-Perception Policy Optimization for Multimodal Emotion Reasoning Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:20:06.533244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-25T21:31:38.450382Z digest=sha256:ecde0a67c0b983f882eb771687de3f7c2f376125b79b473d2bef642a8a73c262

Observation 73f85c7e-294c-4727-b8ef-8ffda2e6fd37 · inbound

Clearer Sight, Fewer Lies: Oriented Pickup Preference Optimization for Multimodal Hallucination Mitigation cites this paper.

Clearer Sight, Fewer Lies: Oriented Pickup Preference Optimization for Multimodal Hallucination Mitigation Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 81

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T06:14:18.706272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-30T06:13:11.013395Z digest=sha256:529c07b33940f840593d3b756ec43bd1a2fc6b49a4ad47a98f3ad929989867d6

Observation a30c4598-20f2-4a9c-88e3-5a1c5d2af6da · inbound

Clearer Sight, Fewer Lies: Oriented Pickup Preference Optimization for Multimodal Hallucination Mitigation cites this paper.

Clearer Sight, Fewer Lies: Oriented Pickup Preference Optimization for Multimodal Hallucination Mitigation Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 81

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T07:05:28.648935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-07-01T07:01:26.911110Z digest=sha256:2bd45ed8344dde3f0e0751951b789a8fdf41c750e50f18642c3a4cd4259a61b7

Observation 61274ef6-e42d-42c0-9004-a2d54a9bb2f6 · inbound

The Ebb and Flow of Multimodal Focus: Scheduling Visual Relay Windows for Grounded VLM Reasoning cites this paper.

The Ebb and Flow of Multimodal Focus: Scheduling Visual Relay Windows for Grounded VLM Reasoning Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 57

Resolution
unresolved
no resolver link, observed 2026-07-14T05:37:19.632869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T05:37:19.632869Z digest=sha256:bdc8ba229b4a0ed4a1ee9cfc255e425c7962e0ae141c6c09acab64b30015a4cd

Observation 51b720e1-8e49-4199-8617-94ac48974886 · inbound

The Ebb and Flow of Multimodal Focus: Scheduling Visual Relay Windows for Grounded VLM Reasoning cites this paper.

The Ebb and Flow of Multimodal Focus: Scheduling Visual Relay Windows for Grounded VLM Reasoning Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T07:01:25.384563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T07:01:25.384563Z digest=sha256:d54e3485b639b1e99831cf201cd2bea06dc23c4afb47d6207840c2b166373de2

Observation b67b8773-e6d1-4d58-94f5-fec2ac192c8e · inbound

ST-Veto: Spatio-Temporal Token Veto for Diffusion MLLMs via Taylor Prediction and Visual Grounding cites this paper.

ST-Veto: Spatio-Temporal Token Veto for Diffusion MLLMs via Taylor Prediction and Visual Grounding Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T16:48:36.743797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:48:36.743797Z digest=sha256:aca8961d41639b80d93c91309a8f3a3404ff94da2c6c6b8b1f669479fba55644

Observation ef72004c-976b-4f46-82b5-678fe5cacc2f · inbound

Role-Break in Attention Heads: Understanding and Detecting Hallucinations in VLMs cites this paper.

Role-Break in Attention Heads: Understanding and Detecting Hallucinations in VLMs Seeing Clearly by Layer Two: Enhancing Attention Heads to Alleviate Hallucination in LVLMs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T07:33:41.915824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T07:33:41.915824Z digest=sha256:f868f424c30665b1ee5cf22f9b90498eae71b0126385e47bd75e981877d3d0c4