Pith. sign in

Paper Citation Record · LEDGER

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models

As of 17 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 0 inbound Pith citation observations for arXiv:2608.11024.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.11024 v1

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T11:55:37.153239Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fb671c46-35b9-47d3-9bf5-8752e60f79bd · outbound

This paper cites Hallucination of Multimodal Large Language Models: A Survey.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models Hallucination of Multimodal Large Language Models: A Survey

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T11:55:37.104914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:55:37.104914Z digest=sha256:45911fb9f411a5add925d46787bf30b33d374e8db6250a9a0dafe5ecf5b4d2e3

Observation 9c238c32-741f-435a-8e85-e9a0ed42b5de · outbound

This paper cites DAMRO: Dive into the attention mechanism of LVLM to reduce object hallucination.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models DAMRO: Dive into the attention mechanism of LVLM to reduce object hallucination

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:55:37.377343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:55:37.118328Z digest=sha256:0d632227035a1c86000e28575531c1d8d8c1a07d051d5423c4a1a2614e1bd5b3

Observation 78a80fb5-4b4e-480c-84d9-2c9f8095b4c7 · outbound

This paper cites A Survey on Hallucination in Large Vision-Language Models.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models A Survey on Hallucination in Large Vision-Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T11:55:37.126061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:55:37.126061Z digest=sha256:bb4cf2b2a7f49a84aa174d0a07df5442b22fb92d2de2a44d877e7e83d87045b0

Observation 08f0bf3b-56fa-432d-8771-f29e05232c63 · outbound

This paper cites Object hallucination in image captioning.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models Object hallucination in image captioning

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:55:37.348675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:55:37.129952Z digest=sha256:8edb061893a7a4608156bd04837f98e10dc69a093ac6a95ddb1a48423379755c

Observation 17bb4e14-4e57-4a74-ab07-0f0cfe258780 · outbound

This paper cites V-DPO: Mitigating hallucination in large vision language models via vision-guided direct preference optimization.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models V-DPO: Mitigating hallucination in large vision language models via vision-guided direct preference optimization

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:55:37.321309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:55:37.137216Z digest=sha256:01e8032cec2888e214ac83904d95047beafac2a9cfe66d65ccdafa402f38fd9c

Observation be0d5add-b79f-46d7-8bc6-acd9d6456e87 · outbound

This paper cites DeCo: Decoupling Token Compression from Semantic Abstraction in Multimodal Large Language Models.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models DeCo: Decoupling Token Compression from Semantic Abstraction in Multimodal Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T11:55:37.141110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:55:37.141110Z digest=sha256:fe0331c5ad2d2e3faa098364083991b4bd42e90bafd91826d372cc9d393dabd0

Observation dcdc9240-19c9-4702-a46f-6a4bccfada68 · outbound

This paper cites A Survey on Multimodal Large Language Models.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models A Survey on Multimodal Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T11:55:37.145641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:55:37.145641Z digest=sha256:02f7f57c3cb29b1d02b7fec35011bd9c9631fd1b828422ee4b74ef814c2b798c

Observation ebe8c51a-9687-492c-b004-68d6bb65c2ee · outbound

This paper cites Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T11:55:37.149500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:55:37.149500Z digest=sha256:87ef36a696141a3f57b5a38e6b90322511877d59cc8fa9d3260300de7ed575b3

Observation a13b1f8d-789c-47e0-a54c-1ae6385dc0b7 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T11:55:37.153239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:55:37.153239Z digest=sha256:d563a3a7c62ddd73fd99e206b3da873e501a96e64439d423d1efe80798e715fa

Observation d1ec341d-33ad-40e9-9be3-15b0c8dcc766 · outbound

This paper cites Mitigating hallucinations in large vision-language models with instruction contrastive decoding.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models Mitigating hallucinations in large vision-language models with instruction contrastive decoding

Reference 2018

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:55:37.334318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:55:37.133458Z digest=sha256:46ee57cc1239088a1fa320996270f33848325502c5460f62f0bb240c2980b49c

Observation e5ea7f70-a83d-4b3c-be08-2fb3cca22165 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-12T11:55:37.114003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:55:37.114003Z digest=sha256:b0eae22a056f7d8921bdfa1921f5b7762215f887941b8984f1ddcfd1ca8ad9c3

Observation 2309cb75-672d-4eb8-b778-c920fcc0c9d7 · outbound

This paper cites SLIM: Sim-to-Real Legged Instructive Manipulation via Long-Horizon Visuomotor Learning.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models SLIM: Sim-to-Real Legged Instructive Manipulation via Long-Horizon Visuomotor Learning

Reference 2022

Resolution
metadata mismatch
local_arxiv, observed 2026-08-12T11:55:37.306108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:55:37.096251Z digest=sha256:d51ebbb36ec55ce6320407efe9c0d2e51ee642dd51d294faddea37f7d0a3fa76

Observation 3a71ccd3-6286-4a4f-81d7-1a33187dc7a8 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-12T11:55:37.109020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:55:37.109020Z digest=sha256:67de3a95d9f7a67a1320ab7dde688bb7f115dd5ecba477dca1e5abea7b20951c

Observation 131cb3de-3d28-4733-8295-c417c8fe4f9a · outbound

This paper cites BLIP- 2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models BLIP- 2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:55:37.362854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:55:37.122322Z digest=sha256:733ecf90c6dfc2c7a54198784bd609d945788ecbd70e174344ef80bdc0710d57

Observation 6bd9bd09-dd96-40b1-92ff-08d0c3a7f8e3 · outbound

This paper cites Qwen2.5-VL Technical Report.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models Qwen2.5-VL Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-12T11:55:37.100522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:55:37.100522Z digest=sha256:e11de5c6d02c464f39716d4765d781e9ebc90bc67c514addf783cb4bbe20226c

Pith citing papers

No inbound Pith citation observations are available.