Pith. sign in

Paper Citation Record · LEDGER

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models

As of 17 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 0 inbound Pith citation observations for arXiv:2608.11024.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.11024 v1

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T11:55:37.153239Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fb671c46-35b9-47d3-9bf5-8752e60f79bd · outbound

This paper cites Hallucination of Multimodal Large Language Models: A Survey.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models Hallucination of Multimodal Large Language Models: A Survey

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T11:55:37.104914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:55:37.104914Z digest=sha256:99dbbd8fe99146afed30643bc9b1d292834f229125d8462e9503fe4919439dda

Observation 9c238c32-741f-435a-8e85-e9a0ed42b5de · outbound

This paper cites DAMRO: Dive into the attention mechanism of LVLM to reduce object hallucination.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models DAMRO: Dive into the attention mechanism of LVLM to reduce object hallucination

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:55:37.377343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:55:37.118328Z digest=sha256:e4d130e7f858ff7ca9810ce76da1a836dbcbdac398746a8f3123bacd8a6b4b09

Observation 78a80fb5-4b4e-480c-84d9-2c9f8095b4c7 · outbound

This paper cites A Survey on Hallucination in Large Vision-Language Models.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models A Survey on Hallucination in Large Vision-Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T11:55:37.126061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:55:37.126061Z digest=sha256:2644e1ebe4c782b5b93a67cb6a267ddd997ec69f1ae6ef17a7d5325772eda85c

Observation 08f0bf3b-56fa-432d-8771-f29e05232c63 · outbound

This paper cites Object hallucination in image captioning.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models Object hallucination in image captioning

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:55:37.348675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:55:37.129952Z digest=sha256:13eff04641a90dbd23549199117f64b872f5f9869fb0ac5b39ba695fe59dddd6

Observation 17bb4e14-4e57-4a74-ab07-0f0cfe258780 · outbound

This paper cites V-DPO: Mitigating hallucination in large vision language models via vision-guided direct preference optimization.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models V-DPO: Mitigating hallucination in large vision language models via vision-guided direct preference optimization

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:55:37.321309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:55:37.137216Z digest=sha256:c31a7279a627c55a13e83b9f5e49516841cc50e5eb3a1b1c4a7f0ebbe88ba989

Observation be0d5add-b79f-46d7-8bc6-acd9d6456e87 · outbound

This paper cites DeCo: Decoupling Token Compression from Semantic Abstraction in Multimodal Large Language Models.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models DeCo: Decoupling Token Compression from Semantic Abstraction in Multimodal Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T11:55:37.141110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:55:37.141110Z digest=sha256:38ac60ef6b6717cc955512dbc9752d22e117b978969c91daf200fbd7e103a759

Observation dcdc9240-19c9-4702-a46f-6a4bccfada68 · outbound

This paper cites A Survey on Multimodal Large Language Models.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models A Survey on Multimodal Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T11:55:37.145641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:55:37.145641Z digest=sha256:ea94ac577b04f9505df0e716ed3e4e65be1610c66d824eb529b89a7504cea291

Observation ebe8c51a-9687-492c-b004-68d6bb65c2ee · outbound

This paper cites Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T11:55:37.149500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:55:37.149500Z digest=sha256:5462e4e039d816fcc8d8fb61654b15e3fc9d4bb20c470e79ae94a77a6f1363ce

Observation a13b1f8d-789c-47e0-a54c-1ae6385dc0b7 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T11:55:37.153239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:55:37.153239Z digest=sha256:0c6c65d0d895630f54dc1cc18e684e437e29bf042c6d6b922beaf4057b5c66e9

Observation d1ec341d-33ad-40e9-9be3-15b0c8dcc766 · outbound

This paper cites Mitigating hallucinations in large vision-language models with instruction contrastive decoding.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models Mitigating hallucinations in large vision-language models with instruction contrastive decoding

Reference 2018

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:55:37.334318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:55:37.133458Z digest=sha256:7959251f0ba947e5f95af4d76eb21f9c3823c9c82ea1088bbc323e36fdaf1570

Observation e5ea7f70-a83d-4b3c-be08-2fb3cca22165 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-12T11:55:37.114003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:55:37.114003Z digest=sha256:3f7d4febc994b4a413378807103c083252da003a15652b16bfb2d7cbffddd2f7

Observation 2309cb75-672d-4eb8-b778-c920fcc0c9d7 · outbound

This paper cites SLIM: Sim-to-Real Legged Instructive Manipulation via Long-Horizon Visuomotor Learning.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models SLIM: Sim-to-Real Legged Instructive Manipulation via Long-Horizon Visuomotor Learning

Reference 2022

Resolution
metadata mismatch
local_arxiv, observed 2026-08-12T11:55:37.306108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:55:37.096251Z digest=sha256:cc3737e1445fe1a05a3f27b56ad2fa112c62826e9ebe9a38f6d800ab33892113

Observation 3a71ccd3-6286-4a4f-81d7-1a33187dc7a8 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-12T11:55:37.109020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:55:37.109020Z digest=sha256:d93a7f6d7c52c2cd2bff23bfaa845c879fe7eeb7b8564d7f6f1e73d454a85b5e

Observation 131cb3de-3d28-4733-8295-c417c8fe4f9a · outbound

This paper cites BLIP- 2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models BLIP- 2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:55:37.362854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T11:55:37.122322Z digest=sha256:c0ac3d4f46de89143d1aba791d6087f7b7eff5f2fadbfd10735daec134326316

Observation 6bd9bd09-dd96-40b1-92ff-08d0c3a7f8e3 · outbound

This paper cites Qwen2.5-VL Technical Report.

When Visual Signals Mislead: A Mechanistic Study of Attribute Hallucination in Vision-Language Models Qwen2.5-VL Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-12T11:55:37.100522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:55:37.100522Z digest=sha256:87a685a7009f31f2c13e4bf51f7766cb4093018ae6df6cb7ecc360c041c4ce6c

Pith citing papers

No inbound Pith citation observations are available.