Pith. sign in

Paper Citation Record · LEDGER

Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2402.15300.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.15300 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T16:44:00.603815Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

3
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 276b6fce-68e5-4ac7-98c1-aff2008ad923 · inbound

Hallucination of Multimodal Large Language Models: A Survey cites this paper.

Hallucination of Multimodal Large Language Models: A Survey Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:33:33.696406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T12:33:32.631346Z digest=sha256:0c96f594e898bee7cd255177d351e5a759fa4c921b317e41eb0ff98a793d9ec3

Observation 19836e82-2759-4388-8322-f5942fcc9886 · inbound

Self-Correcting Decoding with Generative Feedback for Mitigating Hallucinations in Large Vision-Language Models cites this paper.

Self-Correcting Decoding with Generative Feedback for Mitigating Hallucinations in Large Vision-Language Models Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T16:44:00.603815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T16:44:00.603815Z digest=sha256:1b20cc115a3442b24d458399ea92a0422c2efeb002666c625218d8be500d9f55

Observation f5ce18c0-7c9d-47af-861d-958057782d70 · inbound

Mitigating Behavioral Hallucination in Multimodal Large Language Models for Sequential Images cites this paper.

Mitigating Behavioral Hallucination in Multimodal Large Language Models for Sequential Images Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:43:30.548053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:43:30.548053Z digest=sha256:5bbbe49820831fa3f2e705e1aff74b0173c49975383b713ab81d4a36e0073a6e

Observation 5381a6ff-955e-46b2-a247-d81895bd4159 · inbound

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation cites this paper.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.164743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.164743Z digest=sha256:6cc9568571715934ea6715be07d6fede9cbe44f3c7ab8b060c0a341ff85309db

Observation 941699a5-a68d-4549-b759-f20cd207bae8 · inbound

PostAlign: Multimodal Grounding as a Corrective Lens for MLLMs cites this paper.

PostAlign: Multimodal Grounding as a Corrective Lens for MLLMs Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:19.317381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:19.317381Z digest=sha256:23522ce91bb3152e3edd6bd6eb7199c7b24fb82836cf7fc23354f2b2a6a65cbe

Observation b6c7de87-202f-49b0-8533-ae977cac582d · inbound

Energy-Guided Decoding for Object Hallucination Mitigation cites this paper.

Energy-Guided Decoding for Object Hallucination Mitigation Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:39:26.430472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:39:26.430472Z digest=sha256:2dcbf8f5dd170598a92d1bd0bbbbcf5d50a969a64c0ef4e94c5e827349895346

Observation 23f798c5-7cbc-48d7-b969-29589930f9ff · inbound

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey cites this paper.

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding

Reference 192

Resolution
unresolved
no resolver link, observed 2026-08-05T20:29:02.294753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:29:02.294753Z digest=sha256:2cc92dc9f4d32736b8a1ae7bea9b2ba9b1b4a5f171c38d206d9076e4875a06dc

Observation d48f62ed-f031-4991-b33f-ff70f1455e3b · inbound

Controlling Multimodal LLMs via Reward-guided Decoding cites this paper.

Controlling Multimodal LLMs via Reward-guided Decoding Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T19:52:50.201207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:52:50.201207Z digest=sha256:f6c04b388d28f7b54fc65605675a97aa28449de73e8115b3c9981bc1821b09bd

Observation 653d45f7-7fba-41c5-a78e-1ee12a079363 · inbound

Challenges in Understanding Modality Conflict in Vision-Language Models cites this paper.

Challenges in Understanding Modality Conflict in Vision-Language Models Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T11:26:17.281990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:26:17.281990Z digest=sha256:2346f3bab32b86684db40e0714ff294a93cc69da79f08602753b8f805bb29f31

Observation ea2c41b8-dc81-47d6-b52f-5f968b0f6deb · inbound

See Fair, Speak Truth: Equitable Attention Improves Grounding and Reduces Hallucination in Vision-Language Alignment cites this paper.

See Fair, Speak Truth: Equitable Attention Improves Grounding and Reduces Hallucination in Vision-Language Alignment Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T08:26:00.624204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T16:38:51.797197Z digest=sha256:1babf54922ef692692f38f2b543076302eabe2f7ea2accec067e6a52365abe78

Observation 099a0af0-a94e-4112-b53f-20dce5d25ab5 · inbound

Deep Pre-Alignment for VLMs cites this paper.

Deep Pre-Alignment for VLMs Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding

Reference 109

Resolution
verified exact
arxiv_id, observed 2026-05-19T16:27:39.109040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-19T16:26:41.094936Z digest=sha256:ee5ca1613a2680cd1a24934455f40d9b2182f76c7a13eddf8b624770374a05a8

Observation 0dc8bd54-10e0-473b-a2ba-3e324d564d0d · inbound

GAVEL: Grounded Caption Error Verification and Localization cites this paper.

GAVEL: Grounded Caption Error Verification and Localization Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding

Reference 90

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T13:59:52.212267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T04:44:48.346335Z digest=sha256:43ab1cb761946c7381e6428119747f89f9f020d4739b973b78046ae572f830e4

Observation dc442174-122a-4377-9d77-9f60e890fc8b · inbound

Constraint-Anchored Reasoning Traces cites this paper.

Constraint-Anchored Reasoning Traces Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T20:14:07.511863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:14:07.511863Z digest=sha256:f1304376aea8bc2ec15c6e5810b8074655d5c414051658f275ffd442d6a3d26b

Observation 2c2f2ce1-1f90-4cf9-8cb7-9cafd2c99811 · inbound

MissingBench-Verified: Probing Vision-Language Models' Inability to Detect Missing Object Parts cites this paper.

MissingBench-Verified: Probing Vision-Language Models' Inability to Detect Missing Object Parts Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-01T14:44:05.933781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T14:44:05.933781Z digest=sha256:e1e61d14e3bf3e2a3be7d96edf3bae572cdc3e7b3912e73b9e387714b79043d2