Pith. sign in

Paper Citation Record · LEDGER

Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2410.02762.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.02762 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T14:23:23.424231Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T05:54:33.561588Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8738d5ef-ea2c-4835-91b8-743dafce695c · inbound

Hallucination of Multimodal Large Language Models: A Survey cites this paper.

Hallucination of Multimodal Large Language Models: A Survey Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:33:33.852213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-11T12:33:32.631346Z digest=sha256:9b7ee2c332854c2412b14573c7409f6009c94a7ecb9497fcaf7dbe8cccab8c69

Observation a068e6fe-cd88-4107-ba59-448817303d24 · inbound

Devils in Middle Layers of Large Vision-Language Models: Interpreting, Detecting and Mitigating Object Hallucinations via Attention Lens cites this paper.

Devils in Middle Layers of Large Vision-Language Models: Interpreting, Detecting and Mitigating Object Hallucinations via Attention Lens Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T14:23:23.424231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:23:23.424231Z digest=sha256:3cabd7fe76f22ce9e5314d7f59c5f1e5302846b13d281cea5191716235c89c88

Observation 23d4f89a-e093-4483-952c-f7d56133c2f8 · inbound

What's in the Image? A Deep-Dive into the Vision of Vision Language Models cites this paper.

What's in the Image? A Deep-Dive into the Vision of Vision Language Models Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T12:08:51.404123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:08:51.404123Z digest=sha256:4b846f677e8e004835b7e384edf8b48a71f8871331aae631b771df1b8407b091

Observation ff761382-1983-4874-b426-85c06ace8390 · inbound

Beyond Logit Lens: Contextual Embeddings for Robust Hallucination Detection & Grounding in VLMs cites this paper.

Beyond Logit Lens: Contextual Embeddings for Robust Hallucination Detection & Grounding in VLMs Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T10:30:28.856594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T10:30:28.856594Z digest=sha256:678bb043ff5ebbfaabcdc78240e998b6e959b199577fd1c0c6e98fb87a3c4db3

Observation cdf8dc59-77ae-446c-97a4-d1d432a1ea17 · inbound

Debiasing CLIP: Interpreting and Correcting Bias in Attention Heads cites this paper.

Debiasing CLIP: Interpreting and Correcting Bias in Attention Heads Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 11

Resolution
malformed identifier
no resolver link, observed 2026-08-07T14:53:08.673727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:53:08.673727Z digest=sha256:b2a6a1391937407f915f104cebe7f24b8ba009df696620f039987e9fc10da7ff

Observation 1ee8520a-95e1-46a6-ac0b-586354530e1f · inbound

Image Tokens Matter: Mitigating Hallucination in Discrete Tokenizer-based Large Vision-Language Models via Latent Editing cites this paper.

Image Tokens Matter: Mitigating Hallucination in Discrete Tokenizer-based Large Vision-Language Models via Latent Editing Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:27:57.965731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:27:57.965731Z digest=sha256:435844ef57531f9da6cd16436087b43226f8711d849f8e11f51870ee3b38f5d1

Observation 25f00dca-de3e-4cfd-883d-653b0b12c346 · inbound

How Visual Representations Map to Language Feature Space in Multimodal LLMs cites this paper.

How Visual Representations Map to Language Feature Space in Multimodal LLMs Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T01:06:06.969709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:06:06.969709Z digest=sha256:a6e985afc56e855d976970555a06b5a7b7f79f30a8e0efc2ff29b6f6ab354225

Observation 1ec682e6-6488-4a40-b762-042da765fc06 · inbound

SAEmnesia: Erasing Concepts in Diffusion Models with Supervised Sparse Autoencoders cites this paper.

SAEmnesia: Erasing Concepts in Diffusion Models with Supervised Sparse Autoencoders Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T15:42:37.536686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T15:42:37.536686Z digest=sha256:465afcfa1a8a5ae8b0e8ce8973c94fe417af4e30b70ebc3dda3aa5ec05709fa7

Observation ba9802a2-c408-4e9b-9f1a-1e1e868b4c31 · inbound

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models cites this paper.

Evidence Recomposition and Predictive Context Residualization for Visual Attribution in Multimodal Large Language Models Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:37.914746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:58:37.914746Z digest=sha256:dc79bb27ae3386232ce435b673556cf33eb49a828bbde6756f6426ffe51d5d4d

Observation 5c88755e-37fc-4747-9132-0ee6dd3ff9b7 · inbound

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models cites this paper.

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 133

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:40:54.540771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-16T12:39:57.398423Z digest=sha256:90eca10c3b5b28e9006b0c3d8b53a8145dae64a1b9e90b342807b72fc272b5f4

Observation d13a2f25-5820-4bbb-ac3b-b7a3b9c1dc7c · inbound

Relaxing Anchor-Frame Dominance for Mitigating Hallucinations in Video Large Language Models cites this paper.

Relaxing Anchor-Frame Dominance for Mitigating Hallucinations in Video Large Language Models Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:26:03.395976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T14:54:52.710686Z digest=sha256:769e291c25452c4fe776d4f2291dbf42bcae0d33da8e6f198da56ce462fecadf

Observation ccd6de05-7bd9-4d5f-8002-0ffd37638413 · inbound

The Hidden Evolution of Disguised Visual Context inside the VLM cites this paper.

The Hidden Evolution of Disguised Visual Context inside the VLM Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:29:29.139354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-26T18:08:56.044278Z digest=sha256:01c6f7a7dc565de4111bbf80598024beeccaf32c2479b969e549c1f2248cb4e6

Observation 8fa6cb71-373a-41c0-9029-e7d08d52206c · inbound

Vision-Default, Prior-Override: Causal Mechanisms of Perception-Knowledge Conflict in Vision-Language Models cites this paper.

Vision-Default, Prior-Override: Causal Mechanisms of Perception-Knowledge Conflict in Vision-Language Models Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T17:15:52.035788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T03:53:04.984303Z digest=sha256:480811753fa25979bfd186b2faaed8ea5a2fd9bc257bc14b35e3cb40ecab91e4

Observation 67d74a6e-4e64-4164-a191-9bcd82c72c8b · inbound

Analysis-by-Proxy: Localization Signals in VLMs Operating as Condition Encoders cites this paper.

Analysis-by-Proxy: Localization Signals in VLMs Operating as Condition Encoders Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 11

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T05:54:33.562879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-07-08T05:49:25.572256Z digest=sha256:a062e1f5c13e460e7034c459bdaa435ed925f15a1039ec005af98bb692bef9db

Observation 6b1dc2dc-2a6a-4094-b0f2-97b46bbd5f34 · inbound

Verbalizable Representations Form a Global Workspace in Language Models cites this paper.

Verbalizable Representations Form a Global Workspace in Language Models Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T23:15:25.803851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T23:15:25.803851Z digest=sha256:8f507b68c7ebbb72562ca51df513b63025240aeb41ead60778e484068ab0ede6

Observation 316a8dd2-a6fe-4d4c-bdd2-3421c4d4f231 · inbound

TruthLens: Object Hallucination Detection via Self-Evaluating Truthfulness Scores in LVLMs cites this paper.

TruthLens: Object Hallucination Detection via Self-Evaluating Truthfulness Scores in LVLMs Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T05:29:59.247547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T05:29:59.247547Z digest=sha256:094afdffde110f2324ae72e59a01739427198497b70fe452a2c4d9c27afa6cf2

Observation 615a9787-1c90-4108-8fb9-c20c8e3c0411 · inbound

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination cites this paper.

Same Attention, Different Truths: Put Logit-Lens over Visual Attention to Detect and Mitigate LVLM Object Hallucination Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:43.945970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:43.945970Z digest=sha256:f91963b260ad12b06c1a8a89e5854184a7224e45202e4e0e68847677c7708860

Observation 60bfe481-7159-4238-b912-6accdf089da1 · inbound

Multimodal Model Diffing for Feature Discovery and Control cites this paper.

Multimodal Model Diffing for Feature Discovery and Control Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:55.889674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:55.889674Z digest=sha256:5704b7a5ad613b32b830751be9cd3f08e4ac0e6978b0508fef83ff069d45e1a3