Pith. sign in

Paper Citation Record · LEDGER

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations

As of 11 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 0 inbound Pith citation observations for arXiv:2505.17812.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17812 v1

Coverage vector

measured 64 of 64 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:45:32.061891Z

measured 64 of 64 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

64 of 64 outbound references displayed

  • verified exact3
  • verified fuzzy31
  • unresolved29
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 882aa6cd-9854-4dfc-8100-a9006e0ca4f8 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations LLaMA: Open and Efficient Foundation Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:26.861270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:26.861270Z digest=sha256:980dc5af0f9d055380e7ddb99d334cdec483c8b6a61034b1b8d03807c158179f

Observation 737a10df-2d22-40cb-bfce-2668d47766e6 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:26.929012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:26.929012Z digest=sha256:67fe137d7e6c4459339af656fa52cdf4a6caf525ed557006c35cdfa47d3fa42f

Observation 81c64c90-6c7e-4622-bfbb-60dfd6c91d37 · outbound

This paper cites Visual instruction tuning.Adv.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Visual instruction tuning.Adv

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:39.249729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:27.001968Z digest=sha256:3567ff112080e5baad50f4ff34830d8c2036dc975d5568a43ed8e1082d4f40f0

Observation 4455e4e8-22fe-4da9-847f-c24901bf3ee4 · outbound

This paper cites Improved Baselines with Visual Instruction Tuning.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Improved Baselines with Visual Instruction Tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:27.094656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:27.094656Z digest=sha256:e1c97c053b2233edfb250a95d6734e3193d41fef3772dc8951bba09a2bf5c3cd

Observation c8a21e09-834d-476c-9bdd-064d5f45d427 · outbound

This paper cites InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:27.170843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:27.170843Z digest=sha256:1eb7b902b02c7644c131e6067604859ccdf9bc578284152b6492f8e55bc7b9a5

Observation 8f887f16-e771-46b9-8839-be6850db0e26 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:27.251904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:27.251904Z digest=sha256:d629dd5cd728cea40e3b102ec6be31e2f0d86dd80453af4f843d5d14994911ff

Observation 9e209bd6-89d7-4769-830f-56caefcb9c6f · outbound

This paper cites Qwen Technical Report.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Qwen Technical Report

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:27.325359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:27.325359Z digest=sha256:0eaee83048a6a90711df07e47d14b7c441c2e11ae316f9e99a5661d0fb411852

Observation cfe75358-ff73-4c7d-8cbe-b2cfef5b7b6e · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:27.396913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:27.396913Z digest=sha256:f82020a3ca7bb7016afd1ce9d5c4d70c1e9dd05b7186ddf1fc5580c1f47b52f7

Observation f1d4b41e-4c28-4897-82a9-ea75f1ccdfe5 · outbound

This paper cites Hallucination of Multimodal Large Language Models: A Survey.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Hallucination of Multimodal Large Language Models: A Survey

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:27.497184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:27.497184Z digest=sha256:6820509f86e666dc30f1acaf7d38f1b760c1c302ce488c88c908c34504e9dd0b

Observation 66e2ef9a-7aaf-41da-9155-14405f3ba299 · outbound

This paper cites Nullu: Mitigating object hallucinations in large vision-language models via halluspace projection.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Nullu: Mitigating object hallucinations in large vision-language models via halluspace projection

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:39.061415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:27.582588Z digest=sha256:92018c8c57976dec713c518b718e18b9d024ed74cef26a8c1123fc3dddd1b66e

Observation f12fe377-086e-4152-a5f1-8a0c152760c5 · outbound

This paper cites Truthprint: Mitigating lvlm object hallucination via latent truthful-guided pre-intervention.arXiv preprint arXiv:2503.10602, 2025.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Truthprint: Mitigating lvlm object hallucination via latent truthful-guided pre-intervention.arXiv preprint arXiv:2503.10602, 2025

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:27.675017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:27.675017Z digest=sha256:85ac24915f12a588536f05d3a1938a6af75afa4b8f2208947c7d53a1dd1b86ce

Observation 0473397b-fec5-45da-b3c0-13bccfee3515 · outbound

This paper cites Analyzing and mitigating object hallucination in large vision-language models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Analyzing and mitigating object hallucination in large vision-language models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:38.889223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:27.767172Z digest=sha256:dfb151e003b6d50eab249e8acf69f48eddf2a0de044b22ac94f7cd3482fb7aa6

Observation 7bd6d8df-457b-42e7-abf7-ea2e01419e83 · outbound

This paper cites Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:27.838292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:27.838292Z digest=sha256:0228c2df025538b500c4f7df3512b64d514c37b8306ac93d5f7c0b3ed5ff32ee

Observation c1a8a98d-b379-472b-9222-2525fcf3a5c7 · outbound

This paper cites Hallucination augmented contrastive learning for multimodal large language model.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Hallucination augmented contrastive learning for multimodal large language model

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:38.723563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:27.945183Z digest=sha256:f0cc8d28acdd119b60aa16111eb7e2d10d653a2fc194cedc03246297b37ab947

Observation 4c94a0d3-e0d6-4f2f-b22b-47cf3c77c1a1 · outbound

This paper cites Exposing and mitigating spurious correlations for cross-modal retrieval.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Exposing and mitigating spurious correlations for cross-modal retrieval

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:38.528588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:28.024239Z digest=sha256:4cc57b620da11b72cd9d94250a7ea579532153f8a63c024998aec7f764b22a08

Observation 5671c12d-a4aa-4613-a229-587492166ecc · outbound

This paper cites Mitigating object hallucinations in large vision-language models through visual contrastive decoding.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Mitigating object hallucinations in large vision-language models through visual contrastive decoding

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:38.339332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:28.107255Z digest=sha256:fe854b11d46e1af3b259e7c6341176a8e321637cdeddd1aebc349b0e01191968

Observation 196a4003-abf4-40d3-8f5e-25063d5a3327 · outbound

This paper cites Debiasing large visual language models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Debiasing large visual language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:38.147222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:28.181188Z digest=sha256:22eced63985851aa3cc8d7207c2745959c19a222c0d7472cfb2c236bbd38f02d

Observation 65bf3d5c-5958-41c4-ae01-28e195126a20 · outbound

This paper cites Halc: Object hallucination reduction via adaptive focal-contrast decoding.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Halc: Object hallucination reduction via adaptive focal-contrast decoding

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:37.936158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:28.261744Z digest=sha256:6f4a7d695ceedf2046303fc9c9261c7ca7360de3070a2c9feb774aec72f7a409

Observation ad860591-dd20-4f99-9944-9c9f1db318f6 · outbound

This paper cites ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations ICT: Image-Object Cross-Level Trusted Intervention for Mitigating Object Hallucination in Large Vision-Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:28.344201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:28.344201Z digest=sha256:791a28c11e4acbe83abe7bf8be47c2f724a93e940f81bca6db7159dc9c78d40d

Observation 25ab535f-ea73-4a22-8646-e069e69f639a · outbound

This paper cites Reducing hallucinations in large vision-language models via latent space steering.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Reducing hallucinations in large vision-language models via latent space steering

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:37.709424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:28.404847Z digest=sha256:3e4bd49094ace4442da3e93abe1a3926769ba167ea2fc6c5d909fb3697d7dcd8

Observation 0cfe0c03-bf00-4cff-95a9-3fae186b43fb · outbound

This paper cites Where do Large Vision-Language Models Look at when Answering Questions?.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Where do Large Vision-Language Models Look at when Answering Questions?

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:28.465442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:28.465442Z digest=sha256:afaae7dca6db0cc249ab58d14a993656c0034de0a004734f04debad2e442deb3

Observation 49f3903b-7a7a-4ab0-9e11-feaee20efb4a · outbound

This paper cites Lvlm-intrepret: An interpretability tool for large vision-language models, 2024.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Lvlm-intrepret: An interpretability tool for large vision-language models, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:37.516328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:28.517311Z digest=sha256:d35775c89b792c566d587d9ce8b87a4b8c51cff1ae1bd85c27c04c8689aa4320

Observation 26f19368-69be-4297-ad00-dc9f67a84f3a · outbound

This paper cites See What You Are Told: Visual Attention Sink in Large Multimodal Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations See What You Are Told: Visual Attention Sink in Large Multimodal Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:28.622401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:28.622401Z digest=sha256:eee80ce3c5fcce52c724e986a5189ae7b2e92fc4e4c7355008b52c4cd544b8d9

Observation 3f9dcb33-210b-4d5c-bbad-5ea4f33e9034 · outbound

This paper cites Vision Transformers Need Registers.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Vision Transformers Need Registers

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:28.704062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:28.704062Z digest=sha256:16e8c8788dced8b3ba55476eb0f49a056de2da61b475f8e62a245838b31f7d74

Observation bb0f94bf-7c2a-4275-9518-11fbd28bcacf · outbound

This paper cites Massive Activations in Large Language Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Massive Activations in Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:28.759191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:28.759191Z digest=sha256:0d3b228e082f6e4a9b2f0428ee6a5fa1c191fac2d8b17cf20057057c1585dc75

Observation ac162876-881f-4b81-8f13-701f6b5eeb37 · outbound

This paper cites Object Hallucination in Image Captioning.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Object Hallucination in Image Captioning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:28.823154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:28.823154Z digest=sha256:7504beed32680fe2882aab1143ca839886cf4e08a69e4a1193595beb6f3b54c0

Observation d6fcddd8-1821-40b3-b54b-4f0175019e62 · outbound

This paper cites mplug-owl2: Revolutionizing multi-modal large language model with modality collaboration.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations mplug-owl2: Revolutionizing multi-modal large language model with modality collaboration

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:37.289552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:28.894823Z digest=sha256:bdb9f84566ecde5e416ff8883b68047675d099eefe66b2a7086f40dd664b4d63

Observation de3e5d8d-fea9-4b4c-892f-d99344a2f540 · outbound

This paper cites mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:28.975925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:28.975925Z digest=sha256:d3311602ee3eb60cf7aaea876608eb400962c1eb146f566cfd86b14ebc61df70

Observation 9ab16929-c20f-4955-b6bc-ff8ad009c7d1 · outbound

This paper cites Llava-phi: Efficient multi-modal assistant with small language model.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Llava-phi: Efficient multi-modal assistant with small language model

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:37.107790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:29.029305Z digest=sha256:9fe174dfabb6a20e0b424b18da27a8b499b0dbb2cfc98c02248cce9c9975204b

Observation fc8a6a40-c744-42de-9ae0-c7cf38e76412 · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:29.111032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:29.111032Z digest=sha256:18596725344d80c6a073d37bb38292d4043bfb736d05026ebfeb2a61a3ac712c

Observation cdd349ec-b545-4bf8-9229-43c1bfdc52b9 · outbound

This paper cites Detecting and preventing hallucinations in large vision language models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Detecting and preventing hallucinations in large vision language models

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:36.895414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:29.202872Z digest=sha256:ba31d1f9f3a021eb068866d00308e2f88377cee648c1ee72ca2c2fc704767da2

Observation 125f43ae-65de-4b25-be32-532fb592a6ec · outbound

This paper cites Aligning Large Multimodal Models with Factually Augmented RLHF.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Aligning Large Multimodal Models with Factually Augmented RLHF

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:29.287602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:29.287602Z digest=sha256:3238d517f0c6007d09af3e80eb55079531b7daff3798c43bd6d9345f42be2f57

Observation 85fa9ad1-866f-4ebb-b530-46f4fd57c7f6 · outbound

This paper cites Dress: Instructing large vision-language models to align and interact with humans via natural language feedback.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Dress: Instructing large vision-language models to align and interact with humans via natural language feedback

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:36.712636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:29.363088Z digest=sha256:176a7ffa13a5add9f356895462ab7a583779ac9fe9c0d0fa5f81b5961c96266e

Observation 7f9b0304-50f3-4f2e-b0ec-d8cabe019404 · outbound

This paper cites Woodpecker: Hallucination Correction for Multimodal Large Language Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Woodpecker: Hallucination Correction for Multimodal Large Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:29.439784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:29.439784Z digest=sha256:063f6d299cb51a2ded5253b835cdb628e898b97087a7c2497393dae2fddf6261

Observation a0f958ef-718f-482d-b0f9-39b5dd6e23b3 · outbound

This paper cites Mitigating object hallucination in large vision-language models via image-grounded guidance.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Mitigating object hallucination in large vision-language models via image-grounded guidance

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:36.500425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:29.524121Z digest=sha256:12c8af59450561bd0870772482b57b7aab54985b53e01f1f1f1478b052ed39af

Observation ff3e8d54-95f9-4d80-b404-ff620b82690d · outbound

This paper cites Paying more attention to image: A training-free method for alleviating hallucination in lvlms.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Paying more attention to image: A training-free method for alleviating hallucination in lvlms

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:36.337308Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:29.593024Z digest=sha256:9b85572be659dc348b7675db6abedf39813924c8dfd94d0eec07ab483ed8c23d

Observation b6698032-96ff-40d5-a3aa-b68b7600a43a · outbound

This paper cites IBD: Alleviating Hallucinations in Large Vision-Language Models via Image-Biased Decoding.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations IBD: Alleviating Hallucinations in Large Vision-Language Models via Image-Biased Decoding

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:29.672408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:29.672408Z digest=sha256:3bdc15bc75ab012954574811662ee5afecd87b8c2d800ceb8fc43edd9c3f64f9

Observation a24c3d4f-df58-4cf0-bcd3-cb50f4e2da2d · outbound

This paper cites Opera: Alleviating hallucination in multi-modal large language models via over-trust penalty and retrospection-allocation.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Opera: Alleviating hallucination in multi-modal large language models via over-trust penalty and retrospection-allocation

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:36.102530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:29.758823Z digest=sha256:772f687e23dfb268a4e0fd8b1186bb3f7a01dd62c0bbe21c0eafa12667a2bade

Observation 77a176db-9161-4150-9360-879d127009a8 · outbound

This paper cites Multi-modal hallucination control by visual information grounding.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Multi-modal hallucination control by visual information grounding

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:35.874270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:29.857326Z digest=sha256:7086bddfac2a7cf367e5898643daa5736f71d127ac5ef0b0b426b1dde80547c0

Observation a06fcb51-2132-4696-a792-026aa2d63c38 · outbound

This paper cites AlphaEdit: Null-Space Constrained Knowledge Editing for Language Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations AlphaEdit: Null-Space Constrained Knowledge Editing for Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:29.943737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:29.943737Z digest=sha256:cff4f458a310c18b966e158dbb4f854d52afac4d7f8f12bffe0e8e9e95c7f4c4

Observation b2eb7736-2804-407e-9555-9576ea099f9f · outbound

This paper cites Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:30.042461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:30.042461Z digest=sha256:4794af0cbbbdbd05186e0d9210d930053948e326f69474c900cd057fd9837b1c

Observation b4fefc28-d1a6-4b60-ab67-38a9e25413b6 · outbound

This paper cites Grad-cam: Visual explanations from deep networks via gradient-based localization.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Grad-cam: Visual explanations from deep networks via gradient-based localization

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:35.675866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:30.110923Z digest=sha256:1294fa9274d6fd5fdd45a34647b628c8f1a8abcd1a26711eacd9599626799bf4

Observation 847e83b9-e414-4ccd-9632-e2aecb5eb7ab · outbound

This paper cites Grad-cam++: Generalized gradient-based visual explanations for deep convolutional networks.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Grad-cam++: Generalized gradient-based visual explanations for deep convolutional networks

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:35.484548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:30.219446Z digest=sha256:64908ad5a5052eb4bfcc5047f19436bbb71e4a697f0c96c9e0857dd3d449fe61

Observation 3284945c-49ca-4b6d-9391-ef63178ae38a · outbound

This paper cites Generic attention-model explainability for interpreting bi-modal and encoder-decoder transformers.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Generic attention-model explainability for interpreting bi-modal and encoder-decoder transformers

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:35.230180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:30.320579Z digest=sha256:17561976d0f2fd11361ad3fc83195e514c5938626800891c7ec1d12f7448a079

Observation 8272d023-731e-4d73-b974-2645efdb56d7 · outbound

This paper cites Transformer interpretability beyond attention visualization.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Transformer interpretability beyond attention visualization

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:35.077456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:30.399480Z digest=sha256:fe77ce074aaed29f908e5979619d74542fcdd52828a32f46b81d561e9eabbb31

Observation 82447b30-d45b-4a50-a20a-83d9ff7a2806 · outbound

This paper cites Vl- interpret: An interactive visualization tool for interpreting vision-language transformers.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Vl- interpret: An interactive visualization tool for interpreting vision-language transformers

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:34.838376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:30.463305Z digest=sha256:827dd39a32778e721ee286f400805dbe5ed62627cd871148ec849046d94ddf4b

Observation 2c82cd78-bd42-4df1-8c17-3405ca1d6741 · outbound

This paper cites FastRM: An efficient and automatic explainability framework for multimodal generative models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations FastRM: An efficient and automatic explainability framework for multimodal generative models

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:45:32.849467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:30.527268Z digest=sha256:68ac557aec891f087f64170b37203564c61bb3b6dc92de0817a8d59c0fac7fa8

Observation 3519c456-6106-4f20-b0df-74c2d34808f7 · outbound

This paper cites Explaining Multi-modal Large Language Models by Analyzing their Vision Perception.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Explaining Multi-modal Large Language Models by Analyzing their Vision Perception

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:45:32.614357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:30.624312Z digest=sha256:43d5ce810d9ccecd820cda84a0deaf275ea7db934aa42c4dd29512ea0b4d8efd

Observation 7089ab6c-d0a4-4247-8caf-b136bf755174 · outbound

This paper cites From Redundancy to Relevance: Information Flow in LVLMs Across Reasoning Tasks.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations From Redundancy to Relevance: Information Flow in LVLMs Across Reasoning Tasks

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:30.728861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:30.728861Z digest=sha256:dd981b72b8ff7b3318d97015f1d33459e87333dd74f07cd452edc5e8c792211a

Observation bc37ab4e-748e-45aa-aa7f-6b302367a133 · outbound

This paper cites Finding and Editing Multi-Modal Neurons in Pre-Trained Transformers.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Finding and Editing Multi-Modal Neurons in Pre-Trained Transformers

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:45:32.303375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:30.773041Z digest=sha256:4da017c94d6026c54d1c8120d45ba729d5739c6205b82471c0b53fd757922205

Observation 6ccb7e24-eaa5-4926-accd-cfd30a528366 · outbound

This paper cites AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:30.862604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:30.862604Z digest=sha256:8f50a25be7802e06eec0e0231397f73012476c51d82195e71f8da1a26e04d387

Observation 1889192c-40e8-4d30-96c2-56425dc377f8 · outbound

This paper cites Evaluating object hallucination in large vision-language models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Evaluating object hallucination in large vision-language models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:34.654011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:30.949942Z digest=sha256:58a08ddb62bd9ed53a105e014b59c40a91f9602440af0db7029fe39f44646ae5

Observation 3e5446db-90a3-421c-a606-095f94019fde · outbound

This paper cites Aligning large multimodal models with factually augmented rlhf.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Aligning large multimodal models with factually augmented rlhf

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:34.471205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:31.024274Z digest=sha256:1a66a76ca1440ee541ff77d65c5f75b22c1d1c72e64a8ac5bdf212c2e4c59a08

Observation 707ed996-fdbb-4892-94d2-a0dcd63b2e9c · outbound

This paper cites Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training.Adv.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training.Adv

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:34.299454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:31.070353Z digest=sha256:13bb8d72d904bd8f0765f4160d1ec1511c6003d312ff273939c2a935c17c16f6

Observation 6f82f695-ebfd-409f-8796-bc9795f604ae · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:31.150819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:31.150819Z digest=sha256:0fcc04da3158db23da10917da27ef308a45b7654e320c210f4dbf1369c2dffd1

Observation 4fe76543-d92b-4d9d-af14-a3bec28a87d2 · outbound

This paper cites Gqa: A new dataset for real-world visual reasoning and compositional question answering.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Gqa: A new dataset for real-world visual reasoning and compositional question answering

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:34.114986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:31.203608Z digest=sha256:bfc40e60b5bce7670753b0a437e534f1f80ae2d28d0eb438be1fd744c35841dd

Observation ebdf2516-b466-4eaa-ace5-f91f494677cb · outbound

This paper cites Beam Search Strategies for Neural Machine Translation.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Beam Search Strategies for Neural Machine Translation

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:31.289807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:31.289807Z digest=sha256:4da429290355cdfd8b8bcb0b93fa3c46a1d0ff098de05e11cabd6fb1f5bd0246

Observation 4de1a212-ef38-42f1-8a30-cf3c92771f3c · outbound

This paper cites DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:31.400584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:31.400584Z digest=sha256:5cb011c4b416f2d79974adb5398714feccc9e46d98e783bac999859bf563b03e

Observation f8bab591-d094-47bc-931f-2d936769cd25 · outbound

This paper cites Blended diffusion for text-driven editing of natural images.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Blended diffusion for text-driven editing of natural images

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:33.960642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:31.506177Z digest=sha256:7926d94a7b3e517e80705ab8554c6fb2be1ea77beeb54f054407a690a0e25dca

Observation b87fca78-67d1-42ec-9d67-d2cc6903bcab · outbound

This paper cites Unveiling typographic deceptions: Insights of the typographic vulnerability in large vision-language models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Unveiling typographic deceptions: Insights of the typographic vulnerability in large vision-language models

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:33.708879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:31.619307Z digest=sha256:b4151602956bbc399904b8e8e2fe3df0ee43c5447029646d64bb71403d610edb

Observation 2e495d30-7f69-4a63-92fe-e137a6360e93 · outbound

This paper cites Real-time single image and video super-resolution using an efficient sub-pixel convolutional neural network.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Real-time single image and video super-resolution using an efficient sub-pixel convolutional neural network

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:33.567854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:31.774717Z digest=sha256:107cfa2b9672b82899cb6bac339f555d08d0408f5f5687aaede5527019d0f56a

Observation f7264abe-3342-48d1-8953-f240882fb7ca · outbound

This paper cites Towards interpreting visual information processing in vision-language models.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Towards interpreting visual information processing in vision-language models

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:33.440316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T14:45:31.888895Z digest=sha256:be0ded4ace48b8cbac8a847c75d49c57e0d1272c690a9401bae353a61f7c2bb9

Observation bf173a22-c628-4ba6-8068-16487a728e86 · outbound

This paper cites GPT-4 Technical Report.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations GPT-4 Technical Report

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:31.991900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:31.991900Z digest=sha256:49cdf7cf4a58987606e998e11046b46986b4c48baf21f329b2d16b1747edb5c2

Observation e4e60157-5b8f-4138-bf26-7e5b1f2d07bd · outbound

This paper cites Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs

Reference 65

Resolution
malformed identifier
no resolver link, observed 2026-08-07T14:45:32.061891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:32.061891Z digest=sha256:8d0590c7c77de3028403f0b2cded0ae2209541455a08d55677d768d62179d23d

Pith citing papers

No inbound Pith citation observations are available.