Pith. sign in

Paper Citation Record · LEDGER

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models

As of 10 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 6 inbound Pith citation observations for arXiv:2507.00898.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.00898 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:13:15.203711Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-31T23:19:55.558166Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-10T12:15:01.137692Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy28
  • unresolved15
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 54a40df1-2516-4178-a8b7-89b767f966c6 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:13:13.761431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:13:13.761431Z digest=sha256:ecb7b1c8a42d9f8f140ca40b1c750c4fb244f65f25a332e40416d8bb378790a9

Observation cf12b557-b6fb-4513-8b7a-c7c528b116a9 · outbound

This paper cites Hallucination of Multimodal Large Language Models: A Survey.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Hallucination of Multimodal Large Language Models: A Survey

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T21:13:13.787529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:13:13.787529Z digest=sha256:e824cf2545e4376e8f8f884736b75fdc8740546e8ea275eadfc054964d19ef44

Observation f80870c9-3378-4174-a087-7dab42e349b1 · outbound

This paper cites Detecting and Evaluating Medical Hallucinations in Large Vision Language Models.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Detecting and Evaluating Medical Hallucinations in Large Vision Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T21:13:13.824166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:13:13.824166Z digest=sha256:722fe957742112c114dda15bd2e41b389fd3276b420a00b371532c374fd2f184

Observation 99ebc0c0-e415-4a54-88e2-7862ad04a3f1 · outbound

This paper cites An image is worth 1/2 tokens after layer 2: Plug-and-play inference acceleration for large vision-language models.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models An image is worth 1/2 tokens after layer 2: Plug-and-play inference acceleration for large vision-language models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.727760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:13.861387Z digest=sha256:5b15ab23d1d07bf5a437b172c17e45f0578dcf61ed841c9602f08507b7c642d1

Observation 3b4151fa-716d-498f-a31f-ef1c5a242ddb · outbound

This paper cites Mitigating Hallucination in Visual Language Models with Visual Supervision.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Mitigating Hallucination in Visual Language Models with Visual Supervision

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T21:13:13.897579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:13:13.897579Z digest=sha256:dbaefc77e9b651e8f754d77a47f0e0cbd6d5fb42399eee7e1589e41157252825

Observation 5b0ecc14-237b-489e-b3d7-f6d77b201b2c · outbound

This paper cites HALC: Object hallucination reduc- tion via adaptive focal-contrast decoding.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models HALC: Object hallucination reduc- tion via adaptive focal-contrast decoding

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.717883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:13.934353Z digest=sha256:810fa05480b90218e53e3f5871b7eae5aec93370c848567deb5a3d3fd8ed5028

Observation 2be08dd1-ceb6-44cc-ad33-8783f28d5c2d · outbound

This paper cites Gonzalez, Ion Stoica, and Eric P.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Gonzalez, Ion Stoica, and Eric P

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.708025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:13.969904Z digest=sha256:460dc69de39362ba12ff172a68fa971057ed2d11b21ae8f78a27a3c80d6fd396

Observation 256517e9-dd7f-4450-bc3b-50563b81a798 · outbound

This paper cites Glass, and Pengcheng He.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Glass, and Pengcheng He

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.698441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.006179Z digest=sha256:7103629f4fa83b95887ef8ca75c970aeccb194b395397d2779f4d9585c4ba000

Observation 093aa4d0-a7ba-4a80-8db9-18e1612da0c3 · outbound

This paper cites Instructblip: towards general-purpose vision-language models with instruction tuning.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Instructblip: towards general-purpose vision-language models with instruction tuning

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.688615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.032869Z digest=sha256:20057e34b65dacc05093d3e677375cb40becab1763fd75584e4cddce02aef1f7

Observation eb89ff8c-60ac-4469-9836-8e2cee858cf2 · outbound

This paper cites Multi-modal hal- lucination control by visual information grounding.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Multi-modal hal- lucination control by visual information grounding

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.679229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.071610Z digest=sha256:eb6961d51e36e2307326bdf2e81d62e08b06b36d0aeecd167a7cf2e5b1da22d0

Observation 8f0e428b-cbc4-4ed6-9ca8-f403f4b1988d · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T21:13:14.107576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:13:14.107576Z digest=sha256:57dbf725052162233f6be01b8422ad3955f7df28ebac0fad52ec263d3e301d29

Observation 543dcbec-0950-4645-95d6-6b56d924b169 · outbound

This paper cites Opera: Alleviating hallucination in multi- modal large language models via over-trust penalty and retrospection-allocation.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Opera: Alleviating hallucination in multi- modal large language models via over-trust penalty and retrospection-allocation

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.669763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.144828Z digest=sha256:717097d2ee7b40caa1d00a4a5d71cc0da3b68199ba18becf528804645528d3d5

Observation bdaadd7d-9954-4028-bc56-d7841f652ced · outbound

This paper cites Gqa: A new dataset for real-world visual reasoning and compositional question answering.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Gqa: A new dataset for real-world visual reasoning and compositional question answering

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.659908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.170930Z digest=sha256:f6229c5f7233fe2b7bfa280bca57f8ed03c59398300c5b28fd7b38bf87ad196c

Observation ed9b2e5f-978d-46c9-9702-32cb4ec8cae0 · outbound

This paper cites Hallucination augmented contrastive learn- ing for multimodal large language model.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Hallucination augmented contrastive learn- ing for multimodal large language model

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.649937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.205144Z digest=sha256:c1d562bb556386f7bb242f64c8f93fa555745201e865adec6a9b6e71f0da6bae

Observation 583c3cb7-738e-4a56-b79a-fcf6a9558645 · outbound

This paper cites Lisa: Reasoning segmentation via large language model.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Lisa: Reasoning segmentation via large language model

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.639318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.242751Z digest=sha256:7c7c71a3751219858036dbc231077485104214a6d89dc65a3931e7807909a518

Observation 5c442e0b-a577-40d6-a1f1-a09d6a090f20 · outbound

This paper cites Mitigating object hal- lucinations in large vision-language models through visual contrastive decoding.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Mitigating object hal- lucinations in large vision-language models through visual contrastive decoding

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.628080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.278847Z digest=sha256:8cba1d8cb55e2714b3a2d0e3e7b5496329b40153db7cc9336eaf37664ba28daa

Observation b97d9b17-6ab7-4e70-b4ef-87d2540261f7 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.618160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.304998Z digest=sha256:7b3b175813a28a63aeb3a46a38206d035218291f88e0d9e7b20e8bae2fdd27df

Observation 03167bb8-1ec4-4cda-995e-88df1b76f23f · outbound

This paper cites Contrastive decoding: Open-ended text gener- ation as optimization.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Contrastive decoding: Open-ended text gener- ation as optimization

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.608061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.321463Z digest=sha256:3035ef67473bb49fafa7356a9e8ea65d82ac34a08cefce0c7f11bc891382966f

Observation 4fc8d79d-0eb5-4d26-8c4f-b6e7a41b6323 · outbound

This paper cites Evaluating object hallucination in large vision-language models.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Evaluating object hallucination in large vision-language models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.598663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.347124Z digest=sha256:bb4b7c223c6f5bf111c69669bf52134ef38cea6ec6099b4609e4ed626569bb21

Observation 026a6b61-09d3-41cf-8c6c-14f05c8821a8 · outbound

This paper cites Microsoft coco: Common objects in context.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Microsoft coco: Common objects in context

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.589453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.372432Z digest=sha256:ae14f015c8bd56025bbfbb9aeb6879c52d6378ba7766dbb656140cce5d139434

Observation 01eb5120-5060-40ec-bcd6-5ce2d8e0c117 · outbound

This paper cites Visual instruction tuning.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Visual instruction tuning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T21:13:14.397803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:13:14.397803Z digest=sha256:9d6b485f9d43ca68109952a7e195fb8953ebfd15cf54cd6cf4d179b195c7c62a

Observation 109f6128-b6be-4263-95d4-f207f849e1f8 · outbound

This paper cites Improved baselines with visual instruction tuning.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Improved baselines with visual instruction tuning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T21:13:14.433661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:13:14.433661Z digest=sha256:0b3d7270452abf842bb79a2f6297eacc75f814466af52e4951e318203be9ab97

Observation ba36bca9-c390-4210-b01b-1ade9c46a3c8 · outbound

This paper cites Llava-next: Im- proved reasoning, ocr, and world knowledge, 2024.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Llava-next: Im- proved reasoning, ocr, and world knowledge, 2024

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T21:13:14.468003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:13:14.468003Z digest=sha256:83944e4b6726810b618f7e47f86bb5cc0726560988e524fdbb87d4224c47d0e3

Observation 4c482140-d349-47a8-a602-d93e0750a797 · outbound

This paper cites A Survey on Hallucination in Large Vision-Language Models.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models A Survey on Hallucination in Large Vision-Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:13:14.495601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:13:14.495601Z digest=sha256:8a4f170c24688145356b4fb4253b127a11d3241b2efa65404f3591b28a959266

Observation 87fd5bb0-c528-430b-9d3f-d2d99c435ee5 · outbound

This paper cites Mmbench: Is your multi-modal model an all-around player? In European Conference on Computer Vision, pages 216–233.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Mmbench: Is your multi-modal model an all-around player? In European Conference on Computer Vision, pages 216–233

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.564040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.529961Z digest=sha256:18a8fa9c3dc03d070f2cbb4d82b4feb83d575eb102ee3780b17c5ee07e4b40eb

Observation 09e10d18-15f0-413f-b0f6-889c6133e5f1 · outbound

This paper cites Alleviating hallucinations in large vision- language models through hallucination-induced optimiza- tion.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Alleviating hallucinations in large vision- language models through hallucination-induced optimiza- tion

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.552998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.565418Z digest=sha256:f85569272590b2b0d7927cdab2d93973e03539b04fc9bce44fca8a69984c4302

Observation f4c5b4da-0667-4283-a1f5-f45fcb3d51c0 · outbound

This paper cites Learning transferable visual models from natural language supervision.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Learning transferable visual models from natural language supervision

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.543345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.601114Z digest=sha256:2cfcf1e8adc32103274a30502428c2552b6b8f1c82b3d856cdb6f3ad5282872f

Observation c8fada7b-f00a-4b14-afe4-91d329e5d795 · outbound

This paper cites Object hallucination in image captioning.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Object hallucination in image captioning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.533533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.635871Z digest=sha256:5f768c9c11aa26d11a6ee502e5765784203363be842efcc654d3d5c4ad261a42

Observation 66e7c1aa-db3b-4937-afc5-56edd760df57 · outbound

This paper cites A-okvqa: A benchmark for visual question answering using world knowl- edge.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models A-okvqa: A benchmark for visual question answering using world knowl- edge

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.523769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.672037Z digest=sha256:f67806ab9d865b7e1354fd589c701fa215f0778e53669b74b4b91ebfbc31f32b

Observation 5fc4fb16-9bf1-497b-af94-73d1a04bce6c · outbound

This paper cites A mathematical theory of communi- cation.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models A mathematical theory of communi- cation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.514298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.707820Z digest=sha256:e246dd95f7b5e9245264f3eaf4affba10f027200ef12a2044812036d15519c05

Observation 85aafc04-67f3-4324-882c-763922ad0371 · outbound

This paper cites Aligning Large Multimodal Models with Factually Augmented RLHF.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Aligning Large Multimodal Models with Factually Augmented RLHF

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T21:13:14.742673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:13:14.742673Z digest=sha256:a7107f1dac256b65288b9c778cd58d5c675acc4487bf7fbefeea9693c5496a63

Observation e8e30997-6934-4ac4-a3bc-b87b590a2939 · outbound

This paper cites Eyes wide shut? exploring the visual shortcomings of multimodal llms.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Eyes wide shut? exploring the visual shortcomings of multimodal llms

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:13:14.777285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:13:14.777285Z digest=sha256:e4da5c4be829ccf6d4611c287cc231fe8b809e2101a16934d9dede342961d1b4

Observation e4fb4c0d-d185-4227-9d9e-cd9db625a80a · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models LLaMA: Open and Efficient Foundation Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T21:13:14.813564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:13:14.813564Z digest=sha256:a9aa36358f8c7c509aa9844ff7f2345c84b97e93765d0dd8f71004adb01ffddc

Observation f53809e7-6b66-4a09-8b0f-842177dd009a · outbound

This paper cites InstructPart: Task-Oriented Part Segmentation with Instruction Reasoning.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models InstructPart: Task-Oriented Part Segmentation with Instruction Reasoning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T21:13:14.848661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:13:14.848661Z digest=sha256:6413c2fd629b3ac4d9e392dca5b8709582bd9594e62fed768448eb629e23fc0b

Observation e8261dd7-f68c-401e-8e59-cce4a7ec010a · outbound

This paper cites Mitigating hallucinations in large vision-language models with instruction contrastive decoding.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Mitigating hallucinations in large vision-language models with instruction contrastive decoding

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.499202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.884197Z digest=sha256:f4ef277b05a256707b139d1be3c3b6b47bfde4a3c77c13fb97cbe2ea7d5a1986

Observation 7d34cfcb-56d8-4cf1-b158-e7be95a71d82 · outbound

This paper cites Det- toolchain: A new prompting paradigm to unleash detection ability of mllm.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Det- toolchain: A new prompting paradigm to unleash detection ability of mllm

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.489926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.919150Z digest=sha256:c25f6906bd761b1aabbd81e0e55fbcf29ae3dc43ae64ff660fb303f575b9ef1b

Observation 67b05c53-e09d-4b6a-8e09-2fe8c13109b8 · outbound

This paper cites mplug-owl2: Revolutionizing multi-modal large language model with modality collaboration.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models mplug-owl2: Revolutionizing multi-modal large language model with modality collaboration

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.478822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:14.954513Z digest=sha256:a35cf748d438d17a6afea290bdc688c3b940873e73ed17862e590363032dd467

Observation 0c3c586c-1bdf-4b3f-91e5-02a2c06c0dc2 · outbound

This paper cites Woodpecker: Hallucination Correction for Multimodal Large Language Models.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Woodpecker: Hallucination Correction for Multimodal Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T21:13:14.989837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:13:14.989837Z digest=sha256:143202cc1fd53538ee8d88ca437b4e046b7f03aa41380db4cb63034e78edd220

Observation 92c490ae-92a8-4598-8987-ac2a04ccb19a · outbound

This paper cites MM-vet: Evaluating large multimodal models for integrated capabilities.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models MM-vet: Evaluating large multimodal models for integrated capabilities

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.468279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:15.026025Z digest=sha256:ece56eaca42b1e9849acb3f75919454f2c0042b5c892775817c19315e8460658

Observation 83132c60-9fdf-45fa-b3ac-b798d7762f33 · outbound

This paper cites Contextual object detection with multi- modal large language models.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Contextual object detection with multi- modal large language models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.458051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:15.061021Z digest=sha256:b0204e85166b2f290f5c61993759bf278d6747d079071fc33eaecfdf75228ce2

Observation 5ed954af-57ef-463e-b9d5-972cd331b0dd · outbound

This paper cites Incorpo- rating generative feedback for mitigating hallucinations in large vision-language models.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Incorpo- rating generative feedback for mitigating hallucinations in large vision-language models

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.447559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:15.095788Z digest=sha256:b71227d22b73028e44ba4cf5270c70b03772b9509bc57e953c083a80deed5925

Observation c2e08cec-7434-4bf9-98d3-41f949449c21 · outbound

This paper cites Vscan: Rethinking visual token reduction for efficient large vision-language models.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Vscan: Rethinking visual token reduction for efficient large vision-language models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T21:13:15.133064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:13:15.133064Z digest=sha256:4962a3fbc0c238bd0028d577860befa81b59ae50ed67c135952ec2111316c6a2

Observation 4b938062-275a-441e-87ee-3550f9b29499 · outbound

This paper cites Ma, Si- mon Stepputtis, Deva Ramanan, Russ Salakhutdinov, Louis- Philippe Morency, Katia P.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Ma, Si- mon Stepputtis, Deva Ramanan, Russ Salakhutdinov, Louis- Philippe Morency, Katia P

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:13:15.437162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:15.167107Z digest=sha256:cec42f860f5fe6527ffb2cc976031dac54d523e818d73be7c92b3be03376b01c

Observation 296f3db8-8ae8-4659-8d0f-59910f8d5aaa · outbound

This paper cites Is there a {object} in the image?.

ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models Is there a {object} in the image?

Reference 44

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T21:13:15.425541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T21:13:15.203711Z digest=sha256:073510a404c685e38fd1575228201feea6642430a8c6040335cd86431c4c3129

Pith citing papers

Observation a6b45afc-743a-42c1-a244-4009a2519ade · inbound

Revis: Sparse Latent Steering to Mitigate Object Hallucination in Large Vision-Language Models cites this paper.

Revis: Sparse Latent Steering to Mitigate Object Hallucination in Large Vision-Language Models ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:07:20.443129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T05:06:24.973439Z digest=sha256:8e76883fd50c25ac1307a6a08d5a92d13ca32e952071279fe250962ee636e4f3

Observation 6d6f5d86-d98e-401e-969c-655fe4468e46 · inbound

Mitigating Multimodal Hallucination via Phase-wise Self-reward cites this paper.

Mitigating Multimodal Hallucination via Phase-wise Self-reward ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:38:43.287163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T05:10:45.144421Z digest=sha256:5b37f9b409293c6a9bccce28a074998ecca73a0aa2168da51783eb13d474ce71

Observation 2e9493c2-8603-4956-a1e6-7a07beb9c123 · inbound

HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering cites this paper.

HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models

Reference 218

Resolution
verified exact
arxiv_id, observed 2026-05-09T23:54:45.699179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-09T23:51:47.724033Z digest=sha256:cc22a5e89a2adc93d7fee7b9c4993533643e9481a792468b3be2202d8f112577

Observation 8317148d-ad5d-4545-8321-0f26ba1e3ee3 · inbound

Mitigating Action-Relation Hallucinations in LVLMs via Relation-aware Visual Enhancement cites this paper.

Mitigating Action-Relation Hallucinations in LVLMs via Relation-aware Visual Enhancement ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T05:47:21.420670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-13T05:44:05.093926Z digest=sha256:ba8eb257e2ea413316bcbadf08724d926e3f104076e714ebf04ee1bd06633155

Observation 392264aa-5eb5-4da6-a67e-93caf0c54a51 · inbound

VisionPulse: Dynamic Visual Sparsity for Efficient Multimodal Reasoning cites this paper.

VisionPulse: Dynamic Visual Sparsity for Efficient Multimodal Reasoning ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-29T00:12:50.561358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T23:13:15.593994Z digest=sha256:3829bda89408d1f51d3b6a6dc3d8695ab8acb819449d7f8dba6022afb0c303b3

Observation 3b239e49-df4a-4aeb-92f3-5bd34ab9d73b · inbound

Disentangling Semantic Attention from Structural Bias in the Attention Manifold cites this paper.

Disentangling Semantic Attention from Structural Bias in the Attention Manifold ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-31T23:19:55.558166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:19:55.558166Z digest=sha256:5557295f789740167017ce51418c554e2b3aa8c6841d6d41ac5dca56cfdfd639