Pith. sign in

Paper Citation Record · LEDGER

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models

As of 15 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 4 inbound Pith citation observations for arXiv:2411.18659.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.18659 v2

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T11:31:37.008897Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:56:15.066164Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T01:22:15.675661Z

Reference resolution

54 of 54 outbound references displayed

  • verified exact1
  • verified fuzzy3
  • unresolved50
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 499c3666-32cf-4128-9dbe-d92df9f4b219 · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:31:37.971246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T11:31:36.730660Z digest=sha256:bd0bb207e0c89f9fd26f4e29ea6c51e10f499c355b56eb3e026c1935705c602c

Observation d4956da8-ec8b-4377-a435-ed0655e607da · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.736409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.736409Z digest=sha256:622a1a0ea6c4f0c983a788352e7448d4574882f64e6267a91341b8036a61201a

Observation a438c9cd-e55d-426b-b866-61b0be98ada4 · outbound

This paper cites Hallucination Detection in LLMs: Fast and Memory-Efficient Fine-Tuned Models.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Hallucination Detection in LLMs: Fast and Memory-Efficient Fine-Tuned Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.747097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.747097Z digest=sha256:666dcf5ce5035c5e65bfb470981010f3f093b4b43ad9b41f2d3629458f4fedcc

Observation 654dbb21-312c-4056-8b12-c632f9834227 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.752714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.752714Z digest=sha256:a0f3cf4c1975af62bcdc2f76d24a55ecfa2fcd6a2662a54005035dcc6c6d1c81

Observation c750fd4d-36ea-43d5-b263-48f2a893770a · outbound

This paper cites Qwen2.5-VL Technical Report.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Qwen2.5-VL Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.759544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.759544Z digest=sha256:aa186688194b8a36622ae06a96f99f66d943bad89b11a322c4181b61fff25a0b

Observation c7325881-c202-4989-98ae-2fc046f0ce71 · outbound

This paper cites MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.766220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.766220Z digest=sha256:607df384eec8b4a0a493a5867aa4ce0f2a417a4fcbd91ca05bc607ccd661ffd6

Observation 9d6fdff3-ffc7-46fd-a792-50acc5717037 · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.772424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.772424Z digest=sha256:e3e42131137b5fd8b0eb6aa441016052fd699e89bc16792dfef26dc908063bc2

Observation a143ffc7-19df-4326-8e2e-17f312bc2975 · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.778061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.778061Z digest=sha256:1d0d3f44f4eb674017b3644aa076028639972f8ba8435a6f6d9fb1eb257eaff1

Observation 2061ac3d-2903-4089-a301-4f39e00ba40e · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.783076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.783076Z digest=sha256:9eef4a4f42ca53a56f9add7ba43087f5aa21736801d375def12f32fc429f1bc3

Observation 27f0bc06-b78f-48b1-bfad-1c5b608fc1b1 · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:31:37.895672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T11:31:36.788333Z digest=sha256:7aa753c957b650f7dbb8479ef359c43771b917eef19312ced8470e23f23c2b58

Observation 4bc71004-17c6-4eae-a05d-433186372c2e · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:31:37.879324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T11:31:36.793276Z digest=sha256:5e6897c393c31d6863245bf6059c3fbc834c1848703d39d84a60fd5a9e4e676e

Observation af8baed6-0d19-4522-aa02-05b8029979a0 · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.798248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.798248Z digest=sha256:68f1f1a1d64bd38212776cc286de9797b3a8d718c85f4bdb1119e5077a1dd862

Observation 3d81195c-f5e1-404f-9886-7d4a5e3e4a69 · outbound

This paper cites LLaMA-Adapter V2: Parameter-Efficient Visual Instruction Model.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models LLaMA-Adapter V2: Parameter-Efficient Visual Instruction Model

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.803200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.803200Z digest=sha256:4e91c1632807f94fe9ab6539f7b26392807540eb75a7dceca6c307323eb7c079

Observation 44b9cab0-49c9-4914-a435-98f189917aa2 · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:31:37.852216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T11:31:36.808479Z digest=sha256:4b01bb9e535a0b8f088fc58b4989a5eeaba1cb485a208fa91979019d50eaad5b

Observation 7d95d495-6857-446d-9488-5a3ac051e2e9 · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:31:37.836049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T11:31:36.813247Z digest=sha256:e6d391706a4f0168409218ecf19af0a93e3f8dde65a99ca8d2fc117dd4458327

Observation eb01a31b-f810-4f07-a756-da51a011bdc9 · outbound

This paper cites OPERA: Alleviating Hallucination in Multi-Modal Large Language Models via Over-Trust Penalty and Retrospection-Allocation.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models OPERA: Alleviating Hallucination in Multi-Modal Large Language Models via Over-Trust Penalty and Retrospection-Allocation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.818292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.818292Z digest=sha256:db98452136203711f6e9f088229fd350823d5bdb59b5720cfb52e500dc9296e9

Observation cc7177b4-d5f7-4449-8d0e-ea8dd090d4f1 · outbound

This paper cites VCoder: Versatile Vision Encoders for Multimodal Large Language Models.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models VCoder: Versatile Vision Encoders for Multimodal Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.824376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.824376Z digest=sha256:94ce2493c52d44446fee1a998182ff82a3892ef87148db20624ecbf8fec06659

Observation 3ef5dc0f-3d12-42f3-a4d7-17ba2ca8ede0 · outbound

This paper cites Hallucination Augmented Contrastive Learning for Multimodal Large Language Model.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Hallucination Augmented Contrastive Learning for Multimodal Large Language Model

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.829741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.829741Z digest=sha256:acbefb7ddcf1f1ba7f3bf681adf39963c4f904add3f28aeca6f6f8e4542a7c8b

Observation 166c4e59-b61d-48aa-bac3-f2c373f20841 · outbound

This paper cites FaithScore: Fine-grained Evaluations of Hallucinations in Large Vision-Language Models.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models FaithScore: Fine-grained Evaluations of Hallucinations in Large Vision-Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.834922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.834922Z digest=sha256:cbb9b27e5216590021cf3f5dfcc0e48537cd471806b36312afcde3a5d342ddc9

Observation c8167132-eb82-498a-819a-a8f51223f014 · outbound

This paper cites Mitigating Object Hallucinations in Large Vision-Language Models through Visual Contrastive Decoding.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Mitigating Object Hallucinations in Large Vision-Language Models through Visual Contrastive Decoding

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.839992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.839992Z digest=sha256:5228754c5ec6a1d09396629259f705b82d74eccb5f11e01a772a8941fdce9b92

Observation c37c1114-f100-4bc3-a211-1a3d8f2b0762 · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:31:37.819450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T11:31:36.845417Z digest=sha256:e18796aed21630bba4a5ba41b4ab63bd1a7e3886615176670a6a35212efe9db9

Observation 3ad811b1-59ed-42ba-824e-212e5b1c3330 · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:31:37.803587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T11:31:36.850238Z digest=sha256:9d83989d34191eb1f81e5135959c5f179f82aea2ab03b69d626b8e4f289ed08c

Observation d5d72a91-35b8-4408-9c4c-628b2ff4fabc · outbound

This paper cites Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Monkey: Image Resolution and Text Label Are Important Things for Large Multi-modal Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.855085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.855085Z digest=sha256:abc66e183cd749d68016b8d9d139086e2889377ffd762b6e05099f7b9612e760

Observation 593fd2a9-ef8d-4bbf-ae41-381a43a7c741 · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.859995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.859995Z digest=sha256:bc1021ee158b32f5d3da3827d91475aa810d3554f1088c60742754616565f7a4

Observation 3790934f-cc3d-4b3a-b347-e4ed31c7f576 · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.864852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.864852Z digest=sha256:729c929e380f3a2291eff8f255ebcba93589432c85c4ac7a9d86e769a0b35cd8

Observation 7f392159-91f2-4f89-9065-3e7b02ddce29 · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:31:37.749943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T11:31:36.875381Z digest=sha256:14094635002f9a907d1336db17fce6ea9352acfe4ada02277a0ea487bc2dd5f1

Observation 786b42df-4020-422c-b4cf-c25782eb21c5 · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:31:37.733148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T11:31:36.880465Z digest=sha256:e5816d70cad05fdfda2e988e14ee9197a75df75c171be4cd92f8e7345e92d374

Observation 3807862f-ab23-4f51-ae6d-f3a2d1e797cb · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.885575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.885575Z digest=sha256:cf8071e74a01b473a330a0190a608915df8a1c92afd1383a29ad65094afd801d

Observation bd82b105-d4bb-43df-9399-c23c70cb1b3d · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.890534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.890534Z digest=sha256:c0f90ad68dad6d60c23e873824a0c96ec2c016fbb4092be1b0bf1c28f2166a80

Observation 75477e28-e5b3-4fee-82f3-7525647ec6b3 · outbound

This paper cites Evaluation and Enhancement of Semantic Grounding in Large Vision-Language Models.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Evaluation and Enhancement of Semantic Grounding in Large Vision-Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.900914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.900914Z digest=sha256:2878733c6830b655c58846f21d465bddd34decd13b979be578a8265f8890f766

Observation 17883d15-cd73-4f3f-a8e7-af50b1542d5e · outbound

This paper cites s1: Simple test-time scaling.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models s1: Simple test-time scaling

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.906011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.906011Z digest=sha256:76b583abdc6de6bebfc9c51c3e5aebf7be7161d0e3b9e21b39a1766a7586a239

Observation 1982a8d0-afc1-4459-a888-681f726f7f73 · outbound

This paper cites Negative Object Presence Evaluation (NOPE) to Measure Object Hallucination in Vision-Language Models.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Negative Object Presence Evaluation (NOPE) to Measure Object Hallucination in Vision-Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.895687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.895687Z digest=sha256:57cf93600874d12a55aed47393140ac2d5a98422671d6ba913f2e9f4ca70aa57

Observation 676b6f32-42bd-438b-a301-06c58d6b773d · outbound

This paper cites Object Hallucination in Image Captioning.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Object Hallucination in Image Captioning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.916111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.916111Z digest=sha256:96319ce3e27ab96e8b59ea81330927d260e0b3bb1dcb8db22366d6359d83a944

Observation 86cea28f-364f-4dd2-9df7-faab7fb73f0b · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:31:37.683933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T11:31:36.921202Z digest=sha256:a7dc7ef9126a6a10ac5dbbfc8a6b9697af121a40d6b6cbbb1205e095ac799c13

Observation 35aa13d1-f350-4ef1-ad3e-0171c5312499 · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.911327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.911327Z digest=sha256:699ed197907c0672aeac0295115ec34c9170ac157c0fc0c130dbcdcdd1218a65

Observation 51606332-d98d-4404-925f-76b85196ee57 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models LLaMA: Open and Efficient Foundation Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.931316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.931316Z digest=sha256:438a384b95186a7d03b3c0a71d12f181305903f3b6e4c1c17e5ccff5d5fdb6e5

Observation c71d2ebe-3ea6-4936-b779-59a82447ca96 · outbound

This paper cites AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.936383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.936383Z digest=sha256:70ce4f1e0184d59ba83d9bd3f9142a5801dc45181625f38d877890f5e20335af

Observation 0f5edae0-7556-460a-933f-b301f480d74b · outbound

This paper cites Aligning Large Multimodal Models with Factually Augmented RLHF.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Aligning Large Multimodal Models with Factually Augmented RLHF

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.926023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.926023Z digest=sha256:909ca4585c6c0694908ca43bf4f1921391c0320384416437b869e940971f7b59

Observation 0f8ddde5-a7d2-474d-94a6-35baac21fd4e · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:31:37.667022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T11:31:36.946535Z digest=sha256:ac264369a6ed6a3daba8737b3983b9f1ce924bca57182fb812bd8fd7acd35e0b

Observation 56f6a2c3-9332-49b5-b7e6-1495f9ba3d2d · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:31:37.649061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T11:31:36.951544Z digest=sha256:2cb269c459aaa91f8ab963a94fdbf4cd904ef6c74960fe6c301e304bb133a3f4

Observation fc128cf2-f180-4b52-9e0a-faa2596ae868 · outbound

This paper cites Evaluation and Analysis of Hallucination in Large Vision-Language Models.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Evaluation and Analysis of Hallucination in Large Vision-Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.941443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.941443Z digest=sha256:fe9f763867fce0cc160f74ba2878693ed9e0f000cfcc8e5dc22b606f1f298b8a

Observation 63d86a15-42d4-4949-a340-078ffb68f62b · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:31:37.632696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T11:31:36.962098Z digest=sha256:18214ef60fa7e69ee7a55cde7709cfdfc705fe1aadbd486ee5f4d674e08295e2

Observation 425b4d4c-d78a-4510-ac86-016cb18ec212 · outbound

This paper cites HallE-Control: Controlling Object Hallucination in Large Multimodal Models.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models HallE-Control: Controlling Object Hallucination in Large Multimodal Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.966962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.966962Z digest=sha256:96256c3e84950c2f30f8678f4d65b29009e83bceec00907bc9b92fc6fae3234b

Observation 8d3e5fe9-9fa4-44e2-af86-9841feadf88f · outbound

This paper cites Qwen2.5 Technical Report.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Qwen2.5 Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.956287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.956287Z digest=sha256:953c9f7cb6f9261a45019cbaa4081e267aaa1b8dbe95fc05d5551fd09e52a3c8

Observation d0374c8e-1ccc-4433-be04-b9b9293b6c8b · outbound

This paper cites The Blind Men and the Elephant.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models The Blind Men and the Elephant

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:31:37.616264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T11:31:36.978102Z digest=sha256:34e28696847417c7b49b006064eaf15978c5c03e515dc76c7a7fcbfb29b1c512

Observation 66567944-a73f-4492-ae96-3f9f27f2655e · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:31:37.600237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T11:31:36.982950Z digest=sha256:bd8fe31623992b114ca82b85414aadd84841d6e411ea8aebcba30d408944da7f

Observation 7e62cd8c-d72c-4dc2-9caa-b303e60a4058 · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models OPT: Open Pre-trained Transformer Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.972853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.972853Z digest=sha256:65765672694e959ec2af29442c417e6ddd77fcbbba5db38ebfafe8d4683e0d8d

Observation c54f0713-c945-4849-9cb3-d1c482b36fa2 · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.992973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.992973Z digest=sha256:fd59a21c63e17b768803bb3ee5a9129228536f513070434b4c485347c042e2b4

Observation c61a0575-9ec2-4f1a-8c5a-8c1e2a4e33b5 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:36.997741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:36.997741Z digest=sha256:e1991fbcf1b6d8576bfe8c7e60db0e62e8caf48c0e9069462771e88043e4d1e3

Observation 965eda8b-66d5-4f7d-a113-f8666d77d582 · outbound

This paper cites Fighting Fire with Fire (F3): A Training-free and Efficient Visual Adversarial Example Purification Method in LVLMs.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Fighting Fire with Fire (F3): A Training-free and Efficient Visual Adversarial Example Purification Method in LVLMs

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-08-12T11:31:37.188939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T11:31:36.987822Z digest=sha256:125baa70cf7953200300a26f7e1858a60498c32a35e9079186c637129fdecce5

Observation a4b51c0d-94f8-4035-9cef-d6190e43c8cc · outbound

This paper cites zero-cost.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models zero-cost

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:31:37.583911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T11:31:37.008897Z digest=sha256:a63766082eb8a161f89c0618cc6f52ec011564db7cdcf36c052adde53b99351f

Observation 2acead20-36bf-4cc9-817b-67cb41b4d5ff · outbound

This paper cites Connector-S: A Survey of Connectors in Multi-modal Large Language Models.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Connector-S: A Survey of Connectors in Multi-modal Large Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T11:31:37.003070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:31:37.003070Z digest=sha256:ff70705b0ad7c00972f2f97420eebd00f40bf75e33f9ef6f3566306276938e32

Observation 324d2128-39a0-4e1f-b1ad-f0b900a038ff · outbound

This paper cites NeurIPS 35 (2022), 23716–23736.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models NeurIPS 35 (2022), 23716–23736

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:31:37.944425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T11:31:36.741738Z digest=sha256:ea31a65e305ebb51a60d5c4adca0a2ecef64123db32e9722949bb7fc7b6329f2

Observation 04947c63-fcf5-4b56-a84b-1410bae5910f · outbound

This paper cites an unresolved cited work.

DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models Unresolved cited work

Reference 2023

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:31:37.766617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T11:31:36.870232Z digest=sha256:91c1828d9fb53e382a32db6ca6b5b41a63d9171dc69a628b1e33d9d8dae10f68

Pith citing papers

Observation ca405898-fb9b-4d10-be6a-bc60fadf4b13 · inbound

Fighting Fire with Fire (F3): A Training-free and Efficient Visual Adversarial Example Purification Method in LVLMs cites this paper.

Fighting Fire with Fire (F3): A Training-free and Efficient Visual Adversarial Example Purification Method in LVLMs DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:15.066164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:15.066164Z digest=sha256:2a083ce5b55a13ff0eb47da6f1df7d284d7e892bc812423ac28e88ded608c932

Observation a5d68a28-84fa-49d2-b8a5-9c74638b4c88 · inbound

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey cites this paper.

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models

Reference 184

Resolution
unresolved
no resolver link, observed 2026-08-05T20:29:01.463694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:29:01.463694Z digest=sha256:3178b12623980f7034c045e3f9d015b513b259246bda7c4014c92e14794bad61

Observation 96c2936d-b272-44d5-a83f-ba6c9623791d · inbound

MESH -- Understanding Videos Like Human: Measuring Hallucinations in Large Video Models cites this paper.

MESH -- Understanding Videos Like Human: Measuring Hallucinations in Large Video Models DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-04T20:32:57.027292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:32:57.027292Z digest=sha256:8bd7c7474e9732ca3a9cf419f428bd03ca20a63d9491d48b1f2f55ae61130c19

Observation 95833be3-a1a7-4f27-a468-78c6fa321086 · inbound

When to Trust the Answer: Question-Aligned Semantic Nearest Neighbor Entropy for Safer Surgical VQA cites this paper.

When to Trust the Answer: Question-Aligned Semantic Nearest Neighbor Entropy for Safer Surgical VQA DHCP: Detecting Hallucinations by Cross-modal Attention Pattern in Large Vision-Language Models

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:22:15.678535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T01:21:26.854636Z digest=sha256:f9f16f6fe46ab0bbc8cfe6a0ca84b8429d006e2953c9c3876042f97bb39e6e9c