Pith. sign in

Paper Citation Record · LEDGER

VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2406.16338.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.16338 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:43:34.855411Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T10:27:56.151478Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 953d1368-6587-47f8-a72a-d3308781a7c1 · inbound

VidHal: Benchmarking Temporal Hallucinations in Vision LLMs cites this paper.

VidHal: Benchmarking Temporal Hallucinations in Vision LLMs VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-23T16:58:12.093330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T16:57:12.821916Z digest=sha256:683cc6e83ae54cfa096fd4094f1121a3c8ae02ca66d3d40592d661c3b9182c5c

Observation 07f4a18e-c5b7-4a42-880f-b2a1811df3c7 · inbound

Mitigating Behavioral Hallucination in Multimodal Large Language Models for Sequential Images cites this paper.

Mitigating Behavioral Hallucination in Multimodal Large Language Models for Sequential Images VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:43:34.855411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:43:34.855411Z digest=sha256:5189b6cd6060a708b6e20def48c433c8ce481f5d36e74586e3536124002ebef2

Observation d7d201a4-5e4f-4a6d-b11d-e3d890e46e7a · inbound

ARGUS: Hallucination and Omission Evaluation in Video-LLMs cites this paper.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.790898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.790898Z digest=sha256:74f01af6a5a5a541c9a031d80af4d9bb7d16fa3ab284c5419954b6c3f3888381

Observation ec42d65f-0f14-4190-a446-33c46f45a37f · inbound

FIFA: Unified Faithfulness Evaluation Framework for Text-to-Video and Video-to-Text Generation cites this paper.

FIFA: Unified Faithfulness Evaluation Framework for Text-to-Video and Video-to-Text Generation VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:15.872778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:06:15.872778Z digest=sha256:fc006aa941885390f90aadde31361a71a3652ffe5ef1114c4869ff6c2045bcf6

Observation 027c2733-a6de-4e19-b0b2-cc995b089554 · inbound

A comprehensive taxonomy of hallucinations in Large Language Models cites this paper.

A comprehensive taxonomy of hallucinations in Large Language Models VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-06T05:29:20.162727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:29:20.162727Z digest=sha256:521f8e3c471dd0f40e6b79349819206c19dd527549caae88f7a1ac072a4f8762

Observation 67395221-9853-4575-b4c2-3f03ae334068 · inbound

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey cites this paper.

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-05T20:28:49.973347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:28:49.973347Z digest=sha256:e1e510293e8d1558937e8e1ad0ccee88ba5fbcb150f9457add31ef7ce04a8627

Observation 14760259-ed39-4b68-b20a-974d045ae3cb · inbound

MESH -- Understanding Videos Like Human: Measuring Hallucinations in Large Video Models cites this paper.

MESH -- Understanding Videos Like Human: Measuring Hallucinations in Large Video Models VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-04T20:32:56.952259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:32:56.952259Z digest=sha256:b9966cf4884450049f021aae1f53476236a432e5a0abdbf318e9453b09bc9369

Observation d80fcd96-959f-4a72-9985-1e19fcb2a6e0 · inbound

Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models cites this paper.

Enhancing Video Representations with Spatiotemporal-Semantic Residual to Mitigate Hallucinations in Video Large Multimodal Models VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T06:36:26.055855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:36:26.055855Z digest=sha256:6ed9d3e907232aa0f8eeb7aec5f0d6b7d27415de7c11d5f64b4210480f1720b1

Observation c60caf0e-8f36-4278-a306-886a66ec9cd3 · inbound

Relaxing Anchor-Frame Dominance for Mitigating Hallucinations in Video Large Language Models cites this paper.

Relaxing Anchor-Frame Dominance for Mitigating Hallucinations in Video Large Language Models VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:26:03.552575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T14:54:52.710686Z digest=sha256:88846d5ab335a6c354c73861469027e2f6f07798c38a5b5ff88c9165743a281f

Observation dc1ac5f3-e5f9-4bb0-be03-28ec57215932 · inbound

When Text Hijacks Vision: Benchmarking and Mitigating Text Overlay-Induced Hallucination in Vision Language Models cites this paper.

When Text Hijacks Vision: Benchmarking and Mitigating Text Overlay-Induced Hallucination in Vision Language Models VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:46:37.525334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T06:41:59.641410Z digest=sha256:1b190548e8bc1fd53248fafbd650b4d32d201bdb798f16fa4677b7ed962a1582

Observation 233d9c0f-9cb1-4ffa-b4e5-83eee116964d · inbound

Raven: Rethinking Automated Assessment for Scratch Programs via Video-Grounded Evaluation cites this paper.

Raven: Rethinking Automated Assessment for Scratch Programs via Video-Grounded Evaluation VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:10:09.917066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T04:52:42.032751Z digest=sha256:2bf02cc0202c82e0b10b16737b23e19ccb1f7cd4b7903b9f5af0947bda990942

Observation 8ba7f7f4-e5b6-4f15-8692-f2d1bc920c28 · inbound

Video-ToC: Video Tree-of-Cue Reasoning cites this paper.

Video-ToC: Video Tree-of-Cue Reasoning VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:36:09.697431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T01:20:29.374012Z digest=sha256:b0382f6c8f6c5b2c71d60010a1e529a06e2160f861a843f4f8bce8e6fe72d66c

Observation 647867db-5ae3-4a31-b34e-049a3bf89270 · inbound

Towards Temporal Compositional Reasoning in Long-Form Sports Videos cites this paper.

Towards Temporal Compositional Reasoning in Long-Form Sports Videos VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-14T19:27:29.843866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T19:27:29.843866Z digest=sha256:8cdf6b3325386cf95c094cb2dfdc6760e4918c6c7a23f4430d17c1a1799f27ae

Observation 534d4747-67b5-42ac-b4ed-b35df1bc28f7 · inbound

From Priors to Perception: Grounding Video-LLMs in Physical Reality cites this paper.

From Priors to Perception: Grounding Video-LLMs in Physical Reality VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:21:08.287569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T17:41:23.233366Z digest=sha256:f3807f5b3eba49b3b43f84c07ccdf305114f68bb91d49972aebebe29f7250546

Observation e0c4e2a0-238a-45a5-9bf0-b0dd0ba8ad08 · inbound

TOC-Bench: A Temporal Object Consistency Benchmark for Video Large Language Models cites this paper.

TOC-Bench: A Temporal Object Consistency Benchmark for Video Large Language Models VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:31:26.781676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T04:13:21.487431Z digest=sha256:0300383155877678b8aad0bc9d82113745842e067fd3a4f698ae7c4a6c854dfb

Observation d1592a4c-4973-4d8d-9ed4-49b108b48f40 · inbound

TOC-Bench: A Temporal Object Consistency Benchmark for Video Large Language Models cites this paper.

TOC-Bench: A Temporal Object Consistency Benchmark for Video Large Language Models VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:57:28.252738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T06:53:42.726350Z digest=sha256:75ba80f03a42a023e08da6e38fcdbcdb3439aa126caca611557d3e3f8de7f582

Observation d7e427cf-6df5-443c-b8ee-d355a7adc5d1 · inbound

When Vision Speaks for Sound cites this paper.

When Vision Speaks for Sound VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:13:46.830651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T22:12:52.160596Z digest=sha256:1c7ed06d13169cd67daa83a54b44ea0cc1d4799a955ac54f7bd72bb45f15d241

Observation 800a68a7-01b2-4ef0-9d04-a977a9a6ee02 · inbound

OmniHalluc-L: Counterfactual Benchmarking and Modality-Perturbation Reliability Calibration for Long-Form Omni Hallucination cites this paper.

OmniHalluc-L: Counterfactual Benchmarking and Modality-Perturbation Reliability Calibration for Long-Form Omni Hallucination VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-02T06:06:41.663618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T07:31:51.780636Z digest=sha256:1cb21119bd049aebf17ddc99bbe1d2931ef6996b91c91e9ebc59866dbfb7ae7f

Observation 45772a24-4efd-4ff5-8619-d861dc3a8784 · inbound

MultiToP: Learning to Patch Visual Tokens to Mitigate Hallucinations in Video Large Multimodal Models cites this paper.

MultiToP: Learning to Patch Visual Tokens to Mitigate Hallucinations in Video Large Multimodal Models VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:27:56.152859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T10:02:58.341050Z digest=sha256:101816e876a3759f5ee6c794527c5b1fc8b8f54a0fe286a6742b6aab63975f15

Observation 14e98e60-218d-4b8f-b609-c45d9003ed08 · inbound

No Place to Hide: Benchmarking Video Hallucination with Background-Controlled Pairs cites this paper.

No Place to Hide: Benchmarking Video Hallucination with Background-Controlled Pairs VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 71

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:25:41.367812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-01T05:35:08.200219Z digest=sha256:846889896bfd99d559051f3ab370a36a1dbf397677e023f9d413c9d79b2bc8f8

Observation 59a397f0-75c9-4156-bc67-378eb53aaa3b · inbound

MoHallBench: A Benchmark for Motion Hallucination in Video Large Language Models cites this paper.

MoHallBench: A Benchmark for Motion Hallucination in Video Large Language Models VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T13:46:58.702099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-02T13:44:48.250838Z digest=sha256:ec1bba63319e052d2c2670239632df119c04531a8fcbbcacf059975a5a0eb5f1