Pith. sign in

Paper Citation Record · LEDGER

ARGUS: Hallucination and Omission Evaluation in Video-LLMs

As of 18 August 2026, this Paper Citation Record lists 83 of 83 outbound references and 3 inbound Pith citation observations for arXiv:2506.07371.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.07371 v2

Coverage vector

measured 83 of 83 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:41:39.882037Z

measured 86 of 86 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T19:27:29.843866Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-10T12:15:01.137692Z

Reference resolution

83 of 83 outbound references displayed

  • verified exact1
  • verified fuzzy32
  • unresolved49
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 590dd849-9fdc-4cb2-9c2a-a8d657e5fcdf · outbound

This paper cites https://huggingface.co/ blog/smolvlm2, 2025.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs https://huggingface.co/ blog/smolvlm2, 2025

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.567708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.567708Z digest=sha256:5d1a6d074737a3dc361ed2b8668839ed5a7a0712dd1aac9cd113b0e042670f65

Observation 53434356-bcfc-4a62-85aa-714e81b14705 · outbound

This paper cites GPT-4 Technical Report.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs GPT-4 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.572337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.572337Z digest=sha256:67b2e74ede7855d5568087e8447c6277b2725ce28856fff6501546444b81f89c

Observation 97e9437a-4105-4dda-b307-0f1fcec02fb2 · outbound

This paper cites Are language models better at generating an- swers or validating solutions?, 2025.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Are language models better at generating an- swers or validating solutions?, 2025

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.576307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.576307Z digest=sha256:6aba2f73191fe49b1d8a874c129169edb2306795722675a6d066087a8012493d

Observation 594055c8-92bc-44cd-8a98-098970bdff97 · outbound

This paper cites The snli corpus.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs The snli corpus

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.580023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.580023Z digest=sha256:a951572df0b6c90df56642b2624c84421ef070dede5638075881c664a06c3dfd

Observation 23e3db30-430d-4ab3-938e-ffee2a357ce6 · outbound

This paper cites Activitynet: A large-scale video benchmark for human activity understanding.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Activitynet: A large-scale video benchmark for human activity understanding

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:41.013891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.583724Z digest=sha256:84f28345abf6b804e9a66dd97820fcb51de61a1ca83e8d9d980b7a874c584233

Observation b0071cc1-0381-46c1-bc38-2c8f19ae1834 · outbound

This paper cites e-snli: Natural language inference with natural language explanations.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs e-snli: Natural language inference with natural language explanations

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:41.001815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.587569Z digest=sha256:5faa3b0f11a0d1daef86d6bc7cb69cbd2a116d674bc8c027802830fe04075d2b

Observation ea527a28-a167-4066-bada-0ccd3e18f566 · outbound

This paper cites AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs AuroraCap: Efficient, Performant Video Detailed Captioning and a New Benchmark

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.591343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.591343Z digest=sha256:5b70b1d478a9492f25d5564c650d0262a5b0f8a9a98d9f79c4cf41d497581227

Observation de01e0d6-572a-471e-be62-177d0913e4fd · outbound

This paper cites Panda-70m: Captioning 70m videos with multiple cross-modality teachers.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Panda-70m: Captioning 70m videos with multiple cross-modality teachers

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.990292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.595465Z digest=sha256:0224ddd09c40326818540ecf143aef97a792749703b13492e56f14afff10f6b8

Observation e3b75960-d454-47e7-94a1-33433d9612bc · outbound

This paper cites NeuralLog: Natural Language Inference with Joint Neural and Logical Reasoning.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs NeuralLog: Natural Language Inference with Joint Neural and Logical Reasoning

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:41:40.503375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.599610Z digest=sha256:5f5f7b2a4bda8a0a3f6def7b7fc29d611a23f2889c1bf71830a51d5bcddcdb60

Observation 8f774eab-0042-4aef-909f-e623f53af606 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.603674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.603674Z digest=sha256:89b2a20e25eca2cf3e59ead7ef60cebf66a8942e15de19cfc17a7b4ac6601803

Observation daada15a-9842-4e3e-9409-13fce30079e2 · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.607296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.607296Z digest=sha256:235891dd33f57b59b29bd2c3b60abfc856bfe23a4ae818c7c35f88f62ad0e272

Observation 8ff0d009-fb19-4dac-9bd3-b6c8c90a9cfa · outbound

This paper cites VidHal: Benchmarking Temporal Hallucinations in Vision LLMs.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs VidHal: Benchmarking Temporal Hallucinations in Vision LLMs

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.611326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.611326Z digest=sha256:126e699528587adcba7e4e89d26dfc9f68a8b4b29c154d0c665592324d581e55

Observation 8790372f-2667-4003-8322-fb3d0e44679f · outbound

This paper cites Transforming Question Answering Datasets Into Natural Language Inference Datasets.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Transforming Question Answering Datasets Into Natural Language Inference Datasets

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.615131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.615131Z digest=sha256:dd93a1dc4f39ec94539491220b8aa5a28eecd14f45977ea59f8e39081521f5eb

Observation cc096888-3b71-41e3-b9eb-acd30d47cd74 · outbound

This paper cites Sketch, ground, and refine: Top-down dense video caption- ing.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Sketch, ground, and refine: Top-down dense video caption- ing

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.978542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.618862Z digest=sha256:20fa16a323db847b04435eb0c4cc69cbd5870e41f55b18935ccab9a045cc0172

Observation 0f633d05-c3ae-4abc-99e4-cdd89b624e87 · outbound

This paper cites diffusers/shot-categorizer-v0.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs diffusers/shot-categorizer-v0

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.966853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.622535Z digest=sha256:4a11b1d28adda8a1fdae9523261d68bfa074f1c385043fab2e010c56a19c9df7

Observation 11436196-dc89-46e8-992c-55015e1c57bb · outbound

This paper cites Eva: Exploring the limits of masked visual representa- tion learning at scale.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Eva: Exploring the limits of masked visual representa- tion learning at scale

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.955434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.626122Z digest=sha256:b39a249080cf383125cd824259a65c8ecc062022568419ed0e84b1491662b406

Observation f9d3a925-7cf2-41e0-8c54-ce28d147afb5 · outbound

This paper cites TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs TC-Bench: Benchmarking Temporal Compositionality in Text-to-Video and Image-to-Video Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.629628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.629628Z digest=sha256:88a4a175ef86d528f18d281832908cb09a154d8e9112527599878249c5005c29

Observation 7f920541-4fe9-4695-b8a7-0ec9e816b893 · outbound

This paper cites Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.633535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.633535Z digest=sha256:9cb16db7de048f8a60169afb4f47db05463bb7c3295847a658a54241b1f067f5

Observation 604cd840-4811-4ca3-a5b2-7c32a906c890 · outbound

This paper cites TrueTeacher: Learning Factual Consistency Evaluation with Large Language Models.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs TrueTeacher: Learning Factual Consistency Evaluation with Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.637353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.637353Z digest=sha256:9008fe55804b5b65461dcc97b95c217453fd898bf7ddb2a4575420201bc6e467

Observation 853c9b63-4469-4c28-985b-029a6dc028ad · outbound

This paper cites The Llama 3 Herd of Models.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs The Llama 3 Herd of Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.641007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.641007Z digest=sha256:d5d4978b09a6215687f4a55017a10dee98f94ed554cec269b37557f7a02265d1

Observation e5e70f61-5368-4429-8731-bf0fa61fdcde · outbound

This paper cites Ego4d: Around the world in 3,000 hours of egocentric video.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Ego4d: Around the world in 3,000 hours of egocentric video

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.943884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.644523Z digest=sha256:0e8d88fae8e5176d913c46e89453de438e3f0a6988a05ff3d4366593fdb57118

Observation 497078f1-6181-4f4c-8d86-578b877f8d17 · outbound

This paper cites Hallusionbench: an advanced diagnos- tic suite for entangled language hallucination and visual il- lusion in large vision-language models.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Hallusionbench: an advanced diagnos- tic suite for entangled language hallucination and visual il- lusion in large vision-language models

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.932298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.648115Z digest=sha256:bf42ff20fe61fa1461a4d3d7c781039dc8775df27905ab52c5d4601b91055e38

Observation 8e79c9fe-81c7-4c4b-871e-10cfeb9b434c · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.652319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.652319Z digest=sha256:2c7ea3bbd152a6282676efdd5c2e7d6c8314c8e9676d5538ee87aab18e5e0dbc

Observation 78b48fdc-9398-45be-8cdf-acb5b6892294 · outbound

This paper cites CIEM: Contrastive Instruction Evaluation Method for Better Instruction Tuning.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs CIEM: Contrastive Instruction Evaluation Method for Better Instruction Tuning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.656753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.656753Z digest=sha256:d08762420a45d2652d3c89eeeb7fb657998a5530105116f99a4a1961ee6f839c

Observation 9d6a5c6b-4b6b-4c48-a86b-7691c995894f · outbound

This paper cites A Better Use of Audio-Visual Cues: Dense Video Captioning with Bi-modal Transformer.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs A Better Use of Audio-Visual Cues: Dense Video Captioning with Bi-modal Transformer

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.660603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.660603Z digest=sha256:02e1b326d44431d89d6361469c44c9710fa4c8fb7600934ac2f11636ec477cc0

Observation 842241d2-a3ab-4efa-8959-58b73cfcc793 · outbound

This paper cites Multi-modal dense video captioning.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Multi-modal dense video captioning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.920539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.664440Z digest=sha256:25ea047b2fd306545da7d87e1048471786300f231ca21af8c3f0b0ac2e53ccb6

Observation c9a6224b-e33c-4b5d-8107-1ccf937b0b42 · outbound

This paper cites FaithScore: Fine-grained Evaluations of Hallucinations in Large Vision-Language Models.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs FaithScore: Fine-grained Evaluations of Hallucinations in Large Vision-Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.667967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.667967Z digest=sha256:1aee0c6455d1dd345d9074d0597686a4440f44736b8792685b412dd83291663c

Observation bb8401a1-1690-4515-8742-939acb8bcdab · outbound

This paper cites Throne: An object-based hallucination benchmark for the free-form generations of large vision-language models.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Throne: An object-based hallucination benchmark for the free-form generations of large vision-language models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.908766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.672119Z digest=sha256:65eda12b76372f56de6ea42f61a42530e6843e692f9b66f768946ca114924a20

Observation dc8eae4b-c5fa-4bfa-a9b4-4288505245ed · outbound

This paper cites Do you remember? dense video captioning with cross-modal memory retrieval.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Do you remember? dense video captioning with cross-modal memory retrieval

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.897385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.675772Z digest=sha256:1d9c6176662f7b893c8bcb9f1d1ebe5c6a1c049030ebe21bff68b49c93216c23

Observation 0a44e6d2-41bf-4282-94e6-6f4739e841e9 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs LLaVA-OneVision: Easy Visual Task Transfer

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.679348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.679348Z digest=sha256:8cdc996fcb199501275780cd7635f6a023389cdad9d8807478fdf1422fdb7021

Observation 9e51ca5a-fd6e-436d-8e09-b1acf0016e1a · outbound

This paper cites LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.683048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.683048Z digest=sha256:81e7ecc9391ffedf441e2b23fcd66b5668d75da33aca736fb01f39c71480e112

Observation 3235d002-c515-44c3-bca8-56d0f099ac3c · outbound

This paper cites Jointly localizing and describing events for dense video captioning.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Jointly localizing and describing events for dense video captioning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.885454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.687410Z digest=sha256:d0805ac7ec71032cf1f0ce2df96ff8901270a1401f0140b93b6254900f5aff20

Observation 0faea484-9839-4129-a027-1a92b565eff9 · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Evaluating Object Hallucination in Large Vision-Language Models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.691477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.691477Z digest=sha256:570915fad64e0b5c7c080eed469932e8dab4cb74fb6874c5f297cacbb9aeb229

Observation ecbdd126-2c17-4950-8c3c-f65d8f2cea17 · outbound

This paper cites Revisiting the Role of Language Priors in Vision-Language Models.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Revisiting the Role of Language Priors in Vision-Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.695176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.695176Z digest=sha256:7c225eba2395115a7bf1d8bdd1103f9252bf61ef4598db600e6c472500e0a9fa

Observation f391b0fa-d002-4c69-bccb-db675f60a44d · outbound

This paper cites DeepSeek-V3 Technical Report.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs DeepSeek-V3 Technical Report

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.699034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.699034Z digest=sha256:4b475504cdbe21eb76870657a0be4ab3fd22fc4ebf464c42651ee5d54da9c101

Observation 64f98fc9-7c03-41f9-8be1-ef8aa579352d · outbound

This paper cites Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.702539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.702539Z digest=sha256:ba472e06b82331ccb44ea2f33912b0f068112983c7dc42871124fa23c948a5c8

Observation 7ccf135e-3002-481c-b3af-6b83f96d1452 · outbound

This paper cites A Survey on Hallucination in Large Vision-Language Models.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs A Survey on Hallucination in Large Vision-Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.706447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.706447Z digest=sha256:fbd04aaa8e3330ffb3a56439e3990c9087dba73574ae5b33fdded64bca3a3c28

Observation e2b20e15-00a1-4698-bfd0-95a6354a65c2 · outbound

This paper cites Logical Reasoning in Large Language Models: A Survey.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Logical Reasoning in Large Language Models: A Survey

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.710210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.710210Z digest=sha256:3c7d3246d8c0abf9f0ec114a875b85fc2f0d787e6296af57b32c03969328f810

Observation 9330aa65-f803-4b44-b874-5622c5d3b6e6 · outbound

This paper cites Egoschema: A diagnostic benchmark for very long- form video language understanding.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Egoschema: A diagnostic benchmark for very long- form video language understanding

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.873149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.714030Z digest=sha256:b41350c3478189d4613e24f5ed1f8719f4eb4bfca88f2f01f127d6eabaf89a9b

Observation d20a118c-b451-4449-931b-f595327e946e · outbound

This paper cites Streamlined dense video captioning.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Streamlined dense video captioning

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.861710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.717536Z digest=sha256:6ec52fcdc39346fadf904fe1a294db41c5ac89765046de05edacfcb051efd745

Observation 40e8b251-8ace-4e9d-ba64-119b03a54a5b · outbound

This paper cites Neptune: The Long Orbit to Benchmarking Long Video Understanding.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Neptune: The Long Orbit to Benchmarking Long Video Understanding

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.721179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.721179Z digest=sha256:216920797c1a1ef058c6ba71c69fb084dcb8eda8127e796becf56498f6a529e8

Observation d340bd51-3c6c-4172-b062-c3df7d13eb43 · outbound

This paper cites MINERVA: Evaluating Complex Video Reasoning.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs MINERVA: Evaluating Complex Video Reasoning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.724979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.724979Z digest=sha256:3f43602394c2895cb838266c6517a5273f6dee7f5f2eb05a9c52ed636ddf5f22

Observation 28a66fb7-a3a0-4adf-b23a-f80e97d8af3a · outbound

This paper cites an unresolved cited work.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:41:40.849414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.729103Z digest=sha256:a56f50d9db6a324678ab5f3aed2310e96f1835b7ea98321c0f8ef1fb50f47c43

Observation cb402e6e-1dd4-485f-83f1-c6ef69ce98b7 · outbound

This paper cites Per- ception test: A diagnostic benchmark for multimodal video models.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Per- ception test: A diagnostic benchmark for multimodal video models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.838364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.732461Z digest=sha256:70fea07bc62ff646a756a8d2dc6839f2478bd857196de08e5557e3c4bad60e8b

Observation fc36960e-67ce-4338-be28-82def6011f08 · outbound

This paper cites Dense video captioning: A survey of techniques, datasets and evalu- ation protocols.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Dense video captioning: A survey of techniques, datasets and evalu- ation protocols

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.827001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.736074Z digest=sha256:7095f00c8dacb3a8cd770cd4ac69433df23bec124a1021b6e115a8229f600043

Observation d9522d15-3ffc-4202-8232-b29cd8fb2c3f · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Learning transferable visual models from natural language supervi- sion

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.815178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.739626Z digest=sha256:bf398a36f2b1e496e1acea3934072f730ef1a19d36e6aee2d251647ee4d00e44

Observation 8b0b80fd-8d51-4c1e-8171-e57f3938b3aa · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs SAM 2: Segment Anything in Images and Videos

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.743102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.743102Z digest=sha256:022e5b6c3e721fa64082b602720b33393183d8efe1e54e600b9ace39f9edf017

Observation c327895c-27f6-407b-a9e4-89b92372a249 · outbound

This paper cites CinePile: A Long Video Question Answering Dataset and Benchmark.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs CinePile: A Long Video Question Answering Dataset and Benchmark

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.746918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.746918Z digest=sha256:5da269b129c33dc797f61a676a3a6ba02f5eefc7d3d4ea30815769d707ac88dd

Observation 3474a064-b49d-4ea5-a3d0-0a392fdbebd6 · outbound

This paper cites Object Hallucination in Image Captioning.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Object Hallucination in Image Captioning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.750672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.750672Z digest=sha256:84b3841f4152fd8e95834df957554861d7a6a09c572902ce89730afa8194fa13

Observation bfc6ec04-dc45-4cbe-b75c-726f0cd293ea · outbound

This paper cites FENICE: Factuality Evaluation of summarization based on Natural language Inference and Claim Extraction.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs FENICE: Factuality Evaluation of summarization based on Natural language Inference and Claim Extraction

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.754931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.754931Z digest=sha256:a45c0afb73140f4c424f6367cb1ad3273ddcd556f5a653388dc66465791a9ea6

Observation 968c5eb9-0ece-4a9d-b022-39ef207dc3b0 · outbound

This paper cites Aligning Large Multimodal Models with Factually Augmented RLHF.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Aligning Large Multimodal Models with Factually Augmented RLHF

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.758699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.758699Z digest=sha256:ef74ebb9789621784353a3a78fe52dbb0f1ad0b47d9c7061d54b558821c39e35

Observation 032d637b-cfd8-4f28-90b0-4a4a011601a5 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Gemini: A Family of Highly Capable Multimodal Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.762539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.762539Z digest=sha256:b581305bf4511046ab16331b94c262444eaecb79ad8c84dadf4abdf37fb912c7

Observation 78c9d01d-10fd-482b-a514-635336999edc · outbound

This paper cites Laion-aesthetics predictor v1.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Laion-aesthetics predictor v1

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.802807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.766700Z digest=sha256:46ede3b19c9ce7bb053f14d8b96883fc2ac0e1473e9fe461c1f1baf3504324aa

Observation b189f98f-a727-4500-ab51-687af6709944 · outbound

This paper cites Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.777404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.774382Z digest=sha256:f191b82472757c323f4f83c5f1b57c5d9b89faa279b7ac7510a9b3ddd27adad5

Observation beccbad5-5d4a-48a8-9403-44c96a1d2121 · outbound

This paper cites AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.778625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.778625Z digest=sha256:6572ce227b49632fce7c7a79520a84bd66d913a5d9d869d8e30f32b85a3a119f

Observation 272f68dc-8c74-48c3-841d-1b0aea410c5a · outbound

This paper cites End-to-end dense video captioning with parallel decoding.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs End-to-end dense video captioning with parallel decoding

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.764422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.782388Z digest=sha256:6deeb5bd51ed85be2bd326c82a2d17147bb11969bd6b781f8da5facbf05c04fa

Observation 46706f6f-7aae-4888-9b32-f568e6a899ee · outbound

This paper cites LVBench: An Extreme Long Video Understanding Benchmark.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs LVBench: An Extreme Long Video Understanding Benchmark

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.786245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.786245Z digest=sha256:57cceffe2539e085daba922fff87244a9d1ba9d839791975296444c4b8874656

Observation d7d201a4-5e4f-4a6d-b11d-e3d890e46e7a · outbound

This paper cites VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs VideoHallucer: Evaluating Intrinsic and Extrinsic Hallucinations in Large Video-Language Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.790898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.790898Z digest=sha256:b9eddf34564529a49b4dff101aa3da12d5bbdbbf2950e31e9fa405d2713bce65

Observation 237327f6-1c10-4e24-9a56-bc0da2a5377d · outbound

This paper cites A Broad-Coverage Challenge Corpus for Sentence Understanding through Inference.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs A Broad-Coverage Challenge Corpus for Sentence Understanding through Inference

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.795166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.795166Z digest=sha256:35185c04b459e13ad9c35dc9806e0a62e6ecf38512c18f9c457e82b3907fe619

Observation 7121c4c7-ecc8-4b16-883c-ddc3e5f5fa3b · outbound

This paper cites ANLIzing the Adversarial Natural Language Inference Dataset.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs ANLIzing the Adversarial Natural Language Inference Dataset

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.799211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.799211Z digest=sha256:50ed4193935a32ae135c1246e342a52c9992822a7e72be201f7849f384b11ced

Observation 9981f233-237e-49e4-ac79-383070b7792c · outbound

This paper cites Joint event detection and description in continuous video streams.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Joint event detection and description in continuous video streams

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.751725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.802826Z digest=sha256:b08a7c5a1bef04746ec8eae52001250bd055795e416fa775b3eae939ccfdc5ee

Observation c428ba8b-8618-4704-899e-92a1a8d914a6 · outbound

This paper cites Qwen2.5 Technical Report.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Qwen2.5 Technical Report

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.806096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.806096Z digest=sha256:392a20009cb5a0d50cfb82a15bc5bdcd82c20db169dc7d578c8f9357b0e3a4ec

Observation f4cc3595-ad19-4bfd-85b0-93f590950b96 · outbound

This paper cites mplug- owl3: Towards long image-sequence understanding in multi- modal large language models.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs mplug- owl3: Towards long image-sequence understanding in multi- modal large language models

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.739115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.809799Z digest=sha256:eb20762f6fcba3ec545b9db8bfd5ca4ccf62e4538af2affd3cdfc9109364d262

Observation 60c2cd09-2368-4c53-911d-ff2627913e71 · outbound

This paper cites Sigmoid loss for language image pre-training.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Sigmoid loss for language image pre-training

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.726197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.813223Z digest=sha256:ead247ec1a8910ac1142073f98681f38681a14be929dc560ca92ffb29bca95ac

Observation ab331a25-ef26-468e-ad59-7ec1fcd9dcda · outbound

This paper cites Eventhallusion: Diagnosing event hal- lucinations in video llms.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Eventhallusion: Diagnosing event hal- lucinations in video llms

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.816835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.816835Z digest=sha256:d950d43dca660fa148b3d12e2f1cd0f04bb15c8a251cd5b46fc23425c3d4809e

Observation 7b627c57-0fa4-443d-aa50-408e9acd89d6 · outbound

This paper cites Judging llm-as-a-judge with mt-bench and chatbot arena.Advances in Neural Information Processing Systems, 36:46595–46623, 2023.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Judging llm-as-a-judge with mt-bench and chatbot arena.Advances in Neural Information Processing Systems, 36:46595–46623, 2023

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.820651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.820651Z digest=sha256:b6153b7f49f980bb58ed81e16a322117b116b8c9c702bd90cbcfd0b1c0b85564

Observation d81ab276-303d-48da-a435-cf4ae4673279 · outbound

This paper cites Streaming dense video captioning.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Streaming dense video captioning

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.706493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.824330Z digest=sha256:58776b11fc5e974f196be6ba97ef2d496dfd43d4922d0388835def8dcebfd832

Observation 1645717c-29a1-4eb3-8e9f-e7e2978c2325 · outbound

This paper cites Apollo: An Exploration of Video Understanding in Large Multimodal Models.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Apollo: An Exploration of Video Understanding in Large Multimodal Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.827734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.827734Z digest=sha256:fcba052bc612d6bc799399b8a0c7cfb3bb426b5625920f40d772002b3ecca550

Observation 9bc41f06-770e-412a-8654-003eee6d9964 · outbound

This paper cites In our setting, we adapt this score by computing the average aesthetic score across all frames in a video.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs In our setting, we adapt this score by computing the average aesthetic score across all frames in a video

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.695010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.832859Z digest=sha256:49c045017fface9f9fcb648ad29c40dbee0f9b723cdcf88f17c077ce5f1e05a6

Observation a92f3c8c-df99-4e6f-aadd-e7b7c0d94160 · outbound

This paper cites diffusers/shot-categorizer-v0 [15] to identify the lighting type in each frame of a video.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs diffusers/shot-categorizer-v0 [15] to identify the lighting type in each frame of a video

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.683387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.836490Z digest=sha256:fa058ce8de43c0a49dff53cc4760db65d05d57b9e71ede3c5a1481ab5a6f1239

Observation b1979757-13bc-4feb-afa8-5ffe89da3e24 · outbound

This paper cites a person is cooking,.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs a person is cooking,

Reference 72

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T05:41:40.672044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.840081Z digest=sha256:9b11fb0dea97098af25f2d502066713b9f94495eb90870509e36b6ad9b0242d0

Observation b5dc5497-1941-4a4e-a037-0e819dc33584 · outbound

This paper cites an unresolved cited work.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Unresolved cited work

Reference 73

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:41:40.660632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.844299Z digest=sha256:4b262c7eda0182a749f9a792e2fba52737d09d9a61bc75bfb492ac41400c9510

Observation a1c91468-e8e7-4dcd-8dfd-409e2c306c54 · outbound

This paper cites - **Contradiction**: Contains information that directly conflicts and is unsupported by the source_caption.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs - **Contradiction**: Contains information that directly conflicts and is unsupported by the source_caption

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.649092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.848075Z digest=sha256:9e4c3899fdd70d301eff752b07a52c17076509ec6e936093f517c47eaa1af20e

Observation 727e4eac-d3d9-4867-a2e0-5c17aad0d6db · outbound

This paper cites "" {source_caption}.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs "" {source_caption}

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.637407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.851733Z digest=sha256:1c94e5072abb9d09fe5a602727eb5ba3851130312abcc4437d5c9068c25ae759

Observation 85d31883-22d3-4a3e-b194-0a7dc71a8c82 · outbound

This paper cites SUNFEAST PASTA TREAT,.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs SUNFEAST PASTA TREAT,

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.625690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.855285Z digest=sha256:de5b478296c788677d627e712561631537dc8fc3e730a3da03af3af3dd123c2b

Observation 874c3d87-e3e1-44de-ac49-63f4c3ab4484 · outbound

This paper cites an unresolved cited work.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Unresolved cited work

Reference 77

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:41:40.614091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.858986Z digest=sha256:48c738332623c9f830d7a723d324034333adcef67c61a4a0c8300229bf79f536

Observation 3ab7b8ee-7814-4c1e-aec7-01350acc8e0b · outbound

This paper cites Sunfeast Pasta Treat.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Sunfeast Pasta Treat

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.602040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.863928Z digest=sha256:93c100effc83ab8339bbb81f44d631ffe65b9c2f6535883fc3913f390af9373b

Observation 2699120f-4ad9-4311-94d9-262fa15f5fa6 · outbound

This paper cites an unresolved cited work.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Unresolved cited work

Reference 79

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:41:40.588373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.867425Z digest=sha256:17613658a85151425da22bb275373ad4b22f16074a1f88930763ae6668323404

Observation a3a9c83b-57be-44b6-9dae-4d84078b2644 · outbound

This paper cites an unresolved cited work.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Unresolved cited work

Reference 80

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:41:40.576484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.871096Z digest=sha256:ccea43db7031e71bad8fcbc4edcb29ff6d70a5189ac9f23089de3da0debef29e

Observation ed7cf2fa-d875-45eb-9b67-1e041f8e2702 · outbound

This paper cites an unresolved cited work.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:41:40.564348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.874517Z digest=sha256:74706071d0cf08f488745c98acd69b270993122013524c7f45dfbc5652f937af

Observation 87586a80-fd26-450d-8225-09c6a45c0483 · outbound

This paper cites Recipe Card.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Recipe Card

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.552588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.878424Z digest=sha256:6e4affd62952d176971b7908c14a0eaa675f8629fb938e804900e8bffaa65459

Observation fb88127b-ae8a-4e03-a2b4-ac02c886da56 · outbound

This paper cites Quick and Easy.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Quick and Easy

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:41:40.540599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.882037Z digest=sha256:33d7991c065d29ba3f315748332ee04c13ffcd6d2cebcc9a3e2e8206ecd63f47

Observation 278b13a9-a20b-4cf0-8118-c70e186f7894 · outbound

This paper cites an unresolved cited work.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Unresolved cited work

Reference 2022

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:41:40.790525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:41:39.770832Z digest=sha256:35106a35f3ca0596aa097a5473cabeb45d7da5dbb690149707fdb25d4bbc627b

Pith citing papers

Observation 5b0bf116-a3cc-4bbf-a064-779e84b2e2d4 · inbound

Towards Temporal Compositional Reasoning in Long-Form Sports Videos cites this paper.

Towards Temporal Compositional Reasoning in Long-Form Sports Videos ARGUS: Hallucination and Omission Evaluation in Video-LLMs

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-08T21:34:15.466797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-08T12:45:50.422679Z digest=sha256:e332d0cb42296a4b3502d5532f41924df94652c3c5bdcfe8cf3973a5fa9a741c

Observation a51a5976-a148-47b5-b556-9b23dbc5f62a · inbound

Towards Temporal Compositional Reasoning in Long-Form Sports Videos cites this paper.

Towards Temporal Compositional Reasoning in Long-Form Sports Videos ARGUS: Hallucination and Omission Evaluation in Video-LLMs

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-14T19:27:29.843866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T19:27:29.843866Z digest=sha256:d5f4e7bbf267e6ed161f9c6f198a1514c78ac00f4e89b9ee33029bf92c9f970f

Observation 3cd54224-1dff-4d00-bfe9-a76e6a256085 · inbound

VCap: Hypergeometric Rewards for Weak-to-Strong Visual Captioning cites this paper.

VCap: Hypergeometric Rewards for Weak-to-Strong Visual Captioning ARGUS: Hallucination and Omission Evaluation in Video-LLMs

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:23:28.469585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T13:13:57.599970Z digest=sha256:b1df35363faa718d578d3415696c31a1ff526518fa9f0f7966597a42b3f90518