Pith. sign in

Paper Citation Record · LEDGER

Video Question Answering: Datasets, Algorithms and Challenges

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2203.01225.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2203.01225 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T17:16:12.459019Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T19:05:43.003963Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e9706425-97b1-4a35-8d4a-96e7f78f1eaf · inbound

Foundation Models and Adaptive Feature Selection: A Synergistic Approach to Video Question Answering cites this paper.

Foundation Models and Adaptive Feature Selection: A Synergistic Approach to Video Question Answering Video Question Answering: Datasets, Algorithms and Challenges

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T17:16:12.459019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:16:12.459019Z digest=sha256:bef6b50b7659e067a708e04983ed86408fd844bb35a22b2a52ad07f5317ec3b1

Observation f955dc62-daab-4f96-9a33-011b4e8341d0 · inbound

Neptune: The Long Orbit to Benchmarking Long Video Understanding cites this paper.

Neptune: The Long Orbit to Benchmarking Long Video Understanding Video Question Answering: Datasets, Algorithms and Challenges

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-11T16:58:29.918796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T16:58:29.918796Z digest=sha256:7f919be7e2c65a584870bc6cf2e62330339ca166b37b23f62124b0a49f336fa6

Observation f955df89-9382-4154-9cce-d308592807f8 · inbound

The Quest for Visual Understanding: A Journey Through the Evolution of Visual Question Answering cites this paper.

The Quest for Visual Understanding: A Journey Through the Evolution of Visual Question Answering Video Question Answering: Datasets, Algorithms and Challenges

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-10T20:51:53.864871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T20:51:53.864871Z digest=sha256:a721bec2322c3735079e9f09e12272b9325a13b3b71367f508c94726b3a5f09b

Observation 3f4b82f5-3fb3-4486-8689-5137f6397767 · inbound

VIBE: Annotation-Free Video-to-Text Information Bottleneck Evaluation for TL;DR cites this paper.

VIBE: Annotation-Free Video-to-Text Information Bottleneck Evaluation for TL;DR Video Question Answering: Datasets, Algorithms and Challenges

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:41.113713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:41.113713Z digest=sha256:3c20608eeca6e14b32023e96d98c42d9ea5c0a423396a4f81303b10b615a5ab5

Observation f84f3996-e75e-43f0-9d47-412ed472fd7d · inbound

Vid2Coach: Transforming How-To Videos into Task Assistants cites this paper.

Vid2Coach: Transforming How-To Videos into Task Assistants Video Question Answering: Datasets, Algorithms and Challenges

Reference 124

Resolution
unresolved
no resolver link, observed 2026-08-07T12:03:59.081375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:03:59.081375Z digest=sha256:13fb9d2ec221b6a77f58c685e7badc3757bd09a825e0b3365818ec40f9d7a1b5

Observation 8896bb95-5376-4b02-a7d7-7c0f7a5e5aa9 · inbound

Deep Temporal Reasoning in Video Language Models: A Cross-Linguistic Evaluation of Action Duration and Completion through Perfect Times cites this paper.

Deep Temporal Reasoning in Video Language Models: A Cross-Linguistic Evaluation of Action Duration and Completion through Perfect Times Video Question Answering: Datasets, Algorithms and Challenges

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T11:58:08.626602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:58:08.626602Z digest=sha256:c2896c2925ecd6bd4c6a77a1afb0e3e98f776ed4e3cda14667b5a5a326660da7

Observation dda93214-8ef0-4243-b154-ab40607d0c6f · inbound

VidEvent: A Large Dataset for Understanding Dynamic Evolution of Events in Videos cites this paper.

VidEvent: A Large Dataset for Understanding Dynamic Evolution of Events in Videos Video Question Answering: Datasets, Algorithms and Challenges

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:13.914353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:26:13.914353Z digest=sha256:b14eef65b14df59ae713290cc631ac9d843b016d957c384fa57511913dd7f18c

Observation 1e9f3a76-d7a7-4118-adb0-2a0b0df3377f · inbound

MAGNET: A Multi-agent Framework for Finding Audio-Visual Needles by Reasoning over Multi-Video Haystacks cites this paper.

MAGNET: A Multi-agent Framework for Finding Audio-Visual Needles by Reasoning over Multi-Video Haystacks Video Question Answering: Datasets, Algorithms and Challenges

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T05:49:53.335527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:49:53.335527Z digest=sha256:bb49d76fe74a26f3b86f991fa0b95d99eb44b66525bfa20d425b4e50b2b09236

Observation adce09e7-20c2-4d1d-bc7b-33b0047aae42 · inbound

CausalVQA: A Physically Grounded Causal Reasoning Benchmark for Video Models cites this paper.

CausalVQA: A Physically Grounded Causal Reasoning Benchmark for Video Models Video Question Answering: Datasets, Algorithms and Challenges

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T04:45:03.529058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:45:03.529058Z digest=sha256:2eb64c3abaecf680d8a231849fad6657f24ecff929098952149c1791c73c1547

Observation 93e3928d-95fa-4a33-b987-0437019a3569 · inbound

Exploring the Application of Visual Question Answering (VQA) for Classroom Activity Monitoring cites this paper.

Exploring the Application of Visual Question Answering (VQA) for Classroom Activity Monitoring Video Question Answering: Datasets, Algorithms and Challenges

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T11:49:11.706652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:49:11.706652Z digest=sha256:36c8e3429f6e020ffa654a4149d4f7c4c319e873409537b8b57bf2c7c9767bec

Observation 0b428c00-5abe-47ad-9039-1b4c412a6a56 · inbound

AURA: A Fine-Grained Benchmark and Decomposed Metric for Audio-Visual Reasoning cites this paper.

AURA: A Fine-Grained Benchmark and Decomposed Metric for Audio-Visual Reasoning Video Question Answering: Datasets, Algorithms and Challenges

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T22:09:16.203189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:09:16.203189Z digest=sha256:f5c8e7b3f1090339492b8bd75dace4dc97837a897ccc010d8989fea052b3add6

Observation 1d2b31b4-b6d9-4c54-a345-b2b73b09d13d · inbound

Mitigating Easy Option Bias in Multiple-Choice Question Answering cites this paper.

Mitigating Easy Option Bias in Multiple-Choice Question Answering Video Question Answering: Datasets, Algorithms and Challenges

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-05T19:05:43.113848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T19:05:38.903265Z digest=sha256:79b78fa6a58c8d06c614860848baae225bab1e430fb59360b14e464764139fe9