Pith. sign in

Paper Citation Record · LEDGER

VideoAgent2: Enhancing the LLM-Based Agent System for Long-Form Video Understanding by Uncertainty-Aware CoT

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2504.04471.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.04471 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:20:13.504824Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:39:46.773095Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1bfc5510-31fb-46b2-a18e-27f3ea93819f · inbound

DIVE: Deep-search Iterative Video Exploration A Technical Report for the CVRR Challenge at CVPR 2025 cites this paper.

DIVE: Deep-search Iterative Video Exploration A Technical Report for the CVRR Challenge at CVPR 2025 VideoAgent2: Enhancing the LLM-Based Agent System for Long-Form Video Understanding by Uncertainty-Aware CoT

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T22:20:13.504824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:20:13.504824Z digest=sha256:7a5780f2775ce5a26a42f637df9844a12243f0119ac1a12a51425b001ac5e42b

Observation a2adb14f-9d42-4d38-b7e0-922e7efaf176 · inbound

GLANCE: A Global-Local Coordination Multi-Agent Framework for Music-Grounded Non-Linear Video Editing cites this paper.

GLANCE: A Global-Local Coordination Multi-Agent Framework for Music-Grounded Non-Linear Video Editing VideoAgent2: Enhancing the LLM-Based Agent System for Long-Form Video Understanding by Uncertainty-Aware CoT

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:20:51.141817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T19:13:30.400059Z digest=sha256:7d4d3832b09fc8eb00e49684c65fd8d7169deae180e7a770017cc11cca827037

Observation 665262fc-14de-4e6b-8a45-0ff88db0a8e8 · inbound

SVAgent: Storyline-Guided Long Video Understanding via Cross-Modal Multi-Agent Collaboration cites this paper.

SVAgent: Storyline-Guided Long Video Understanding via Cross-Modal Multi-Agent Collaboration VideoAgent2: Enhancing the LLM-Based Agent System for Long-Form Video Understanding by Uncertainty-Aware CoT

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:00:48.959183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T20:20:08.590407Z digest=sha256:b56a570922303f31530807c46c60bf1ea8fa88033b4eb2400b23cd9a078435ef

Observation f6717910-b0a6-45b5-aab0-a00dfcb494e7 · inbound

SagaQA: A Multi-hop Reasoning Benchmark for Long-form Narrative Understanding in TV Series cites this paper.

SagaQA: A Multi-hop Reasoning Benchmark for Long-form Narrative Understanding in TV Series VideoAgent2: Enhancing the LLM-Based Agent System for Long-Form Video Understanding by Uncertainty-Aware CoT

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T03:16:33.481200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T10:14:48.103490Z digest=sha256:92bc00045ce81cf1fcefef3eb37c7105079803c08a92687820e71b26e13075e4

Observation 7c727cf3-ca7f-48f2-8f7b-6cc87e0d11ea · inbound

Watch, Remember, Reason: Human-View Video Understanding with MLLMs cites this paper.

Watch, Remember, Reason: Human-View Video Understanding with MLLMs VideoAgent2: Enhancing the LLM-Based Agent System for Long-Form Video Understanding by Uncertainty-Aware CoT

Reference 182

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:27:15.702666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T22:00:28.350003Z digest=sha256:28ac822ea73890aed0a7979f7a61dce4a2b583fb80be5ff4955d384d079f131e

Observation 96c1d961-ebf1-443c-916e-2f4b09575ae8 · inbound

CoVStream: Edge-Cloud Collaboration for Understanding of Long Video Streams cites this paper.

CoVStream: Edge-Cloud Collaboration for Understanding of Long Video Streams VideoAgent2: Enhancing the LLM-Based Agent System for Long-Form Video Understanding by Uncertainty-Aware CoT

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:39:46.774494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T09:42:58.861599Z digest=sha256:69b70ca8e72535b05f6c7dbac1665d4b48be878427cf400d8ff09a322fa708cf

Observation 7c456548-7f5e-42d0-bab2-7fc368bcd735 · inbound

EventCoT: Event-centric Video Chain-of-thought for Reasoning Temporal Localization cites this paper.

EventCoT: Event-centric Video Chain-of-thought for Reasoning Temporal Localization VideoAgent2: Enhancing the LLM-Based Agent System for Long-Form Video Understanding by Uncertainty-Aware CoT

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-11T12:10:52.100410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T12:10:52.100410Z digest=sha256:592775557d3ffc6ad28798a34d6e449fe629eee9cc688cab0cedc191f27bbe0a

Observation 9845b1b6-cfa6-4b06-80cf-186d8d6f6ff4 · inbound

AgenticVAU: Multi-Agent Explore-Verify Reasoning for Video Anomaly Understanding cites this paper.

AgenticVAU: Multi-Agent Explore-Verify Reasoning for Video Anomaly Understanding VideoAgent2: Enhancing the LLM-Based Agent System for Long-Form Video Understanding by Uncertainty-Aware CoT

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T12:48:10.918578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:48:10.918578Z digest=sha256:6275dc2f579bdf5773bf8b5dbc1f1a7e74ed03a246ce5935b0ae31cda3b33311