Pith. sign in

Paper Citation Record · LEDGER

HourVideo: 1-Hour Video-Language Understanding

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2411.04998.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.04998 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:45:26.874122Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T16:03:08.061136Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation cbe8f332-5b99-4c76-8ab7-33c1155b89e5 · inbound

VideoChat-Flash: Hierarchical Compression for Long-Context Video Modeling cites this paper.

VideoChat-Flash: Hierarchical Compression for Long-Context Video Modeling HourVideo: 1-Hour Video-Language Understanding

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:02:43.666934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T04:02:43.261543Z digest=sha256:e52a891fab648dfd6c892b551429999bc61b614985155697eb4ec490c901c9c4

Observation c04387ad-e5a2-4f71-b3a1-cb412a049b00 · inbound

InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling cites this paper.

InternVideo2.5: Empowering Video MLLMs with Long and Rich Context Modeling HourVideo: 1-Hour Video-Language Understanding

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-17T02:52:20.694510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T02:52:20.643070Z digest=sha256:006e365bfc9362736a3d3ddcf133354cf8898d3aa948351715d60bed20cbf4ad

Observation 414f0402-51a4-494a-af35-088dcae167d2 · inbound

WorldSense: Evaluating Real-world Omnimodal Understanding for Multimodal LLMs cites this paper.

WorldSense: Evaluating Real-world Omnimodal Understanding for Multimodal LLMs HourVideo: 1-Hour Video-Language Understanding

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:53:26.359538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T05:53:26.066674Z digest=sha256:3f16b075b5ccebb993699ec8064966d3bc773c574f6843b6a161534ee887c88d

Observation ed00242b-b453-4eac-8022-3632c09b33c5 · inbound

CausalVQA: A Physically Grounded Causal Reasoning Benchmark for Video Models cites this paper.

CausalVQA: A Physically Grounded Causal Reasoning Benchmark for Video Models HourVideo: 1-Hour Video-Language Understanding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:44:58.871782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:44:58.871782Z digest=sha256:92ec423fda2cf2662deb5a5f56a999a4c8162c8294fe581be52ae6a86d405ab9

Observation 1be4414a-8ae4-4d38-8c76-e76aaa204ed4 · inbound

VideoConviction: A Multimodal Benchmark for Human Conviction and Stock Market Recommendations cites this paper.

VideoConviction: A Multimodal Benchmark for Human Conviction and Stock Market Recommendations HourVideo: 1-Hour Video-Language Understanding

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T10:45:26.874122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:45:26.874122Z digest=sha256:fa8c893b8dec7c819b3dd6eff249a40de974380e2c925c2e2c781ad7709cd8e9

Observation 01bcc5ab-e0c7-4408-a60e-48774759a8fe · inbound

Infinite Video Understanding cites this paper.

Infinite Video Understanding HourVideo: 1-Hour Video-Language Understanding

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:09:18.211447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:09:18.211447Z digest=sha256:8a8f5fb80c10423b57ec8cea84153e16e6fb2edc084df5861b59c9705a2b4551

Observation 4f25828d-e529-40a3-a226-d630579dc041 · inbound

EgoEverything: A Benchmark for Human Behavior Inspired Long Context Egocentric Video Understanding in AR Environment cites this paper.

EgoEverything: A Benchmark for Human Behavior Inspired Long Context Egocentric Video Understanding in AR Environment HourVideo: 1-Hour Video-Language Understanding

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:36:00.359960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T17:35:34.760659Z digest=sha256:6820b6472bc8d5989b0dbba617090dcc926deb2cc4aad9f8d2795818f13f5462

Observation 7d037126-b102-4049-b78a-c9a19bdd683b · inbound

EgoEverything: A Benchmark for Human Behavior Inspired Long Context Egocentric Video Understanding in AR Environment cites this paper.

EgoEverything: A Benchmark for Human Behavior Inspired Long Context Egocentric Video Understanding in AR Environment HourVideo: 1-Hour Video-Language Understanding

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T00:18:16.657829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:18:16.657829Z digest=sha256:ea98405878cfea24a0a3ced2637242056ab9e037407ae95a8415845ef2d8b0d4

Observation 864588dc-4e06-40b1-b981-5da984992834 · inbound

Minerva-Ego: Spatiotemporal Hints for Egocentric Video Understanding cites this paper.

Minerva-Ego: Spatiotemporal Hints for Egocentric Video Understanding HourVideo: 1-Hour Video-Language Understanding

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-19T16:03:08.063406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T16:02:53.887605Z digest=sha256:90069ca2655d94886976ea09ce5be341fa0489e805dfb875d663ab1fbca9a5cd