Pith. sign in

Paper Citation Record · LEDGER

Video ReCap: Recursive Captioning of Hour-Long Videos

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2402.13250.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.13250 v6

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T12:23:40.491200Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0de19784-c60f-4729-b61c-ea3f715db1b5 · inbound

InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output cites this paper.

InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output Video ReCap: Recursive Captioning of Hour-Long Videos

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-17T10:46:28.890064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T10:46:28.447347Z digest=sha256:a63fd38c463acc937b730de8d733ec31c42942e6b291f90b86c6c3c0e4e1bde9

Observation 6e619aac-8417-4420-a274-05fc15f6c823 · inbound

LLaVA-Video: Video Instruction Tuning With Synthetic Data cites this paper.

LLaVA-Video: Video Instruction Tuning With Synthetic Data Video ReCap: Recursive Captioning of Hour-Long Videos

Reference 215

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:20:33.166464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T23:20:32.330351Z digest=sha256:d3692a9ee54d03a9f94945b83ba2097a8f8e4e1dbd66ad30ee08e6aea2f67197

Observation e1ddd6bc-bfa5-4fc8-9098-56725910e111 · inbound

LV-XAttn: Distributed Cross-Attention for Long Visual Inputs in Multimodal Large Language Models cites this paper.

LV-XAttn: Distributed Cross-Attention for Long Visual Inputs in Multimodal Large Language Models Video ReCap: Recursive Captioning of Hour-Long Videos

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T12:23:40.491200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T12:23:40.491200Z digest=sha256:98d1f1f34cb146ea4b1fc523377e0bc6fae19f678382ee2e4da7cc8d9bab8694

Observation c69bb0d7-baac-4ee6-80ee-9b06a3803f67 · inbound

VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering cites this paper.

VideoForest: Person-Anchored Hierarchical Reasoning for Cross-Video Question Answering Video ReCap: Recursive Captioning of Hour-Long Videos

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T04:49:38.388595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:49:38.388595Z digest=sha256:d4c4228152e5c7fa2ee9a40471a3300bfe650e1c82b54e44b15f31808ecf82c9

Observation bd7c0c34-4049-49dc-a022-2133816c0c6c · inbound

Rethinking RAG in Long Videos: What to Retrieve and How to Use It? cites this paper.

Rethinking RAG in Long Videos: What to Retrieve and How to Use It? Video ReCap: Recursive Captioning of Hour-Long Videos

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:28:33.973977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T06:30:33.428489Z digest=sha256:ec51da9d1a64b7475165d69d01fb513ce78998ac275486a11b41abe4f69597f2

Observation 766da962-601e-430c-8d41-ccd02614570e · inbound

Bridging the Usability Gap: Lessons from Interpreting Studies for Machine Interpreting Design cites this paper.

Bridging the Usability Gap: Lessons from Interpreting Studies for Machine Interpreting Design Video ReCap: Recursive Captioning of Hour-Long Videos

Reference 121

Resolution
verified exact
arxiv_id, observed 2026-06-27T03:40:27.860940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T03:34:37.373928Z digest=sha256:39dbdde360ee0eed5b68349d7b0736a1774c2e435a72e99a55bec665d0d1a2f9