Pith. sign in

Paper Citation Record · LEDGER

DreamFrame: Enhancing Video Understanding via Automatically Generated QA and Style-Consistent Keyframes

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2403.01422.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.01422 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T00:23:08.106301Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T06:39:37.586156Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation eec8055e-32e7-4197-bc63-f3613be186f8 · inbound

MLVU: Benchmarking Multi-task Long Video Understanding cites this paper.

MLVU: Benchmarking Multi-task Long Video Understanding DreamFrame: Enhancing Video Understanding via Automatically Generated QA and Style-Consistent Keyframes

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-14T19:55:26.445584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T19:55:26.333923Z digest=sha256:70b97600f479713cacfa95e812bc96228879d37ba8b0daf27f084cc56b4305c7

Observation a3bf9cf7-8906-4a9c-acc7-b93a7426b70e · inbound

VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding cites this paper.

VideoLLaMA 3: Frontier Multimodal Foundation Models for Image and Video Understanding DreamFrame: Enhancing Video Understanding via Automatically Generated QA and Style-Consistent Keyframes

Reference 149

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:20:00.076582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T01:19:59.603343Z digest=sha256:d9200c1e0ff2be855e493844ac3030ca0398e550a9c2f80ec47bac1e8040d6a5

Observation 3f50a790-a285-458a-a68b-6cf1940de7c4 · inbound

UniForm: A Unified Multi-Task Diffusion Transformer for Audio-Video Generation cites this paper.

UniForm: A Unified Multi-Task Diffusion Transformer for Audio-Video Generation DreamFrame: Enhancing Video Understanding via Automatically Generated QA and Style-Consistent Keyframes

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T00:23:08.106301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:23:08.106301Z digest=sha256:2f14cf6a2913905b536a6c49c8bf0dbe6b278391c78a4d959bda7025c81fdd70

Observation 2abf08f3-9364-4881-ab45-3467b07caf6a · inbound

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey cites this paper.

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey DreamFrame: Enhancing Video Understanding via Automatically Generated QA and Style-Consistent Keyframes

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:30:57.084788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T16:36:33.264166Z digest=sha256:53cf0eb23c26ebb81b64c804b2d273d58b92058594cf1de4f8921c40d3ac6f43

Observation 5dde9bc3-9171-4c61-bc3a-f7ffe8c86a11 · inbound

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey cites this paper.

Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey DreamFrame: Enhancing Video Understanding via Automatically Generated QA and Style-Consistent Keyframes

Reference 60

Resolution
unresolved
no resolver link, observed 2026-07-12T22:04:31.302192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T22:04:31.302192Z digest=sha256:5febe6ff83ea78d869c59c13b18bd1e7a51da1a15e48be9160dcc41269a6a816

Observation ce518f8b-60c1-4ac8-9543-4306ad6383f2 · inbound

Towards Effective Long-Video Event Prediction via Multi-Level Event Semantics Mining cites this paper.

Towards Effective Long-Video Event Prediction via Multi-Level Event Semantics Mining DreamFrame: Enhancing Video Understanding via Automatically Generated QA and Style-Consistent Keyframes

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T00:02:50.348153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T23:16:42.001361Z digest=sha256:72712b788430dd9c4d1d199f94432e009d20cda057627066fa0eee8b5cc4033a

Observation 7c586a91-264b-4629-b4b8-0e6c2741cbfc · inbound

HPP: Hierarchical Programmatic Probing for Long Video Understanding by Decoupling Perception and Reasoning cites this paper.

HPP: Hierarchical Programmatic Probing for Long Video Understanding by Decoupling Perception and Reasoning DreamFrame: Enhancing Video Understanding via Automatically Generated QA and Style-Consistent Keyframes

Reference 152

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T06:39:37.587883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T14:19:53.450263Z digest=sha256:d9eb88e0749bd256b24f4a93bd891d4754ac7762d4226a00288f302c4496985a

Observation 0dcbb900-7f09-48eb-9572-e28596bcc381 · inbound

Antigen-specific Antibody Multi-modal Foundation Model for Functional Antibody Design cites this paper.

Antigen-specific Antibody Multi-modal Foundation Model for Functional Antibody Design DreamFrame: Enhancing Video Understanding via Automatically Generated QA and Style-Consistent Keyframes

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-01T10:59:17.682021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:59:17.682021Z digest=sha256:2a277166facad645abb295ffcc71057822c44e84ab144d560f8482cfd69c555d