Pith. sign in

Paper Citation Record · LEDGER

End-to-End Video Question Answering with Frame Scoring Mechanisms and Adaptive Sampling

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2407.15047.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.15047 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T15:56:37.569807Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T10:48:12.855049Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 11511695-2036-4371-bb90-7c06e4173f08 · inbound

ReasVQA: Advancing VideoQA with Imperfect Reasoning Process cites this paper.

ReasVQA: Advancing VideoQA with Imperfect Reasoning Process End-to-End Video Question Answering with Frame Scoring Mechanisms and Adaptive Sampling

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T15:56:37.569807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:56:37.569807Z digest=sha256:53e5e51bbcba757b566bbd5935aa71c589dff7a9ee17fb36ebfd19cd7e8a76c7

Observation c5b6ce89-4f84-4b2b-bdbd-d2ca24eee092 · inbound

Towards Effective Long Video Understanding of Multimodal Large Language Models via One-shot Clip Retrieval cites this paper.

Towards Effective Long Video Understanding of Multimodal Large Language Models via One-shot Clip Retrieval End-to-End Video Question Answering with Frame Scoring Mechanisms and Adaptive Sampling

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T00:11:23.171904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-17T00:10:01.596422Z digest=sha256:4cb0b9efe60af84db4ca405231421666c8ee880c76e8d5d10879a0bcc77f19a7

Observation ec4b263b-f997-4731-b04d-6a7916ed4723 · inbound

ForestPrune: High-ratio Visual Token Compression for Video Multimodal Large Language Models via Spatial-Temporal Forest Modeling cites this paper.

ForestPrune: High-ratio Visual Token Compression for Video Multimodal Large Language Models via Spatial-Temporal Forest Modeling End-to-End Video Question Answering with Frame Scoring Mechanisms and Adaptive Sampling

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T00:58:25.539956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-15T00:56:47.841355Z digest=sha256:70b613ff3f135b50345c19cdc121d1300649ac7e4c677ce3d98c499bb4b94790

Observation 005c2245-a687-447f-985f-a480817be04c · inbound

Scaling Video Understanding via Compact Latent Multi-Agent Collaboration cites this paper.

Scaling Video Understanding via Compact Latent Multi-Agent Collaboration End-to-End Video Question Answering with Frame Scoring Mechanisms and Adaptive Sampling

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:21:09.542544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-09T20:11:11.410051Z digest=sha256:7041597472ab27f0c3f297c34494ad23465ab601cbad79a0010080e54e716e4e

Observation b08716df-c478-4c33-97e5-9bd74edc60af · inbound

CRAFT: Critic-Refined Adaptive Key-Frame Targeting for Multimodal Video Question Answering cites this paper.

CRAFT: Critic-Refined Adaptive Key-Frame Targeting for Multimodal Video Question Answering End-to-End Video Question Answering with Frame Scoring Mechanisms and Adaptive Sampling

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:48:12.856768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-20T10:45:46.597723Z digest=sha256:78b7688ab5def16419f7eca121705cf7b3a2aba4377c1fd9e87cb6c67025b698