Pith. sign in

Paper Citation Record · LEDGER

Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2503.12496.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.12496 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:33:58.078291Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T23:39:04.701981Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation eab857a9-1f34-4e9a-82cf-91757979e69b · inbound

Mitigating Object Hallucinations via Sentence-Level Early Intervention cites this paper.

Mitigating Object Hallucinations via Sentence-Level Early Intervention Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-25T08:35:32.671787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-25T08:31:24.173135Z digest=sha256:1756b9598db3badc542d3b8090122655ff870d8a8bd18cedf921cd455248d891

Observation d074ebb3-c178-4196-b9c9-61f8004e1eb0 · inbound

VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning cites this paper.

VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:58.078291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:33:58.078291Z digest=sha256:1fdc54126cf10b0a1a7c9e9a4539eee4f6f908d5d9d1ff2923dcf1cce52ac081

Observation 8be86950-d6de-4d84-bf9c-243eb57add5d · inbound

EgoSelf: From Memory to Personalized Egocentric Assistant cites this paper.

EgoSelf: From Memory to Personalized Egocentric Assistant Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:06:03.813089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T02:23:21.119521Z digest=sha256:08dd2355353db5d9e1c73e21d21843740322f328e420327d86e745bb82bc5507

Observation fb71b523-eaff-48d3-87a1-a6f35c8a1649 · inbound

Uni-OPD: Unifying On-Policy Distillation with a Dual-Perspective Recipe cites this paper.

Uni-OPD: Unifying On-Policy Distillation with a Dual-Perspective Recipe Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:26:16.162141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T17:00:49.448352Z digest=sha256:0757eb5861729e26563783e9c905f5e9fb439761bf1dd2ba86f9c37c2581d355

Observation a0eb97c8-35c1-4a53-b269-e86c34f5582c · inbound

Video-Zero: Self-Evolution Video Understanding cites this paper.

Video-Zero: Self-Evolution Video Understanding Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-06-30T21:35:04.432104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-30T21:32:16.939563Z digest=sha256:cb97783cc0cabb361697a9e274c93923dd03b3cb8aea406b67190afe698dc152

Observation d483c9cc-90a3-4602-9dbd-c336ec92fc5f · inbound

How Well Can Your Video Model Remember? Measuring Memory-Budget Trade-offs in Long Video Understanding cites this paper.

How Well Can Your Video Model Remember? Measuring Memory-Budget Trade-offs in Long Video Understanding Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T23:39:04.705018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-26T21:51:04.050833Z digest=sha256:41efe1bf973b224ec482265710aa0a67364f60de77baf1e68bc23e64e5bba145