Pith. sign in

Paper Citation Record · LEDGER

Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2503.12496.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.12496 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:33:58.078291Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T23:39:04.701981Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation eab857a9-1f34-4e9a-82cf-91757979e69b · inbound

Mitigating Object Hallucinations via Sentence-Level Early Intervention cites this paper.

Mitigating Object Hallucinations via Sentence-Level Early Intervention Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-25T08:35:32.671787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-25T08:31:24.173135Z digest=sha256:65d566a2e18e9b32557978af9a7abb01bb8b3f6c508751b3dee7f3bba6f8f47a

Observation d074ebb3-c178-4196-b9c9-61f8004e1eb0 · inbound

VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning cites this paper.

VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T16:33:58.078291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:33:58.078291Z digest=sha256:1fdc54126cf10b0a1a7c9e9a4539eee4f6f908d5d9d1ff2923dcf1cce52ac081

Observation 8be86950-d6de-4d84-bf9c-243eb57add5d · inbound

EgoSelf: From Memory to Personalized Egocentric Assistant cites this paper.

EgoSelf: From Memory to Personalized Egocentric Assistant Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:06:03.813089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T02:23:21.119521Z digest=sha256:5e11be147114c946f4398217d81b225797143a15b1caf3e45ca260c7277cadf5

Observation fb71b523-eaff-48d3-87a1-a6f35c8a1649 · inbound

Uni-OPD: Unifying On-Policy Distillation with a Dual-Perspective Recipe cites this paper.

Uni-OPD: Unifying On-Policy Distillation with a Dual-Perspective Recipe Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:26:16.162141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-07T17:00:49.448352Z digest=sha256:c23baed38e58e713ae7a85a9fe54156a6f03430b02c92bd3020fb1b5232096ad

Observation a0eb97c8-35c1-4a53-b269-e86c34f5582c · inbound

Video-Zero: Self-Evolution Video Understanding cites this paper.

Video-Zero: Self-Evolution Video Understanding Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-06-30T21:35:04.432104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T21:32:16.939563Z digest=sha256:6785831cb7383b18d1dc687f7ebbf754fde5a62fb46086af32df39cd328758b8

Observation d483c9cc-90a3-4602-9dbd-c336ec92fc5f · inbound

How Well Can Your Video Model Remember? Measuring Memory-Budget Trade-offs in Long Video Understanding cites this paper.

How Well Can Your Video Model Remember? Measuring Memory-Budget Trade-offs in Long Video Understanding Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T23:39:04.705018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-26T21:51:04.050833Z digest=sha256:5d0c6792e4424e2c75a973b4b90ae433cd55fccf88556e6d39f43d495e413669