Pith. sign in

Paper Citation Record · LEDGER

PVC: Progressive Visual Token Compression for Unified Image and Video Processing in Large Vision-Language Models

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2412.09613.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.09613 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:59:39.604858Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-04T23:46:52.445239Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5e1fa2e5-5a96-4760-904e-70b11fe8fe7a · inbound

LaViDa: A Large Diffusion Language Model for Multimodal Understanding cites this paper.

LaViDa: A Large Diffusion Language Model for Multimodal Understanding PVC: Progressive Visual Token Compression for Unified Image and Video Processing in Large Vision-Language Models

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T14:59:39.604858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:59:39.604858Z digest=sha256:6191fa33161702b1e5c483df84603a8eb460b57ccd8ceee19ec5e396ef55fdf2

Observation 28511721-b1c7-4ccf-b01b-d516d5dfe9c3 · inbound

HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes cites this paper.

HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes PVC: Progressive Visual Token Compression for Unified Image and Video Processing in Large Vision-Language Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-05T19:03:07.915605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:03:07.915605Z digest=sha256:82e667c5772cbd1c9f1a959fb7c38d4a894481d6b71b15757cf63f2d33cde6f3

Observation a55a2557-5730-4d44-8d43-a7dfa88ea0b1 · inbound

MESH -- Understanding Videos Like Human: Measuring Hallucinations in Large Video Models cites this paper.

MESH -- Understanding Videos Like Human: Measuring Hallucinations in Large Video Models PVC: Progressive Visual Token Compression for Unified Image and Video Processing in Large Vision-Language Models

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-04T20:32:56.967932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:32:56.967932Z digest=sha256:6886df8246f912856bfc6a1d39526fcdf91e3bbd0a1b56775accc642988124ab

Observation e0bfad12-8a05-46ca-a1dc-50349f089c34 · inbound

CRAFT: Compression via Recursive Adaptive Fusion of Video Tokens for Vision-Language Models cites this paper.

CRAFT: Compression via Recursive Adaptive Fusion of Video Tokens for Vision-Language Models PVC: Progressive Visual Token Compression for Unified Image and Video Processing in Large Vision-Language Models

Reference 15

Resolution
metadata mismatch
local_arxiv, observed 2026-08-04T23:46:52.450235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-04T23:46:51.570775Z digest=sha256:8845573bf38ef4f20ed381667defa73b226fc10d38bd9fbac9f5326056a67f80