Pith. sign in

Paper Citation Record · LEDGER

VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2410.23317.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.23317 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T23:09:25.101407Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T01:36:44.125857Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b031d2b5-81c3-4933-942f-d478c5a449db · inbound

FrameFusion: Combining Similarity and Importance for Video Token Reduction on Large Vision Language Models cites this paper.

FrameFusion: Combining Similarity and Importance for Video Token Reduction on Large Vision Language Models VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T23:09:25.101407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:09:25.101407Z digest=sha256:9e339e1760ad354a48ab42eb658e224b7285cec176f419fa1ec0b4f2c6f4f2dd

Observation 90c47060-c235-431c-96c6-da62c80e40bc · inbound

EffiVLM-BENCH: A Comprehensive Benchmark for Evaluating Training-Free Acceleration in Large Vision-Language Models cites this paper.

EffiVLM-BENCH: A Comprehensive Benchmark for Evaluating Training-Free Acceleration in Large Vision-Language Models VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:08.527752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:09:08.527752Z digest=sha256:43519f92578e2a5a668e430fedb4db7a562c724d1ad35f4f2c12fd319f9f3e8f

Observation 5640b0b4-4a97-406b-bbbe-6fce424be439 · inbound

Blink: Dynamic Visual Token Resolution for Enhanced Multimodal Understanding cites this paper.

Blink: Dynamic Visual Token Resolution for Enhanced Multimodal Understanding VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-03T17:09:37.841750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T17:09:37.841750Z digest=sha256:dd11bb4c3efec26f39709e1084dec573fa8d7315a08bee663977282beec88c98

Observation 0eaae785-0c23-4276-af5c-453416f6e029 · inbound

FreqCache: Accelerating Embodied VLN Models with Adaptive Frequency-Guided Token Caching cites this paper.

FreqCache: Accelerating Embodied VLN Models with Adaptive Frequency-Guided Token Caching VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:16:14.386310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-08T03:10:29.477937Z digest=sha256:2e406f8ed98f6cfa9d049affe3e72ac272fbf98bc68159ca4f88902a487728d2

Observation 3badd050-e228-468c-b125-9876c8b2ff81 · inbound

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models cites this paper.

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:31:16.372723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-07T16:50:26.044591Z digest=sha256:66ec7554c01830362507a607d8798d3f00b5ee80d4e53b97e7e4c24cd2b54fc1

Observation 2542ebda-0081-4fd9-b658-cd54ead99e4f · inbound

Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction cites this paper.

Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:41:26.551067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-12T05:02:25.513351Z digest=sha256:a756f9d4aad69e28944e38a937a8e6331a56242030283b2e078e0ccd3dff4b99

Observation 6046ae1d-bb82-41cc-8c7a-e866a4f8f604 · inbound

KVCapsule: Efficient Sequential KV Cache Compression for Vision-Language Models with Asymmetric Redundancy cites this paper.

KVCapsule: Efficient Sequential KV Cache Compression for Vision-Language Models with Asymmetric Redundancy VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:58:59.047015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-20T19:55:35.051834Z digest=sha256:1e1c5d38d377d319a0445709bb9b944d21407eda4ae84e8c6fdede1ff2ac1dd1

Observation f29b59e7-a302-4fff-965d-7748e7d2375b · inbound

Rotation-Aligned Key Channel Pruning for Efficient Vision-Language Model Inference cites this paper.

Rotation-Aligned Key Channel Pruning for Efficient Vision-Language Model Inference VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:43:23.805107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-20T07:43:18.828740Z digest=sha256:08f31ec6eedee0561655f6aab73f42123546bb275c021e6657cef0afe9dd3cde

Observation 82acc9df-1f2a-40cf-bd23-5dad3beb21df · inbound

CIVIC: End-to-End Sequence Compactness for Efficient Vision-Language Models cites this paper.

CIVIC: End-to-End Sequence Compactness for Efficient Vision-Language Models VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:13:26.457127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T12:12:51.867760Z digest=sha256:a49ce3ceff2b77774d6fae560eba3ededd420dfb9fd1594bc2aee36a03019c71

Observation 7b166343-3228-4f89-8eb1-e3d06546c103 · inbound

AsymVLM: Asymmetric Token Pruning for Efficient Vision-Language Model Inference cites this paper.

AsymVLM: Asymmetric Token Pruning for Efficient Vision-Language Model Inference VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:53:15.966617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:48:58.583657Z digest=sha256:022e4ad14c501b01b255995e625861dc79e44c211275d2439efdbd4c1089d6b9

Observation 7e68414f-3ad7-4120-bf96-bf65e130301d · inbound

AURA: Action-Gated Memory for Robot Policies at Constant VRAM cites this paper.

AURA: Action-Gated Memory for Robot Policies at Constant VRAM VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:16:24.296427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-28T14:30:49.006075Z digest=sha256:32b51bb0011c628eb28287c3fd3bd4bfe3ec46d172608dca9dcf7bf663fdf501

Observation a76f5f52-7d7a-4b1d-a24b-8f239a7457e6 · inbound

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents cites this paper.

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 115

Resolution
verified exact
local_arxiv, observed 2026-07-10T01:36:44.126991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-07-10T01:26:59.421158Z digest=sha256:418d3747ad65bce0672867ebffc3dd2c0817ebad0ff135cca0f1ec6a17c8dfde

Observation a8c0ecb3-1caa-4907-9a68-1e4757205293 · inbound

Reflex: Real-Time VLA Control through Streaming Inference cites this paper.

Reflex: Real-Time VLA Control through Streaming Inference VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T01:26:09.793123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:26:09.793123Z digest=sha256:f81c7baa4afd41e2d17f374dab4abcdddb3f4ec89dafc089f4a448433a09d68b

Observation a074e4b0-1be7-4cee-9531-71a8ab901f67 · inbound

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs cites this paper.

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-04T13:43:56.846247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T13:43:56.846247Z digest=sha256:f6cb911597f96050e4a30ac1fd7b0c9b32f0fa0198b78f6c841b93f0b6bffd5c

Observation 0f2b95f0-7e76-412d-a8bf-6831a46cd8b1 · inbound

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs cites this paper.

PhyCheck: Fine-Grained Evidence-Grounded Dataset for Physical Law Understanding in Video-LLMs VL-Cache: Sparsity and Modality-Aware KV Cache Compression for Vision-Language Model Inference Acceleration

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-07T00:11:49.128823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:11:49.128823Z digest=sha256:41edd69444e213bc338334a7e75c7915c67cf98eccdb6308b3395a3ba97d4867