Pith. sign in

Paper Citation Record · LEDGER

DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2412.14838.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.14838 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:57:15.839816Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 20efee29-b8eb-44cc-a975-33186125f2c7 · inbound

CaliDrop: KV Cache Compression with Calibration cites this paper.

CaliDrop: KV Cache Compression with Calibration DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T13:57:15.839816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:57:15.839816Z digest=sha256:f692c63928529bbb49ba6057a85a0627f8fb233e74f2f03b252091d0494be73b

Observation bb8e1ff1-35dd-4b7e-bb30-414854bdeb30 · inbound

TPLA: Tensor Parallel Latent Attention for Efficient Disaggregated Prefill and Decode Inference cites this paper.

TPLA: Tensor Parallel Latent Attention for Efficient Disaggregated Prefill and Decode Inference DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T17:55:08.261078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:55:08.261078Z digest=sha256:e84f99049f3e16c19c7feb0bec47d7bac92db87bfeb760959ef6e468113d0bf7

Observation ea16eb07-5299-44a1-9870-516b83da6ff3 · inbound

CompilerKV: Risk-Adaptive KV Compression via Offline Experience Compilation cites this paper.

CompilerKV: Risk-Adaptive KV Compression via Offline Experience Compilation DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-21T13:34:11.370570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T13:32:15.598047Z digest=sha256:30b92b53093723356a90f9e9f988ce302a339a3676d37793e5b940acf206aa9e

Observation af94e715-099d-4c2d-a69c-569597dc15ef · inbound

CompilerKV: Risk-Adaptive KV Compression via Offline Experience Compilation cites this paper.

CompilerKV: Risk-Adaptive KV Compression via Offline Experience Compilation DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T03:16:19.908039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:16:19.908039Z digest=sha256:72b4318c1009dedf0ecd4ae8f325a65f574c76f41a48229a5139105d111c047e

Observation 3dee4f86-afda-4614-8bc4-3ce7403f73b3 · inbound

From Static Inference to Dynamic Interaction: A Survey of Streaming Large Language Models cites this paper.

From Static Inference to Dynamic Interaction: A Survey of Streaming Large Language Models DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T16:20:09.970164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T16:16:15.819622Z digest=sha256:480bae34c226ee8f6f7805dfbdc88aa353c9ad73633bdac10589d5e1ce335340

Observation fd5408ee-5436-4d65-8215-22112e7fc78e · inbound

Don't Waste Bits! Adaptive KV-Cache Quantization for Lightweight On-Device LLMs cites this paper.

Don't Waste Bits! Adaptive KV-Cache Quantization for Lightweight On-Device LLMs DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:55:48.123886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T19:31:17.420639Z digest=sha256:757302221b900558f755326f569d6bc4916a6804fb527c05a3b84524e43d74b8

Observation ea482e67-46c4-4806-bebd-45fefa3c12a7 · inbound

SAGE: Selective Attention-Guided Extraction for Token-Efficient Document Indexing cites this paper.

SAGE: Selective Attention-Guided Extraction for Token-Efficient Document Indexing DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:32:52.113523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T08:32:02.222528Z digest=sha256:d45c3d8c675fac21b5f934df2e8fa28e1349ac242548e788b2aa1640cef41df0

Observation 91571be8-3c29-4401-8a00-d1c468f90051 · inbound

HieraSparse: Hierarchical Semi-Structured Sparse KV Attention cites this paper.

HieraSparse: Hierarchical Semi-Structured Sparse KV Attention DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-10T07:16:54.213976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T07:15:19.184970Z digest=sha256:d69b67c6b485a940e6a1369b92532755c5651fc233a9fd9e96e1ed53bf796e9e

Observation 2bd6a4be-be7c-437f-a69b-bf517eedde51 · inbound

Network Edge Inference for Large Language Models: Principles, Techniques, and Opportunities cites this paper.

Network Edge Inference for Large Language Models: Principles, Techniques, and Opportunities DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 211

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:16:10.305729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T09:45:57.201837Z digest=sha256:87cde39a0fd32550c5b4e66416289efc712ea10bfc5ddae84c9f3a22cd97376a

Observation 4ea9d3cb-195e-4348-b96e-e9815ba6a8d0 · inbound

ReST-KV: Robust KV Cache Eviction with Layer-wise Output Reconstruction and Spatial-Temporal Smoothing cites this paper.

ReST-KV: Robust KV Cache Eviction with Layer-wise Output Reconstruction and Spatial-Temporal Smoothing DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:26:30.323223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T02:52:33.076123Z digest=sha256:0b537317efafe06ebdaafc967f35443bb1f0be0efcead2e2df2bc16017432369

Observation 2b30aa87-2826-4855-bf1a-f389a5bccd30 · inbound

YouZhi: Towards High-Concurrency Financial LLMs via Adaptive GQA-to-MLA Transition cites this paper.

YouZhi: Towards High-Concurrency Financial LLMs via Adaptive GQA-to-MLA Transition DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T13:06:59.429187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T01:33:54.383507Z digest=sha256:a877d0e5eb56bb253e1cd424c1b909fa2570cc1fc7dbee7c5e117ed1898c980f

Observation 2e9c34ce-26cb-40d9-90aa-e3a271609d4f · inbound

SpectrumKV: Per-Token Mixed-Precision KV Cache Transfer for Prefill-Decode Disaggregated LLM Serving cites this paper.

SpectrumKV: Per-Token Mixed-Precision KV Cache Transfer for Prefill-Decode Disaggregated LLM Serving DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 22

Resolution
malformed identifier
arxiv_id, observed 2026-07-02T22:27:26.309060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T18:50:50.066713Z digest=sha256:2884e3425376d068016159c86668575f66c467e2ce809c063ca5cdaa33c2cb0b

Observation 5b49ead8-949a-43ce-a411-79add8557bad · inbound

IntentKV: Cross-Turn Intent-Aware KV Cache Pruning for Agent Inference cites this paper.

IntentKV: Cross-Turn Intent-Aware KV Cache Pruning for Agent Inference DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:57:22.975598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T20:04:44.703782Z digest=sha256:1defb7ce5150f382e49f1d4a7f5a0dd7f5556c174256792a94b3a9d723290114

Observation fc489183-a0d4-4e26-94cd-a9ffddf66d45 · inbound

ReasonAlloc: Hierarchical Decoding-Time KV Cache Budget Allocation for Reasoning Models cites this paper.

ReasonAlloc: Hierarchical Decoding-Time KV Cache Budget Allocation for Reasoning Models DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T05:37:40.355882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T13:09:37.979051Z digest=sha256:61cb3127bb6d6f93f708451b239c0cfe016a5dd16c87152abe44b23a27543360

Observation 2b1e6053-7804-403d-860e-cf5810fad4c0 · inbound

AnchorKV: Safety-Aware KV Cache Compression via Soft Penalty with a Refusal Anchor cites this paper.

AnchorKV: Safety-Aware KV Cache Compression via Soft Penalty with a Refusal Anchor DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:18:57.886130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T01:23:23.488555Z digest=sha256:db106fead64027c179d8c88ba75d195bc9570140980b0151062f1084a16ce4d0

Observation 04608168-8d5e-441b-8da9-e1fb4b2613fd · inbound

Finetuning Vision-Language-Action Models Requires Fewer Layers Than You Think cites this paper.

Finetuning Vision-Language-Action Models Requires Fewer Layers Than You Think DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-04T04:29:34.949974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T16:59:28.243515Z digest=sha256:5f8e678d570a962ff4b6f5085946308f635521920c84faf22a5e79fabd6885a2

Observation a2b0cd05-6a7d-4054-aa8b-ca6dc843c250 · inbound

MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference cites this paper.

MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-14T10:41:08.967261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T10:41:08.967261Z digest=sha256:568c673daf11d3df2372c002b86502c62c3e30a7b61431aac7aab4c4ff60ca9e