Pith. sign in

Paper Citation Record · LEDGER

DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2412.14838.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.14838 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:57:15.839816Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 20efee29-b8eb-44cc-a975-33186125f2c7 · inbound

CaliDrop: KV Cache Compression with Calibration cites this paper.

CaliDrop: KV Cache Compression with Calibration DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T13:57:15.839816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:57:15.839816Z digest=sha256:b0c39ba56257b3deeec2065a88d92beb56050c6129bccc01d84c10cc8845cef1

Observation bb8e1ff1-35dd-4b7e-bb30-414854bdeb30 · inbound

TPLA: Tensor Parallel Latent Attention for Efficient Disaggregated Prefill and Decode Inference cites this paper.

TPLA: Tensor Parallel Latent Attention for Efficient Disaggregated Prefill and Decode Inference DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T17:55:08.261078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:55:08.261078Z digest=sha256:e84f99049f3e16c19c7feb0bec47d7bac92db87bfeb760959ef6e468113d0bf7

Observation ea16eb07-5299-44a1-9870-516b83da6ff3 · inbound

CompilerKV: Risk-Adaptive KV Compression via Offline Experience Compilation cites this paper.

CompilerKV: Risk-Adaptive KV Compression via Offline Experience Compilation DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-21T13:34:11.370570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T13:32:15.598047Z digest=sha256:7fbac2187afb63125a95cdc4844d5c7c1a21af693b26c54c94d3a43fc74c8e6c

Observation af94e715-099d-4c2d-a69c-569597dc15ef · inbound

CompilerKV: Risk-Adaptive KV Compression via Offline Experience Compilation cites this paper.

CompilerKV: Risk-Adaptive KV Compression via Offline Experience Compilation DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T03:16:19.908039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:16:19.908039Z digest=sha256:72b4318c1009dedf0ecd4ae8f325a65f574c76f41a48229a5139105d111c047e

Observation 3dee4f86-afda-4614-8bc4-3ce7403f73b3 · inbound

From Static Inference to Dynamic Interaction: A Survey of Streaming Large Language Models cites this paper.

From Static Inference to Dynamic Interaction: A Survey of Streaming Large Language Models DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T16:20:09.970164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T16:16:15.819622Z digest=sha256:b264693b37aca73d246189400c8c6dc4230efd44c50c47da62439e60e35082a9

Observation fd5408ee-5436-4d65-8215-22112e7fc78e · inbound

Don't Waste Bits! Adaptive KV-Cache Quantization for Lightweight On-Device LLMs cites this paper.

Don't Waste Bits! Adaptive KV-Cache Quantization for Lightweight On-Device LLMs DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:55:48.123886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T19:31:17.420639Z digest=sha256:fc6c5c9a900a55e01aba74a1e6be9b364d3c7e5fa335a10cc4dd745dd63781a1

Observation ea482e67-46c4-4806-bebd-45fefa3c12a7 · inbound

SAGE: Selective Attention-Guided Extraction for Token-Efficient Document Indexing cites this paper.

SAGE: Selective Attention-Guided Extraction for Token-Efficient Document Indexing DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:32:52.113523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T08:32:02.222528Z digest=sha256:3785e09559a38bd4f73c31e85d94c1d43c93305d70fb1627810d0b89f49d55de

Observation 91571be8-3c29-4401-8a00-d1c468f90051 · inbound

HieraSparse: Hierarchical Semi-Structured Sparse KV Attention cites this paper.

HieraSparse: Hierarchical Semi-Structured Sparse KV Attention DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-10T07:16:54.213976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T07:15:19.184970Z digest=sha256:7430df28f3e9c2d62dcb8046847ae721528c2a61a03bccedc5b04764e381e09d

Observation 2bd6a4be-be7c-437f-a69b-bf517eedde51 · inbound

Network Edge Inference for Large Language Models: Principles, Techniques, and Opportunities cites this paper.

Network Edge Inference for Large Language Models: Principles, Techniques, and Opportunities DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 211

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:16:10.305729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T09:45:57.201837Z digest=sha256:6624795483e56b98ef6bee5b50b94ec51a6cd683e9bb1c66fb993322f2fbe913

Observation 4ea9d3cb-195e-4348-b96e-e9815ba6a8d0 · inbound

ReST-KV: Robust KV Cache Eviction with Layer-wise Output Reconstruction and Spatial-Temporal Smoothing cites this paper.

ReST-KV: Robust KV Cache Eviction with Layer-wise Output Reconstruction and Spatial-Temporal Smoothing DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:26:30.323223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T02:52:33.076123Z digest=sha256:839040a457c7602e2edb28848e61682b34898d8bb94b7725a8838381d29b83e8

Observation 2b30aa87-2826-4855-bf1a-f389a5bccd30 · inbound

YouZhi: Towards High-Concurrency Financial LLMs via Adaptive GQA-to-MLA Transition cites this paper.

YouZhi: Towards High-Concurrency Financial LLMs via Adaptive GQA-to-MLA Transition DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T13:06:59.429187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T01:33:54.383507Z digest=sha256:a920bc0f44ec973c5adfc71a423f53389a2d795076ee2527b822bff2d65a9190

Observation 2e9c34ce-26cb-40d9-90aa-e3a271609d4f · inbound

SpectrumKV: Per-Token Mixed-Precision KV Cache Transfer for Prefill-Decode Disaggregated LLM Serving cites this paper.

SpectrumKV: Per-Token Mixed-Precision KV Cache Transfer for Prefill-Decode Disaggregated LLM Serving DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 22

Resolution
malformed identifier
arxiv_id, observed 2026-07-02T22:27:26.309060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T18:50:50.066713Z digest=sha256:98ca79d70744fec0ee03b46fd6bc5537ff2f89df63244a41e5c55cf9ab49fec8

Observation 5b49ead8-949a-43ce-a411-79add8557bad · inbound

IntentKV: Cross-Turn Intent-Aware KV Cache Pruning for Agent Inference cites this paper.

IntentKV: Cross-Turn Intent-Aware KV Cache Pruning for Agent Inference DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:57:22.975598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T20:04:44.703782Z digest=sha256:9f835d0e74c9f1310863eea836d612a0dec9a0a93eb82b44e3453ca4f5089cf4

Observation fc489183-a0d4-4e26-94cd-a9ffddf66d45 · inbound

ReasonAlloc: Hierarchical Decoding-Time KV Cache Budget Allocation for Reasoning Models cites this paper.

ReasonAlloc: Hierarchical Decoding-Time KV Cache Budget Allocation for Reasoning Models DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T05:37:40.355882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T13:09:37.979051Z digest=sha256:069e51faf169a90b0d2dcd910f685fead2b5c35b63c827d5c7ff148679401bde

Observation 2b1e6053-7804-403d-860e-cf5810fad4c0 · inbound

AnchorKV: Safety-Aware KV Cache Compression via Soft Penalty with a Refusal Anchor cites this paper.

AnchorKV: Safety-Aware KV Cache Compression via Soft Penalty with a Refusal Anchor DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:18:57.886130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T01:23:23.488555Z digest=sha256:93ce01a4a87f40524c32247093dd054adac960410eb755d903c7339eff0e617e

Observation 04608168-8d5e-441b-8da9-e1fb4b2613fd · inbound

Finetuning Vision-Language-Action Models Requires Fewer Layers Than You Think cites this paper.

Finetuning Vision-Language-Action Models Requires Fewer Layers Than You Think DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-04T04:29:34.949974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T16:59:28.243515Z digest=sha256:772020a96b894d62c0e9a467483e47039ae055e89311a850de08e3bf29323dad

Observation a2b0cd05-6a7d-4054-aa8b-ca6dc843c250 · inbound

MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference cites this paper.

MemDecay: Region-Aware KV Cache Eviction for Efficient LLM Agent Inference DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-14T10:41:08.967261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T10:41:08.967261Z digest=sha256:568c673daf11d3df2372c002b86502c62c3e30a7b61431aac7aab4c4ff60ca9e