Pith. sign in

Paper Citation Record · LEDGER

ZipCache: Accurate and Efficient KV Cache Quantization with Salient Token Identification

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2405.14256.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.14256 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T20:34:01.234695Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T21:27:24.350095Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b8b6f30e-69ab-4186-804d-ceb69798c826 · inbound

Cache Me If You Must: Adaptive Key-Value Quantization for Large Language Models cites this paper.

Cache Me If You Must: Adaptive Key-Value Quantization for Large Language Models ZipCache: Accurate and Efficient KV Cache Quantization with Salient Token Identification

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-09T20:34:01.234695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:34:01.234695Z digest=sha256:7f488e9ca26d57e06ad6b883a678497d14662c2d5401ebe173797ad694e682bc

Observation 274878eb-54ab-4a3f-b47c-6cb7c37c1241 · inbound

HACK: Homomorphic Acceleration via Compression of the Key-Value Cache for Disaggregated LLM Inference cites this paper.

HACK: Homomorphic Acceleration via Compression of the Key-Value Cache for Disaggregated LLM Inference ZipCache: Accurate and Efficient KV Cache Quantization with Salient Token Identification

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T04:32:01.192677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:32:01.192677Z digest=sha256:8ea43c9468251d2ee2626cca4a8da245b25570a6f60e21faaef5967f9a4d9673

Observation 735a74b6-9c5a-4958-af5b-36956bc96a53 · inbound

NSNQuant: A Double Normalization Approach for Calibration-Free Low-Bit Vector Quantization of KV Cache cites this paper.

NSNQuant: A Double Normalization Approach for Calibration-Free Low-Bit Vector Quantization of KV Cache ZipCache: Accurate and Efficient KV Cache Quantization with Salient Token Identification

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:42:20.783045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:42:20.783045Z digest=sha256:92a2e83befcae1dd394f64c35ebd2cc5c65c9a22e505671258fec9b23bcc8e4e

Observation e9cccbfa-a369-4d8f-bf85-2dddd7291d5f · inbound

Memory-Efficient Visual Autoregressive Modeling with Scale-Aware KV Cache Compression cites this paper.

Memory-Efficient Visual Autoregressive Modeling with Scale-Aware KV Cache Compression ZipCache: Accurate and Efficient KV Cache Quantization with Salient Token Identification

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:13.857590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:16:13.857590Z digest=sha256:e465a07c2243cb5a5b70c61b61b8b6b6a8a5a841aa7b1c9bd4f9f7d6aa6f4af1

Observation b39b5d56-55ca-4faf-bff3-a4be71881975 · inbound

Curse of High Dimensionality Issue in Transformer for Long-context Modeling cites this paper.

Curse of High Dimensionality Issue in Transformer for Long-context Modeling ZipCache: Accurate and Efficient KV Cache Quantization with Salient Token Identification

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T13:23:12.619787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:23:12.619787Z digest=sha256:ac9a5516ad7eb6b384c878abd8b98c371994a426fafc9639d8f97db612656e5f

Observation 6ce893f5-77e9-4c12-bed1-0e2015e2d6a7 · inbound

Through the Stealth Lens: Attention-Aware Defenses Against Poisoning in RAG cites this paper.

Through the Stealth Lens: Attention-Aware Defenses Against Poisoning in RAG ZipCache: Accurate and Efficient KV Cache Quantization with Salient Token Identification

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-25T08:15:34.148556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-25T08:10:31.473565Z digest=sha256:40f0f0b239e4db57930da161bf59491d0536557dbf8946de350f5e4fa6cd8d1e

Observation 18d9c572-7fdc-4e8b-bcc0-0d3fa60daa9b · inbound

SpindleKV: A Novel KV Cache Reduction Method Balancing Both Shallow and Deep Layers cites this paper.

SpindleKV: A Novel KV Cache Reduction Method Balancing Both Shallow and Deep Layers ZipCache: Accurate and Efficient KV Cache Quantization with Salient Token Identification

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T19:10:01.179964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:10:01.179964Z digest=sha256:d7e10b5b5965a5c93a170673bcbf42d2e770da79b502541581aa5144d73d2245

Observation 83187b8a-959e-40c0-a79a-35d4d6a8652f · inbound

CaliDrop: KV Cache Compression with Calibration cites this paper.

CaliDrop: KV Cache Compression with Calibration ZipCache: Accurate and Efficient KV Cache Quantization with Salient Token Identification

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T13:57:15.706240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:57:15.706240Z digest=sha256:7a2c854715abd6f5f0ebd0c2b2b810b23f9b5764669fd331fa57a13fe40ba2f9

Observation 9e8dd8f3-f008-4cfd-921e-a6a2ef78b2fe · inbound

PagedEviction: Structured Block-wise KV Cache Pruning for Efficient Large Language Model Inference cites this paper.

PagedEviction: Structured Block-wise KV Cache Pruning for Efficient Large Language Model Inference ZipCache: Accurate and Efficient KV Cache Quantization with Salient Token Identification

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T10:16:13.067336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:16:13.067336Z digest=sha256:82d0398dcde339849794e527cd5f64bf55a4586db8883c729d1064a69a82e634

Observation c303e582-96bf-4f7d-8dce-cf5b0a35b1a1 · inbound

SPHERICAL KV: Angle-Domain Attention and Rate-Distortion Retention for Efficient Long-Context Inference cites this paper.

SPHERICAL KV: Angle-Domain Attention and Rate-Distortion Retention for Efficient Long-Context Inference ZipCache: Accurate and Efficient KV Cache Quantization with Salient Token Identification

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T20:33:43.329715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T20:30:33.208386Z digest=sha256:6180a2dfea265329b493e134c0e55b93dcb93d98ca9709c229a1d30ec1eed471

Observation ffc92941-7216-447a-aa98-099f1a121f88 · inbound

Meta-Soft: Leveraging Composable Meta-Tokens for Context-Preserving KV Cache Compression cites this paper.

Meta-Soft: Leveraging Composable Meta-Tokens for Context-Preserving KV Cache Compression ZipCache: Accurate and Efficient KV Cache Quantization with Salient Token Identification

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-22T05:11:06.161392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T05:10:53.133346Z digest=sha256:dab060e85bf8f3bc17a7a02188222a906a24c8d12a98cb91702f76b97886f4ce

Observation c0329f1f-75f5-48fa-b904-0ac5f5c29a03 · inbound

Meta-Soft: Leveraging Composable Meta-Tokens for Context-Preserving KV Cache Compression cites this paper.

Meta-Soft: Leveraging Composable Meta-Tokens for Context-Preserving KV Cache Compression ZipCache: Accurate and Efficient KV Cache Quantization with Salient Token Identification

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-30T17:34:57.865491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T17:28:08.750049Z digest=sha256:dafe8f376f70dc91c1c9bbcc999cd5424a06ff5c9de6c208ce0a0d7f53ef0508

Observation fb7cab7c-7a72-46ce-87c6-645405c7dd02 · inbound

IndexMem: Learned KV-Cache Eviction with Latent Memory for Long-Context LLM Inference cites this paper.

IndexMem: Learned KV-Cache Eviction with Latent Memory for Long-Context LLM Inference ZipCache: Accurate and Efficient KV Cache Quantization with Salient Token Identification

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:03:59.876490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T22:03:07.685253Z digest=sha256:3602663a72e30ff51eed7063cd1f5041632c80a364b75bca5a290fed62d2700d

Observation c802f100-cb57-4803-b045-e45fde5d62df · inbound

HACK++: Towards More Effective Head-Aware Key-Value Compression for Efficient Visual Autoregressive Modeling cites this paper.

HACK++: Towards More Effective Head-Aware Key-Value Compression for Efficient Visual Autoregressive Modeling ZipCache: Accurate and Efficient KV Cache Quantization with Salient Token Identification

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:27:24.351589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T19:46:43.514413Z digest=sha256:cadabbc7ebf2a051bb5f4ef1aecf852468e8de6174fe44554812e93cba41d063

Observation a378ae46-075a-44fa-b980-7c0b09bb624d · inbound

GSRQ: Gain-Shape Residual Quantization for Sub-1-bit KV Cache cites this paper.

GSRQ: Gain-Shape Residual Quantization for Sub-1-bit KV Cache ZipCache: Accurate and Efficient KV Cache Quantization with Salient Token Identification

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T15:57:06.394332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-02T15:55:40.177742Z digest=sha256:9797cb300e6f6bd54a1c29f9f47304575c40f07468437d8a09fd26f538e048fd

Observation ddb8a54c-fef2-44d7-bffc-93230c7d777b · inbound

LOCKS: Page-Local Compact Key Summaries for Efficient Long-Context Decoding cites this paper.

LOCKS: Page-Local Compact Key Summaries for Efficient Long-Context Decoding ZipCache: Accurate and Efficient KV Cache Quantization with Salient Token Identification

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-31T11:49:11.655479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T11:49:11.655479Z digest=sha256:f719c44433e375e5447030318d2ea027b3ae16594b92d37fc93b2c3c42d77410