Pith. sign in

Paper Citation Record · LEDGER

SKVQ: Sliding-window Key and Value Cache Quantization for Large Language Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2405.06219.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.06219 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T14:13:38.019187Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T00:19:13.103948Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 00067a60-19fe-4def-932b-f7cb71f6d53c · inbound

Position: Episodic Memory is the Missing Piece for Long-Term LLM Agents cites this paper.

Position: Episodic Memory is the Missing Piece for Long-Term LLM Agents SKVQ: Sliding-window Key and Value Cache Quantization for Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T14:13:38.019187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:13:38.019187Z digest=sha256:303d7d584b45a3be1e7e782ea1bf6aebae66bfee856307b60b8c2b71a01128f0

Observation 0bb28117-42f1-45b9-97d9-8d2d9dc4e9b6 · inbound

PM-KVQ: Progressive Mixed-precision KV Cache Quantization for Long-CoT LLMs cites this paper.

PM-KVQ: Progressive Mixed-precision KV Cache Quantization for Long-CoT LLMs SKVQ: Sliding-window Key and Value Cache Quantization for Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:34:25.438387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:34:25.438387Z digest=sha256:21432ac6bd8cbb6bba15c28606a8ce0659ffc6ab951503c671875666cee6dc28

Observation 68e19fd9-cb02-46b3-95a7-b72eb6c75299 · inbound

Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache cites this paper.

Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache SKVQ: Sliding-window Key and Value Cache Quantization for Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T01:10:37.672974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T01:10:37.672974Z digest=sha256:16e8f2c7ed0f756c5383201d051603a60cbde48a2790e266d773e9773db7d413

Observation 839a30a5-2e41-4d6d-9ad0-3da9a4a62d9d · inbound

Which Heads Matter for Reasoning? RL-Guided KV Cache Compression cites this paper.

Which Heads Matter for Reasoning? RL-Guided KV Cache Compression SKVQ: Sliding-window Key and Value Cache Quantization for Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T10:50:09.099521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:50:09.099521Z digest=sha256:7f789ef690bf9712eeab09ba1c70b0361967d90e4d177bde5b8bd8a8eb5f8440

Observation 88730d19-f797-4fbf-b89d-8087c0c71659 · inbound

Quant VideoGen: Auto-Regressive Long Video Generation via 2-Bit KV-Cache Quantization cites this paper.

Quant VideoGen: Auto-Regressive Long Video Generation via 2-Bit KV-Cache Quantization SKVQ: Sliding-window Key and Value Cache Quantization for Large Language Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-16T08:00:44.687440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T07:58:24.456859Z digest=sha256:89d0c3474e408317d7fc4bf256179cf2332400a7a628690f32fdcb93d7294b3e

Observation a1391a59-a702-484a-b249-363a435a4e3a · inbound

Harmonia: Algorithm-Hardware Co-Design for Memory- and Compute-Efficient BFP-based LLM Inference cites this paper.

Harmonia: Algorithm-Hardware Co-Design for Memory- and Compute-Efficient BFP-based LLM Inference SKVQ: Sliding-window Key and Value Cache Quantization for Large Language Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-03T04:37:29.496918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:37:29.496918Z digest=sha256:abcecd63b4a08c5d856fc0e1edd8fe4a9688b11881d3ffce76fdc6ccb0e19cde

Observation f0a8787f-f301-4d6d-ad0b-aba83ec3c8d5 · inbound

WindowQuant: Mixed-Precision KV Cache Quantization based on Window-Level Similarity for VLMs Inference Optimization cites this paper.

WindowQuant: Mixed-Precision KV Cache Quantization based on Window-Level Similarity for VLMs Inference Optimization SKVQ: Sliding-window Key and Value Cache Quantization for Large Language Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:36:07.942065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-09T16:06:26.450483Z digest=sha256:1ca84ba2d21efc476309130a9e48a8d3f4ab666f318d6fc429bdd77e49a72db6

Observation 0bdfee88-ddf9-497d-8696-1b0eb739d4f0 · inbound

KVServe: Service-Aware KV Cache Compression for Communication-Efficient Disaggregated LLM Serving cites this paper.

KVServe: Service-Aware KV Cache Compression for Communication-Efficient Disaggregated LLM Serving SKVQ: Sliding-window Key and Value Cache Quantization for Large Language Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:42:31.163430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T17:42:06.870534Z digest=sha256:bdca87b1f89f7b6c2397e3fb75cca18efce6dca63d7841ba93af7c4bb4c6ad8d

Observation 8f723870-2de1-4691-9464-f77518348e75 · inbound

OScaR: The Occam's Razor for Extreme KV Cache Quantization in LLMs and Beyond cites this paper.

OScaR: The Occam's Razor for Extreme KV Cache Quantization in LLMs and Beyond SKVQ: Sliding-window Key and Value Cache Quantization for Large Language Models

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:58:07.629942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T07:57:51.032025Z digest=sha256:e91b3ad36ef33e759bdba34ffc2d5fa2ce6eb2fa6a2e74849085b53f166eb64b

Observation d70827e0-5817-47d5-8500-025b85dd4f4f · inbound

Dual Dimensionality for Local and Global Attention cites this paper.

Dual Dimensionality for Local and Global Attention SKVQ: Sliding-window Key and Value Cache Quantization for Large Language Models

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:19:13.105685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T21:22:05.095209Z digest=sha256:0d08730618f95ab7fe85389f25b20771ffaa0ae6fc96597fdf2f420687368c8e