Pith. sign in

Paper Citation Record · LEDGER

KVSink: Understanding and Enhancing the Preservation of Attention Sinks in KV Cache Quantization for LLMs

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2508.04257.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.04257 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T17:04:24.573502Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T11:14:37.612699Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f0444d80-a47e-4336-80b8-18b549fa9bb7 · inbound

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models cites this paper.

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models KVSink: Understanding and Enhancing the Preservation of Attention Sinks in KV Cache Quantization for LLMs

Reference 286

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:40:54.782557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-16T12:39:57.398423Z digest=sha256:2d52773a77a3803b9725d9055f63925471b326cb4e26627e19df289c8e00d4cc

Observation 941a7b1b-0002-4dab-a2d8-811c82e3db2e · inbound

SnapMLA: Efficient Long-Context MLA Decoding via Hardware-Aware FP8 Quantized Pipelining cites this paper.

SnapMLA: Efficient Long-Context MLA Decoding via Hardware-Aware FP8 Quantized Pipelining KVSink: Understanding and Enhancing the Preservation of Attention Sinks in KV Cache Quantization for LLMs

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-16T06:00:40.756876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-16T05:58:03.113220Z digest=sha256:68640da1d316edef6d95c67ba1efd5a0c6f4660f1018d06c8a08f2f0b2ebc261

Observation 27420f8b-2d0c-41f3-a13d-df94dd1219a3 · inbound

When Sinks Help or Hurt: Unified Framework for Attention Sink in Large Vision-Language Models cites this paper.

When Sinks Help or Hurt: Unified Framework for Attention Sink in Large Vision-Language Models KVSink: Understanding and Enhancing the Preservation of Attention Sinks in KV Cache Quantization for LLMs

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:23:26.812390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-13T23:20:51.899127Z digest=sha256:9cd83e07504b678fc2b4638321af033420ab576db891972c464d8bdf9bf92d11

Observation bf77deb5-b070-4ebe-9370-dace9342f59a · inbound

When Sinks Help or Hurt: Unified Framework for Attention Sink in Large Vision-Language Models cites this paper.

When Sinks Help or Hurt: Unified Framework for Attention Sink in Large Vision-Language Models KVSink: Understanding and Enhancing the Preservation of Attention Sinks in KV Cache Quantization for LLMs

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T17:04:24.573502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T17:04:24.573502Z digest=sha256:fcfbfd86f6d6e5f9afb2436054d0a7a8a00b91a5c946061cb67644e1a860352c

Observation ade11d7c-0e74-4014-a1cb-322d12922c7a · inbound

OScaR: The Occam's Razor for Extreme KV Cache Quantization in LLMs and Beyond cites this paper.

OScaR: The Occam's Razor for Extreme KV Cache Quantization in LLMs and Beyond KVSink: Understanding and Enhancing the Preservation of Attention Sinks in KV Cache Quantization for LLMs

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:58:07.564169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-20T07:57:51.032025Z digest=sha256:64d97731075df1118a016b60482143b4ecce3a1afb659fc838e3c846bc099141

Observation a921ef1b-1ac5-4a90-b9bc-1717c8e0e6a9 · inbound

Inference Time Optimization with Confidence Dynamics cites this paper.

Inference Time Optimization with Confidence Dynamics KVSink: Understanding and Enhancing the Preservation of Attention Sinks in KV Cache Quantization for LLMs

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:14:37.614076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T11:12:46.757296Z digest=sha256:bd4e541c49fab3e08c78ae23d4c60cac82590d461aac21138532e0a4e30ad657