Pith. sign in

Paper Citation Record · LEDGER

SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2406.15486.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.15486 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T11:11:17.839813Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e00000cd-e7e5-4938-b067-d560714c60f6 · inbound

Speculative Prefill: Turbocharging TTFT with Lightweight and Training-Free Token Importance Estimation cites this paper.

Speculative Prefill: Turbocharging TTFT with Lightweight and Training-Free Token Importance Estimation SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-09T11:11:17.839813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T11:11:17.839813Z digest=sha256:09e51400e86a041ccc7af9ee1207932d587c857b05f6df5de187f1bd134d2267

Observation 880c0f54-d757-48f2-b888-07afa42268bf · inbound

SALE : Low-bit Estimation for Efficient Sparse Attention in Long-context LLM Prefilling cites this paper.

SALE : Low-bit Estimation for Efficient Sparse Attention in Long-context LLM Prefilling SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:04.130346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:04.130346Z digest=sha256:83389396f991e24e05966d24f8978c010386f5becd8697128191e9119a93d8ff

Observation 1e48076b-dbe4-4a8a-82fc-b0250b397dc8 · inbound

Breaking the Boundaries of Long-Context LLM Inference: Adaptive KV Management on a Single Commodity GPU cites this paper.

Breaking the Boundaries of Long-Context LLM Inference: Adaptive KV Management on a Single Commodity GPU SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T23:00:11.244159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:00:11.244159Z digest=sha256:ceb223b420cfa4ec3cd9f0d2fe7f5bcf732777637b7192922e70f14c13979d5e

Observation 11eee389-6bd1-4de9-84ad-83e1ceebfab3 · inbound

Unveiling Effective In-Context Configurations for Image Captioning: An External & Internal Analysis cites this paper.

Unveiling Effective In-Context Configurations for Image Captioning: An External & Internal Analysis SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-06T19:21:13.532463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:21:13.532463Z digest=sha256:1ed080ace72a33255e651f11036f919fc30bf5718dc1bcae9f2f9cbdd7db647c

Observation f07f11f8-9fad-4f27-8450-67999036c943 · inbound

vAttention: Verified Sparse Attention cites this paper.

vAttention: Verified Sparse Attention SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T11:21:10.344773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:21:10.344773Z digest=sha256:d625973395a02913115c927dca3a4045d9607fd69ac675850944a500796c5987

Observation e129d356-e531-4ea7-b65f-203fa9d92b89 · inbound

CSAttention: Centroid-Scoring Attention for Accelerating LLM Inference cites this paper.

CSAttention: Centroid-Scoring Attention for Accelerating LLM Inference SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:53:02.152558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T21:52:25.148104Z digest=sha256:25fda8c514f7b6b1a82a8d7571998c8ea321716520ad62343d9f832f3e9ee6f5

Observation fd98a068-3906-4903-9d7c-a58fd7ffb68b · inbound

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving cites this paper.

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:38:23.097181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T00:37:08.671539Z digest=sha256:a67128009df515339b8f431842d7faed4dd33eca909eb6be41fb168f6d73f04b

Observation 736369e9-1bae-4162-a0af-46f3c09eb51d · inbound

Sparse Attention as a Range Searching Problem: Towards an Inference-Efficient Index for KV Cache cites this paper.

Sparse Attention as a Range Searching Problem: Towards an Inference-Efficient Index for KV Cache SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:15:50.583185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T01:15:35.871863Z digest=sha256:1a2aad4b648cde6eba458f614340e4f0386e7f5e7ba6812485314a0d78bf5d62

Observation 502066ef-d2a9-4853-8c26-1d7d1db5d78a · inbound

SIFT: Selective-Index For Fast Compute of RAG Prefill by Exploiting Attention Invariance cites this paper.

SIFT: Selective-Index For Fast Compute of RAG Prefill by Exploiting Attention Invariance SampleAttention: Near-Lossless Acceleration of Long Context LLM Inference with Adaptive Structured Sparse Attention

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:37:31.226802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T16:24:31.109508Z digest=sha256:44d14a1be8f4486d9de26326a70844adbec43c0634981b8db3cafa85f725db51