Pith. sign in

Paper Citation Record · LEDGER

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing

As of 9 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 3 inbound Pith citation observations for arXiv:2505.19578.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19578 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:16:40.355183Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-23T03:59:58.634512Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T04:02:30.324254Z

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1dc9b640-3ba3-49b7-98d6-99fe94da9097 · outbound

This paper cites online" 'onlinestring :=.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:38.469493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:38.469493Z digest=sha256:0ac2b4c303b0c1df60c37d3dee53b262f8b7c6b206a13ba57aeda5d5db4d505c

Observation 827d8313-2989-4717-a290-4af0085fbbd4 · outbound

This paper cites write newline.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:38.543759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:38.543759Z digest=sha256:7ca482bf1fe0b1a3b1cc5ef2d03fd1d85fa92e310b17748b4724d299f09cd9e2

Observation 6915cbc0-5938-4780-b9e4-470887c3aa27 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:42.431606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:16:38.615637Z digest=sha256:7c6710e6dc533e6fcc8be074553b97ee9f03f5d5470f68225eebee68b0153457

Observation 213b999d-51af-45e2-b971-ca71fd5341a7 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:42.237866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:16:38.682780Z digest=sha256:e468f33da2adb981b481785dcda11e7305e0fb6ab17143903708eabfddcf0f29

Observation e4159931-1ccb-400a-bba9-23c45a77975f · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:42.027748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:16:38.747418Z digest=sha256:22f2f0156ee19a4323c1c00cee82dd1c00d314153bfa0fd6137acca7c2547641

Observation bf3a50cf-0d8b-4ffe-b341-7bea06ba24ca · outbound

This paper cites Challenges in Deploying Long-Context Transformers: A Theoretical Peak Performance Analysis.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Challenges in Deploying Long-Context Transformers: A Theoretical Peak Performance Analysis

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:38.844275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:38.844275Z digest=sha256:bb1c3258fb1d9056e600c15c811722e5bc9a019cbb2076cb8bcf747005a788e7

Observation 3dc423ee-3c08-4c77-8182-bfee253b9632 · outbound

This paper cites Chain-of-Thought Hub: A Continuous Effort to Measure Large Language Models' Reasoning Performance.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Chain-of-Thought Hub: A Continuous Effort to Measure Large Language Models' Reasoning Performance

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:38.906618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:38.906618Z digest=sha256:d31b6445f71a9458bf75798a3c8bc90bd3046f11ecdce27dd47171e8b79f917c

Observation 1338f882-ee10-40dc-8718-847d1628ef14 · outbound

This paper cites SeerAttention: Learning Intrinsic Sparse Attention in Your LLMs.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing SeerAttention: Learning Intrinsic Sparse Attention in Your LLMs

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:38.982877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:38.982877Z digest=sha256:7165d2513bab65c44dbcd0913cf49ee348339ac22b5868541a1b9910fd1484a4

Observation 6c284909-fc44-408e-8529-3d4c55203ece · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:41.825903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.069843Z digest=sha256:9f8b8129889c368517ed265ab8af81bce257bfa18bd0efdf7d2fd592bbdce124

Observation 9e4143fe-b934-46ba-a450-5ca26e49ac93 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:41.593628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.137625Z digest=sha256:8c35f5cf78f59d76e5f185e20e31ecc9404a408f2365fcffe2db64edd9d091b8

Observation ba716828-7304-4818-a5b0-98e495fb9fbd · outbound

This paper cites MoBA: Mixture of Block Attention for Long-Context LLMs.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing MoBA: Mixture of Block Attention for Long-Context LLMs

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:39.237793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:39.237793Z digest=sha256:252f6a8fd8d308b11ca30565b271b43f678692f213e6a47756b1c5ece33e8b8c

Observation a840e619-b27d-4529-a078-8c29998a3e3d · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:41.433549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.305185Z digest=sha256:67e661118d4d810ee2cc2f2b37e84632044df0a0c5be35b00f34f52c73d8d4f2

Observation 7fe57c19-3de6-4f68-8951-80766f1a1c2a · outbound

This paper cites Rae, Anna Potapenko, Siddhant M.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Rae, Anna Potapenko, Siddhant M

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:41.256030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.374466Z digest=sha256:d2403184e7b55cc03661cfea2a1e05842dc2f4d0971ac5eb1456a270f5a4f24f

Observation ba8af17f-dfd1-4ac8-b226-bc342feab119 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:39.492504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:39.492504Z digest=sha256:db604e65d7b28d505e520d6624b5fd79e72e359b19eefbba7104d6d589d8183d

Observation 156a87a1-6eee-47df-8e8b-6c232526fe15 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:39.578526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:39.578526Z digest=sha256:b4b085f3062ca3b05d6c8b06ffe5d00f9866d89f295d661c2c422ee058893d88

Observation d1cb9a39-2d09-4316-ba70-0d18eee7356d · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:39.677520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:39.677520Z digest=sha256:85495a9c648fcdb7d6044ca2b1fdf1eb8acd7c0ba4bb92ed17b958461f56f41c

Observation f17f4dc1-b674-4621-81e6-37dd411cba22 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:41.053429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.816195Z digest=sha256:4e02bb8f6aa9ccb68c6c6792aa74c9ac5f41e663d3894adf4ed8d354e3666e01

Observation 86afec40-3f91-4afe-9d08-80d28665fd63 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:40.874801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.924333Z digest=sha256:e4c5554cb46bd78735e587f3cb413877e2bdd818d1fa7e8b97b40c575db31f91

Observation 0c668210-0348-4e32-97f0-15d24d57d5c1 · outbound

This paper cites Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:40.017957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:40.017957Z digest=sha256:d61404e69ebca2f5c61add458468a4768dd31f94519900d16a62a55ea54642ac

Observation 153761b0-0eff-4f28-99c2-a3794f7c0422 · outbound

This paper cites Multi-agent Architecture Search via Agentic Supernet.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Multi-agent Architecture Search via Agentic Supernet

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:40.125801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:40.125801Z digest=sha256:6f575a3eca95640dc76b5e4b5fbcb357fd91878b90659e2c4605234a31f4753d

Observation 426a03a4-1ef5-42d8-a68b-01905b6e7ef9 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:40.676404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:16:40.247475Z digest=sha256:b28c584606a8de02a509a2398bd8f3153f01a13707487748a59488be3aad1452

Observation cfd5149f-e717-47e3-b0a7-3d786cb52dd1 · outbound

This paper cites Migrating Code At Scale With LLMs At Google.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Migrating Code At Scale With LLMs At Google

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:40.355183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:40.355183Z digest=sha256:e22949a41564b1ccb78f5d0d3c7615a5ba7b17b334e45ea87fd26800ad6252f6

Pith citing papers

Observation 39dad653-1567-44cb-8789-578909a41159 · inbound

FastKV: Decoupling of Context Reduction and KV Cache Compression for Prefill-Decoding Acceleration cites this paper.

FastKV: Decoupling of Context Reduction and KV Cache Compression for Prefill-Decoding Acceleration Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:02:30.327548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-23T03:59:58.634512Z digest=sha256:35057680339bc114ed04e09bb5fccfa2850a3fe4d8d141a507e126239b95ddcc

Observation 80c0fe47-b306-47a8-bb3c-4cb78229bbed · inbound

EchoKV: Efficient KV Cache Compression via Similarity-Based Reconstruction cites this paper.

EchoKV: Efficient KV Cache Compression via Similarity-Based Reconstruction Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-15T01:09:36.929964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T01:09:07.983785Z digest=sha256:367aa804ba6bd6c977aba4a3d00d8fb145e3ea0f68ea20bd3e43d4cd321fd9ad

Observation 7637f781-4ef2-450b-b8d1-c632a0af0ef3 · inbound

Flux Attention: Context-Aware Hybrid Attention for Efficient LLMs Inference cites this paper.

Flux Attention: Context-Aware Hybrid Attention for Efficient LLMs Inference Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:25:58.269821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T17:39:18.387881Z digest=sha256:4f94c7bda8965b0984ff8cd4711dfdfd401438d382eb64ab282c984e2b175123