Pith. sign in

Paper Citation Record · LEDGER

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing

As of 21 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 3 inbound Pith citation observations for arXiv:2505.19578.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19578 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:16:40.355183Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-23T03:59:58.634512Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T04:02:30.324254Z

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1dc9b640-3ba3-49b7-98d6-99fe94da9097 · outbound

This paper cites online" 'onlinestring :=.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:38.469493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:38.469493Z digest=sha256:6aa5bb9dc329c301b006049e05adac0fc53700c7f65fb20a480013b350a4304e

Observation 827d8313-2989-4717-a290-4af0085fbbd4 · outbound

This paper cites write newline.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:38.543759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:38.543759Z digest=sha256:827eba036dc3514d55ace4eb18bc31070bed80a006684a75fecb07234c1c50c1

Observation 6915cbc0-5938-4780-b9e4-470887c3aa27 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:42.431606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T14:16:38.615637Z digest=sha256:b69353b723b43b6b4b5d4c5972577565aabefe5ed3344b901bd341768b378ce3

Observation 213b999d-51af-45e2-b971-ca71fd5341a7 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:42.237866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T14:16:38.682780Z digest=sha256:fa5d7e23a4d4339d50f573ba50e6690e94489153ad378640e93b72b85c20e8b8

Observation e4159931-1ccb-400a-bba9-23c45a77975f · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:42.027748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T14:16:38.747418Z digest=sha256:fe8c14bceddda2825d06f54bc28cf351ab081cd20848fdf007338d79c4eb68b0

Observation bf3a50cf-0d8b-4ffe-b341-7bea06ba24ca · outbound

This paper cites Challenges in Deploying Long-Context Transformers: A Theoretical Peak Performance Analysis.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Challenges in Deploying Long-Context Transformers: A Theoretical Peak Performance Analysis

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:38.844275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:38.844275Z digest=sha256:b1d8b2626c5d090dcc0c4c690fc6b87fd403ad3edc3c1019fb0934e866d3706d

Observation 3dc423ee-3c08-4c77-8182-bfee253b9632 · outbound

This paper cites Chain-of-Thought Hub: A Continuous Effort to Measure Large Language Models' Reasoning Performance.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Chain-of-Thought Hub: A Continuous Effort to Measure Large Language Models' Reasoning Performance

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:38.906618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:38.906618Z digest=sha256:e0004a62e4c1afa344881ef470a182f2dfaf0a9e55ad51aba2512998daa51613

Observation 1338f882-ee10-40dc-8718-847d1628ef14 · outbound

This paper cites SeerAttention: Learning Intrinsic Sparse Attention in Your LLMs.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing SeerAttention: Learning Intrinsic Sparse Attention in Your LLMs

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:38.982877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:38.982877Z digest=sha256:a0423fc267356dfb93372eda271ae2531b9a72aba4ebca71c3f46806975e5140

Observation 6c284909-fc44-408e-8529-3d4c55203ece · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:41.825903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.069843Z digest=sha256:dae75ae34b3fc14ba4cda195948291e627ec1b5c8fd80d4b64eee40acfe65a68

Observation 9e4143fe-b934-46ba-a450-5ca26e49ac93 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:41.593628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.137625Z digest=sha256:0d80a58eb9d26dc0122b7599b4cb5d54233f24a10d27261fb0b724ee4d6eeb48

Observation ba716828-7304-4818-a5b0-98e495fb9fbd · outbound

This paper cites MoBA: Mixture of Block Attention for Long-Context LLMs.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing MoBA: Mixture of Block Attention for Long-Context LLMs

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:39.237793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:39.237793Z digest=sha256:eac4226ae00dedeb7350afe95c2ebb1bc5db4cac4cd814a6856cea4a7f4b70b3

Observation a840e619-b27d-4529-a078-8c29998a3e3d · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:41.433549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.305185Z digest=sha256:5bc37111750e01a7db72c6d29f205fabf2618e978d92ac2b21ad6a27d5045767

Observation 7fe57c19-3de6-4f68-8951-80766f1a1c2a · outbound

This paper cites Rae, Anna Potapenko, Siddhant M.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Rae, Anna Potapenko, Siddhant M

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:41.256030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.374466Z digest=sha256:cbb80f3b66a342efae4f0141fd597a83aa5598c3cc22ce03d9124bfe249dee05

Observation ba8af17f-dfd1-4ac8-b226-bc342feab119 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:39.492504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:39.492504Z digest=sha256:923b17ee85fadec4a78e95ff08d5f53a8a2a420a1e8cada95c9fc18081c20498

Observation 156a87a1-6eee-47df-8e8b-6c232526fe15 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:39.578526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:39.578526Z digest=sha256:85c61fb5ac09e2f83b4c921155f6589e1339beb7e0dcc6d4c4e940f1e7a08991

Observation d1cb9a39-2d09-4316-ba70-0d18eee7356d · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:39.677520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:39.677520Z digest=sha256:cecb512e4269182d132d08bbdcd4d85b862f4b8d9c34252a4a84825d662d23c6

Observation f17f4dc1-b674-4621-81e6-37dd411cba22 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:41.053429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.816195Z digest=sha256:e4f7f5fad5ca3f08c14a2c75904b382b48dc2be80e69b0a2c22fb343537e98e4

Observation 86afec40-3f91-4afe-9d08-80d28665fd63 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:40.874801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.924333Z digest=sha256:e14314008aed8360ddacdc2caefb88f8c44def545ff21e8a098cdc804444f570

Observation 0c668210-0348-4e32-97f0-15d24d57d5c1 · outbound

This paper cites Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:40.017957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:40.017957Z digest=sha256:3b15b4ff573145e830d32e3ce1fc68801cc0fe134e9ffaeea86d60d861c2bbfa

Observation 153761b0-0eff-4f28-99c2-a3794f7c0422 · outbound

This paper cites Multi-agent Architecture Search via Agentic Supernet.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Multi-agent Architecture Search via Agentic Supernet

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:40.125801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:40.125801Z digest=sha256:596d24fd6abd31faa98a4277e83aff19db9954720362c789cc4a175581765fe4

Observation 426a03a4-1ef5-42d8-a68b-01905b6e7ef9 · outbound

This paper cites an unresolved cited work.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:40.676404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-07T14:16:40.247475Z digest=sha256:a96140346b72bf784c8dcdb0c1df35802bf032d102156b04f4c74ddb910a62b2

Observation cfd5149f-e717-47e3-b0a7-3d786cb52dd1 · outbound

This paper cites Migrating Code At Scale With LLMs At Google.

Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing Migrating Code At Scale With LLMs At Google

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:40.355183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:40.355183Z digest=sha256:a16dd37cb294485e4bea2b78e1d3046d929337a314f8724d6a07fa2a63a877cc

Pith citing papers

Observation 39dad653-1567-44cb-8789-578909a41159 · inbound

FastKV: Decoupling of Context Reduction and KV Cache Compression for Prefill-Decoding Acceleration cites this paper.

FastKV: Decoupling of Context Reduction and KV Cache Compression for Prefill-Decoding Acceleration Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:02:30.327548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-23T03:59:58.634512Z digest=sha256:fbba2c464a507c512b545f150d51053ed4d6076a5d32ed406194d94b55d25b7f

Observation 80c0fe47-b306-47a8-bb3c-4cb78229bbed · inbound

EchoKV: Efficient KV Cache Compression via Similarity-Based Reconstruction cites this paper.

EchoKV: Efficient KV Cache Compression via Similarity-Based Reconstruction Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-15T01:09:36.929964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T01:09:07.983785Z digest=sha256:8334464c5c320ff00190d6d59f573fe6f31d7f37c60f2d0ea4b66070a84fe787

Observation 7637f781-4ef2-450b-b8d1-c632a0af0ef3 · inbound

Flux Attention: Context-Aware Hybrid Attention for Efficient LLMs Inference cites this paper.

Flux Attention: Context-Aware Hybrid Attention for Efficient LLMs Inference Accelerating Prefilling for Long-Context LLMs via Sparse Pattern Sharing

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:25:58.269821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T17:39:18.387881Z digest=sha256:ebd18d5b782ad610ce37960c93ff0496e6de326baee4c1ca1d48045bae9af9a0