Pith. sign in

Paper Citation Record · LEDGER

Efficient Attentions for Long Document Summarization

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 29 inbound Pith citation observations for arXiv:2104.02112.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2104.02112 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 29 of 29 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:43:52.626266Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T21:57:26.175240Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 10964c7f-fb85-46cd-923d-444d78d92895 · inbound

Retentive Network: A Successor to Transformer for Large Language Models cites this paper.

Retentive Network: A Successor to Transformer for Large Language Models Efficient Attentions for Long Document Summarization

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:29:59.827872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-11T20:29:59.633357Z digest=sha256:3fffe0098042599b5b11dd9ad002b4b9759e9129421ac1a2dfd79ae1e9ef9fa8

Observation 58304fc4-6592-4cf3-916f-6b12d957eeb9 · inbound

Ada-KV: Optimizing KV Cache Eviction by Adaptive Budget Allocation for Efficient LLM Inference cites this paper.

Ada-KV: Optimizing KV Cache Eviction by Adaptive Budget Allocation for Efficient LLM Inference Efficient Attentions for Long Document Summarization

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-17T11:16:32.060641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-17T11:16:31.904921Z digest=sha256:3fdfe712716eb285851c70c47708da72dd680dd8cd4f197841a14a5cf494fc82

Observation 4f359a4d-f3ed-447d-92e0-caefd4b71304 · inbound

Breaking the Stage Barrier: A Novel Single-Stage Approach to Long Context Extension for Large Language Models cites this paper.

Breaking the Stage Barrier: A Novel Single-Stage Approach to Long Context Extension for Large Language Models Efficient Attentions for Long Document Summarization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T19:11:06.092188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T19:11:06.092188Z digest=sha256:5a25a5cea553c38bee68275ca9a00c212ec7b639395e15b6b1e76363e15cde6f

Observation 326f13d4-abc0-4fe5-a144-9b06272784dc · inbound

More Tokens, Lower Precision: Towards the Optimal Token-Precision Trade-off in KV Cache Compression cites this paper.

More Tokens, Lower Precision: Towards the Optimal Token-Precision Trade-off in KV Cache Compression Efficient Attentions for Long Document Summarization

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T13:52:48.695396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T13:52:48.695396Z digest=sha256:a8a71d409c8ad44382f73dbc7f892ecc12e228d0b95bbbc2c4448c224f70fa08

Observation a0c11acb-7f8d-4164-b65d-acf589380943 · inbound

Fourier Position Embedding: Enhancing Attention's Periodic Extension for Length Generalization cites this paper.

Fourier Position Embedding: Enhancing Attention's Periodic Extension for Length Generalization Efficient Attentions for Long Document Summarization

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T05:18:54.484670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:18:54.484670Z digest=sha256:3fc8367acef0473a50686ea24efb9614febd25e4c100dc461d09104e75f92eb9

Observation 4bb28842-85ce-43bb-9aad-aecc76921dad · inbound

CaseSumm: A Large-Scale Dataset for Long-Context Summarization from U.S. Supreme Court Opinions cites this paper.

CaseSumm: A Large-Scale Dataset for Long-Context Summarization from U.S. Supreme Court Opinions Efficient Attentions for Long Document Summarization

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T23:05:07.292596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T23:05:07.292596Z digest=sha256:0ab5e6ab1929ee05794793e89ca3f7272c2379ce74b5764686203d652c3173de

Observation 1e410ddf-31ce-4697-a539-4964f08141cf · inbound

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference cites this paper.

AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference Efficient Attentions for Long Document Summarization

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:51.258739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:51.258739Z digest=sha256:ca0e19ff245c3a6f039d2eaca907d7abadda2c474e601ac4c62e906c28ce64f1

Observation 895f4727-556e-4266-8a88-567002d11549 · inbound

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads cites this paper.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads Efficient Attentions for Long Document Summarization

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.122694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.122694Z digest=sha256:c221af46f758bfe7933cc638d9b7edbc8cdc3be0032355fdb3caa0aa78d57f8f

Observation f5345240-a8a7-4fda-be5b-148b51d37ebe · inbound

CriticalKV: Optimizing KV Cache Eviction from an Output Perturbation Perspective cites this paper.

CriticalKV: Optimizing KV Cache Eviction from an Output Perturbation Perspective Efficient Attentions for Long Document Summarization

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T00:45:56.863397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T00:45:56.863397Z digest=sha256:c3e5364febc0ce9780f5acb4214fb6ad1bafd8515f836e92363a5051519adc92

Observation 6c19549a-e3fa-422d-8858-65e281fd969a · inbound

HalluMix: A Task-Agnostic, Multi-Domain Benchmark for Real-World Hallucination Detection cites this paper.

HalluMix: A Task-Agnostic, Multi-Domain Benchmark for Real-World Hallucination Detection Efficient Attentions for Long Document Summarization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:52.626266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:52.626266Z digest=sha256:a90146a0617811ab42315102662a108a79ccf4ffe8061bf65da00f327f5b0126

Observation b142c41e-b56c-49c1-9beb-a3eb0de83948 · inbound

A Split-then-Join Approach to Abstractive Summarization for Very Long Documents in a Low Resource Setting cites this paper.

A Split-then-Join Approach to Abstractive Summarization for Very Long Documents in a Low Resource Setting Efficient Attentions for Long Document Summarization

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:06.000675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T22:35:06.000675Z digest=sha256:d40eafa460d1291c1a632f0655af6e0a4d5d654d0fb1f67a1fd2d8ed7306a5ba

Observation 402ad735-96af-4d01-bd92-776a4aac9fba · inbound

Alignment-Augmented Speculative Decoding with Alignment Sampling and Conditional Verification cites this paper.

Alignment-Augmented Speculative Decoding with Alignment Sampling and Conditional Verification Efficient Attentions for Long Document Summarization

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T20:22:04.424586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:22:04.424586Z digest=sha256:4f0e34d498242e9a9143c245626f659f5ad05d4685fd9b5280e9e170725b0043

Observation a6377163-eeb6-429d-a227-4de38537b4b9 · inbound

PagedEviction: Structured Block-wise KV Cache Pruning for Efficient Large Language Model Inference cites this paper.

PagedEviction: Structured Block-wise KV Cache Pruning for Efficient Large Language Model Inference Efficient Attentions for Long Document Summarization

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T10:16:13.078259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:16:13.078259Z digest=sha256:68f54e1e00a690bb0604f081dfc474cddaa19a8a218c064a9b7ec7b2e77c7d0c

Observation c574ba02-1743-4dc4-85c6-245e7c1ffe8f · inbound

HoPE: Hyperbolic Rotary Positional Encoding for Stable Long-Range Dependency Modeling in Large Language Models cites this paper.

HoPE: Hyperbolic Rotary Positional Encoding for Stable Long-Range Dependency Modeling in Large Language Models Efficient Attentions for Long Document Summarization

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T05:36:53.837356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T05:36:53.837356Z digest=sha256:62193d25148d5f9bbb9a6e7ab64fa4061a44fe80091cfd194dffcbb903a818b8

Observation 3cf5adb1-59b7-46cb-b6ee-be469b6c2134 · inbound

EvolKV: Evolutionary KV Cache Compression for LLM Inference cites this paper.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Efficient Attentions for Long Document Summarization

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.091290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.091290Z digest=sha256:ab79228455343ca3240141414e0008bdfd3e0f8167a3b26a3b6910c0a9419777

Observation 73d24d71-cbd5-46ac-b7a9-ee8255c9b99c · inbound

CompilerKV: Risk-Adaptive KV Compression via Offline Experience Compilation cites this paper.

CompilerKV: Risk-Adaptive KV Compression via Offline Experience Compilation Efficient Attentions for Long Document Summarization

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-21T13:34:11.434119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T13:32:15.598047Z digest=sha256:841656c89903d39ecb3a4e08ed89c3d8313cdd441e5ad26088c354f9a1b257ff

Observation 28a89545-aa60-4a94-950b-458d7654da3b · inbound

CompilerKV: Risk-Adaptive KV Compression via Offline Experience Compilation cites this paper.

CompilerKV: Risk-Adaptive KV Compression via Offline Experience Compilation Efficient Attentions for Long Document Summarization

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T03:16:19.880944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:16:19.880944Z digest=sha256:6d3fbf6cef9e02cc3fdebbba05e38bc8e11c6b972b3d7c087eb2a29a67772814

Observation f7662b0b-ce84-4628-af95-f4be3d9bf2f4 · inbound

From Global to Local: Learning Context-Aware Graph Representations for Document Classification and Summarization cites this paper.

From Global to Local: Learning Context-Aware Graph Representations for Document Classification and Summarization Efficient Attentions for Long Document Summarization

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T04:56:23.067350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:56:23.067350Z digest=sha256:f3d0d354d22d1c17a2bd97929e17b443d5130d2b18c70ad0120af66f1f3c80e9

Observation bf15c77a-b126-42c8-9175-e740885c74e6 · inbound

SparKV: Overhead-Aware KV Cache Loading for Efficient On-Device LLM Inference cites this paper.

SparKV: Overhead-Aware KV Cache Loading for Efficient On-Device LLM Inference Efficient Attentions for Long Document Summarization

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:46:07.739539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-08T14:09:30.821354Z digest=sha256:3ee437b4bc3dc0a70bae652c989d9d517b5754c0cf60b3f57d5e12daa33c7924

Observation 743a815a-e7ec-46bc-bfc5-be887273ba95 · inbound

DepthKV: Layer-Dependent KV Cache Pruning for Long-Context LLM Inference cites this paper.

DepthKV: Layer-Dependent KV Cache Pruning for Long-Context LLM Inference Efficient Attentions for Long Document Summarization

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T21:56:12.778516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-08T03:49:01.267556Z digest=sha256:daba32a314ef1569209c933e1d97a01989e19bc80275df3a9ffd5e3876c52178

Observation 467e96e9-9e8b-48ed-b76d-56ada5ccb4a5 · inbound

FlashRT: Towards Computationally and Memory Efficient Red-Teaming for Prompt Injection and Knowledge Corruption cites this paper.

FlashRT: Towards Computationally and Memory Efficient Red-Teaming for Prompt Injection and Knowledge Corruption Efficient Attentions for Long Document Summarization

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:16:28.252669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-07T06:49:56.316472Z digest=sha256:ecd02b18fe8f90d2a74df636aaf03cc2503238278b512706289ac5527c58d7e6

Observation d560fc7d-073d-4c2d-b73c-3c99f851fcb6 · inbound

How to Compress KV Cache in RL Post-Training? Shadow Mask Distillation for Memory-Efficient Alignment cites this paper.

How to Compress KV Cache in RL Post-Training? Shadow Mask Distillation for Memory-Efficient Alignment Efficient Attentions for Long Document Summarization

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:25:57.314115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-11T01:25:30.458137Z digest=sha256:e9d6cb32f8da52823359f511c65673f6a9cc4188f1f64bf9f9ebd955554f237d

Observation f06a8140-4e99-4091-ad4d-147156ba33a2 · inbound

Reformulating KV Cache Eviction Problem for Long-Context LLM Inference cites this paper.

Reformulating KV Cache Eviction Problem for Long-Context LLM Inference Efficient Attentions for Long Document Summarization

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:10:53.455240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-11T02:37:52.545943Z digest=sha256:04d8a3c91f280330d26c3c0e2485e9eead32606a955fa5c088b171be60431a9b

Observation 4c13e76e-bb39-4d91-8dd1-34a3841f5f1d · inbound

Larch: Learned Query Optimization for Semantic Predicates cites this paper.

Larch: Learned Query Optimization for Semantic Predicates Efficient Attentions for Long Document Summarization

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-02T21:57:26.176647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-27T19:21:22.978335Z digest=sha256:343f3b840e897dccc966c9de1874a2a570f469261fb5eec10e67bb1f6ddddf35

Observation b184685a-4cee-4247-b016-892215fc08e1 · inbound

Coverage-Driven KV Cache Eviction for Efficient and Improved Inference of LLM cites this paper.

Coverage-Driven KV Cache Eviction for Efficient and Improved Inference of LLM Efficient Attentions for Long Document Summarization

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:24:21.377682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-30T07:19:48.530272Z digest=sha256:95133bdab76c3c9f1074d3e9ad1d56803469c7a0d6e783e118c85327c9d1d951

Observation d68bd663-9252-47ea-98fe-446a8d2f2f6c · inbound

Demystifying the Design Space and Best Practices for Heterogeneous LLM Inference and Serving cites this paper.

Demystifying the Design Space and Best Practices for Heterogeneous LLM Inference and Serving Efficient Attentions for Long Document Summarization

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-06-30T14:04:45.313977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T05:37:13.211613Z digest=sha256:382a1caa6d356fcfda8b7c1e97df5d081e41d8aabdab59da47ef0496c0951c6e

Observation 5c30acc4-9429-4ffb-934b-b7c5183fe016 · inbound

Demystifying the Design Space and Best Practices for Heterogeneous LLM Inference and Serving cites this paper.

Demystifying the Design Space and Best Practices for Heterogeneous LLM Inference and Serving Efficient Attentions for Long Document Summarization

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-01T08:55:35.166941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-01T07:06:53.318182Z digest=sha256:bc19b67d58191a6153fe08b67ff68babe8ae4cf070ee1325c54640f7c9675535

Observation 746124be-e2f0-4d8c-9150-8c2d77111e8e · inbound

RED-PIM: Reducing Data Movement for Transformers using Processing-in-Memory cites this paper.

RED-PIM: Reducing Data Movement for Transformers using Processing-in-Memory Efficient Attentions for Long Document Summarization

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T06:59:51.397427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T06:59:51.397427Z digest=sha256:88dd934cc1a973c0a510ada062505ac06433769802e11f656053fa87acd9506d

Observation f16433bc-b32b-4c31-a56a-a5c617c6fc45 · inbound

When Local Variance Optimality Is Not Enough: RoPE-Aligned Q/K Rotations for Dynamic 4-Bit Quantisation cites this paper.

When Local Variance Optimality Is Not Enough: RoPE-Aligned Q/K Rotations for Dynamic 4-Bit Quantisation Efficient Attentions for Long Document Summarization

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-14T12:54:26.515463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T12:54:26.515463Z digest=sha256:4aa749cdec14562bb51c3de3c4fae56c00c814f830a302ce49cebe86ad400ab0