Pith. sign in

Paper Citation Record · LEDGER

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads

As of 19 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 2 inbound Pith citation observations for arXiv:2501.15113.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.15113 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T14:41:05.147543Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T13:46:11.383126Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T09:05:58.128419Z

Reference resolution

40 of 40 outbound references displayed

  • verified exact1
  • verified fuzzy4
  • unresolved35
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2f1ccdae-e7b5-4cef-8e8b-f67e4c96942d · outbound

This paper cites A Survey on In-context Learning.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads A Survey on In-context Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:04.955353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:04.955353Z digest=sha256:74c08a766aa31e736cec36c3300bafd09d51e1896bd70f9d42d25d3ed7294ba6

Observation 91d1ecf5-b8b0-4265-b4ed-29df5387584e · outbound

This paper cites Tool Learning with Foundation Models.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads Tool Learning with Foundation Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:04.960913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:04.960913Z digest=sha256:81dbb03180833240b70352c6526f2ad7416ff3da64326330ec7c6fc5ec2ee953

Observation fdc862a5-ee84-42e3-912a-004898cd8962 · outbound

This paper cites Multi-News: a Large-Scale Multi-Document Summarization Dataset and Abstractive Hierarchical Model.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads Multi-News: a Large-Scale Multi-Document Summarization Dataset and Abstractive Hierarchical Model

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:04.966169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:04.966169Z digest=sha256:fc0b202ffa9bbcfbf2a299de0041329d315559146549465927a1bddef20eee4b

Observation 14533b86-b51e-4d82-addf-15368f9c0ee1 · outbound

This paper cites StreamingDialogue: Prolonged Dialogue Learning via Long Context Compression with Minimal Losses.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads StreamingDialogue: Prolonged Dialogue Learning via Long Context Compression with Minimal Losses

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-08-10T14:41:05.756449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T14:41:04.971128Z digest=sha256:ff680214e643d1a0ca30afa3ff5cd90662a30836bc8ca1a2b0ee3ec5d1248d28

Observation e6a7c9b2-2294-41b9-89ab-29a9ed5c17a6 · outbound

This paper cites Retrieval-Augmented Generation for Large Language Models: A Survey.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads Retrieval-Augmented Generation for Large Language Models: A Survey

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:04.975704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:04.975704Z digest=sha256:8b6a6997106209a2b683b1e454e3313dd546586d53c417b9d852c9c2a497de22

Observation 270c1694-f7a4-4dfa-9034-7966d0d99f3f · outbound

This paper cites BioRAG: A RAG-LLM Framework for Biological Question Reasoning.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads BioRAG: A RAG-LLM Framework for Biological Question Reasoning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:04.980455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:04.980455Z digest=sha256:427bf438f172f855d53b39fef468a99f4bcaced7730eba6d3bace5c837ab8270

Observation 50930c29-0d24-495e-8bad-822699b412e0 · outbound

This paper cites H2o: Heavy-hitter oracle for efficient generative inference of large language models,.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads H2o: Heavy-hitter oracle for efficient generative inference of large language models,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:41:05.898375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T14:41:04.985502Z digest=sha256:a9b888be545c90c3848ed76907bfce5fe625a0cac0149fe3103f7cf3b4c4408b

Observation 187e3ac1-45a8-4e0c-bf85-be68c19a8dea · outbound

This paper cites LongHeads: Multi-Head Attention is Secretly a Long Context Processor.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads LongHeads: Multi-Head Attention is Secretly a Long Context Processor

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:04.989763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:04.989763Z digest=sha256:251add314cff6173d2cb4dfe0e3bdd76e9efb5d0e6e33bd78f59f4c82cb7b19a

Observation 64b5f45f-4193-4d29-8e0b-0f620fa8a31a · outbound

This paper cites Efficient Streaming Language Models with Attention Sinks.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads Efficient Streaming Language Models with Attention Sinks

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:04.994236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:04.994236Z digest=sha256:616fab726ef577a28b493fe198b97a1c753fb18211cc07f64486a61ecf6496a2

Observation 57cdb415-6c07-429c-8499-23a0f071df84 · outbound

This paper cites PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:04.999148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:04.999148Z digest=sha256:a7fd14b5693657e9ac54908babf6e451fa55ef41d40bdd7805138cd6c32f330c

Observation 46fdbb1a-83a3-4239-8abf-ed34e63c0914 · outbound

This paper cites PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.004324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.004324Z digest=sha256:13fc8dd53e05c317d479d1d5e1ebcddaaaa351decf5582613db48a73aef98b19

Observation 01f823a4-458b-4548-ba99-01702a53a02d · outbound

This paper cites SnapKV: LLM Knows What You are Looking for Before Generation.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads SnapKV: LLM Knows What You are Looking for Before Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.009013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.009013Z digest=sha256:e68f76818abf1819c07f24508bf917450cceb4b1de16622a6894e61538b3b850

Observation de681a8d-578e-4497-80dd-db5200ce5d35 · outbound

This paper cites MiniCache: KV Cache Compression in Depth Dimension for Large Language Models.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads MiniCache: KV Cache Compression in Depth Dimension for Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.013818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.013818Z digest=sha256:50d6430c9b3101c0498a55b2e8f0ab125a3d9be019c38d5dbf3b730bad5b4e51

Observation 3ea337a1-1de8-4e1d-afe8-e223adcffd65 · outbound

This paper cites RazorAttention: Efficient KV Cache Compression Through Retrieval Heads.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads RazorAttention: Efficient KV Cache Compression Through Retrieval Heads

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.018480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.018480Z digest=sha256:c7fc91e42b814f226bb97f3f66d4cca7643c59e117d25a69695beea90121e5d6

Observation 1456b600-29f7-40ca-a27a-e6e4fdaa9894 · outbound

This paper cites Ada-KV: Optimizing KV Cache Eviction by Adaptive Budget Allocation for Efficient LLM Inference.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads Ada-KV: Optimizing KV Cache Eviction by Adaptive Budget Allocation for Efficient LLM Inference

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.023298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.023298Z digest=sha256:7f4a88bd295912cdf4d2f86b92886f5451f15b8c94dc44c40f04409e22a9eba6

Observation 235de114-2b4d-43c0-abb4-a832ba9db4c8 · outbound

This paper cites DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.028215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.028215Z digest=sha256:4b87d5d852b876c2f86841babb5ba8de951ac971534c4619cd0be4a0a14288bf

Observation 607948dc-9359-4ae5-865f-3496095887fc · outbound

This paper cites Not all heads matter: A head-level kv cache compression method with integrated retrieval and reasoning,.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads Not all heads matter: A head-level kv cache compression method with integrated retrieval and reasoning,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.032990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.032990Z digest=sha256:d006e9a9074bd7f9a3f9a59a74bf8c40fd66a9a58e027806929479dbbf04736c

Observation 789e0f0c-50b4-4f78-83da-a4edb2d57cdc · outbound

This paper cites Retrieval Head Mechanistically Explains Long-Context Factuality.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads Retrieval Head Mechanistically Explains Long-Context Factuality

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.037675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.037675Z digest=sha256:da623e0d5a278af46f52090386e93b53ceb21b0c4aef8500c436f73c1b94cfbb

Observation 2f11d050-6f92-435a-866d-782b30f4ade7 · outbound

This paper cites Contributions of Transformer Attention Heads in Multi- and Cross-lingual Tasks.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads Contributions of Transformer Attention Heads in Multi- and Cross-lingual Tasks

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.045154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.045154Z digest=sha256:8a5d0166ec9b93f6ca0f25899e85063f33733c8f6e6bcfb12732dcca93f97d41

Observation ed73b555-b77c-477f-b1a9-17c3bbea4cdd · outbound

This paper cites A closer look at transformer attention for multilingual translation,.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads A closer look at transformer attention for multilingual translation,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:41:05.882178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T14:41:05.053105Z digest=sha256:cd5f2037b4328c060295a2f3e672a038d0b68bf4f6a6dd254f12a26c8d73ee48

Observation ee253154-3d1c-4f8f-8b0e-8b4d4cb8fe7f · outbound

This paper cites LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.057591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.057591Z digest=sha256:d41088c4a33828e3c82c9dcaf07b7e2c44666fbe032ab9c4fc7a55ba7daccfaf

Observation f0ba6493-e27c-406f-9d64-b46f41054683 · outbound

This paper cites LooGLE: Can Long-Context Language Models Understand Long Contexts?.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads LooGLE: Can Long-Context Language Models Understand Long Contexts?

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.062461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.062461Z digest=sha256:4a859bfb4176add11ace3c5b9d7e389949ceaf867cc2eee28ba3969b573da885

Observation 99270b41-2d38-46ef-8d8b-74b50b6801f7 · outbound

This paper cites GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.067310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.067310Z digest=sha256:02184ea0b2ca4c19eaebc30eee3e68dccf58f87ee68331dc3bfba32f6946186a

Observation 6110563c-1c2d-486b-b5ed-1d8de218eb40 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.072104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.072104Z digest=sha256:6dc7291bf404480fc78a6f967c814b41e560a9063787e64a613be334c3b055a8

Observation 3b580b22-f8f7-4a8e-b9e6-df4079c13f8e · outbound

This paper cites Mistral 7B.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads Mistral 7B

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.076675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.076675Z digest=sha256:1ea93e4f05873eacb68ccc36c1d0fc322b7de68adb35edead67d0892d552407b

Observation bab3992c-02fc-4d60-9a6d-428c0c80f4a7 · outbound

This paper cites InfLLM: Training-Free Long-Context Extrapolation for LLMs with an Efficient Context Memory.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads InfLLM: Training-Free Long-Context Extrapolation for LLMs with an Efficient Context Memory

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.081285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.081285Z digest=sha256:c39243c86b65761fb5c129ac208f818935d6738f30f2bb10c31124c767d9b679

Observation 867bfa47-cabe-46a9-ba94-c207f0001ed8 · outbound

This paper cites Model Tells You What to Discard: Adaptive KV Cache Compression for LLMs.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads Model Tells You What to Discard: Adaptive KV Cache Compression for LLMs

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.086383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.086383Z digest=sha256:a647dd144d320a682b6d13727e05c2f87d04c9508028c8e080e1ae43175ac61f

Observation 786dca31-1e67-458f-b3b7-48a54f0c473b · outbound

This paper cites Principal components analysis (pca),.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads Principal components analysis (pca),

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.091154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.091154Z digest=sha256:9b9819f6bae11e2ac22019d3e4a0ebe5995f3f7ea165e9cb9ebfbd2f1b731b33

Observation 7c41285c-af47-4c36-b609-d3a2d7919af6 · outbound

This paper cites Attention is all you need,.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads Attention is all you need,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.095345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.095345Z digest=sha256:10e64633033ef0c224713efeb8f6d4f1a20de0a70c28f346e320bbca886f9542

Observation 29f70fda-990d-4de0-8647-57ce9d70a351 · outbound

This paper cites LM-Infinite: Zero-Shot Extreme Length Generalization for Large Language Models.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads LM-Infinite: Zero-Shot Extreme Length Generalization for Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.099612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.099612Z digest=sha256:9e42e8d9c7153d5f21e23439e330cbb00cac8b14af0aeb8b0007adb13aa1b6d7

Observation 55e611ed-2dce-4163-ace3-02824524d6e8 · outbound

This paper cites A Dataset of Information-Seeking Questions and Answers Anchored in Research Papers.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads A Dataset of Information-Seeking Questions and Answers Anchored in Research Papers

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.103927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.103927Z digest=sha256:ecc6e58f03113cb9f1ec5457258b78fe05751103c9157bbd03464caf1115f769

Observation b405bba8-c603-490a-8238-e3786c687bc9 · outbound

This paper cites HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.108203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.108203Z digest=sha256:69912d4111d5fcd9edeb5813cb98ccba374765859ffb0958ecfe036836f8e9fc

Observation e52b249a-740e-4685-9fd1-f4255a219cc3 · outbound

This paper cites Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.112914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.112914Z digest=sha256:05cd4ce65cb8c9db687db64b3f5b6ee22173d1fb98a7ec939e29b47f06931589

Observation 82fa0b58-9acb-4d82-833b-baeb2cc29240 · outbound

This paper cites QMSum: A New Benchmark for Query-based Multi-domain Meeting Summarization.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads QMSum: A New Benchmark for Query-based Multi-domain Meeting Summarization

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.117333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.117333Z digest=sha256:31e30e63599150e0d06b87b15321fbbf4a623ea4dc63be6be5d774e35667b7d4

Observation 895f4727-556e-4266-8a88-567002d11549 · outbound

This paper cites Efficient Attentions for Long Document Summarization.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads Efficient Attentions for Long Document Summarization

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.122694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.122694Z digest=sha256:c221af46f758bfe7933cc638d9b7edbc8cdc3be0032355fdb3caa0aa78d57f8f

Observation 350bfaf1-8983-4426-912c-d06967c507ac · outbound

This paper cites TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.127858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.127858Z digest=sha256:61b1cd7a864430ad627fcf676aaf78151356f54b8883729d630d8c184fc6e89a

Observation 1e7a16d4-c105-48c2-a9e3-28962695ed2d · outbound

This paper cites Learning question classifiers,.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads Learning question classifiers,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:41:05.846991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T14:41:05.132863Z digest=sha256:930885ace8f10c130402517ffc0c824eea20a82f9198d9b85863d2e9a974ca43

Observation ea4a1dda-b874-4ca7-9ea9-4d339c9e065f · outbound

This paper cites Exploiting semantic resources for large scale text categorization,.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads Exploiting semantic resources for large scale text categorization,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T14:41:05.830603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-10T14:41:05.137793Z digest=sha256:b0fe63b6ee337a1bec7762c5375516ac0208fb309d78b544b643762f41377f0c

Observation 7a138ed5-73a8-4825-976e-86791e37726a · outbound

This paper cites RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.142575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.142575Z digest=sha256:ba5ec046c9dbeaec02193517fab81f749afb18ff301f849b3b3ef704866ad2bd

Observation c233d579-5980-4642-81ad-37da6ff46449 · outbound

This paper cites Longcoder: A long- range pre-trained language model for code completion,.

Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads Longcoder: A long- range pre-trained language model for code completion,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T14:41:05.147543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:41:05.147543Z digest=sha256:f3a14ca6178f3f0f1effcdeb818914aedc966bb01338289ed87197200f7cf0fa

Pith citing papers

Observation 1b23ff28-9c93-44f8-b1d1-d7b36a38705d · inbound

Unveiling Simplicities of Attention: Adaptive Long-Context Head Identification cites this paper.

Unveiling Simplicities of Attention: Adaptive Long-Context Head Identification Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T13:46:11.383126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:46:11.383126Z digest=sha256:280cd3def9ab101009be2c0ea98dfd93c31a2b80c00c452578f3d26abf9b0b57

Observation 84eace44-4527-4779-9b34-87224001a7b8 · inbound

Attention Sink in Transformers: A Survey on Utilization, Interpretation, and Mitigation cites this paper.

Attention Sink in Transformers: A Survey on Utilization, Interpretation, and Mitigation Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:05:58.130976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T16:17:09.834609Z digest=sha256:68f2379edab3e8551b6f1d9bedf88709b444b3ebd538f7788937d66b4388c06f