Pith. sign in

Paper Citation Record · LEDGER

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning

As of 19 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 1 inbound Pith citation observation for arXiv:2605.23200.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.23200 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-25T05:22:47.198185Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-31T11:49:11.776915Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

54 of 54 outbound references displayed

  • verified exact27
  • verified fuzzy16
  • unresolved6
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 47d7dafa-e7e1-49c3-a1e4-23504f6b75df · outbound

This paper cites H2O: Heavy-hitter oracle for efficient generative inference of large language models.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning H2O: Heavy-hitter oracle for efficient generative inference of large language models

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T11:56:56.485249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:066d434fba27e852263f2a74fb263420cc43b5304ec6aa0c99a10dd33a214b5a

Observation fb1e026e-cfb4-49f4-8a0e-915ee2b7bb2a · outbound

This paper cites Efficient Memory Management for Large Language Model Serving with PagedAttention , booktitle =.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Efficient Memory Management for Large Language Model Serving with PagedAttention , booktitle =

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-25T05:25:22.801192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:27510f1e8515680dfb63c37353cf4ddc228e4e85e22f940b1d34bb99c2dd9492

Observation 4bfb4e11-c9c5-404c-906c-7a5c0666f1d5 · outbound

This paper cites KVQuant: Towards 10 Million Context Length LLM Inference with KV Cache Quantization.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning KVQuant: Towards 10 Million Context Length LLM Inference with KV Cache Quantization

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:25:23.488548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:e3f3d9fd1ebdc3028169a7989f22de6bf8081749b736e1726c91720f958406fd

Observation d3eec471-38d6-4772-8839-2152578a6869 · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-25T05:25:23.493343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:45d2eb070985b5af9f6c160c3133fa4d5678f8573d1aca932e272d7457a4ba02

Observation 3edd48f8-8567-48c5-b3cb-31e363df5aeb · outbound

This paper cites R-KV: Redundancy-aware KV cache compression for reasoning models.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning R-KV: Redundancy-aware KV cache compression for reasoning models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:25:22.808246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:2d17f39f156627ba3c06585db0bc1d17248f8d5636dca2785eddda89d9622a95

Observation ec7e0a2d-cbb2-4d58-b567-5a53d0c9d029 · outbound

This paper cites Reasoning path compression: Compressing generation trajectories for efficient LLM reasoning.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Reasoning path compression: Compressing generation trajectories for efficient LLM reasoning

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T11:56:56.499401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:b07e6e4cb9fe1efe94600bff49a649a515793e386913e2ede7552e85750868b1

Observation 872c8772-af5a-427b-baf2-199cdcba217d · outbound

This paper cites NeurIPS 2025 Poster.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning NeurIPS 2025 Poster

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T11:56:56.522328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:f76f21327ba363e491add663e840602554acfe15f4c097e1988cc232a49407d9

Observation 585c9d86-b34e-410e-87fb-15c9d8aeb1d6 · outbound

This paper cites TriAttention: Efficient Long Reasoning with Trigonometric KV Compression.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning TriAttention: Efficient Long Reasoning with Trigonometric KV Compression

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-25T05:25:23.483565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:95f4607d9a935f80d9caf59bfdbb34f568e17cfa843eeab04b436d1056ccd443

Observation 9ec922cb-5ab8-452e-87f7-cce24f22df58 · outbound

This paper cites TriAttention: Efficient Long Reasoning with Trigonometric KV Compression.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning TriAttention: Efficient Long Reasoning with Trigonometric KV Compression

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-05-25T05:25:22.790851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:9f34f173544ec26e3d25cd0baee3afe50b18e73ec33d6239728074a57d0f6046

Observation a236cba2-8f0b-45aa-9f3f-5b057acc0283 · outbound

This paper cites FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-05-25T05:25:22.796195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:4bba263bb223feafbbf099bca773f1d88ab7e5394d67abcd73e2592f5cff156a

Observation 6b004f1c-b16b-4823-b4ba-b3d9833e2108 · outbound

This paper cites Efficient Streaming Language Models with Attention Sinks.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Efficient Streaming Language Models with Attention Sinks

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-25T05:25:23.500159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:609ee7f72f2c4dec6d3ee2f5c06c79e264605e1fc18de52ee7a2b85c94e7999d

Observation deaddc38-fd8b-4a00-b0a9-a450bee7d478 · outbound

This paper cites Transformers are multi-state RNNs.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Transformers are multi-state RNNs

Reference 12

Resolution
verified exact
doi, observed 2026-05-25T05:25:22.844753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:2ac397e91335bad74746dc0cac5b5b27e0fb414a1455268ce19a041b33d69b36

Observation 66bd86df-7bcf-4680-bd7d-f0fe7b4158ef · outbound

This paper cites an unresolved cited work.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-05-25T11:56:56.474810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:be2fd018434a273ca96e8096043a2fcd30d6d53989ac36f937ce8fa2421089db

Observation 39be823e-c8cd-4e40-9f5d-f9be27e03428 · outbound

This paper cites PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-25T05:25:22.830790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:21d17e092366b604521a8ebbad324e45e9448d251bb79665c378cc8828269da0

Observation 77d77e80-9eaf-4657-bc06-0d1495c20bd8 · outbound

This paper cites Omnikv: Dynamic context selection for efficient long-context LLMs.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Omnikv: Dynamic context selection for efficient long-context LLMs

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T11:56:56.516073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:81088917e6d3e6781d6a0a2239247ec4f026719df1a975fd1303388284d6a957

Observation fdc12e03-b86d-439d-8fa7-f8f5c6a9afc0 · outbound

This paper cites Lacache: Ladder-shaped KV caching for efficient long-context modeling of large language models.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Lacache: Ladder-shaped KV caching for efficient long-context modeling of large language models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T11:56:56.502962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:e0b3b3c1abd761d6932d7771140293caf2079aaccb4f326404c8e937287ae06a

Observation f6ef1f4f-8a40-4a43-8f63-a5272cdff75d · outbound

This paper cites Ada-KV: Optimizing KV Cache Eviction by Adaptive Budget Allocation for Efficient LLM Inference.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Ada-KV: Optimizing KV Cache Eviction by Adaptive Budget Allocation for Efficient LLM Inference

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-25T05:25:22.824928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:83431197e9b9e26960ca1f4618d9e397084247896a12a12e84056040ff60d6d6

Observation 04a61047-c70c-49ff-9da9-265f8a944d75 · outbound

This paper cites Duoattention: Efficient long-context LLM inference with retrieval and streaming heads.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Duoattention: Efficient long-context LLM inference with retrieval and streaming heads

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T11:56:56.535094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:67e3e9b80de1b6df77bc5485d8be17c307f974a8b86ccb6671db30b6a15d4b09

Observation 4411b3e0-230f-4d3a-9aee-7c98427b3f52 · outbound

This paper cites Not all heads matter: A head-level KV cache compression method with integrated retrieval and reasoning.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Not all heads matter: A head-level KV cache compression method with integrated retrieval and reasoning

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:25:22.836196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:3676010248bd8e1e02f6ff268fc3de45f57f3ee671b55184ab153521d1977292

Observation f0765777-e841-4e2a-9689-e3fad1c48fef · outbound

This paper cites Razorattention: Efficient kv cache compression through retrieval heads.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Razorattention: Efficient kv cache compression through retrieval heads

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T11:56:56.538152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:7fd99b0660d883a6953e5f9a0254fda3f92662ded6e1983f3559c049490502fe

Observation 76bdaba2-da55-4d83-a34e-de2eb30fc050 · outbound

This paper cites SnapKV: LLM Knows What You are Looking for Before Generation.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning SnapKV: LLM Knows What You are Looking for Before Generation

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-25T05:25:22.925216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:2e1358d3a911fde853efa93afcb78930f329ee6b534d30758b4dba18eefad445

Observation d0127d42-256f-4fa7-9068-dd540bef59e3 · outbound

This paper cites Sablock: Semantic-aware KV cache eviction with adaptive compression block size.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Sablock: Semantic-aware KV cache eviction with adaptive compression block size

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:25:22.813646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:4de832a9c3eb56d5192f7783e8602ecee791cb5a9bc45057c56445a751cec61c

Observation 21e7a736-04de-4a3f-a0cd-e564770cedfe · outbound

This paper cites ClusterKV: Manipulating LLM KV Cache in Semantic Space for Recallable Compression.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning ClusterKV: Manipulating LLM KV Cache in Semantic Space for Recallable Compression

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:25:22.818782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:bd1f37a68c954d05990d7aac0e2fd82d6893d69ea6ae57a7037f30a022e09003

Observation b56acc4e-0235-4b0f-82c4-ac8b0f226949 · outbound

This paper cites Protokv: Long-context knowledges are already well-organized before your query.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Protokv: Long-context knowledges are already well-organized before your query

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T11:56:56.509669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:9646e857490c4d017ebf086cd1d3aaeb80602f6f40c919f34a5ee0a508a99454

Observation 7a0673d6-ba0f-47f9-ba61-5c114be5d4de · outbound

This paper cites TreeKV: Smooth Key-Value Cache Compression with Tree Structures.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning TreeKV: Smooth Key-Value Cache Compression with Tree Structures

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:25:22.841372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:072ab9c37f4847b9cc04f07fdfcc6dbb15ea2dbf016e22e84af8f115df741097

Observation 3e02c168-b4cd-47a0-9f10-bc9227dd97f8 · outbound

This paper cites HeteroCache: A Dynamic Retrieval Approach to Heterogeneous KV Cache Compression for Long-Context LLM Inference.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning HeteroCache: A Dynamic Retrieval Approach to Heterogeneous KV Cache Compression for Long-Context LLM Inference

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-25T05:25:22.915182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:e70e4caf75b1ca23a2ab36e3d69873e8ff40a70d9ff3832d24e93ccb8d9db7e1

Observation 099d8e43-74be-400a-b109-d68ff2d18a07 · outbound

This paper cites Expected attention: KV cache compression by estimating attention from future queries distribution.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Expected attention: KV cache compression by estimating attention from future queries distribution

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:25:22.906104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:2ed732f6cec6c517179b0a705715d04f25b1a3caa3172349bc09e9730f3840cc

Observation 70f175fd-eaf1-4b77-9e30-30a292b444f7 · outbound

This paper cites Chunkkv: Semantic-preserving KV cache compression for efficient long-context LLM inference.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Chunkkv: Semantic-preserving KV cache compression for efficient long-context LLM inference

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:25:22.900743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:34974f1f75f90d6736726a97ccec15450d98a8fd7e0579ee8a4333cbc56c6b9c

Observation f80f96d5-4bdc-451d-8af6-57f8d28139a8 · outbound

This paper cites ReST-KV: Robust KV cache eviction with layer-wise output reconstruction and spatial-temporal smoothing.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning ReST-KV: Robust KV cache eviction with layer-wise output reconstruction and spatial-temporal smoothing

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T11:56:56.495550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:43109766b2ba20e5c209e630eb3deae58006c54ab2088a87a338afd8d69c4161

Observation 4c3e7acf-4664-40f4-9dac-e7392da40686 · outbound

This paper cites Kediff: Key similarity-based KV cache eviction for long-context LLM inference in resource-constrained environments.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Kediff: Key similarity-based KV cache eviction for long-context LLM inference in resource-constrained environments

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:25:22.910954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:8cd2962e1bd5ae231624e180a4109e5f79bbfefc358d5a626c57d827ff458625

Observation cc45516e-0e19-40b4-9b17-2300cc497529 · outbound

This paper cites SCOPE: Optimizing Key-Value Cache Compression in Long-context Generation.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning SCOPE: Optimizing Key-Value Cache Compression in Long-context Generation

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:25:22.920136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:26e976e6e9ea660e5c4a7d6758bf47357dd1cb91fcac76657e2e03b9175d08f9

Observation afa7b4e2-11d3-4e6d-bb09-7afe42e3635e · outbound

This paper cites G-KV: Decoding-time KV cache eviction with global attention.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning G-KV: Decoding-time KV cache eviction with global attention

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:25:22.895992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:337479e110652222d9db7712f1a5e24b6ff7888459455985d971bba6c34edaac

Observation 8d02ed1a-8487-4593-ac03-aa78281f361c · outbound

This paper cites Scissorhands: Exploiting the persistence of importance hypothesis for llm kv cache compression at test time.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Scissorhands: Exploiting the persistence of importance hypothesis for llm kv cache compression at test time

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T11:56:56.548601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:3f65b148bb894865c2dabe02dcb6b3dc58c5a75aef2f0dd43c6603f8362ab5c8

Observation cdcd8753-33e6-4825-a0c9-53568346a642 · outbound

This paper cites KV-Compress: Paged KV-Cache Compression with Variable Compression Rates per Attention Head.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning KV-Compress: Paged KV-Cache Compression with Variable Compression Rates per Attention Head

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:25:22.891170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:dbc0283f3d86499718d4888e7efb108c7a2d4b8d53a2519a788c79806e3d5814

Observation ba2d0acf-50a1-4ab3-82db-b89108d65011 · outbound

This paper cites Liu, Kevin Lin, John Hewitt, Ashwin Paranjape, Michele Bevilacqua, Fabio Petroni, and Percy Liang.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Liu, Kevin Lin, John Hewitt, Ashwin Paranjape, Michele Bevilacqua, Fabio Petroni, and Percy Liang

Reference 35

Resolution
verified exact
doi, observed 2026-05-25T05:25:22.885884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:5c481f48ca7d3e07a93e97eb94d888bca821b101263dd01ae5b84bc89425445b

Observation 286f1144-6852-4b20-9d6d-c98dfba57b4b · outbound

This paper cites The curious case of neural text degeneration.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning The curious case of neural text degeneration

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T11:56:56.513042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:a89c1006a50d24d50ecd31c0e10975d735977c850d4443c56bec1f4731f64351

Observation 653dd828-8a57-47f0-9efd-106439da8f25 · outbound

This paper cites Kevin Zhou, and Xike Xie.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Kevin Zhou, and Xike Xie

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:25:22.871512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:f13ef8ae40a983c38c8c5f1e686fe4618fe19f275262d421900dc31bf7cacb32

Observation d10bef2c-5053-48fa-b8c4-1ab0b0f55879 · outbound

This paper cites LongFlow: Efficient KV Cache Compression for Reasoning Models.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning LongFlow: Efficient KV Cache Compression for Reasoning Models

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-05-25T05:25:22.879156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:1ced0f7ac1399a25fde73d1b22f83988908fa48de6f61158e81c835a40884c7a

Observation 08965108-e33e-4a4d-a554-c7b6ab65dd56 · outbound

This paper cites LongFlow: Efficient KV Cache Compression for Reasoning Models.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning LongFlow: Efficient KV Cache Compression for Reasoning Models

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-25T05:25:23.505112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:cfdf09241e0db7f4761582a5ffde47b0f762130a0b1ff9da3859081d361bb3ed

Observation 9c59b3d2-70db-46ba-8099-0942ca04da86 · outbound

This paper cites Lethe: Layer- and time-adaptive kv cache pruning for reasoning-intensive llm serving.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Lethe: Layer- and time-adaptive kv cache pruning for reasoning-intensive llm serving

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:25:22.859630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:d7bbad3539cd2e87340ff3ae9cb3b2dacdeee567bbeb306486d0082fdd5e17e7

Observation 3f3a449f-4149-4e1b-bcdc-abcf6df48112 · outbound

This paper cites ThinKV: Thought-Adaptive KV Cache Compression for Efficient Reasoning Models.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning ThinKV: Thought-Adaptive KV Cache Compression for Efficient Reasoning Models

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-05-25T05:25:22.865447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:14d73b8c1d0e82925b93c913bbbaefc4b87bd793dbc20ae70704475d1b86621e

Observation d1050df4-6db3-466b-a9b8-ae92b4c0f6ab · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Measuring Mathematical Problem Solving With the MATH Dataset

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-05-25T05:25:22.853864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:95639aca4ae457b0908b6ed00e2a3a3f7da358a1c9525e7d957754daa2748615

Observation 432c82e3-8083-4a43-a185-2ed756185742 · outbound

This paper cites Let’s verify step by step.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Let’s verify step by step

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T11:56:56.478442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:c02a986049c17b5561f9220e28e448e5dd7554822275cc9ae7e392383226a95f

Observation 64968117-e6c0-4ccd-b9f1-12d4e657f983 · outbound

This paper cites an unresolved cited work.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-05-25T11:56:56.488768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:dc57a51d2cf7ee8ab825629ef94a0bfc528e06fd100687c593a2d34507a057c1

Observation 62f6a1d4-beae-4c40-ae8e-e9c9d68e6590 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Training Verifiers to Solve Math Word Problems

Reference 45

Resolution
metadata mismatch
local_arxiv, observed 2026-05-25T05:25:22.849059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:9895f0ab6fb3aac87a8ad349ca6d2972454ebe57538524df5d39fc60523bdbde

Observation 13b99926-8d00-4aee-b18f-a55847821a1f · outbound

This paper cites Deepseek-r1-distill-qwen-7b.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Deepseek-r1-distill-qwen-7b

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T11:56:56.528849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:226c02ceb29795855807a20154cf559142e435e4b3b1c9ebf2c37e9fbaa1e725

Observation 3db0515e-d7d5-4465-994f-dac415d5d83a · outbound

This paper cites Deepseek-r1-distill-qwen-32b.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Deepseek-r1-distill-qwen-32b

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T11:56:56.481650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:b72abe234b1bd1ad3f78a80e4b59cd89464086c4b5cfdb59dd62002fc71b7818

Observation 2990f303-9a6e-4acf-b933-5d75301c31ec · outbound

This paper cites lost in the middle.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning lost in the middle

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T11:56:56.506203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:7eb6167d6e1169ef3c09eec30c18a4611bc646ed22d9f08bdfb7528e31f0e9a0

Observation 644d0bba-ab12-4ef0-b2b6-c610d1954223 · outbound

This paper cites an unresolved cited work.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-05-25T11:56:56.492078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:b6f9e554079ab8a78b3097f078e77ab63256498694d627e9d81731a592803320

Observation eb63684a-c8eb-4416-a27d-1d00ee7445e8 · outbound

This paper cites Pass@1 is computed from metric_main / num_samples in the csv.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Pass@1 is computed from metric_main / num_samples in the csv

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T11:56:56.531866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:54c33861b3336da180385243c893cf9688979ee5cfaebd7e2ce4a0ab82a02cc1

Observation 98763174-6935-4f39-ba41-d8a3f7b95619 · outbound

This paper cites an unresolved cited work.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-05-25T11:56:56.525650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:2615c569d802f81c1532e9983f72315ba8cd53509f8962209ab468ae0f169b5f

Observation 8e26c9b0-c4da-4aa4-a04a-23ed96a65df0 · outbound

This paper cites an unresolved cited work.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-05-25T11:56:56.519133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:7166a628052dac4ee250f87c8a50235dae456a758744ba9668586ec78ad17730

Observation 8545ba2d-a254-42f5-bac8-77da825a6c3a · outbound

This paper cites an unresolved cited work.

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-05-25T11:56:56.540889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:7936fbb5abe9d437df5b2fea30b8fe649242aa6d84162eeeed0f07b04a058fa9

Observation 450c1818-f43d-492b-82f3-197d02763788 · outbound

This paper cites Simplify (u+ 4)(u−1)−(u+ 4)(u−1).

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning Simplify (u+ 4)(u−1)−(u+ 4)(u−1)

Reference 54

Resolution
malformed identifier
raw_fallback, observed 2026-05-25T11:56:56.545357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T05:22:47.198185Z digest=sha256:a2f8b2ecbdf48fc685d4c7ad26ee440d5a08d41636b59f19d9e619be1a91e391

Pith citing papers

Observation fb304ee9-3e1a-4e57-9269-b67b116de627 · inbound

LOCKS: Page-Local Compact Key Summaries for Efficient Long-Context Decoding cites this paper.

LOCKS: Page-Local Compact Key Summaries for Efficient Long-Context Decoding Adaptive Mass-Segmented KV Compression for Long-Context Reasoning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-07-31T11:49:11.776915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T11:49:11.776915Z digest=sha256:d12dcf034f6c6b9fd6b385cd078f84ad68be225cfc34e30ff0b16722128972a5