Pith. sign in

Paper Citation Record · LEDGER

EvolKV: Evolutionary KV Cache Compression for LLM Inference

As of 19 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2509.08315.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.08315 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T20:53:30.792140Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

54 of 54 outbound references displayed

  • verified exact7
  • verified fuzzy1
  • unresolved46
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4e95db61-9d48-40b4-87ce-6ed386ad51f7 · outbound

This paper cites URL: " 'urlintro :=.

EvolKV: Evolutionary KV Cache Compression for LLM Inference URL: " 'urlintro :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:26.291319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:26.291319Z digest=sha256:23e4cf5a883b60531291713f269bdf40adfafb74051699064f59bcdb151bc2c9

Observation 038f5414-d74b-428b-88ff-b2c5ea509870 · outbound

This paper cites write newline.

EvolKV: Evolutionary KV Cache Compression for LLM Inference write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:26.358653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:26.358653Z digest=sha256:9514f58b5b4ad0ce04c62588766fb7401485f7e91bdbb680dc79d00e4ab0782a

Observation 9e27fdca-227a-41f1-882f-7a0ad92a7a99 · outbound

This paper cites Genetic Algorithm: Reviews, Implementations, and Applications.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Genetic Algorithm: Reviews, Implementations, and Applications

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-04T20:53:32.231576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-04T20:53:26.518273Z digest=sha256:a9b32089f21f05abc233a454b205f3dd98947f219ada554188c7e54d50026151

Observation 593b2dde-4a6e-472f-9c65-39b467271ec6 · outbound

This paper cites Qwen Technical Report.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Qwen Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:26.644643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:26.644643Z digest=sha256:6f74f89992265411f6765f74a906eac3f9df09a74a103d1976f5f3fdcbaa34ce

Observation 74b606d7-a608-4507-9008-6a668a2b8a28 · outbound

This paper cites LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding.

EvolKV: Evolutionary KV Cache Compression for LLM Inference LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:26.772769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:26.772769Z digest=sha256:f4dd8db483ee56fe9d0348fe9da3d239914d9f1b9cf23d597f3984c85a02dced

Observation 806b7f12-12a1-4b0f-a31a-73fabaec4ad6 · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 6

Resolution
verified exact
doi, observed 2026-08-04T20:53:31.588316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-04T20:53:26.981819Z digest=sha256:4fa73335f3923081a4ffd4ac05ae5f6b8b9bcfaa668c116402a5f6f340510421

Observation 37e36d34-34d6-458f-b35b-8db03f0729a8 · outbound

This paper cites Longformer: The Long-Document Transformer.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Longformer: The Long-Document Transformer

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:27.205605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:27.205605Z digest=sha256:0285ef0ea118be372f342204ec9c4e4d3e55426c3bf373154e5b971b27c6d745

Observation ec5e6bb3-5d9a-4eb8-be3f-f54dad7cc057 · outbound

This paper cites PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling.

EvolKV: Evolutionary KV Cache Compression for LLM Inference PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:27.311694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:27.311694Z digest=sha256:f9b21abc10a2e00a83f60bd08a811beed8affa35a5b4a3d423c90564351aef32

Observation 8d1def59-fa32-485f-b2c1-cfb3e04dcd89 · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 9

Resolution
verified exact
doi, observed 2026-08-04T20:53:31.437565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-04T20:53:27.372470Z digest=sha256:f992c1eaa1a156cc34af1878e083db1d94f319bffe91202ee67ebd4a720a9ad5

Observation 9ef49c92-b34f-4faa-9cfa-c4499f9355f9 · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:27.452212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:27.452212Z digest=sha256:87ce09c6df8e96ffe6329f679c83d17d3ecf2719ab17e1b7c3605063c4b413d0

Observation a2b9a39e-aff6-4078-bc8b-a51f1dffcb31 · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 11

Resolution
verified exact
doi, observed 2026-08-04T20:53:31.302934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-04T20:53:27.526927Z digest=sha256:15dba1507520480bd91363659046dc57f838f96b4d38b277f3252262d62c9648

Observation 227e76f9-5a1e-42ce-a729-3f38e495cddf · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 12

Resolution
verified exact
doi, observed 2026-08-04T20:53:31.172668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-04T20:53:27.618861Z digest=sha256:4a1527ecac2e351d235e4c09418acc5c0533bbc13db8b794b0daf125bf1ef20f

Observation dc81beab-0131-4406-a83d-f52ac40a97be · outbound

This paper cites Generating Long Sequences with Sparse Transformers.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Generating Long Sequences with Sparse Transformers

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:27.664702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:27.664702Z digest=sha256:64f6a0bf1f19c94da2d265aab6e575ca3a6fde7403d9ba3d8b266f4973688fd3

Observation b2a0d275-84b9-494b-a5c7-97df1e1d456b · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Training Verifiers to Solve Math Word Problems

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:27.718150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:27.718150Z digest=sha256:04136f40f9ee0096ef1c1e444045a637c025d10231c888a1bd61ce2ef97fbea2

Observation b2669a8f-751b-4fc0-8631-0cfca693f10f · outbound

This paper cites FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness.

EvolKV: Evolutionary KV Cache Compression for LLM Inference FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:27.863745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:27.863745Z digest=sha256:938f051f11bf0a4004b9bb1340859e5a504071bcb3047812aab6aa5761f8b533

Observation 97110207-a1d0-4143-a3e6-5a3e99adea6f · outbound

This paper cites A Dataset of Information-Seeking Questions and Answers Anchored in Research Papers.

EvolKV: Evolutionary KV Cache Compression for LLM Inference A Dataset of Information-Seeking Questions and Answers Anchored in Research Papers

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:27.975894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:27.975894Z digest=sha256:55074dfd1bad58867df74c4a6de53e8f5323de0aeeb6de64670a80721572767c

Observation d479713b-ab9a-4d4a-af37-fc942267fce2 · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:28.025096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:28.025096Z digest=sha256:040ab66448d8084d92b715039509388f4777a961e606a01691bb3a3d855c9e6f

Observation 89d44375-6ca3-4c16-b863-3d71f108746a · outbound

This paper cites Model Tells You What to Discard: Adaptive KV Cache Compression for LLMs.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Model Tells You What to Discard: Adaptive KV Cache Compression for LLMs

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:28.117122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:28.117122Z digest=sha256:0464e140fd3f2ffecdc50b35a607166438a6524b4167ef924c6bb0bb0fdbda08

Observation ffd10026-e907-405f-b4ef-93185d0e7f7b · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:28.177042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:28.177042Z digest=sha256:d300882729d2734b3912871b1aff7eec4c77f6433b1b757c0e029adb92d4f4c7

Observation 25f304d6-cb41-4a7e-b292-c6c5c8907899 · outbound

This paper cites The Llama 3 Herd of Models.

EvolKV: Evolutionary KV Cache Compression for LLM Inference The Llama 3 Herd of Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:28.280449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:28.280449Z digest=sha256:7133cb14abab1b475efc6bc54a69d2de0471f49e6d72bc58ccb952b65e391d25

Observation cce07c45-05bf-4860-8809-92152efa642c · outbound

This paper cites LongCoder: A Long-Range Pre-trained Language Model for Code Completion.

EvolKV: Evolutionary KV Cache Compression for LLM Inference LongCoder: A Long-Range Pre-trained Language Model for Code Completion

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:28.374598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:28.374598Z digest=sha256:4dc2fb03d51652c4b5c1cadc05f99902706f2071010b745e4fd3c81c36ee473f

Observation c7da493d-c064-4e38-aab2-bb054bd88822 · outbound

This paper cites Müller, and Petros Koumoutsakos.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Müller, and Petros Koumoutsakos

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:28.492763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:28.492763Z digest=sha256:45148d4565df0a42228d1652bbc0292ff447921e7b7409f8904d12729ea71c58

Observation 88d67d04-10e9-4b9e-91fb-5cbfd1af2c05 · outbound

This paper cites Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Constructing A Multi-hop QA Dataset for Comprehensive Evaluation of Reasoning Steps

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:28.604509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:28.604509Z digest=sha256:bb7e61bd7065da2cff5dde31e3f88be014adb1a96f85ccc2a85e3b5eafb4e529

Observation 02bce1f9-b549-4454-89d1-0fae7b300aff · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:28.708057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:28.708057Z digest=sha256:f6905bfe533682b143cd8a8b65af921f5a9c503fa0e3a5d36ec20ac9e0abbf70

Observation 2aa56bfe-9d3d-4cfb-8a1b-2fa388d93e0b · outbound

This paper cites RULER: What's the Real Context Size of Your Long-Context Language Models?.

EvolKV: Evolutionary KV Cache Compression for LLM Inference RULER: What's the Real Context Size of Your Long-Context Language Models?

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:28.828244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:28.828244Z digest=sha256:34abf9286546afa872121704c0914ca7d13f32a8169663af527200918faa27ef

Observation dd278312-35d1-4ccc-a742-c2ccea7cd444 · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.006808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.006808Z digest=sha256:e78e04d8d95cae7a0190e25ccdec078ca818304099ad228f3ce6d8bc4005155b

Observation 3cf5adb1-59b7-46cb-b6ee-be469b6c2134 · outbound

This paper cites Efficient Attentions for Long Document Summarization.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Efficient Attentions for Long Document Summarization

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.091290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.091290Z digest=sha256:ab79228455343ca3240141414e0008bdfd3e0f8167a3b26a3b6910c0a9419777

Observation 1f9b3a52-e930-4d2c-88a6-387b61f4d6e5 · outbound

This paper cites Mistral 7B.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Mistral 7B

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.222206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.222206Z digest=sha256:98d23ba15c9c5e99ef6af9568b9311775a0c3174f7fe64b2db8786980975e9cf

Observation 3aff6643-38d1-4803-a35e-cdf8c451fdf9 · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.329672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.329672Z digest=sha256:614f134b7120681375f634a77d8ecd06437f28d895aa324537a612af9132c82b

Observation 28bc5f7d-8fc1-495e-87c5-bde3063a1d2b · outbound

This paper cites Kennedy and R.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Kennedy and R

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.384326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.384326Z digest=sha256:c1c4c1d239591a5487adb8f4f2f0656cd4ffe2a9f60f38f3c010387f92440f48

Observation 217b08b5-b4f7-47c9-acf3-a86a751fc34e · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-04T20:53:32.773075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-04T20:53:29.443038Z digest=sha256:993063cd83cd6be74e27e8add97a6f0b7b0b107dc2e45adfd710ce401448d3c9

Observation 710b9a84-f8e7-4dee-b917-7c5ebf46d614 · outbound

This paper cites The NarrativeQA Reading Comprehension Challenge.

EvolKV: Evolutionary KV Cache Compression for LLM Inference The NarrativeQA Reading Comprehension Challenge

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.536470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.536470Z digest=sha256:055d16d8a163df783abbd9ab8d4145bb9f49792d61e7ac32430feeadaa003435

Observation 007b7138-efe7-469d-a343-b3a8c91df1dc · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-04T20:53:32.569127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-04T20:53:29.595373Z digest=sha256:b5919902de64f7e6ec94dea7302bc74abd287b308b335fbf8f226caee3410006

Observation a721037a-51e1-4aea-bf3e-cc4a3655fb36 · outbound

This paper cites SnapKV: LLM Knows What You are Looking for Before Generation.

EvolKV: Evolutionary KV Cache Compression for LLM Inference SnapKV: LLM Knows What You are Looking for Before Generation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.655289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.655289Z digest=sha256:250a1f9b20a8c23752b0ddd61ddb90879a78debe2869602d1740639460ceeba0

Observation 620f96f9-9081-457b-90e4-5d9cb31461cb · outbound

This paper cites RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems.

EvolKV: Evolutionary KV Cache Compression for LLM Inference RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.703129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.703129Z digest=sha256:562d5f20c933aadbfcb4d4b3d1176e2b271fe7b068f6a43cdcce3f794e5ca770

Observation 11cc42c4-5729-4e86-9a55-7f13fbb5c03b · outbound

This paper cites Scissorhands: Exploiting the Persistence of Importance Hypothesis for LLM KV Cache Compression at Test Time.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Scissorhands: Exploiting the Persistence of Importance Hypothesis for LLM KV Cache Compression at Test Time

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.751152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.751152Z digest=sha256:2ead79b75c56d17961fe556e90d87bc7181737965e8b70897586da0e4a93089a

Observation 85e4721f-70b0-40ec-9c85-57f601f6a962 · outbound

This paper cites StarCoder 2 and The Stack v2: The Next Generation.

EvolKV: Evolutionary KV Cache Compression for LLM Inference StarCoder 2 and The Stack v2: The Next Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.804774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.804774Z digest=sha256:5a42fdca99791a7a7f2bcc3c1453492221e67c59b63296cdea951d6da0526c77

Observation 3881c240-8c37-4ffd-86b0-e6b050091b4c · outbound

This paper cites GPT-4 Technical Report.

EvolKV: Evolutionary KV Cache Compression for LLM Inference GPT-4 Technical Report

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.842871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.842871Z digest=sha256:b0e8acb5f1472ff9e2ddf911b4c87e7c729a866aa18989de3870fe441d56cdaa

Observation f833dff9-0183-4bd6-ac51-7b99147bfd01 · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.906574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.906574Z digest=sha256:06142ec256a6e3fc86a015d20fb8f215e0996f2b4193d16da13c71b9780d8fdc

Observation a2ab1ab1-e008-4d84-959a-dbcb3bf958eb · outbound

This paper cites Language Modelling with Pixels.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Language Modelling with Pixels

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:29.956650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:29.956650Z digest=sha256:893a82a4253ac3b1b188ea69d51bee14dfe273d7b408edd365cdf19f5a01692e

Observation 411d3c69-623e-4cf4-a814-2ca78ee496be · outbound

This paper cites Fu, Zhiqiang Xie, Beidi Chen, Clark W.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Fu, Zhiqiang Xie, Beidi Chen, Clark W

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T20:53:32.356461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-04T20:53:30.052751Z digest=sha256:a9490ec3bebc3ea849392e3277b0489573498f72a1900501353107a65972c437

Observation 64453fe3-4c28-4191-8b70-bc5a7136d4cb · outbound

This paper cites Keep the Cost Down: A Review on Methods to Optimize LLM' s KV-Cache Consumption.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Keep the Cost Down: A Review on Methods to Optimize LLM' s KV-Cache Consumption

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.115688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.115688Z digest=sha256:f4aaa7146eb3a474f5e253f20c096bccf9d2592314125ebbbf3d6667b507f56c

Observation d8738f86-aa29-4df8-809c-ad338e73fd5c · outbound

This paper cites Layer by Layer: Uncovering Hidden Representations in Language Models.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Layer by Layer: Uncovering Hidden Representations in Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.156275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.156275Z digest=sha256:8537c333015e13e141d12896a6953de72f72c91f3d05908e85825d1d6135b48f

Observation 4c26b3d9-2503-4607-9a8c-225d265e964c · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.208315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.208315Z digest=sha256:fffc7c6aa455cf62b5043436afb2a94f54effbac79c08911c3739983896ee986

Observation 2fb67ddc-9f16-427a-8fe6-591cdc4b8232 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.257649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.257649Z digest=sha256:15994b927665cdb62b29c369e6f5dff084cb89fddeb1265f01ac93ddc2c18964

Observation 8264b62f-82ed-4454-8939-f2b0d50d5de0 · outbound

This paper cites MuSiQue: Multihop Questions via Single-hop Question Composition.

EvolKV: Evolutionary KV Cache Compression for LLM Inference MuSiQue: Multihop Questions via Single-hop Question Composition

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.299251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.299251Z digest=sha256:212a626ed2f12d6df3f1678bd8420dd84919a700374428df8781cb37a11fc996

Observation 514806e7-e6b7-442c-8cc4-2365e51547d8 · outbound

This paper cites an unresolved cited work.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Unresolved cited work

Reference 47

Resolution
verified exact
doi, observed 2026-08-04T20:53:31.024481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-04T20:53:30.361487Z digest=sha256:a9286839953bc5aff70cba66ecb3755c39681258eb6bc2bd4b363e5c45d5af2e

Observation 118dd673-bb33-4469-9a45-f1d5d9bf49c2 · outbound

This paper cites Rethinking the Value of Transformer Components.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Rethinking the Value of Transformer Components

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-04T20:53:31.795752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-04T20:53:30.417248Z digest=sha256:0f2b916595a810fcb6a8a4fa1c8d793a01b623089a96cac1b95fe086954d3062

Observation e9f72f18-0519-4052-abec-3a59ada6d253 · outbound

This paper cites Efficient Streaming Language Models with Attention Sinks.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Efficient Streaming Language Models with Attention Sinks

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.488876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.488876Z digest=sha256:2b8e4bc9aaa77c7f3606c4804e6e6f4179048cd9849a79b4241da36de8237c22

Observation 9ac1f9ac-caa9-43a1-85fd-33854434a692 · outbound

This paper cites PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference.

EvolKV: Evolutionary KV Cache Compression for LLM Inference PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.523155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.523155Z digest=sha256:33e1b23604a3fda65d5375f589aceb4c39097cba4879c8856bd6736ce6b5d207

Observation 03a53764-862b-4270-90a3-bc611a1c7698 · outbound

This paper cites HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering.

EvolKV: Evolutionary KV Cache Compression for LLM Inference HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.564217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.564217Z digest=sha256:472a1a3fe49850524ce8f7c45dbf00213a6fa7d7c4d20241c4d21c8a7ad37a39

Observation acec151c-263f-4e90-81a4-32562f1924b2 · outbound

This paper cites Investigating Layer Importance in Large Language Models.

EvolKV: Evolutionary KV Cache Compression for LLM Inference Investigating Layer Importance in Large Language Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.630054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.630054Z digest=sha256:38b7fb0ad8a252ffbc2aa5e88b011dccd27ec855fc1cd6d5582d1ad44b25ee9a

Observation 451a3792-123c-43ed-b4ce-a1a3e8bebde6 · outbound

This paper cites H$_2$O: Heavy-Hitter Oracle for Efficient Generative Inference of Large Language Models.

EvolKV: Evolutionary KV Cache Compression for LLM Inference H$_2$O: Heavy-Hitter Oracle for Efficient Generative Inference of Large Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.705096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.705096Z digest=sha256:5e4641d54e9317d4935f464236ac6883018d576669a476332e6630556f1cc54d

Observation e4422975-3d6f-4edc-a58b-11aa2d9e8ac6 · outbound

This paper cites QMSum: A New Benchmark for Query-based Multi-domain Meeting Summarization.

EvolKV: Evolutionary KV Cache Compression for LLM Inference QMSum: A New Benchmark for Query-based Multi-domain Meeting Summarization

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T20:53:30.792140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T20:53:30.792140Z digest=sha256:7fe27317aa6bfa675bf7eb82d62613be0b38b2321bbada28c6e623be93e0b16f

Pith citing papers

No inbound Pith citation observations are available.