Pith. sign in

Paper Citation Record · LEDGER

Compute Or Load KV Cache? Why Not Both?

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2410.03065.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.03065 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:22.542928Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T01:36:26.266282Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0773a48c-07e3-4efb-8310-2e9b15e8aa42 · inbound

From Human Memory to AI Memory: A Survey on Memory Mechanisms in the Era of LLMs cites this paper.

From Human Memory to AI Memory: A Survey on Memory Mechanisms in the Era of LLMs Compute Or Load KV Cache? Why Not Both?

Reference 135

Resolution
verified exact
arxiv_id, observed 2026-05-17T11:05:09.878154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T11:05:09.588491Z digest=sha256:7619575dbba84b2af0f735f4a17362cc6394210cf5a766067ef0661f205a167b

Observation a974abcd-187d-46b9-9315-854d348d651f · inbound

CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration cites this paper.

CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Compute Or Load KV Cache? Why Not Both?

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:22.542928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:22.542928Z digest=sha256:e7810316b5703fef6e0629be0acf28697eb45933cc895bee4ba1e609393afea7

Observation a34eb7be-896c-4b89-b6ba-31cd8f00ef2a · inbound

Learn from the Past: Fast Sparse Indexing for Large Language Model Decoding cites this paper.

Learn from the Past: Fast Sparse Indexing for Large Language Model Decoding Compute Or Load KV Cache? Why Not Both?

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:38:38.180039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:38:38.180039Z digest=sha256:96d88ebaf2438d53991a582ed632aaf7089ce0840cc78fa7c3c87eb97eb5401e

Observation 41843f4d-a06b-4137-95cb-584887bcee05 · inbound

Krul: Efficient State Restoration for Multi-turn Conversations with Dynamic Cross-layer KV Sharing cites this paper.

Krul: Efficient State Restoration for Multi-turn Conversations with Dynamic Cross-layer KV Sharing Compute Or Load KV Cache? Why Not Both?

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:46:57.628851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:46:57.628851Z digest=sha256:b61850edf2bce205fadcddf1fb4b20badac8281b1d2eb7bcffb31c07410bc6c3

Observation 1faccbb3-415a-48b9-beb6-c7686a55f92f · inbound

SparKV: Overhead-Aware KV Cache Loading for Efficient On-Device LLM Inference cites this paper.

SparKV: Overhead-Aware KV Cache Loading for Efficient On-Device LLM Inference Compute Or Load KV Cache? Why Not Both?

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:46:07.696919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T14:09:30.821354Z digest=sha256:b1872942d04e93259a3bc7b452e1dbfc7c2d0d3c9294310fc20d3da30a580268

Observation 520a9206-8d88-4a4b-a550-a6029e536c10 · inbound

CachePrune: Privacy-Aware and Fine-Grained KV Cache Sharing for Efficient LLM Inference cites this paper.

CachePrune: Privacy-Aware and Fine-Grained KV Cache Sharing for Efficient LLM Inference Compute Or Load KV Cache? Why Not Both?

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:15:19.855517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T04:14:44.387842Z digest=sha256:9d622ee028656011628aab5ad28b0915153a2f07f41b26e4957eea33d235fd19

Observation 58bd1c11-4136-41e3-954a-415e369a8051 · inbound

Leyline: KV Cache Directives for Agentic Inference cites this paper.

Leyline: KV Cache Directives for Agentic Inference Compute Or Load KV Cache? Why Not Both?

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:36:14.557072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T16:51:08.114092Z digest=sha256:009644ace20ea127bbd74014db37d800895df96bf90c541440f7102690a9ce4a

Observation 46890d96-4721-4e6b-af7f-8748365ee318 · inbound

MomentKV: Closing the Directional Gap in KV Cache Eviction for Long-Context Inference cites this paper.

MomentKV: Closing the Directional Gap in KV Cache Eviction for Long-Context Inference Compute Or Load KV Cache? Why Not Both?

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:56:15.399276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T16:07:22.196982Z digest=sha256:fa64758b24599c14acb2471025296e14bedc85ec6044fd87d0b830ac30aeceb9

Observation 9a519988-9bd0-4579-942f-4ea34a84875e · inbound

Multi-Segment Attention: Enabling Efficient KV-Cache Management for Faster Large Language Model Serving cites this paper.

Multi-Segment Attention: Enabling Efficient KV-Cache Management for Faster Large Language Model Serving Compute Or Load KV Cache? Why Not Both?

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-02T01:36:26.270553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T11:38:05.435505Z digest=sha256:4e87cdbbb7b73e82822225b95812d07e6b65a6ff9eede4c39bdb715873e4add4