Pith. sign in

Paper Citation Record · LEDGER

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization

As of 15 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 2 inbound Pith citation observations for arXiv:2505.19586.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19586 v2

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:16:42.984625Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:45:12.553567Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T17:55:13.328095Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0009c9d3-4a69-441a-84cb-52e148f9e73e · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:46.286667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:16:38.963314Z digest=sha256:7287047261e0f2446edcedfddabebdca996ca142adbd7041719493309322361b

Observation e61ec203-ee94-4a34-acf9-b5da69754873 · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:46.145966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.013330Z digest=sha256:372280f50a2a69733d8357d195a774e343677d6c10e299287139399ebcdcba93

Observation 56da942a-5538-4a0f-801a-95e2e1113a05 · outbound

This paper cites GPT-4 Technical Report.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization GPT-4 Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:39.105730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:39.105730Z digest=sha256:d66310884a19fb2a43089074a90714ceb8eeb4e9ba4514ce382262b43f2b6c2b

Observation bff1fdef-e70e-4e89-ae0f-f8f661f4b18d · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:39.190085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:39.190085Z digest=sha256:384940407f480767812ba5d20625e74cb5ffb3942186bdbf16d0ac92207e5297

Observation 3da24f0b-5d7d-4529-b2cc-5e6e828b937e · outbound

This paper cites PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:39.266790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:39.266790Z digest=sha256:2b0655991d0cb4de86956fe64868dfc95b45745a19999f9a6cd20bd967531d31

Observation b0700a1d-b0be-4d48-a286-c8dc4035002d · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:45.978753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.358085Z digest=sha256:13b9e00e7f4bcd94ee155debb7bc07fa89cf21d7a0cb1121440c049412ddeb7a

Observation c358e741-9b03-4389-adaa-4897f3af5a9c · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:39.438404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:39.438404Z digest=sha256:10b07031ea7d83dcea5ff9913dabf4c45c741038ab2b9134b4a2cfdba852d2af

Observation 7917efbd-44f6-4c9b-adca-9457f47d1654 · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:39.530091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:39.530091Z digest=sha256:b4eebbcbc0ba0531783f23f3f09730283e1d12283eaa9c29560ee18fc9cf4700

Observation 050c77f7-e884-40c9-8c1c-5b1a9e8b8769 · outbound

This paper cites The Llama 3 Herd of Models.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization The Llama 3 Herd of Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:39.619435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:39.619435Z digest=sha256:15652df062e084260b6d37a5599ad37cd7c9d827268c4c3557c58dffb48ff723

Observation be925673-3d0b-4f96-ba40-04959777713c · outbound

This paper cites Ada-KV: Optimizing KV Cache Eviction by Adaptive Budget Allocation for Efficient LLM Inference.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Ada-KV: Optimizing KV Cache Eviction by Adaptive Budget Allocation for Efficient LLM Inference

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:39.736272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:39.736272Z digest=sha256:52592dc4449cfd4f3863b8d2e93a5241ec1744788881001f81ae4e64be551091

Observation e5b8f505-a5b4-486f-add0-f3ca2a95bc1c · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:45.756652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:16:39.882527Z digest=sha256:e5b12de9ec0fb44b37b71a48417e0ec1c2dbaee60c804a8c399f5e6efdaffaef

Observation 75d60ba6-1429-42b5-9726-e7d47c12acd8 · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:39.957230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:39.957230Z digest=sha256:9da5b245992cced087286fdd9d8698a03f8fc24982264e68fecf863e48c64122

Observation 1d546562-dc02-475e-9b47-73266fb0ab03 · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:40.059217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:40.059217Z digest=sha256:2bf30e1113eef7587e694fcfba7b71d52fd569f0187384f1b1a55d6d7a1a7176

Observation 4d09a508-089b-47bb-a5f2-97700cf23728 · outbound

This paper cites Abdi, Dongsheng Li, Chin-Yew Lin, Yuqing Yang, and Lili Qiu.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Abdi, Dongsheng Li, Chin-Yew Lin, Yuqing Yang, and Lili Qiu

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:40.185668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:40.185668Z digest=sha256:f14d50153dde2c81cf9aade7aeada3e71bfdc0603bd97c190266894c93de912a

Observation 7e68ccd2-51c8-47b1-a411-fa6b04f267a7 · outbound

This paper cites GEAR: An Efficient KV Cache Compression Recipe for Near-Lossless Generative Inference of LLM.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization GEAR: An Efficient KV Cache Compression Recipe for Near-Lossless Generative Inference of LLM

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:40.300614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:40.300614Z digest=sha256:e321868088ed72a10bb13fe7b6169760c13780b7300b622a2410ba5f9fe27783

Observation 83393d5b-049a-4ca5-bb72-3ce241205685 · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:40.399838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:40.399838Z digest=sha256:9900a71daa908b443358a84c3200b21b6ebd1d2fbad6eb7132c4e0c354528c70

Observation 0ba4e595-50ed-412f-a995-e05f8a7aeaea · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:40.533731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:40.533731Z digest=sha256:a33f3c9d2315e1f895859666861dd29b8976e3f629f9a3787791fc33a45d935f

Observation 9c533e52-2d32-4895-af8c-64dc29d1e3eb · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:40.602920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:40.602920Z digest=sha256:696675bb3a6ffceef0ff088ef74723dc8af7e1d875659e3275817df9e4a6da4d

Observation 2a066f05-082b-49cb-b01b-6ee0622ab53b · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:40.733388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:40.733388Z digest=sha256:5654cc6b55279537afb17a50f421012f479f5333a16fb8ccf0b4e7d3c8b44107

Observation fd7275c9-816b-45ff-96c9-3afbc57e709d · outbound

This paper cites DeepSeek-V3 Technical Report.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization DeepSeek-V3 Technical Report

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:40.840599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:40.840599Z digest=sha256:a9eff23dd2b3c84dbb9ca9f26711ed96631543a5ad1e5bec912dede05a017370

Observation 9a04388b-1d84-45e3-b2ab-1a747ca1e589 · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:45.426399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:16:40.929427Z digest=sha256:de99804e5ea0b38751c6406d82f899edf712f109096dcdeb21721c48bd319df0

Observation 8cd1bf44-9bf8-486c-96b6-084e41673e6f · outbound

This paper cites RetrievalAttention: Accelerating Long-Context LLM Inference via Vector Retrieval.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization RetrievalAttention: Accelerating Long-Context LLM Inference via Vector Retrieval

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:40.998220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:40.998220Z digest=sha256:108e8cffe1d06f406d0c754c698d994d2b25c34157c4bc735a4eb9a39fd539b1

Observation f66f4c85-15b0-439d-8ff4-9ce75c957d13 · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:45.178131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:16:41.097932Z digest=sha256:ac68efea405263a5b5d27190304845cd901b21261302675189add7f1985934be

Observation 663a592d-2f60-412c-b8b6-82a99ede768d · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:44.909157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:16:41.196983Z digest=sha256:d69d29f8164a9392d855119300612b6ea1c7acb4213cf74c87da8d7293b6e84a

Observation a079a4d3-391c-45f9-b92e-54d21c51123b · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:44.754722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:16:41.296524Z digest=sha256:668631c17438cd9a3b5dc9102fdc57102a8a67a4cfc151030e15296bbe320070

Observation 1a8fe988-6039-4b96-bfdf-7a2269b3c748 · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:44.592866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:16:41.372786Z digest=sha256:0edb3eafb6b658ff6ae792b868c9fb43ca95519efdb58e10a223bd07233fb610

Observation 52e8d0bc-9369-454a-a3b1-c7267ef526cd · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:44.434457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:16:41.464856Z digest=sha256:0e4ab7d18b2a9d5c6a8e413b69559e54dbd34126c6657d3329df9a5212458481

Observation 83f3bebf-fe43-41b5-ae85-56d670b907ed · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:44.304156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:16:41.553509Z digest=sha256:249653f5bc055c454e8bcbb549622b322a656b52802b10dd1fe7033175091c3c

Observation b100b888-b2b6-4049-839b-082a781e0eac · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization LLaMA: Open and Efficient Foundation Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:41.649120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:41.649120Z digest=sha256:21e3117d48802229d41f4d803256f091c2d53b4ba53eb5a69f5d1557ffba9aa9

Observation 868b6744-535c-4cc2-8824-f3fd5b1dad29 · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:44.083353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:16:41.769077Z digest=sha256:4179c10cee80db1dc7a2f6a35cbb5bb3684a77c3fb4773ed9a9355528b142dce

Observation c4e447d9-8d09-4019-8676-3864b2e16da2 · outbound

This paper cites Deep Graph Library: A Graph-Centric, Highly-Performant Package for Graph Neural Networks.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Deep Graph Library: A Graph-Centric, Highly-Performant Package for Graph Neural Networks

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:41.882697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:41.882697Z digest=sha256:c748bf270eb040e921b6fe121ea89eadcb0c4e26570fba43a774d9ba7f57605b

Observation 7924559a-cd26-4bca-b90b-b394ea4b6781 · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:43.891966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:16:41.976279Z digest=sha256:de4c53e037b44b1500909b4e800ae030041d15da7a22efd8413aa4f704d57002

Observation 0fa3dd3c-f37c-4e1f-a563-75bc8879d05f · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:16:43.696181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:16:42.074404Z digest=sha256:a4af27b7d4de02650a38029f4450edd6dba8199592ab3e8bc97f11ea16d70bae

Observation e3a1e404-48ed-442b-abc8-6981acdd4d24 · outbound

This paper cites No Token Left Behind: Reliable KV Cache Compression via Importance-Aware Mixed Precision Quantization.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization No Token Left Behind: Reliable KV Cache Compression via Importance-Aware Mixed Precision Quantization

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:42.177025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:42.177025Z digest=sha256:92f1fde265c4101b0cb35b685aded6cd5108e80dfbc5d3182ab678f975392b08

Observation 8bae7c40-8cad-4de8-b326-6a4e6348b460 · outbound

This paper cites Post-Training Sparse Attention with Double Sparsity.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Post-Training Sparse Attention with Double Sparsity

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:42.302352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:42.302352Z digest=sha256:787ddd3d19962fbcbeac7ca8489b8d0624b35737a939b9a571f74c365a7d39e0

Observation d06da1e1-b521-4fa2-8675-91a428ba9431 · outbound

This paper cites PQCache: Product Quantization-based KVCache for Long Context LLM Inference.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization PQCache: Product Quantization-based KVCache for Long Context LLM Inference

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:42.391215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:42.391215Z digest=sha256:765659a2ad7a506d554845dece0b7e96c201d33d73ce113d821519675487ad84

Observation 32ac6e75-29cb-4ae5-8a7b-7d439bd44850 · outbound

This paper cites Dovetail: A CPU/GPU Heterogeneous Speculative Decoding for LLM inference.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Dovetail: A CPU/GPU Heterogeneous Speculative Decoding for LLM inference

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:42.476185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:42.476185Z digest=sha256:834b5f7be29a72fe577748bcc9c4cc2a5cf129aef29c416e97a10e6ed573042d

Observation 11eb4041-968b-4b28-860f-0d80e8b8c153 · outbound

This paper cites an unresolved cited work.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:42.573762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:42.573762Z digest=sha256:504c3c6cef09536e93003866cd79d35aa1644193b8f235fd9e6f5b1a7baf4e78

Observation ceee1b52-8eca-48c7-8c1e-6345ac942997 · outbound

This paper cites LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:42.683728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:42.683728Z digest=sha256:ad59ba39d8a7d656ce9f4b9aebba192d795067d62a057eba377def41cc30baf6

Observation ef746939-1c9f-47a3-9895-200ca4c889e3 · outbound

This paper cites Barrett, Zhangyang Wang, and Beidi Chen.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization Barrett, Zhangyang Wang, and Beidi Chen

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:16:43.494267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-08-07T14:16:42.793624Z digest=sha256:725c526eb71a3c56661639b18eff6cff7b1e7095ea44a48bf80e8c2cdbd21c14

Observation ba9f6c0b-ec83-4895-932f-6a81fdad01d7 · outbound

This paper cites online" 'onlinestring :=.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization online" 'onlinestring :=

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:42.896424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:42.896424Z digest=sha256:bb8e1a1b74967e28faccf82e1e4912677c6ba2196f6f19f96fc0b8ed8060edc1

Observation 5a44742e-afbb-46f6-8f28-6abae5b91aea · outbound

This paper cites write newline.

TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization write newline

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:16:42.984625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:16:42.984625Z digest=sha256:c02310fa5a186e720418942c9d0eacd63ae7ee7e8337bf52a601f82c9269f17d

Pith citing papers

Observation c81e1b24-b203-471f-98e7-755364f2c47d · inbound

LOOM-Scope: a comprehensive and efficient LOng-cOntext Model evaluation framework cites this paper.

LOOM-Scope: a comprehensive and efficient LOng-cOntext Model evaluation framework TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T19:45:12.553567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:45:12.553567Z digest=sha256:a47c8ebbfbbb58aa947e5e69e208142194b4e25ed4208084ae97079c34d327bb

Observation 06792c92-3127-4413-b66c-d107532327c0 · inbound

TPLA: Tensor Parallel Latent Attention for Efficient Disaggregated Prefill and Decode Inference cites this paper.

TPLA: Tensor Parallel Latent Attention for Efficient Disaggregated Prefill and Decode Inference TailorKV: A Hybrid Framework for Long-Context Inference via Tailored KV Cache Optimization

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-08-05T17:55:13.365091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T17:55:10.372511Z digest=sha256:e4691bccd49363bc02e92e50bf04c84c1ab705510a74bfccb2c48d2cb6dd34ea