Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 30 inbound Pith citation observations for arXiv:2405.12532.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T11:37:07.674267Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T01:36:44.048156Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 1d273cea-bb4c-45d8-9942-948c028785f0 · inbound
PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b3e5f2d2-b669-4ca8-81eb-47de6ea97466 · inbound
LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 7ced9945-c15e-4e04-939c-711b8dbb8902 · inbound
MiniKV: Pushing the Limits of LLM Inference via 2-Bit Layer-Discriminative KV Cache PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06d812ad-d052-4c96-b8a5-43a226a96667 · inbound
ZigZagkv: Dynamic KV Cache Compression for Long-context Modeling based on Layer Uncertainty PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcae64ab-e48c-4d7b-98f6-450c255989ef · inbound
SepLLM: Accelerate Large Language Models by Compressing One Segment into One Separator PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2ec4fe1-747a-4a2b-8a12-5f7720a90062 · inbound
More Tokens, Lower Precision: Towards the Optimal Token-Precision Trade-off in KV Cache Compression PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f328f19a-bac3-4371-a5a2-5be7ba46d374 · inbound
TreeKV: Smooth Key-Value Cache Compression with Tree Structures PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46fdbb1a-83a3-4239-8abf-ed34e63c0914 · inbound
Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35e38757-05ef-4846-815e-4eba4dbf94ba · inbound
Semantic Integrity Matters: Benchmarking and Preserving High-Density Reasoning in KV Cache Compression PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation cd6b8efe-e3fa-4474-a515-8ec7e9cc7aae · inbound
HACK: Homomorphic Acceleration via Compression of the Key-Value Cache for Disaggregated LLM Inference PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2239970b-c879-4ee0-a22c-4c9f7a1fa120 · inbound
Efficient-vDiT: Efficient Video Diffusion Transformers With Attention Tile PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d88e456d-ea6b-40d7-aaea-51187e3df623 · inbound
Position: Episodic Memory is the Missing Piece for Long-Term LLM Agents PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 102
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 827a1ae8-2935-4ac9-96d7-895870c2b1be · inbound
Memory-Efficient Visual Autoregressive Modeling with Scale-Aware KV Cache Compression PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 923033ba-0edb-4294-9ebd-9e6604817a56 · inbound
AnchorAttention: Difference-Aware Sparse Attention with Stripe Granularity PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58f35e4a-a465-400d-b3a3-ceb2549dc751 · inbound
MadaKV: Adaptive Modality-Perception KV Cache Eviction for Efficient Multimodal Long-Context Inference PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72268e6b-f74b-4d8c-bf88-5a2f34360f62 · inbound
Breaking the Boundaries of Long-Context LLM Inference: Adaptive KV Management on a Single Commodity GPU PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c385049-5d20-4389-bb81-fc98a8695fd0 · inbound
Multi-Granular Spatio-Temporal Token Merging for Training-Free Acceleration of Video LLMs PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c455c61-e622-48db-b715-3f0ab16003a7 · inbound
Think Clearly: Improving Reasoning via Redundant Token Pruning PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08f65a1c-9488-4181-974d-2413965dcaed · inbound
KV-Latent: Dimensional-level KV Cache Reduction with Frequency-aware Rotary Positional Embedding PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1280fc6c-c0e7-48ed-90fa-0b80692b91d2 · inbound
CaliDrop: KV Cache Compression with Calibration PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a133e17-f04e-46f2-a209-9d66bbdc268e · inbound
GraphKV: Breaking the Static Selection Paradigm with Graph-Based KV Cache Eviction PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ac1f9ac-caa9-43a1-85fd-33854434a692 · inbound
EvolKV: Evolutionary KV Cache Compression for LLM Inference PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec0782cc-b44f-435c-b9b1-4cb23cfdf3bd · inbound
ClusterFusion++: Expanding Cluster-Level Fusion to Full Transformer-Block Decoding PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 7adfc9b3-a699-43a6-8f06-1303ddf5a6ee · inbound
Sparse Prefix Caching for Hybrid and Recurrent LLM Serving PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0a7177e5-8175-4beb-8971-d857ab6ec030 · inbound
Reformulating KV Cache Eviction Problem for Long-Context LLM Inference PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e3b5b830-d887-42bb-9dc9-50d00a1e4f32 · inbound
ReST-KV: Robust KV Cache Eviction with Layer-wise Output Reconstruction and Spatial-Temporal Smoothing PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 81b036f3-9b64-4a9b-93be-9e841b39eecb · inbound
Protection Is (Nearly) All You Need: Structural Protection Dominates Scoring in Globally Capped KV Eviction PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 102101e7-0036-4788-aa5f-aeec894e557d · inbound
ProactiveLLM: Learning Active Interaction for Streaming Large Language Models PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 112
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f4f64744-e14b-46f3-9e34-bc9bd6626770 · inbound
From Rigid to Dynamic: Entropy-Guided Adaptive Inference for Long-Context LLMs PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ec731856-38ec-41c6-b8ce-74bd7eec91d6 · inbound
What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents PyramidInfer: Pyramid KV Cache Compression for High-throughput LLM Inference
Reference 135
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.