Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:2405.14366.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T16:20:32.412901Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
3
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation a54f002f-c183-4e53-88e0-19641b7e1091 · inbound
Hymba: A Hybrid-head Architecture for Small Language Models MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39ffc95a-dca4-42ef-829d-d6bc784359bd · inbound
MiniKV: Pushing the Limits of LLM Inference via 2-Bit Layer-Discriminative KV Cache MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab5fa997-8eea-4f97-b365-6962333cce6e · inbound
Compressing KV Cache for Long-Context LLM Inference with Inter-Layer Attention Similarity MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b35376c6-1d3e-4277-9e74-9f3fa2bdc621 · inbound
DynamicKV: Task-Aware Adaptive KV Cache Compression for Long Context LLMs MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a09a16a1-07cc-4ba0-a737-ae241f0805eb · inbound
HashEvict: A Pre-Attention KV Cache Eviction Strategy using Locality-Sensitive Hashing MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51faa30f-39bb-4c62-ba21-5637acb71c24 · inbound
A Silver Bullet or a Compromise for Full Attention? A Comprehensive Study of Gist Token-based Context Compression MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 437891ab-6e45-49eb-b1ab-69b9315df6f1 · inbound
A Survey on Large Language Model Acceleration based on KV Cache Management MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 102
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de681a8d-578e-4497-80dd-db5200ce5d35 · inbound
Task-KV: Task-aware KV Cache Optimization via Semantic Differentiation of Attention Heads MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8100fd37-f4e5-42b6-8a23-f126bcce6ec9 · inbound
Cache Me If You Must: Adaptive Key-Value Quantization for Large Language Models MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c27c169-0ae3-49b1-9f7d-caf7525136ff · inbound
Semantic Integrity Matters: Benchmarking and Preserving High-Density Reasoning in KV Cache Compression MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation cff4f956-d691-4937-a2a9-a02e8aa7f41c · inbound
Position: Episodic Memory is the Missing Piece for Long-Term LLM Agents MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10b2764a-bae7-4d44-83c7-958a0cdd247b · inbound
TransMLA: Multi-Head Latent Attention Is All You Need MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5509a30-74a1-4411-b871-35b83934b4b2 · inbound
LogQuant: Log-Distributed 2-Bit Quantization of KV Cache with Superior Accuracy Preservation MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 46aec037-8f7c-40db-bf10-3ec3a2194891 · inbound
Memory-Efficient Visual Autoregressive Modeling with Scale-Aware KV Cache Compression MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efa4715b-447e-4500-b65e-e60faca4728b · inbound
Semantic Scheduling for LLM Inference MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 338b3950-8251-4c0f-99be-28abbd3377f5 · inbound
SpindleKV: A Novel KV Cache Reduction Method Balancing Both Shallow and Deep Layers MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 958f7d1e-9fa0-4a05-91c7-7d69f4b8c3dd · inbound
S2O: Early Stopping for Sparse Attention via Online Permutation MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 20a2a60e-c8b6-45a4-b02f-46fc08f9e349 · inbound
HACK++: Towards More Effective Head-Aware Key-Value Compression for Efficient Visual Autoregressive Modeling MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6aa01486-410a-4af5-afd6-ddaead8d8ea1 · inbound
RoPE-Aware Bit Allocation for KV-Cache Quantization MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 90c45ba2-3c61-48ec-823f-22d5f61ed95d · inbound
FreqDepthKV: Frequency-Guided Depth Sharing for Robust KV Cache Compression in Long-Context LLM Inference MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1912ef61-47c5-439e-be22-b44e35444151 · inbound
DepthWeave-KV: Token-Adaptive Cross-Layer Residual Factorization for Long-Context KV Cache Compression MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 103
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8c1ab5a5-eceb-4b81-b1dc-7e73b170419a · inbound
What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9f43cf07-6576-47ee-b798-22ec9e9db6d5 · inbound
Looped Latent Attention: Cross-Loop KV Compression for Looped Transformers MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb39eb81-8397-4bdb-985c-5218d34daf3d · inbound
SOS-LoRA: Static Orthogonal-Subspace Low-Rank Adaptation with Fixed Multi-Scale Scaling MiniCache: KV Cache Compression in Depth Dimension for Large Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.