Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T01:10:40.475653Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 1 inbound Pith citation observation for arXiv:2506.11886.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T01:10:40.475653Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-08T03:16:00.868964Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T22:11:14.344298Z
37 of 37 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d32e9be9-ef2c-4430-b9bf-712937b1c710 · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02d1bb4c-cd57-4c9d-8c14-2494a8bd2552 · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fb386be-fb77-4d39-a4b4-f5461bb4da0b · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77689f63-3157-4247-bc97-f1a0c49a8c57 · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Palu: Compressing KV-Cache with Low-Rank Projection
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6253145d-a773-4e12-acdc-ac8697952030 · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6da607bf-3b81-4706-8acf-a729cf6e5bbc · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b7b21fae-2aca-4a3a-b3b9-908a6c333cce · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 68e19fd9-cb02-46b3-95a7-b72eb6c75299 · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache SKVQ: Sliding-window Key and Value Cache Quantization for Large Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5812b41b-1140-4cc9-be14-481be40d733f · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache The Llama 3 Herd of Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea84603b-c5ee-4a70-9602-28b9142a0bb8 · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Challenges in Deploying Long-Context Transformers: A Theoretical Peak Performance Analysis
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99434401-507e-4f4f-b858-2bcab6d94012 · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c5521ef6-9edd-4b9c-ac07-2dcd278e4c3c · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7240aae5-b597-4d7a-bb9a-2b6eae481e31 · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Fourier Transformer: Fast Long Range Modeling by Removing Sequence Redundancy with FFT Operator
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8f7e5ad-7493-4c4b-bf6d-c01e6c2334aa · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache KVQuant: Towards 10 Million Context Length LLM Inference with KV Cache Quantization
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9fb54cf-ac7a-4c2d-b7a3-31786ecf5f9d · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache RULER: What's the Real Context Size of Your Long-Context Language Models?
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46ba7888-c3be-4b12-8c05-e67657da7bb9 · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 967d1b8f-b801-46da-aba6-4b36a56336f4 · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67428f7a-63c7-4a04-9878-4d2b2088ef6b · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache SnapKV: LLM Knows What You are Looking for Before Generation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adf38444-e51f-42c3-8966-7231c91222ec · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff632d91-d771-47fd-9d85-475ea280ab99 · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 983d1a8a-e6be-4680-8311-397d95ee450b · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e7937253-4c7a-41ec-a915-1dfb34afc9f8 · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache KIVI: A Tuning-Free Asymmetric 2bit Quantization for KV Cache
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fe3750d-4734-40d0-9fae-ead55c535cea · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 65c08443-1b49-4777-a6b1-5d70180bcccb · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e0282b44-b765-4fb5-93e1-2489826ff66e · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache GPT-4 Technical Report
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ef5c277-ceb2-4d3a-bfd5-2dad65f6b045 · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b6e55613-4df6-4cc7-995d-c7e0a46085b8 · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95c0db8f-a94d-4cb8-afae-3c29ab962c30 · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Eigen Attention: Attention in Low-Rank Space for KV Cache Compression
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97ed5edf-350d-439e-8607-810faf00b50a · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Unresolved cited work
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 061adc09-550a-45d4-b832-eb41ae823006 · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd2bb41e-d2cc-4fd1-b1fc-e96016993a50 · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Unresolved cited work
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4246cb51-bf02-4602-aaf7-32b841f3ddb0 · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e068253c-41e4-4b36-8de6-4d1079fbc21c · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache PyramidDrop: Accelerating Your Large Vision-Language Models via Pyramid Visual Redundancy Reduction
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a93da625-3f42-4cc3-8641-7cc93e498190 · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05e59751-bd21-4913-9283-e5589a22469d · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache Unresolved cited work
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ae29bc8-1ccb-45ae-bb41-224307ebc67d · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache online" 'onlinestring :=
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa1ff873-8178-4325-85f1-4dc287ea332c · outbound
Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache write newline
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c197b14-e90d-4b7e-87a7-c398ac0bdf04 · inbound
When Quantization Is Free: An int4 KV Cache That Outruns fp16 on Apple Silicon Beyond Homogeneous Attention: Memory-Efficient LLMs via Fourier-Approximated KV Cache
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.