Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T00:33:27.738595Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 0 inbound Pith citation observations for arXiv:2608.01126.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T00:33:27.738595Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
22 of 22 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e2eeb969-5e33-495d-82f2-6c6302f4bac7 · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework Efficient memory management for large language model serving with PagedAttention,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 10c5b6ea-ce36-4dd6-bf5a-46eeb3f1a314 · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework Prompt cache: Modular attention reuse for low-latency inference,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2aaec917-3eff-47d3-8152-afd367a0b045 · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework SGLang: Efficient execution of structured language model programs,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16cb85a0-263c-4a50-83b4-20f1f9a89a89 · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework Cost-efficient large language model serving for multi-turn conversations with CachedAttention,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 519c1dc4-193f-4fe9-8c12-6d544378171d · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework Preble: Effi- cient distributed prompt scheduling for LLM serving,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8ed69102-1abf-4a82-988c-ccbeda51369b · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework Mooncake: A KVCache-centric Disaggregated Architecture for LLM Serving
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6cca0e0-e550-412a-a0e0-4fc291032f64 · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework Shadowserve: Interference-free KV cache fetching for distributed prefix caching,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e55ad204-30a7-4234-b30d-e3f1a12614a9 · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework Accelerating Local LLMs on Resource-Constrained Edge Devices via Distributed Prompt Caching
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c7746b22-b539-4273-8750-7f785c0c6357 · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework Caching placement in stochastic wireless caching helper networks: Channel selection diversity via caching,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 868fb934-095d-42f9-b513-5b3b03dd0315 · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework Probabilistic caching in wireless D2D networks: Cache hit optimal versus throughput optimal,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a046a7f6-bf6a-4b31-836e-aecac1c03b85 · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework A tractable approach to coverage and rate in cellular networks,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23e284de-6843-49b2-91c0-0029b9daa1ca · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework Haenggi,Stochastic Geometry for Wireless Networks
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88564840-d7b8-4f20-b6bd-5a0b7403e602 · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework Keyformer: KV cache reduction through key tokens selection for efficient generative inference,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1997ebb7-4474-4584-8fbe-1213f32c63b6 · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework Marconi: Prefix Caching for the Era of Hybrid LLMs
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acff936a-9c9a-4132-9cfa-0c7c262eb75a · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework Distserve: Disaggregating prefill and decoding for goodput- optimized large language model serving,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 979d09f1-5178-4063-8666-bfd273e32561 · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework Splitwise: Efficient generative LLM inference using phase splitting
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f74fc455-52c0-43e4-8a25-266a3d78e117 · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework Taming Throughput-Latency Tradeoff in LLM Inference with Sarathi-Serve
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dee33baa-89dd-4005-8575-75202281c9fe · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework Llumnix: Dynamic Scheduling for Large Language Model Serving
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 251d395c-dede-4ccd-bd4c-523d8ccf8a65 · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bd48454-58e4-4421-bc62-59bdf493b9ad · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework KVDirect: Distributed Disaggregated LLM Inference
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8f82b89-0ba8-445c-93ef-2f2041318699 · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework CacheGen: KV Cache Compression and Streaming for Fast Large Language Model Serving
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2b26b4b-9773-47df-b952-99168949a7c0 · outbound
Spatial Prefix Caching for Wireless Edge LLM Inference: A Stochastic-Geometry and Queueing Framework Towards Distributed Inference of LLMs on a P2P Network
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
No inbound Pith citation observations are available.