Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2406.01698.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T10:10:22.757378Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-11T01:57:50.618032Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation bd779b54-fef4-4499-b02c-d0d4df4e3041 · inbound
Memory Offloading for Large Language Model Inference with Latency SLO Guarantees Demystifying AI Platform Design for Distributed Inference of Next-Generation LLM models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e72a6acf-d675-41d8-8f5c-aa33806f719b · inbound
MIST: A Co-Design Framework for Heterogeneous, Multi-Stage LLM Inference Demystifying AI Platform Design for Distributed Inference of Next-Generation LLM models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fd25a10f-cd5d-43a4-96c0-d32870f5f889 · inbound
Scaling Intelligence: Designing Data Centers for Next-Gen Language Models Demystifying AI Platform Design for Distributed Inference of Next-Generation LLM models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5da1ad3d-e9c4-4aa5-927c-c81c802741e6 · inbound
Sim-FA: A GPGPU Simulator Framework for Fine-Grained Asynchronous Pipeline Analysis Demystifying AI Platform Design for Distributed Inference of Next-Generation LLM models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a5d85336-3a48-40c7-9833-bbd5d7ed03a1 · inbound
Sim-FA: A GPGPU Simulator Framework for Fine-Grained Asynchronous Pipeline Analysis Demystifying AI Platform Design for Distributed Inference of Next-Generation LLM models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d73e722-a1b4-46b1-b429-697183779ab8 · inbound
SOLAR: AI-Powered Speed-of-Light Performance Analysis Demystifying AI Platform Design for Distributed Inference of Next-Generation LLM models
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f1cfe1ea-0cbd-4e32-ab63-b576b25c3c1e · inbound
MLSYSIM: First-Principles Infrastructure Modeling for Machine Learning Systems Demystifying AI Platform Design for Distributed Inference of Next-Generation LLM models
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f22cb1a-6aeb-4615-b333-8e06171bedab · inbound
Think Before You Grid-Search: Floor-First Triage for LLM Serving Demystifying AI Platform Design for Distributed Inference of Next-Generation LLM models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4e5a5394-2bf8-4750-b027-16d66ad1e45b · inbound
Think Before You Grid-Search: Floor-First Triage for LLM Serving Demystifying AI Platform Design for Distributed Inference of Next-Generation LLM models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c51afc1d-95f4-4a48-b5dd-c68499f93faf · inbound
TileSight: A First-Principles Tile-Centric Analytical GPU Performance Model from Cores to Clusters Demystifying AI Platform Design for Distributed Inference of Next-Generation LLM models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ecb6663-a376-4b03-b5ae-600d7647461c · inbound
LLMET: Enabling Cross-Layer Evaluation of Emerging M3D Memories for Energy-Efficient LLM Serving Demystifying AI Platform Design for Distributed Inference of Next-Generation LLM models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9474e2ac-8847-448d-a2c5-33784756ba3a · inbound
SLIM: Saturation-Aware Lightweight Performance Modeling for LLM Serving Demystifying AI Platform Design for Distributed Inference of Next-Generation LLM models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4b1a9e1-72b9-4d48-981d-c4f5cd7b2c87 · inbound
RAG-Stack: Co-Optimizing RAG Serving Performance and Quality Demystifying AI Platform Design for Distributed Inference of Next-Generation LLM models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.