Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2402.11131.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T20:43:27.532056Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-23T05:52:37.526395Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation a02f28c2-4398-40ae-b504-4adb62a11546 · inbound
Small Language Models (SLMs) Can Still Pack a Punch: A survey (updated 2026) Speculative Streaming: Fast LLM Inference without Auxiliary Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 55c28658-f78e-4556-8a76-5f94d086f314 · inbound
Judge Decoding: Faster Speculative Sampling Requires Going Beyond Model Alignment Speculative Streaming: Fast LLM Inference without Auxiliary Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2abcd8ac-4e7f-44c2-a606-6e08b53cde8c · inbound
CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing Speculative Streaming: Fast LLM Inference without Auxiliary Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d7e73a3-b19b-4020-8b21-4c906ee9a86a · inbound
M2R2: Mixture of Multi-Rate Residuals for Efficient Transformer Inference Speculative Streaming: Fast LLM Inference without Auxiliary Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c04e46d-7b85-4a4a-89a6-1e6ba3cd289f · inbound
QuantSpec: Self-Speculative Decoding with Hierarchical Quantized KV Cache Speculative Streaming: Fast LLM Inference without Auxiliary Models
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31054bfa-44b0-4869-8a8e-e466a8545f10 · inbound
Your LLM Knows the Future: Uncovering Its Multi-Token Prediction Potential Speculative Streaming: Fast LLM Inference without Auxiliary Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a3dc938-8149-45cf-8772-de4aea22ac62 · inbound
CSR: Infinite-Horizon Real-Time Policies with Massive Cached State Representations Speculative Streaming: Fast LLM Inference without Auxiliary Models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 529d8384-7001-4f9c-b503-3d4eb6bbac29 · inbound
Leaky Language Models: Stealing Architecture and Inference Optimizations via Per-Token Timing Speculative Streaming: Fast LLM Inference without Auxiliary Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.