Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T13:15:07.306486Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 2 inbound Pith citation observations for arXiv:2412.00061.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T13:15:07.306486Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:16:26.218271Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-28T19:02:34.639740Z
28 of 28 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7dafc0a5-e07b-4b38-9f4c-963008ed5712 · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81e64fd7-4189-4b95-a1f8-8a6d545713fe · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Hydra: Sequentially-Dependent Draft Heads for Medusa Decoding
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39c6ce24-f594-4281-973d-e1b8eceb7b43 · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Medusa: Simple framework for accelerating llm generation with multiple decoding heads, 2023
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation cdc6218a-69f1-4979-98a0-ea0fcf3165cc · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Accelerating Large Language Model Decoding with Speculative Sampling
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbb37119-5ca2-498f-9683-4fe5acd61444 · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Cascade Speculative Drafting for Even Faster LLM Inference
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec52fe2a-7e90-4c52-a0e6-0d1e81cbad69 · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bb13f95-7222-43db-8910-0a07a5aa976d · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Training Verifiers to Solve Math Word Problems
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 680c2458-c6e9-49e8-bac1-419f3190087c · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Tutorial on directed acyclic graphs
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation beed94b1-721e-44ad-8974-102fac87eac9 · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 746dfe29-8e6b-4a38-a806-c19cbc1ca527 · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration SPEED: Speculative Pipelined Execution for Efficient Decoding
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67459f86-ae2a-4eef-a242-e0fb2c7f3ad4 · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Mistral 7B
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e41ad1c-e367-4eed-a1f7-21f74b60999f · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Fast inference from transformers via speculative decoding
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 035dd2d0-7333-48d4-8ba3-e175db4f4d79 · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration EAGLE: Speculative Sampling Requires Rethinking Feature Uncertainty
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 139f8768-1ac8-486b-a60b-659023c6dd8d · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration SpecInfer: Accelerating Generative Large Language Model Serving with Tree-based Speculative Inference and Verification
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b955a8cc-9cb4-4edf-b883-f1225addf738 · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Accelerating LLM Inference with Staged Speculative Decoding
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcf09a45-24f9-42ea-aa36-08e6cec69b61 · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Insertion transformer: Flexible sequence generation via insertion operations
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8554d824-62b3-4d4b-a860-74ce9a8650a1 · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Blockwise parallel decoding for deep autoregressive models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f7555fba-55e2-4eb3-9abc-3605e48fd5ec · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Instantaneous Grammatical Error Correction with Shallow Aggressive Decoding
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46a3c2dd-d6ee-4b52-b306-a8a8358e1e58 · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Spectr: Fast speculative decoding via optimal transport
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69cd8bde-5543-4faa-8d6b-5828699f8564 · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration An introduction to conditional random fields
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e5fb8e5f-0bc6-472f-9433-e9c514a86e46 · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration LLaMA: Open and Efficient Foundation Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74eb00b7-fb34-424d-898d-79eeedc215a4 · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b351093e-64fc-4cbd-a61f-1569b67f3936 · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Speculative decoding: Exploiting speculative execution for accelerating seq2seq generation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d116a5fe-8de5-42a1-9943-686b508e1d1d · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Unlocking Efficiency in Large Language Model Inference: A Comprehensive Survey of Speculative Decoding
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aeb6f735-149c-4e40-895f-ca6eae19d0ba · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Predictive Pipelined Decoding: A Compute-Latency Trade-off for Exact LLM Decoding
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0e5002c-af63-483c-9901-77ee3353e1d0 · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Draft & Verify: Lossless Large Language Model Acceleration via Self-Speculative Decoding
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8a9214b-1cd2-4e06-881f-14e16d536fba · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration Judging llm-as-a-judge with mt-bench and chatbot arena
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e8bcaf73-917d-478c-8fad-6b701a3a0a92 · outbound
Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration DistillSpec: Improving Speculative Decoding via Knowledge Distillation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfaa0113-e027-4ed0-9e7f-fb27dead30ad · inbound
Advancing Decoding Strategies: Enhancements in Locally Typical Sampling for LLMs Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c1fd930-b4d6-41aa-ae7c-cadfb9313ffb · inbound
DREAM-S: Speculative Decoding with Searchable Drafting and Target-Aware Refinement for Multimodal Generation Speculative Decoding with CTC-based Draft Model for LLM Inference Acceleration
Reference 112
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.