Pith. sign in

Paper Citation Record · LEDGER

Fast Large Language Model Collaborative Decoding via Speculation

As of 10 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 6 inbound Pith citation observations for arXiv:2502.01662.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.01662 v2

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T19:34:42.923184Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:53:58.892122Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T08:15:32.154273Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact0
  • verified fuzzy14
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 49871604-9da7-4e22-9b0d-e314f0068ade · outbound

This paper cites GPT-4 Technical Report.

Fast Large Language Model Collaborative Decoding via Speculation GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.718344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.718344Z digest=sha256:6f8abf5274ba385e9ae860ed4462157c008ccf0a68afcba8b60bdda1a91efd30

Observation 7aa4dc56-7a14-49d8-bfd5-489dc2304b41 · outbound

This paper cites Optimized multi-token joint decoding with auxiliary model for llm inference.

Fast Large Language Model Collaborative Decoding via Speculation Optimized multi-token joint decoding with auxiliary model for llm inference

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.573707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.724002Z digest=sha256:f9507b23d8f11cb26e09aaa6843f90762f1d30426c279dcc7504d6accee6f06c

Observation ae8786af-4db5-4dae-ac23-998c72282bbc · outbound

This paper cites Judge decoding: Faster speculative sampling requires going beyond model alignment.

Fast Large Language Model Collaborative Decoding via Speculation Judge decoding: Faster speculative sampling requires going beyond model alignment

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.557785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.729049Z digest=sha256:cabcf896c17b1e000ee420de7c63ec73f16915bb939a7ea3e650b0fd81105c5a

Observation 04428c34-ab39-4d2a-ba7b-65186a07b5c2 · outbound

This paper cites Qwen Technical Report.

Fast Large Language Model Collaborative Decoding via Speculation Qwen Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.734335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.734335Z digest=sha256:4873ec0f42e9e96639e3d24b452aaf24888a7faadec2742165f773a862db6e4a

Observation 8a3a0a72-5f5f-4ef9-b8db-465d01b864f3 · outbound

This paper cites D., Chen, D., and Dao, T.

Fast Large Language Model Collaborative Decoding via Speculation D., Chen, D., and Dao, T

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.541076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.739899Z digest=sha256:e02eee44801a98caa8f29c1add9ff53c6a2541ff322837326b2fde7d4dda0e63

Observation d5b976c1-93b1-48db-85ef-b1d084624ca6 · outbound

This paper cites FrugalGPT: How to Use Large Language Models While Reducing Cost and Improving Performance.

Fast Large Language Model Collaborative Decoding via Speculation FrugalGPT: How to Use Large Language Models While Reducing Cost and Improving Performance

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.751817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.751817Z digest=sha256:0ca86d0c386b71db894f613700d7b5213fd91fe76f40d06762c0ee438606b5a9

Observation cf78432b-2df7-41ec-9ecd-33183fcdfe82 · outbound

This paper cites an unresolved cited work.

Fast Large Language Model Collaborative Decoding via Speculation Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.756975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.756975Z digest=sha256:2d352da2c8024cb51b59c231fa7045603bd0b9527858ec97a78779dab83148f8

Observation a4739842-7586-40bf-8294-ccc7371d628c · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Fast Large Language Model Collaborative Decoding via Speculation Training Verifiers to Solve Math Word Problems

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.761775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.761775Z digest=sha256:3eee9dd3980570a8bb07c91fcc1d3adf34e17ebf9f2e79f74ec84a5f7c2d3e48

Observation eec7d799-9a97-44c5-bc9f-91d2f33b4537 · outbound

This paper cites Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing.

Fast Large Language Model Collaborative Decoding via Speculation Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.766560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.766560Z digest=sha256:734a50e4ac3d8e57387208143aa1751d4b621bed5dce168d403dd6a1dc405e95

Observation 252909aa-7469-4e0b-bd55-20a0f166ef01 · outbound

This paper cites The Llama 3 Herd of Models.

Fast Large Language Model Collaborative Decoding via Speculation The Llama 3 Herd of Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.771328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.771328Z digest=sha256:f149322561514ca8de16c6b6d2d18fcf5353a6626970ee59e3101d5cbfed0911

Observation 23dd7218-7081-46fd-b675-6f638150924e · outbound

This paper cites Layerskip: Enabling early exit inference and self-speculative decoding.

Fast Large Language Model Collaborative Decoding via Speculation Layerskip: Enabling early exit inference and self-speculative decoding

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.514791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.776079Z digest=sha256:30e312ad363e6631ebe20dcc35edc2c31ed389e176f7f210989b5684a55c94a8

Observation 748291ee-edbe-449d-a385-b5fc5f86673f · outbound

This paper cites Llm-topla: Efficient llm ensemble by maximising diversity.

Fast Large Language Model Collaborative Decoding via Speculation Llm-topla: Efficient llm ensemble by maximising diversity

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.498782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.780629Z digest=sha256:5aaf88892ad1b566aaea826b5391016804584f58c0b7f12d3c38e1a1710b3676

Observation d40c8953-5ab6-4f82-bfdb-b9a654dced8d · outbound

This paper cites Graph-structured speculative decoding.

Fast Large Language Model Collaborative Decoding via Speculation Graph-structured speculative decoding

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.482971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.785156Z digest=sha256:fcaf2b47b8bcf5c92d53ab72e6ac6d1d38cdcf92f0da05607c775d3c265512ce

Observation a33534c7-7f12-43c4-8123-8d23427725b5 · outbound

This paper cites Measuring massive multitask language understanding.

Fast Large Language Model Collaborative Decoding via Speculation Measuring massive multitask language understanding

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.789550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.789550Z digest=sha256:d44f6a44aa35856bfa775e78f3b3d48993d351ca435fe9606579d54116f094e6

Observation 75fa5e1c-755e-483e-92e0-46360856478e · outbound

This paper cites and Huang, H.

Fast Large Language Model Collaborative Decoding via Speculation and Huang, H

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.454728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.794003Z digest=sha256:1aa364560d8ec160bd6969cb8da9af66225caacef48844453449d87858e49e10

Observation b6bf4570-b47d-4935-b25b-c2f6001300c1 · outbound

This paper cites an unresolved cited work.

Fast Large Language Model Collaborative Decoding via Speculation Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.798452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.798452Z digest=sha256:29e73310a3e0b05b4bb19187caa24babd50b419713b8c592c8cb5877dcf55979

Observation fb7fb685-553c-47a5-a86f-386059aa2797 · outbound

This paper cites an unresolved cited work.

Fast Large Language Model Collaborative Decoding via Speculation Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-09T19:34:43.427208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.803010Z digest=sha256:795523e2009a0c4104cbbbf30fa17eb98f0a33f4d4a84bf7f8e16dcd98a87537

Observation fb261483-57e6-4297-a5e1-62adc71cd95d · outbound

This paper cites Fast inference from transformers via speculative decoding.

Fast Large Language Model Collaborative Decoding via Speculation Fast inference from transformers via speculative decoding

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.411677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.807337Z digest=sha256:dce9554a1f187fd916d1ddd4dfc34e1c6e28393099185b63cc2b4c5dc15d551c

Observation e412d72e-990f-4846-b9d3-f9aedcdc5b8e · outbound

This paper cites Purifying Large Language Models by Ensembling a Small Language Model.

Fast Large Language Model Collaborative Decoding via Speculation Purifying Large Language Models by Ensembling a Small Language Model

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.813056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.813056Z digest=sha256:81c35e79e77a6349b73efc5bc6e46802c30bdbd3b56dbd62c974a441cf5ccf6b

Observation 0e2fa5d1-4094-48c8-b91f-d876e514fcb6 · outbound

This paper cites L., Holtzman, A., Fried, D., Liang, P., Eisner, J., Hashimoto, T.

Fast Large Language Model Collaborative Decoding via Speculation L., Holtzman, A., Fried, D., Liang, P., Eisner, J., Hashimoto, T

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.394793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.818122Z digest=sha256:12b981bd492db2501aa9acbfda7fc79b413e4d1dccc8b794a0524d7512bf7467

Observation 67c3f56f-e88c-46de-993f-27dfadf9262b · outbound

This paper cites Eagle: Speculative sampling requires rethinking feature uncertainty.

Fast Large Language Model Collaborative Decoding via Speculation Eagle: Speculative sampling requires rethinking feature uncertainty

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.379561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.822769Z digest=sha256:f28ce27815a9fefdac2af77cd3cf635bc352dfa0b2900ddbbad7010b3c0c98ec

Observation 4b49196a-229e-40f8-82bd-750fbf7533be · outbound

This paper cites Eagle-2: Faster inference of language models with dynamic draft trees.

Fast Large Language Model Collaborative Decoding via Speculation Eagle-2: Faster inference of language models with dynamic draft trees

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.364261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.827314Z digest=sha256:b0a2c147eade1f770ceea98e592084cd991e9011f2d289a3233777aa4e22c40b

Observation 48bfff0f-49b9-468d-981c-56d669518c22 · outbound

This paper cites DeepSeek-V3 Technical Report.

Fast Large Language Model Collaborative Decoding via Speculation DeepSeek-V3 Technical Report

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.831914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.831914Z digest=sha256:d9e5c5910daa31c61094495ae840a8c768d0ac36f14815972c38ee79855919d1

Observation 02f97edf-a985-47ca-a27f-00e711e5f0be · outbound

This paper cites Merge, Ensemble, and Cooperate! A Survey on Collaborative Strategies in the Era of Large Language Models.

Fast Large Language Model Collaborative Decoding via Speculation Merge, Ensemble, and Cooperate! A Survey on Collaborative Strategies in the Era of Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.836845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.836845Z digest=sha256:84a57c4e8db837eb2fddf44a5f95623d1b57fbb675ac706f776dff9bf85c5f56

Observation d2b8a880-4c2a-4ee9-a89f-fe0377bb164d · outbound

This paper cites Routing to the Expert: Efficient Reward-guided Ensemble of Large Language Models.

Fast Large Language Model Collaborative Decoding via Speculation Routing to the Expert: Efficient Reward-guided Ensemble of Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.841795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.841795Z digest=sha256:f06c9b53376028417c7e058b98bac6c48214cbefe58ade4e6d1b08affbf863f8

Observation 9bdd421b-f85f-437f-a567-40302a9d35e4 · outbound

This paper cites Blending Is All You Need: Cheaper, Better Alternative to Trillion-Parameters LLM.

Fast Large Language Model Collaborative Decoding via Speculation Blending Is All You Need: Cheaper, Better Alternative to Trillion-Parameters LLM

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.846637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.846637Z digest=sha256:bad0fc600d25d00bf539138f6d19489d387a7d237ab510a7d6338f82bee4fa48

Observation 72939cfb-f673-4b09-9f62-2d5fa45a280a · outbound

This paper cites an unresolved cited work.

Fast Large Language Model Collaborative Decoding via Speculation Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-09T19:34:43.348336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.851824Z digest=sha256:4b2818ab0fa0789fb1926f8c3a56ae4e373d6a8ccbc23f196bdaa0776adcfe61

Observation 77f8a8d3-87c7-455c-954d-c70c8b7efd55 · outbound

This paper cites Accelerating Large Language Model Decoding with Speculative Sampling.

Fast Large Language Model Collaborative Decoding via Speculation Accelerating Large Language Model Decoding with Speculative Sampling

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.856186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.856186Z digest=sha256:c97679ed54a4afcac512a460368d6a794ec0a3f8128ed8ad85c2bde64b5abbed

Observation 03a5b853-f377-4768-98d3-6931b2cf520e · outbound

This paper cites J., and Manning, C.

Fast Large Language Model Collaborative Decoding via Speculation J., and Manning, C

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.860706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.860706Z digest=sha256:bb2afcdfe65c423322cadb3251e0dc30605e26ea161672bc950571e22119c359

Observation 27ebf683-fdb2-41a6-ac28-69759eb4949c · outbound

This paper cites Large language model routing with benchmark datasets.

Fast Large Language Model Collaborative Decoding via Speculation Large language model routing with benchmark datasets

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.332303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.865522Z digest=sha256:b1e770d1fca087a03860dbd4f97a9fffc3437d0da0f087a5d3ae39093c66d5b2

Observation b2c7f695-12f1-424e-8de9-df01264d1df0 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Fast Large Language Model Collaborative Decoding via Speculation Gemini: A Family of Highly Capable Multimodal Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.870039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.870039Z digest=sha256:a29ee98d478008cccf08806b2a04d660b663ed5725f329f79b52281647bb66a1

Observation 05fe3f58-a276-4ef0-8fb0-82ff86f795d5 · outbound

This paper cites Qwen2.5: A party of foundation models, September 2024.

Fast Large Language Model Collaborative Decoding via Speculation Qwen2.5: A party of foundation models, September 2024

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.874887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.874887Z digest=sha256:9ae3c3bddb68a58ca8043bfa2030571e6e418ec787f0beac3b43ead68e367445

Observation 090bae2a-49ad-4bf2-8af7-c14f962c974b · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Fast Large Language Model Collaborative Decoding via Speculation Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.879635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.879635Z digest=sha256:eb26fc5ae69d100c97bd0ec03ecd5d33535cd1d275b52acac82c38c6605632d9

Observation a18cd2b1-401e-46a9-a9ed-6ccf3ad43396 · outbound

This paper cites MLLM can see? Dynamic Correction Decoding for Hallucination Mitigation.

Fast Large Language Model Collaborative Decoding via Speculation MLLM can see? Dynamic Correction Decoding for Hallucination Mitigation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.884194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.884194Z digest=sha256:48fd73ebc2d410c295729ad6921e9bba00cfbf2fc91cc935971160688bc19fe1

Observation 2ea745e2-d4ab-4b09-9eec-151ba1c3f0d7 · outbound

This paper cites Generation Meets Verification: Accelerating Large Language Model Inference with Smart Parallel Auto-Correct Decoding.

Fast Large Language Model Collaborative Decoding via Speculation Generation Meets Verification: Accelerating Large Language Model Inference with Smart Parallel Auto-Correct Decoding

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.889215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.889215Z digest=sha256:17df83200aa26fe6c9defc723f5ef0638b57ec2e4989fb12923266ee22c7aae1

Observation 1127ae5f-789a-4bfa-bd84-6e73b7bd718e · outbound

This paper cites C., Ziqi, Y., Yucheng, C., and Li, Y.-S.

Fast Large Language Model Collaborative Decoding via Speculation C., Ziqi, Y., Yucheng, C., and Li, Y.-S

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.894227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.894227Z digest=sha256:205a960d4dd0396847028d054b10beb404f1d7cf2d61028222864610e0a27c8f

Observation eb5975f4-8907-4908-ab98-6f337af84a9f · outbound

This paper cites Speculative contrastive decoding.

Fast Large Language Model Collaborative Decoding via Speculation Speculative contrastive decoding

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.899070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.899070Z digest=sha256:b2a7a49c7642f14b1d7b200f7906999787aa79d41efdca028c5206a34c7b02de

Observation b9de50db-3336-40d3-915a-dcebf95d5daf · outbound

This paper cites Draft&verify: Lossless large language model acceleration via self-speculative decoding.

Fast Large Language Model Collaborative Decoding via Speculation Draft&verify: Lossless large language model acceleration via self-speculative decoding

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.305285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.903885Z digest=sha256:b498e1adb28019b65db6921933da07cbd491664eeb27ff98427a97fa38ad42fa

Observation fd20c50e-dc9e-4d74-bfff-c787c4675911 · outbound

This paper cites V., Mihaylov, T., Ott, M., Shleifer, S., Shuster, K., Simig, D., Koura, P.

Fast Large Language Model Collaborative Decoding via Speculation V., Mihaylov, T., Ott, M., Shleifer, S., Shuster, K., Simig, D., Koura, P

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.908917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.908917Z digest=sha256:41f0271a71b04668baeb4faa174c16e8031e069614a94d50661cf82b19ae9c9b

Observation 30691e11-0381-4fe2-9c08-9a63eecc2e8a · outbound

This paper cites P., Zhang, H., Gonzalez, J.

Fast Large Language Model Collaborative Decoding via Speculation P., Zhang, H., Gonzalez, J

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.913554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.913554Z digest=sha256:26b682281dfe45f23952cfee28fee688c1591730eb4dc0837e51de65cbb7c535

Observation 11d3e7c9-1227-4859-ae1a-6560d4efc7e6 · outbound

This paper cites S., Menon, A., Rostamizadeh, A., Kumar, S., Kagy, J.-F., and Agarwal, R.

Fast Large Language Model Collaborative Decoding via Speculation S., Menon, A., Rostamizadeh, A., Kumar, S., Kagy, J.-F., and Agarwal, R

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:34:43.267902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:34:42.918529Z digest=sha256:9519de6f15625efa365edbfaefee39ddb288f54f427735746a7d6c21e4f945bf

Observation c9fcbd5e-9ba1-411f-8df0-3a967ae10fac · outbound

This paper cites write newline.

Fast Large Language Model Collaborative Decoding via Speculation write newline

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.923184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.923184Z digest=sha256:65bb14c7e9f739aa1e717af5bf6b21e995ecc96076e1055bced14dc141537604

Pith citing papers

Observation 571f5654-e954-4b8c-aeb2-2456940e7102 · inbound

Multimodal Tabular Reasoning with Privileged Structured Information cites this paper.

Multimodal Tabular Reasoning with Privileged Structured Information Fast Large Language Model Collaborative Decoding via Speculation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T10:53:58.892122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:53:58.892122Z digest=sha256:38d2b0f5d7b63d92bd1c57e10de7d30d6dce140303a0d38b1b8280933985a928

Observation 54d16537-58a6-40ee-901d-fa6a88501dbf · inbound

SignAligner: Harmonizing Complementary Pose Modalities for Coherent Sign Language Generation cites this paper.

SignAligner: Harmonizing Complementary Pose Modalities for Coherent Sign Language Generation Fast Large Language Model Collaborative Decoding via Speculation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:10:11.741454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:10:11.741454Z digest=sha256:285135d81af700484999e09f560f695aadbcc907d2ca4916ed9a4867ab2b9d02

Observation 684a0bb9-e8f3-4fb3-98b2-7ac8e6982251 · inbound

AnchorSeg: Language Grounded Query Banks for Reasoning Segmentation cites this paper.

AnchorSeg: Language Grounded Query Banks for Reasoning Segmentation Fast Large Language Model Collaborative Decoding via Speculation

Reference 58

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T09:43:49.444527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T05:10:44.608959Z digest=sha256:a3a66268dc370bf90d87be8c5c1d263fd74486df10866ce90288716a80e0c3f1

Observation 7ed78080-faf4-480a-b385-14755324412a · inbound

SpecFed: Accelerating Federated LLM Inference with Speculative Decoding and Compressed Transmission cites this paper.

SpecFed: Accelerating Federated LLM Inference with Speculative Decoding and Compressed Transmission Fast Large Language Model Collaborative Decoding via Speculation

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:31:17.251029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-07T15:12:12.450979Z digest=sha256:af28fb06d7bf32e89c434bda7a3dd5deff0ac6f0ef22301156f7616a0a762dae

Observation 4c1f544c-6ebb-4f79-ab83-08e5ecbad181 · inbound

Rethinking LLM Ensembling from the Perspective of Mixture Models cites this paper.

Rethinking LLM Ensembling from the Perspective of Mixture Models Fast Large Language Model Collaborative Decoding via Speculation

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-01T08:15:32.157274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-01T08:07:09.032028Z digest=sha256:d3addb53d9d245be37510c9110c6e0c0a65dafdebe003f5936bd96c91f064a40

Observation 9f275b1b-19c2-4763-b4dc-173b72f56d06 · inbound

Revisiting Lossy Verification in Speculative Decoding: Mechanisms, Trade-offs, and Failure Modes cites this paper.

Revisiting Lossy Verification in Speculative Decoding: Mechanisms, Trade-offs, and Failure Modes Fast Large Language Model Collaborative Decoding via Speculation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-01T12:07:01.364373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T12:07:01.364373Z digest=sha256:dc8092920361c358399becabb3ca57817fc312da88bbfe607edb2e2b455fa404