Pith. sign in

Paper Citation Record · LEDGER

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding

As of 8 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2502.08020.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.08020 v2

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T11:11:44.679976Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact2
  • verified fuzzy2
  • unresolved31
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f1c88063-e920-424e-9a23-aca66e854004 · outbound

This paper cites GPT-4 Technical Report.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.544426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.544426Z digest=sha256:f95481c5818061d4018b577d159767e7de1d473c26abddbaa19a47db022dc7f5

Observation fe2f73eb-d4f4-4090-87c3-bd7602e0a2d2 · outbound

This paper cites The Llama 3 Herd of Models.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding The Llama 3 Herd of Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.566293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.566293Z digest=sha256:b3226ffe0746933de42fcea543a4a7c4bb980915768dfc94a89cc12c70c24520

Observation 02401361-2175-4902-a5cc-ac6e24047b05 · outbound

This paper cites LayerSkip: Enabling Early Exit Inference and Self-Speculative Decoding.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding LayerSkip: Enabling Early Exit Inference and Self-Speculative Decoding

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.570486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.570486Z digest=sha256:df2d05270a07611f99c358fcd030c97fdba371065037a7d9798889bfce8aa99f

Observation ef4643f8-036c-425e-a74f-17ed7d626ae7 · outbound

This paper cites Arcee's MergeKit: A Toolkit for Merging Large Language Models.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding Arcee's MergeKit: A Toolkit for Merging Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.574179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.574179Z digest=sha256:8699bfefc8ef481d1c236d45d720af9c7fe10bf6810d85db37f517e84021e5b1

Observation 74effaa4-f223-4aff-b597-3a67b4a55f62 · outbound

This paper cites CharED: Character-wise Ensemble Decoding for Large Language Models.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding CharED: Character-wise Ensemble Decoding for Large Language Models

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-08T11:11:44.990466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:11:44.578176Z digest=sha256:1a049c18e4b85d5d8065210740c16f439ff302fe558d2d5e74d2d5be25ab3d90

Observation cd165eb8-2c2f-48a2-941d-18167b7bed3d · outbound

This paper cites DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.581783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.581783Z digest=sha256:d21c863305a7a2e2a816924a4fc0643ac6169030d7a4176b3eb41675cb2c4da2

Observation 8294b9c4-0386-42d3-bf90-a250a678b6ab · outbound

This paper cites REST: Retrieval-Based Speculative Decoding.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding REST: Retrieval-Based Speculative Decoding

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.585657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.585657Z digest=sha256:db22917f351d54452ebf1c69788516e73688647420d33a376478ec7103498774

Observation 5e26a74a-f867-4f6f-b5ee-c3e1726addef · outbound

This paper cites Measuring Massive Multitask Language Understanding.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding Measuring Massive Multitask Language Understanding

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.589369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.589369Z digest=sha256:52e4983a52303475357517577e22d10246928aec6103a18fb5b0788ff675af63

Observation 58036a07-7914-4ce7-82b4-06e2650c5d25 · outbound

This paper cites Mistral 7B.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding Mistral 7B

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.597320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.597320Z digest=sha256:f481f4cdf36299609c6bcff519e7d7c86dae2f1977f796e43dea423c57268020

Observation f5d946a5-ce6d-4e39-ad76-8bec83e82c11 · outbound

This paper cites SOLAR 10.7B: Scaling Large Language Models with Simple yet Effective Depth Up-Scaling.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding SOLAR 10.7B: Scaling Large Language Models with Simple yet Effective Depth Up-Scaling

Reference 15

Resolution
malformed identifier
no resolver link, observed 2026-08-08T11:11:44.600880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.600880Z digest=sha256:f54ec256b535d5aec2c35e71825d72758e01cd023a804a3a1bb51bb5366142a6

Observation 8155c263-900f-463b-bca9-60a5013e2078 · outbound

This paper cites Nearest Neighbor Speculative Decoding for LLM Generation and Attribution.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding Nearest Neighbor Speculative Decoding for LLM Generation and Attribution

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.604539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.604539Z digest=sha256:36e831e889ca53475200088df78eec31098a41b8b69899f2d2755836398fdecd

Observation b08be297-1449-4dfd-9263-120d214869a6 · outbound

This paper cites TruthfulQA: Measuring How Models Mimic Human Falsehoods.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding TruthfulQA: Measuring How Models Mimic Human Falsehoods

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.608838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.608838Z digest=sha256:3439a297214d4cd23205b38d24eb410772916e111dcf1932f55f96907bfec340

Observation c33c23c0-a3e2-46e5-bad9-3c707f678f44 · outbound

This paper cites WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.612539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.612539Z digest=sha256:b1f72db7e3222a06f1d24046d3fa5dd8b96ee10a44988942374615902520bc98

Observation d3459bcc-ba0c-4930-9bdb-c49e0dd33885 · outbound

This paper cites SpecInfer: Accelerating Generative Large Language Model Serving with Tree-based Speculative Inference and Verification.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding SpecInfer: Accelerating Generative Large Language Model Serving with Tree-based Speculative Inference and Verification

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.616267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.616267Z digest=sha256:6cdf6ccdaef1f237fc8d624292b77a101537a5c65c1bf3089d7063fede9ee7ec

Observation 318c6d8a-22e0-450d-9c13-d5dbcfaef349 · outbound

This paper cites RouteLLM: Learning to Route LLMs with Preference Data.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding RouteLLM: Learning to Route LLMs with Preference Data

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.619720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.619720Z digest=sha256:fc02516f395174e534af8fc1380e5c30c4fed52d99d46eeafc8ccd90efbebfb1

Observation b7e3a1d0-600e-4ddb-a448-84fe9dbb1dcd · outbound

This paper cites tinyBenchmarks: evaluating LLMs with fewer examples.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding tinyBenchmarks: evaluating LLMs with fewer examples

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.623525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.623525Z digest=sha256:404eb86821ed7a4eb6478975dde3d0afd982be8c78e1b31ff6e2f14e9c9910e9

Observation e738b2c6-01bb-42f5-99ca-b0dbfa6622b1 · outbound

This paper cites Learning to Decode Collaboratively with Multiple Language Models.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding Learning to Decode Collaboratively with Multiple Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.627016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.627016Z digest=sha256:71921f89332cb70d025038265052eeb1deef142385f68e748c4064a38aadb4ad

Observation 8e349a1d-dae5-4151-971b-bc1dd2e68b9e · outbound

This paper cites Knowledge Fusion of Large Language Models.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding Knowledge Fusion of Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.634199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.634199Z digest=sha256:cd5e601cace19f97deff1013e6326de74a4a500d11482d3d793a9293463b3a3c

Observation 7e61f48a-c519-4174-886d-2ad60b69a50e · outbound

This paper cites Federated Learning with Matched Averaging.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding Federated Learning with Matched Averaging

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.638040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.638040Z digest=sha256:f243a96fdcf08d88c7e9b770175fcae48441b5006efa2e64dc70973e4bd9258b

Observation 40e10578-8a23-49d9-b8a1-79251b34e779 · outbound

This paper cites Fusing Models with Complementary Expertise.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding Fusing Models with Complementary Expertise

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.641988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.641988Z digest=sha256:ad2b2674d6cef1b77b1c786528a72564df607bfd73f63fdb82cc7f92c3d10293

Observation b34b07c6-e9dd-4733-afff-3155b3b24259 · outbound

This paper cites LLaMA Pro: Progressive LLaMA with Block Expansion.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding LLaMA Pro: Progressive LLaMA with Block Expansion

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.645850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.645850Z digest=sha256:38ad45d3e94c8aa50978aeaaff7389e851dc3fc86518e5e870f6ec4a2420bf2d

Observation 40f3ccc8-a460-4563-9ac9-7a62da980de5 · outbound

This paper cites Speculative decoding: Exploiting speculative execu- tion for accelerating seq2seq generation.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding Speculative decoding: Exploiting speculative execu- tion for accelerating seq2seq generation

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:11:45.079656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:11:44.649541Z digest=sha256:41d6581307c89d96f09fbc3969a5f6255b295792f18f04a1d804362d34d9b10e

Observation 4b6a643e-a163-4897-818a-5763be994129 · outbound

This paper cites Unlocking Efficiency in Large Language Model Inference: A Comprehensive Survey of Speculative Decoding.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding Unlocking Efficiency in Large Language Model Inference: A Comprehensive Survey of Speculative Decoding

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.653199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.653199Z digest=sha256:32e3c3cbc66f5d1fa4cb5434d2e11b86ff7faefab292d6c2a87f686aa323fbf0

Observation bf853922-cfcb-4687-a7de-72a0d0581b14 · outbound

This paper cites OpenMoE: An Early Effort on Open Mixture-of-Experts Language Models.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding OpenMoE: An Early Effort on Open Mixture-of-Experts Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.656931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.656931Z digest=sha256:65e4943bee5ce2a96bdf609d5090a80b272b943b34393ce80a0ee05ab0fc214e

Observation cd65d410-14f6-4eaa-9746-915eb3f41109 · outbound

This paper cites Predictive Pipelined Decoding: A Compute-Latency Trade-off for Exact LLM Decoding.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding Predictive Pipelined Decoding: A Compute-Latency Trade-off for Exact LLM Decoding

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.660620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.660620Z digest=sha256:9d10bcf440e36da6b86099cbbbae2f3987549185d7e78359125c9ea813888fc3

Observation 64e82450-681d-4c0e-a5d8-ab5eaf12af0a · outbound

This paper cites HellaSwag: Can a Machine Really Finish Your Sentence?.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding HellaSwag: Can a Machine Really Finish Your Sentence?

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.664644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.664644Z digest=sha256:5a38dd507fdea9ece9361eb766ce2e11a4612cd6883b6ad374c292d6404c89df

Observation df442e90-9acb-4880-8538-07ae015a6601 · outbound

This paper cites TinyLlama: An Open-Source Small Language Model.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding TinyLlama: An Open-Source Small Language Model

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.672626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.672626Z digest=sha256:bb59ff1f3865f46de59a9e8abb92adffb0c6e2447b34839162a25b458213b59a

Observation c13cc53f-76d5-4775-be7a-e2b7ecdc3e84 · outbound

This paper cites DistillSpec: Improving Speculative Decoding via Knowledge Distillation.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding DistillSpec: Improving Speculative Decoding via Knowledge Distillation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.676440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.676440Z digest=sha256:981891a085c9fa45f0ab46e88f45ead0f93a723a45cd0346d21412c60ebc0cf9

Observation 8f6bc54d-51fa-4d30-a039-f45eec127b6c · outbound

This paper cites LLaMA-MoE: Building Mixture-of-Experts from LLaMA with Continual Pre-training.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding LLaMA-MoE: Building Mixture-of-Experts from LLaMA with Continual Pre-training

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.679976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.679976Z digest=sha256:3901465abefc6a4d1777cc1ffcdc19ddfb381570708b1da7557d70f9271b54e7

Observation 72c7d422-84b5-46d2-83e1-2774e256cb9a · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.630374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.630374Z digest=sha256:41cf70debbf35d7a43eb2554cb05938e7fda6272b83d6163e1a48f605d4daefb

Observation d6ce2fff-505c-46e5-915d-2f6178d92b1e · outbound

This paper cites Draft & Verify: Lossless Large Language Model Acceleration via Self-Speculative Decoding.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding Draft & Verify: Lossless Large Language Model Acceleration via Self-Speculative Decoding

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.668944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.668944Z digest=sha256:11ae4a49a7e23fbcd0742add49da50f223ff2f885a048486588dd267a3ebf122

Observation 4e180fbc-8f5e-45de-a96f-b3b859ab0f9f · outbound

This paper cites Analysis of Linear Mode Connectivity via Permutation-Based Weight Matching: With Insights into Other Permutation Search Methods.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding Analysis of Linear Mode Connectivity via Permutation-Based Weight Matching: With Insights into Other Permutation Search Methods

Reference 2020

Resolution
verified exact
local_arxiv, observed 2026-08-08T11:11:44.940264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:11:44.593522Z digest=sha256:1a6e4a6bdde2f0040808ff46bc490c7dcd6755c6e0d49ea64874816872e46808

Observation c9f857e3-9ea8-4b2c-a035-cf3ef17643d1 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding Training Verifiers to Solve Math Word Problems

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.562217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.562217Z digest=sha256:885cd1bca05a5be10f97fdabe0381a2e84af6b06adb5281d16f009529c16295f

Observation 7b4bd154-b336-4246-bad2-fefd95285e18 · outbound

This paper cites J., Ganesh, E.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding J., Ganesh, E

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:11:45.092327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:11:44.553936Z digest=sha256:bd8ad76da5e84b7578a4ba2d699077271aae41299d813ab6976f0faa05ccfa2e

Observation dbea834e-fc5d-4e3b-bb79-8f9cd4b472e5 · outbound

This paper cites Git Re-Basin: Merging Models modulo Permutation Symmetries.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding Git Re-Basin: Merging Models modulo Permutation Symmetries

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.549448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.549448Z digest=sha256:aec43794c083071ad34b2c14ecee453c7f6e472ed07ada3779d515221841c925

Observation b4213b6b-6b4b-4ef6-ab25-734519ee4d56 · outbound

This paper cites Accelerating Large Language Model Decoding with Speculative Sampling.

Speculate, then Collaborate: Fusing Knowledge of Language Models during Decoding Accelerating Large Language Model Decoding with Speculative Sampling

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-08T11:11:44.557990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:11:44.557990Z digest=sha256:17d58c77335037008c3378841a312adad365b9d514f41d38dcfc057e59bfe69d

Pith citing papers

No inbound Pith citation observations are available.