Pith. sign in

Paper Citation Record · LEDGER

Reward-Guided Speculative Decoding for Efficient LLM Reasoning

As of 10 August 2026, this Paper Citation Record lists 67 of 67 outbound references and 34 inbound Pith citation observations for arXiv:2501.19324.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.19324 v3

Coverage vector

measured 67 of 67 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T20:42:44.485144Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 34 of 34 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:23:03.610978Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

67 of 67 outbound references displayed

  • verified exact1
  • verified fuzzy6
  • unresolved60
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 7c57154e-b9fa-46a7-a2f3-1212ae786b4a · outbound

This paper cites write newline.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:43.964524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:43.964524Z digest=sha256:8b1683d1353e54a3601e037ae7d549ab521763bc3e155b4379cb305c1bad354b

Observation 21f442d9-1119-4632-925a-78a8109bf0f0 · outbound

This paper cites Claude 3.5 sonnet model card addendum.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Claude 3.5 sonnet model card addendum

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:43.975595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:43.975595Z digest=sha256:3dd28d06465aab260dd4b91f08cbab419ac60938731d7e7e955969e5a79d7dc6

Observation 1766472c-2f05-4475-b9ab-dc161d5463dc · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:43.982848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:43.982848Z digest=sha256:38e767f0b7ecc826355ed631a14f4ce035a093ce0f2fd8e1e3494c1b00239346

Observation 33331469-0c8b-40c8-ac85-f16ff7f5d3a3 · outbound

This paper cites D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:43.990044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:43.990044Z digest=sha256:2f644dd2733af1267ad18c38f74b5236fab689b05cdfe7e55e3dd70270daa9bd

Observation 20457001-abdd-4814-8ffd-43ed7ca435e4 · outbound

This paper cites Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.000347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.000347Z digest=sha256:a0b8bfe4eac4e5a1878ef2ef12a860b660983ee4f5f5106d5cf5ee9f0d338afb

Observation 6252f64e-0fd0-46a2-8f2e-812491ad2b79 · outbound

This paper cites Accelerating Large Language Model Decoding with Speculative Sampling.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Accelerating Large Language Model Decoding with Speculative Sampling

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.016298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.016298Z digest=sha256:1d746663da3f1b2fdb8ff562c31fc8f5a52ccab09ba37d7b53088edb7635fc56

Observation af617fc0-5954-4655-8222-a5b29688ce9e · outbound

This paper cites AlphaMath Almost Zero: Process Supervision without Process.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning AlphaMath Almost Zero: Process Supervision without Process

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.023438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.023438Z digest=sha256:945200a875eb0d6e7f032a1888da3a5e24724e466c09978deadfb76a5482c2ac

Observation 0bc4f55b-fcdc-4c32-a090-e565eb9668dd · outbound

This paper cites Cascade Speculative Drafting for Even Faster LLM Inference.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Cascade Speculative Drafting for Even Faster LLM Inference

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.031679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.031679Z digest=sha256:95ec78b61ac8ef00ea76586913e420871fbf918307fd2529fedf840079e1d7c7

Observation 44775d27-2243-4374-b8be-bba113f3069f · outbound

This paper cites Sequoia: Scalable, Robust, and Hardware-aware Speculative Decoding.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Sequoia: Scalable, Robust, and Hardware-aware Speculative Decoding

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.038874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.038874Z digest=sha256:d2372837c0d4afb4960ee29986254633a3ab723c52d8d9f67e901e1aaf138422

Observation 7f688bd6-eb15-4e4b-acf1-f7d027636715 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Training Verifiers to Solve Math Word Problems

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.053221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.053221Z digest=sha256:3ff7e415ef855dcacd9059d9bcb436ab59d7fd202bad21d2572812bfb374c6e3

Observation 70fe6d90-bff1-48da-b1fb-187b7491ba83 · outbound

This paper cites RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.060352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.060352Z digest=sha256:95e14046f0cdaa56d6e9736c09e0ba743798bc692bedba9430cf88e93d16ae5a

Observation d6a60e9f-6edc-4a6c-b1e6-0c27b324570b · outbound

This paper cites RLHF Workflow: From Reward Modeling to Online RLHF.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning RLHF Workflow: From Reward Modeling to Online RLHF

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.066469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.066469Z digest=sha256:76d38059077a5e5908f36ac0257ec4f1efe319c03e5aa10a27a020117fc41a6a

Observation d7451f83-a507-4cda-9735-8bda0461a28e · outbound

This paper cites an unresolved cited work.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.079172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.079172Z digest=sha256:c75a31759186ca6c15b398bc6f62c1dbfdfb4d00cc997fd34cdb507aa6504ac9

Observation c3e2f23e-be78-4702-b332-81a9b9099ce6 · outbound

This paper cites LayerSkip: Enabling Early Exit Inference and Self-Speculative Decoding.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning LayerSkip: Enabling Early Exit Inference and Self-Speculative Decoding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.084818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.084818Z digest=sha256:f118cd83fd26983855b9cd6b93e489d54cb01946ff03cd7ebda177f46a17e81f

Observation 16b3c45a-7ba2-4f75-938b-5ded5372334a · outbound

This paper cites GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.090452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.090452Z digest=sha256:7d635f65ad08348dceadfa8843e33dac80f724b9974cb877c8d8509644bfb24f

Observation 09b45938-ff0d-42b4-a847-3a7b3579bf19 · outbound

This paper cites Break the Sequential Dependency of LLM Inference Using Lookahead Decoding.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Break the Sequential Dependency of LLM Inference Using Lookahead Decoding

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.096964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.096964Z digest=sha256:867dcedc7264419a7d163b838ce56de4ee2d777357a6246c0453ed972a2bfba5

Observation ea697bf9-2400-4e9d-bf62-128b821e5170 · outbound

This paper cites Arcee's mergekit: A toolkit for merging large language models.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Arcee's mergekit: A toolkit for merging large language models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T20:42:46.128713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T20:42:44.103934Z digest=sha256:74f25d7b10ade24c2f4a19c73b8d09fe3502cbf0aa11dce88c974cdf1fbd5ceb

Observation 034a3c22-e44b-4f95-a3b3-ef365da94d8d · outbound

This paper cites The Llama 3 Herd of Models.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning The Llama 3 Herd of Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.110085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.110085Z digest=sha256:0cc5b4b8b0379cad434dc905681a6b350e266180f8771b0c9c98ec7d0c331868

Observation fca8ae2d-bc2d-493c-b18f-db9a31e63317 · outbound

This paper cites OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.116312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.116312Z digest=sha256:64c01581b38528d018fd83cd6c7949d2a26c882d79862c65c164984826b3ac7d

Observation c45b9cfd-433c-4866-87cf-19a8228acbff · outbound

This paper cites Measuring Massive Multitask Language Understanding.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Measuring Massive Multitask Language Understanding

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.122919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.122919Z digest=sha256:1fa5d9639408733180ec8fe659391911f88cc872ece93032c10d71cb54c55fe2

Observation 294021f5-4e61-4677-ab53-4f67cd2c0b42 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Measuring Mathematical Problem Solving With the MATH Dataset

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.130145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.130145Z digest=sha256:150ae20e0f4bbed557d24fdd05eec5de8f2e92d73ec9e3a7996ba7d0e3ab8dae

Observation bf924908-1438-465d-af60-d03b883e0f49 · outbound

This paper cites Deep Learning Scaling is Predictable, Empirically.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Deep Learning Scaling is Predictable, Empirically

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.136644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.136644Z digest=sha256:3a4828c0cd050e90de71c2f6fbe56ecfd9ad32a7f0a9ade911c6c8c2596d22f0

Observation 3a4ce8fd-c225-44ca-8ca6-a13cb908d185 · outbound

This paper cites Training Compute-Optimal Large Language Models.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Training Compute-Optimal Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.144341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.144341Z digest=sha256:8e40050ed73361d47889ed7ba0e66ac42151183cda7a994d24ef36d7a3b20f89

Observation 2a8bf702-4455-4fb0-83a3-6c2f5260842a · outbound

This paper cites The curious case of neural text degeneration.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning The curious case of neural text degeneration

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.150908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.150908Z digest=sha256:4fe9048eba34287a074fa5b9a8a5b587056bfb50e7297130f49d21e1e40165b6

Observation 9a9bf614-f949-4dcb-b71f-8a2c2b47877b · outbound

This paper cites GPT-4o System Card.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning GPT-4o System Card

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.156789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.156789Z digest=sha256:9e1a2eaff3d7d8ca22992616bf61f7c2b5007cee8df69670e4ebfd80cda7a67a

Observation ba44915f-09b0-4659-b8d0-c9f2045e1e4e · outbound

This paper cites MindStar: Enhancing Math Reasoning in Pre-trained LLMs at Inference Time.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning MindStar: Enhancing Math Reasoning in Pre-trained LLMs at Inference Time

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.162517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.162517Z digest=sha256:cd7894d228d42f915e5cb8c79cb3b4b8d3a833457aa461386cd094663e674548

Observation d5e6f325-4651-43d4-bb87-cb9cbfb2c9ee · outbound

This paper cites Scaling Laws for Neural Language Models.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Scaling Laws for Neural Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.168357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.168357Z digest=sha256:cfc263d39d287f580015e5049f84954c0d0d147eedf93dbb9fd6f43da3ff2c9d

Observation 4b10a533-dc69-42eb-a020-9ce9b3988a8d · outbound

This paper cites H., Gonzalez, J., Zhang, H., and Stoica, I.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning H., Gonzalez, J., Zhang, H., and Stoica, I

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.173811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.173811Z digest=sha256:c00b526611fa3389b9737ffb31e1e89f880e8eb40a17c9d3c6ce29fe6122aa3e

Observation 0f8361c1-3841-490f-88a5-23a22c4bcae9 · outbound

This paper cites Fast inference from transformers via speculative decoding.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Fast inference from transformers via speculative decoding

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.179626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.179626Z digest=sha256:56bfe312398e02d62723faef05e0786f139e61c747d218d501e67ab4f6e3ace8

Observation 7890918e-caaf-4cf6-b430-bca9f6162af8 · outbound

This paper cites V., Slone, A., Anil, C., Schlag, I., Gutman - Solo, T., Wu, Y., Neyshabur, B., Gur - Ari, G., and Misra, V.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning V., Slone, A., Anil, C., Schlag, I., Gutman - Solo, T., Wu, Y., Neyshabur, B., Gur - Ari, G., and Misra, V

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T20:42:46.068396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T20:42:44.188322Z digest=sha256:67dd6e436b807da0bc33e58db09f374cc948dbf6184c065b3383bd39dc64513c

Observation aaae243e-3fa8-489b-b733-e5e6419876bf · outbound

This paper cites an unresolved cited work.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.199802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.199802Z digest=sha256:2820d50791bf3cdb7e9475a0aa3e846a0dcc309ef23e13f5e5333c7aae0dead1

Observation 37154e07-8711-46fd-8ce9-3a7606eed1a6 · outbound

This paper cites SnapKV: LLM Knows What You are Looking for Before Generation.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning SnapKV: LLM Knows What You are Looking for Before Generation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.207816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.207816Z digest=sha256:89a17ab20de62851fb9de743e386f8c766b0f9eb5e9e2b21c43257e8f412982c

Observation 3ba03539-c786-4ab2-86b7-3e11edf2d45b · outbound

This paper cites EAGLE: Speculative Sampling Requires Rethinking Feature Uncertainty.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning EAGLE: Speculative Sampling Requires Rethinking Feature Uncertainty

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.215750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.215750Z digest=sha256:572c46033fe27ecfe548b77062acd6c25e1a4bf867fdfeab92971b6d09fdf3d7

Observation a6c5d164-233c-4de2-b146-8b1f9a3bf2ed · outbound

This paper cites 3-in-1: 2D Rotary Adaptation for Efficient Finetuning, Efficient Batching and Composability.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning 3-in-1: 2D Rotary Adaptation for Efficient Finetuning, Efficient Batching and Composability

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-09T20:42:44.570676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T20:42:44.222281Z digest=sha256:6fdf933f303185aec0d3a21f1f5e30f362d35c3e58552beb3b816f7785dea091

Observation c88f4ff8-826a-40e8-9d58-2afecb657ab0 · outbound

This paper cites MARIO: MAth Reasoning with code Interpreter Output -- A Reproducible Pipeline.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning MARIO: MAth Reasoning with code Interpreter Output -- A Reproducible Pipeline

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.228570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.228570Z digest=sha256:3e8b4850676fde76457b88a7c619ad198a11728da055a32c503db5043aa7875d

Observation 5272d4cf-a94a-4bc4-9fac-c9b81664380d · outbound

This paper cites Let's Verify Step by Step.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Let's Verify Step by Step

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.234941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.234941Z digest=sha256:97a3764b1884c19e1a4cd3fa4b72cef9ec38095737573d25a5be48eb5d200fb6

Observation 806e61ad-878d-4c21-8a33-b118f93b95e9 · outbound

This paper cites Awq: Activation-aware weight quantization for on-device llm compression and acceleration.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Awq: Activation-aware weight quantization for on-device llm compression and acceleration

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.241668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.241668Z digest=sha256:e6d1b5371fde6ace2a8030263df44fbaab763981fe000042525f6b06e79132ba

Observation 259d1ad3-0597-4ee6-a5ff-9af1e849bba6 · outbound

This paper cites Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.251912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.251912Z digest=sha256:10fe218f8e413884955d5d6a0025dfdec50b5cfeba64586fa8591687f6e839cd

Observation d7665727-c85a-4025-b2ee-498f7361fd6d · outbound

This paper cites an unresolved cited work.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.258159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.258159Z digest=sha256:22dc9265802337a7c5dbd864cf45be2e8db458496218ee47664aaab4fc6c1cee

Observation 8d04ec9e-89c5-4e7e-900d-537f67887d17 · outbound

This paper cites Skywork-o1 open series.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Skywork-o1 open series

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.263839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.263839Z digest=sha256:7259d1fcb88d9d16598fcbee5570ab0adae613da544964be73ca5312ddd52c8c

Observation 1801d9b3-fd31-4a80-b40c-201c127abba8 · outbound

This paper cites Learning to reason with llms, 2024.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Learning to reason with llms, 2024

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.272726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.272726Z digest=sha256:711ee678c29ab5ac07806526d4105784bd22d413657c63cffdd902ef1d1402f1

Observation 9d01e379-b894-410f-92a7-9a121291dfa8 · outbound

This paper cites Carbon Emissions and Large Neural Network Training.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Carbon Emissions and Large Neural Network Training

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.280816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.280816Z digest=sha256:e9a7a2de64b60bd74feaa75811dabdfa8b2e594b1f0d631b37f038cd71f00c91

Observation 7af0daca-458b-4d09-bdab-3bd5a8633d16 · outbound

This paper cites Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.286290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.286290Z digest=sha256:76f35e2a73e975e95978b54ccee57cd3fd6fe1adff9f2829bc26948b9972dda8

Observation 1882eb5d-2c89-42a5-91e4-91d8c98e94e8 · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.292259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.292259Z digest=sha256:9408460a89864e381b2d19b5aa540915ebedfc340c6058b3f1418d6ae72d2b18

Observation 9790a47a-a798-41f8-9841-2c0039784b25 · outbound

This paper cites Beyond Chinchilla-Optimal: Accounting for Inference in Language Model Scaling Laws.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Beyond Chinchilla-Optimal: Accounting for Inference in Language Model Scaling Laws

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.299119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.299119Z digest=sha256:0009d7f7304d71598497a406a6ddc9c11defd891278c703e6bcb2cf78a9d6853

Observation 460091e8-bb92-4268-92d0-d340a5a8ba69 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.308717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.308717Z digest=sha256:eb4f72ce2a9126cebb63a426898d73a85b0889d1643cf8ffacd2cd64ab228c61

Observation 1ffd8a92-6085-4617-8bca-d5105949050a · outbound

This paper cites Blockwise parallel decoding for deep autoregressive models.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Blockwise parallel decoding for deep autoregressive models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.316584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.316584Z digest=sha256:2204d765e107c759ed16dfb85e9310c7c67205911bc2e655d64192f87032bce7

Observation ca42c7bf-7a16-4c9a-9880-30bb22f15db9 · outbound

This paper cites TriForce: Lossless Acceleration of Long Sequence Generation with Hierarchical Speculative Decoding.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning TriForce: Lossless Acceleration of Long Sequence Generation with Hierarchical Speculative Decoding

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.326192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.326192Z digest=sha256:19c7bc5b074eb31158a745650045bda92a48ff2ae4ff67cc093e7e874bd07a85

Observation c9f81045-d925-4f46-8b77-0632616fb7b2 · outbound

This paper cites A Simple and Effective Pruning Approach for Large Language Models.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning A Simple and Effective Pruning Approach for Large Language Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.340013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.340013Z digest=sha256:1b434fc2c531af9a9f5e86a3427a421d3adb3aa692e2bd1ba595f3a8ac4c83b9

Observation 949f76bd-36b3-4259-aed6-51136094ed54 · outbound

This paper cites T., Ro, J.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning T., Ro, J

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T20:42:45.957902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T20:42:44.346930Z digest=sha256:ebc1a9790f8d2cd6432962b179f7093d65b93f814e1b957632d2b2d664844c7f

Observation f9e9c874-d97f-4d90-8736-71cf62e774c3 · outbound

This paper cites Mathscale: Scaling instruction tuning for mathematical reasoning.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Mathscale: Scaling instruction tuning for mathematical reasoning

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T20:42:45.932348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T20:42:44.353442Z digest=sha256:e7f89ae4857d3165e1c8f395ffddab50a3720ca88d6a2033f40ad8fda5c63a46

Observation e5b55482-202e-445e-9856-66c7ba8342c7 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.368701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.368701Z digest=sha256:269a16e784f09f76009527649b8144e811b256bdbeb10f85116c705024bcc19a

Observation 9bdd4c00-a5d2-4cf6-bcae-7eca24c93b5a · outbound

This paper cites N., Kaiser, L., and Polosukhin, I.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning N., Kaiser, L., and Polosukhin, I

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T20:42:45.912643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T20:42:44.374476Z digest=sha256:89cb83c7e7878a51a6a63800a61fce1017a20746649e11a4d35be70d1722f7a4

Observation 627a8fc0-d445-45d1-bb99-665028e04d2d · outbound

This paper cites Math-shepherd: Verify and reinforce llms step-by-step without human annotations.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Math-shepherd: Verify and reinforce llms step-by-step without human annotations

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T20:42:45.894536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T20:42:44.385548Z digest=sha256:e9e837282bd035c34cdb69046600c34e2e5d10dca66534a9d0833ed2ab60dcf5

Observation d4d370f4-c46f-424b-98f5-db247d902e32 · outbound

This paper cites Unlocking Efficiency in Large Language Model Inference: A Comprehensive Survey of Speculative Decoding.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Unlocking Efficiency in Large Language Model Inference: A Comprehensive Survey of Speculative Decoding

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.390745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.390745Z digest=sha256:40b80438ed94fc2b46aa1b67a10a5e13a4bbf61ee1f8a3799ddd16a00f0bd021

Observation 57a1196a-9508-4ac8-8f35-504f2110c370 · outbound

This paper cites QA-LoRA: Quantization-Aware Low-Rank Adaptation of Large Language Models.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning QA-LoRA: Quantization-Aware Low-Rank Adaptation of Large Language Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.396702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.396702Z digest=sha256:a98898010225bf22ba82a4ac6ce87388db43318654bd099c050f0e9f1d45a932

Observation 259cc8dc-bee7-487a-8a9a-8d6126e9d80b · outbound

This paper cites ThinK: Thinner Key Cache by Query-Driven Pruning.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning ThinK: Thinner Key Cache by Query-Driven Pruning

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.405709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.405709Z digest=sha256:24bcf928757f4607aa3ebd8ade077ef0aa4b3f80e0096438f0ee748af3a65fb3

Observation 2bb10878-8b67-4146-9962-38cf55ce3c7d · outbound

This paper cites What Matters for Model Merging at Scale?.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning What Matters for Model Merging at Scale?

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.413356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.413356Z digest=sha256:10ed0ac65cd3fd9fd67306a05cd66c337180f3124e968c5712d9bfbff5092bd1

Observation 3893f11d-fad5-4fe8-a75e-64f520860562 · outbound

This paper cites Qwen2.5 Technical Report.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Qwen2.5 Technical Report

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.425099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.425099Z digest=sha256:c9adf98a4e1c391f016432ec883d74421395962fd72c7ea841fc345cbcfdf434

Observation 3a701c5d-f690-42eb-b5ce-eb0c4b4e7d75 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.433332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.433332Z digest=sha256:9914648faca593a4ee210ef284cfea06bcea599466835610cfd6414aeda69cda

Observation ac747d20-58d0-4e16-9744-e5bf63d89c1d · outbound

This paper cites Tree of thoughts: Deliberate problem solving with large language models.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Tree of thoughts: Deliberate problem solving with large language models

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.442735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.442735Z digest=sha256:ae0ec9d09e9a7727de5ac4fcaecd194c65c76f6107f1377ac095b2d713488b72

Observation deb6154c-f8e9-43c0-b3a3-4d8d2e812720 · outbound

This paper cites OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.448540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.448540Z digest=sha256:59cdafd204f3af903d853943dbe84254db232a5faf4f6b2dd43e2507b9a74c81

Observation c801eed3-7e36-479b-94ce-f91b886d7a0c · outbound

This paper cites Draft & Verify: Lossless Large Language Model Acceleration via Self-Speculative Decoding.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning Draft & Verify: Lossless Large Language Model Acceleration via Self-Speculative Decoding

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.456802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.456802Z digest=sha256:3a910b238523c4b3c092821a36c440952203d90b67cbbe4b174b9b801d1ea907

Observation 61de3e2e-1767-48a4-a500-7a1fae3b4be5 · outbound

This paper cites H2o: Heavy-hitter oracle for efficient generative inference of large language models.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning H2o: Heavy-hitter oracle for efficient generative inference of large language models

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.463803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.463803Z digest=sha256:983fea7bbaf09fb03480575e63345f563d3fc9a804611ef5e7d1affb054706eb

Observation e32c95fe-10d1-4937-a910-9a0b37dbe291 · outbound

This paper cites The Lessons of Developing Process Reward Models in Mathematical Reasoning.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning The Lessons of Developing Process Reward Models in Mathematical Reasoning

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.472297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.472297Z digest=sha256:60e4b46939743f36de0c4715cf0d7c67664f2d1314e192745f4a7d9b4831fb15

Observation 60007197-a309-43ee-b7cf-157f0fdc42e9 · outbound

This paper cites ProcessBench: Identifying Process Errors in Mathematical Reasoning.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning ProcessBench: Identifying Process Errors in Mathematical Reasoning

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.478833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.478833Z digest=sha256:a6f20ee185415106168013d9e56d9e82857a4b3d8e3ae005f3a84fba4cdabbbe

Observation 97823314-e74b-498a-8f40-6686fe05f016 · outbound

This paper cites P., Wang, K., Chang, J., Gao, Z., Kallus, N., Weinberger, K.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning P., Wang, K., Chang, J., Gao, Z., Kallus, N., Weinberger, K

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.485144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.485144Z digest=sha256:682f1aa11296de82c0050cf85ec81396a82414d131120849db40c8d0a461f7e7

Pith citing papers

Observation bcec02f8-e6ff-473a-8f3b-0cbd1a804409 · inbound

Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models cites this paper.

Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-05-14T01:29:57.246050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-14T01:29:56.480020Z digest=sha256:e4799b8d798f4c9b920eb80427b386c9bf66ac3b91f71b905097d0541d8be004

Observation 611322b8-23fc-4a20-81c5-a72412382bc8 · inbound

When to Continue Thinking: Adaptive Thinking Mode Switching for Efficient Reasoning cites this paper.

When to Continue Thinking: Adaptive Thinking Mode Switching for Efficient Reasoning Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:23:03.610978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:23:03.610978Z digest=sha256:d0e26f9d5ecc8c1b6e31c33f6d9028047673c0e0f082c2105f5070a3e016f521

Observation c7ed6251-5e1d-4de9-a306-712159e8ad51 · inbound

Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs cites this paper.

Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:35:13.236480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T05:35:13.118221Z digest=sha256:f4de5c71da970e1bc065670ee6bd36b13f5d00d9bfa3e2a0f7a35fddde336170

Observation 32740862-8c87-4e4b-861e-6b4eee0e0851 · inbound

VeriThinker: Learning to Verify Makes Reasoning Model Efficient cites this paper.

VeriThinker: Learning to Verify Makes Reasoning Model Efficient Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:42:12.665171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:42:12.665171Z digest=sha256:3745e01f8eab89507be9a51f826685bd8d71cbb4e21e594ba5b77dd5622d3ac3

Observation 038d5e65-a888-4af9-9498-e1d5a6d4bae4 · inbound

Think Before You Accept: Semantic Reflective Verification for Faster Speculative Decoding cites this paper.

Think Before You Accept: Semantic Reflective Verification for Faster Speculative Decoding Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:32:41.836423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:32:41.836423Z digest=sha256:8f6cf3f6a9a88246ca37a8e8a417e6ad33ad9e14839a968cab251ee212c71c4d

Observation c0ce5cfc-5745-4f52-bb92-fee9490dc8af · inbound

What makes Reasoning Models Different? Follow the Reasoning Leader for Efficient Decoding cites this paper.

What makes Reasoning Models Different? Follow the Reasoning Leader for Efficient Decoding Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:52:18.444052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:52:18.444052Z digest=sha256:15b59324d082fd0e1165f739c53f904a51a569aae09fa931cebf9644d31d2a6d

Observation 1a1662d7-9fd7-4836-a74c-7d216dbed935 · inbound

How Far Are We from Optimal Reasoning Efficiency? cites this paper.

How Far Are We from Optimal Reasoning Efficiency? Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:49:36.717578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:49:36.717578Z digest=sha256:980d7d9c07f8add208728dece8645196a6cbbb6223fb4b2839f9fcdf69ad4c85

Observation c7301252-0f20-4fba-a57c-29f0ce93276f · inbound

Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency cites this paper.

Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:16.691688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:16.691688Z digest=sha256:d613d8ac4315e6a582eb86f7c906a3b72f09093443da011ae9634c2f354a1991

Observation 388b0d75-619c-4211-984d-b25b88b294e2 · inbound

PREMISE: Scalable and Strategic Prompt Optimization for Efficient Mathematical Reasoning in Large Models cites this paper.

PREMISE: Scalable and Strategic Prompt Optimization for Efficient Mathematical Reasoning in Large Models Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T04:26:29.108342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:26:29.108342Z digest=sha256:0caba441770c332df20aff0021f707848f6410238d3075f679058449963dfd83

Observation 35cbcd96-7426-4654-aa51-2ea55b5f6645 · inbound

Efficient Reasoning Through Suppression of Self-Affirmation Reflections in Large Reasoning Models cites this paper.

Efficient Reasoning Through Suppression of Self-Affirmation Reflections in Large Reasoning Models Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T00:57:01.524906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:57:01.524906Z digest=sha256:b6dfa5c1a09c2d04e055a8c827f2b47eea7a8cf372bf95080d695b1dcf274f9f

Observation 87c8092e-3e82-4317-b042-40fe4076d637 · inbound

AALC: Large Language Model Efficient Reasoning via Adaptive Accuracy-Length Control cites this paper.

AALC: Large Language Model Efficient Reasoning via Adaptive Accuracy-Length Control Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T23:00:45.946584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:00:45.946584Z digest=sha256:b8a91a47cd81a6a408d09158363c549fc315d330d86004c2013cd0b4aa665069

Observation ed7eaab6-c0a9-410b-84ce-95380584986d · inbound

CoRE: Enhancing Metacognition with Label-free Self-evaluation in LRMs cites this paper.

CoRE: Enhancing Metacognition with Label-free Self-evaluation in LRMs Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T19:16:31.326240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:16:31.326240Z digest=sha256:449b8ef907d6ade50b969393ed130b0fbcc8a7eda2a40c613464dc5def0196cb

Observation c7e65240-89f6-420b-8825-cc760c589dcf · inbound

Logit Arithmetic Elicits Long Reasoning Capabilities Without Training cites this paper.

Logit Arithmetic Elicits Long Reasoning Capabilities Without Training Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T16:45:58.173827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:45:58.173827Z digest=sha256:c3c720f017a7ce039d108dd2d94649932345ca5dd2437f3590b1bd0ca0c69d93

Observation bbf519a6-9680-4607-8070-a3251649d9d9 · inbound

Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities cites this paper.

Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T16:34:25.083207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:34:25.083207Z digest=sha256:5eb09a964ff561f96c496462762b23dc6becaa6f250af69e51d3351317d384da

Observation b879e10c-fcb3-4518-83f2-40c8eacad096 · inbound

Collaborative Inference and Learning between Edge SLMs and Cloud LLMs: A Survey of Algorithms, Execution, and Open Challenges cites this paper.

Collaborative Inference and Learning between Edge SLMs and Cloud LLMs: A Survey of Algorithms, Execution, and Open Challenges Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T15:06:47.952749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:06:47.952749Z digest=sha256:1b23b9f9d4322e7db30da16da3f225221e0048cfe5304b6bddedaa5a6af756b8

Observation ea256494-2030-4bbb-ba13-a7d696c27044 · inbound

From Long to Short: LLMs Excel at Trimming Own Reasoning Chains cites this paper.

From Long to Short: LLMs Excel at Trimming Own Reasoning Chains Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T00:04:00.084214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:04:00.084214Z digest=sha256:235cacc3dcb8fd78ec7ab3da5eae7b2956939b6bd2dfbec159637708dd529bd9

Observation f0e449cb-e897-4b6a-9b53-f4b6d1e66884 · inbound

Rethinking Visual Autoregressive Sampling with Information-Grounding Guidance cites this paper.

Rethinking Visual Autoregressive Sampling with Information-Grounding Guidance Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T14:43:13.834755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:43:13.834755Z digest=sha256:47f8539ef580058ac9fe57bc2cfbaea5487a0c9865a1b79fc71805717544bf22

Observation b28f02d8-d9ba-475a-8e33-0fd3b860cb39 · inbound

Simultaneous Multi-objective Alignment Across Verifiable and Non-verifiable Rewards cites this paper.

Simultaneous Multi-objective Alignment Across Verifiable and Non-verifiable Rewards Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T13:15:44.660507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T13:15:44.660507Z digest=sha256:289f5ccd338989d9855d4be7538f5ecb6597e34c35fca9dec2b4c8da0db757b5

Observation 6fcb612f-1c20-47a5-96de-8b4e4f5eda65 · inbound

MixReasoning: Switching Modes to Think cites this paper.

MixReasoning: Switching Modes to Think Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T11:16:36.269089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:16:36.269089Z digest=sha256:d5cd55e8f4b40033957e7efdd6d46fcde77ffaa86c15f90303d9f4d289a797de

Observation 8a082f96-ef59-4865-b01f-3914f9c8a059 · inbound

Policy-Guided Stepwise Model Routing for Cost-Effective Reasoning cites this paper.

Policy-Guided Stepwise Model Routing for Cost-Effective Reasoning Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:01:11.844961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T10:29:23.538806Z digest=sha256:08fb56ffcaf80c4f071fd60391e1ba58c618df3197e207e16972cb077b2cb686

Observation beb4d729-6445-4488-a429-aa9471f96554 · inbound

Post Reasoning: Improving the Performance of Non-Thinking Models at No Cost cites this paper.

Post Reasoning: Improving the Performance of Non-Thinking Models at No Cost Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 115

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:06:09.780452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-08T10:19:08.451445Z digest=sha256:310ad397c8655b61a6d856f1d899d529634e1f29fd8bff86ec8e09d7d71b5673

Observation b1bc2428-7e1f-4d2d-a8f2-886b8997ceac · inbound

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration cites this paper.

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:51:30.110828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T03:51:52.375703Z digest=sha256:7d100f1b45f0f36880c1e4a51e0351f2d4ef8cec78dfbefbede029049c15c8fe

Observation b9aa2bb9-de64-4d15-9a03-810c6695c4bd · inbound

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration cites this paper.

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:15:03.414244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T05:11:32.053440Z digest=sha256:13f2548481a6ae95fd30a3743bad1bd7998510ba38525a4fde2b7fdf5b052cd3

Observation 6b4b70f5-f254-4976-8fd3-020ac22a6c20 · inbound

Pause and Reflect: Conformal Aggregation for Chain-of-Thought Reasoning cites this paper.

Pause and Reflect: Conformal Aggregation for Chain-of-Thought Reasoning Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:23:31.576800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-15T02:23:22.298606Z digest=sha256:cd602db83f717f2db7d7dd9c43a1c495f206940e05b8c108a5910df377c93f2f

Observation 6528dc9d-a31c-44d6-937c-719ad32bb994 · inbound

Performance-Driven Policy Optimization for Speculative Decoding with Adaptive Windowing cites this paper.

Performance-Driven Policy Optimization for Speculative Decoding with Adaptive Windowing Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-19T16:42:39.750479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T16:40:17.370100Z digest=sha256:69a71867bcd41ea0882238b904ef13ab548d36f49e66d52f9523f92618483b2f

Observation 8c66803d-ef90-423c-94bd-8205ca26b35d · inbound

ThoughtFold: Folding Reasoning Chains via Introspective Preference Learning cites this paper.

ThoughtFold: Folding Reasoning Chains via Introspective Preference Learning Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:26:28.717883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T10:07:16.700499Z digest=sha256:982fbf458d9c6599262e0d8bd8703136c191f5a51a3f9fe2f14de09aa4e2fe8a

Observation 0486a096-6862-4eef-87df-9433fe9830e1 · inbound

KCSAT-ML: Probing Reasoning Models with Nationwide-Cohort Human Difficulty cites this paper.

KCSAT-ML: Probing Reasoning Models with Nationwide-Cohort Human Difficulty Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-06-27T13:10:55.947824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T13:07:59.509660Z digest=sha256:5ed2674d4a75841aaa07489eb8b869bfaec537185f9129871897e91152b10ddb

Observation d2fb637d-0b15-42bd-9b2f-e1d868ee854d · inbound

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes cites this paper.

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 145

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:57:41.512993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T12:59:51.091008Z digest=sha256:a37c4e54f8ba554520745bbdaa4d3b02520dad1b17338f195d673fa8558c49ad

Observation c7a64538-8131-4f38-9c2b-110847bc4108 · inbound

Efficient and Trainable Language Model Test-Time Scaling via Local Branch Routing cites this paper.

Efficient and Trainable Language Model Test-Time Scaling via Local Branch Routing Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:20:06.711514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-25T21:29:40.564852Z digest=sha256:f46883a1d1237e8d751adf4579deee84dccd5c7871497032ffa410638862ea9f

Observation aac4b144-e9fe-4ae3-a728-0a9b1dbe4f9d · inbound

Efficient and Trainable Language Model Test-Time Scaling via Local Branch Routing cites this paper.

Efficient and Trainable Language Model Test-Time Scaling via Local Branch Routing Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-01T06:35:29.567228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-01T06:33:45.824727Z digest=sha256:131bb31a65c2a13ebe72f24b4d76532f86988e34f4df46659738e8875d24bb3d

Observation 928758a2-7325-4ee5-a6b3-a370bab1e0a7 · inbound

Before Thinking, Learn to Decide: Proactive Routing for Efficient Visual Reasoning cites this paper.

Before Thinking, Learn to Decide: Proactive Routing for Efficient Visual Reasoning Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T13:34:41.134711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T05:55:19.083517Z digest=sha256:3ee06e48694d33ff9d6164db72614a1d64c1552f8625b331d4de774c6921d39d

Observation 84169b7c-5669-40d0-9f03-9c4d03678213 · inbound

Multi-Head Latent Control: A Unified Interface for LLM Agent Decision Making cites this paper.

Multi-Head Latent Control: A Unified Interface for LLM Agent Decision Making Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T02:42:03.779702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T02:42:03.779702Z digest=sha256:b69315fb320c62d2a601812f77321a07948a2aca906615cea4253bda0da38e32

Observation 8ba2f2ae-deba-4dbb-a4f7-8e2ef50281cd · inbound

When to Plan: Learning to Select Between Reactive Control and Deliberative Planning cites this paper.

When to Plan: Learning to Select Between Reactive Control and Deliberative Planning Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T21:04:09.724503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:04:09.724503Z digest=sha256:98ab7934ea3a3560cc8eddbabc7efd10a03addd862573a5f2727effc65f45393

Observation 14bc6ab4-22d8-4176-9126-858b46c5137c · inbound

Is Your Model Thinking or Just Stagnating? PUMA: Diagnosing Reasoning Pathology via Phase-Momentum Alignment cites this paper.

Is Your Model Thinking or Just Stagnating? PUMA: Diagnosing Reasoning Pathology via Phase-Momentum Alignment Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T18:49:28.872386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:49:28.872386Z digest=sha256:2b4496a1735ac20513b047973c8e1fba15602cdb52c889b902d1e06dea4a36e7