Pith. sign in

Paper Citation Record · LEDGER

Reward-Guided Speculative Decoding for Efficient LLM Reasoning

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 31 inbound Pith citation observations for arXiv:2501.19324.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.19324 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 31 of 31 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:52:18.444052Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bcec02f8-e6ff-473a-8f3b-0cbd1a804409 · inbound

Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models cites this paper.

Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-05-14T01:29:57.246050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T01:29:56.480020Z digest=sha256:0a01e603628ea89a2f5f94cb4af8405235ee9d0e628a28981c45e9fa7b0636fc

Observation c7ed6251-5e1d-4de9-a306-712159e8ad51 · inbound

Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs cites this paper.

Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:35:13.236480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T05:35:13.118221Z digest=sha256:b794e76ee82c5c96f0a09b9857ed80c28fdcdd66fb8ebb4fae0331c637d3c046

Observation c0ce5cfc-5745-4f52-bb92-fee9490dc8af · inbound

What makes Reasoning Models Different? Follow the Reasoning Leader for Efficient Decoding cites this paper.

What makes Reasoning Models Different? Follow the Reasoning Leader for Efficient Decoding Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:52:18.444052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:52:18.444052Z digest=sha256:f93b2bade212b01b8a773100ef442b8e8294c5e516bc3d86ff7871e4c1ca3882

Observation 1a1662d7-9fd7-4836-a74c-7d216dbed935 · inbound

How Far Are We from Optimal Reasoning Efficiency? cites this paper.

How Far Are We from Optimal Reasoning Efficiency? Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:49:36.717578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:49:36.717578Z digest=sha256:bc8ee77c646889de8e04730581b50150bc7b894cab37d9d97199d08fa6f9c443

Observation c7301252-0f20-4fba-a57c-29f0ce93276f · inbound

Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency cites this paper.

Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:16.691688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:16.691688Z digest=sha256:3dc31893f1c33fdb9940c9f98bb57221cf03cee9f8454a0029c522ae98b13575

Observation 388b0d75-619c-4211-984d-b25b88b294e2 · inbound

PREMISE: Scalable and Strategic Prompt Optimization for Efficient Mathematical Reasoning in Large Models cites this paper.

PREMISE: Scalable and Strategic Prompt Optimization for Efficient Mathematical Reasoning in Large Models Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T04:26:29.108342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:26:29.108342Z digest=sha256:cc4f65049696c44c481092d415d63d4f2049513fc966825d2c968a45370ecdb6

Observation 35cbcd96-7426-4654-aa51-2ea55b5f6645 · inbound

Efficient Reasoning Through Suppression of Self-Affirmation Reflections in Large Reasoning Models cites this paper.

Efficient Reasoning Through Suppression of Self-Affirmation Reflections in Large Reasoning Models Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T00:57:01.524906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:57:01.524906Z digest=sha256:4f0b1842cc0b8d71ca4c21bf4d6040a6ca5c1b5b69f576afc1f070eb0d379b35

Observation 87c8092e-3e82-4317-b042-40fe4076d637 · inbound

AALC: Large Language Model Efficient Reasoning via Adaptive Accuracy-Length Control cites this paper.

AALC: Large Language Model Efficient Reasoning via Adaptive Accuracy-Length Control Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T23:00:45.946584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:00:45.946584Z digest=sha256:6de6fde35b03f7a5bea651d345a8396c3ad1dfa2e83cbec5f2cd25d1ffede198

Observation ed7eaab6-c0a9-410b-84ce-95380584986d · inbound

CoRE: Enhancing Metacognition with Label-free Self-evaluation in LRMs cites this paper.

CoRE: Enhancing Metacognition with Label-free Self-evaluation in LRMs Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T19:16:31.326240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:16:31.326240Z digest=sha256:221e3ba8f1f48964c5d46f99bdbc8b0fab254c6cdc60bbb6e1a7c1c7db3b7e62

Observation c7e65240-89f6-420b-8825-cc760c589dcf · inbound

Logit Arithmetic Elicits Long Reasoning Capabilities Without Training cites this paper.

Logit Arithmetic Elicits Long Reasoning Capabilities Without Training Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T16:45:58.173827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:45:58.173827Z digest=sha256:827a1b79a82fa26f0868bfe9822fb3695644368ba7cad17b0daa4d6be9ec238b

Observation bbf519a6-9680-4607-8070-a3251649d9d9 · inbound

Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities cites this paper.

Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T16:34:25.083207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:34:25.083207Z digest=sha256:ff3c96ec5234a285930222fca1b9a6db59a6e3ec5d6a6a0db1b24e6eff5e21b6

Observation b879e10c-fcb3-4518-83f2-40c8eacad096 · inbound

Collaborative Inference and Learning between Edge SLMs and Cloud LLMs: A Survey of Algorithms, Execution, and Open Challenges cites this paper.

Collaborative Inference and Learning between Edge SLMs and Cloud LLMs: A Survey of Algorithms, Execution, and Open Challenges Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T15:06:47.952749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:06:47.952749Z digest=sha256:65334db12517de69f151b2d9fb40d44c5a73ca8e7dab3b8d8f96be01cecbc383

Observation ea256494-2030-4bbb-ba13-a7d696c27044 · inbound

From Long to Short: LLMs Excel at Trimming Own Reasoning Chains cites this paper.

From Long to Short: LLMs Excel at Trimming Own Reasoning Chains Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T00:04:00.084214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:04:00.084214Z digest=sha256:050e62e352fa63d72a328c309aa7cf58f257b151649b1bd904b92758f364eba5

Observation f0e449cb-e897-4b6a-9b53-f4b6d1e66884 · inbound

Rethinking Visual Autoregressive Sampling with Information-Grounding Guidance cites this paper.

Rethinking Visual Autoregressive Sampling with Information-Grounding Guidance Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T14:43:13.834755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:43:13.834755Z digest=sha256:e5903575a8116ee2219cee3139d66d5e8ed2b110ed2e43bc270d54c8d9c79828

Observation b28f02d8-d9ba-475a-8e33-0fd3b860cb39 · inbound

Simultaneous Multi-objective Alignment Across Verifiable and Non-verifiable Rewards cites this paper.

Simultaneous Multi-objective Alignment Across Verifiable and Non-verifiable Rewards Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T13:15:44.660507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T13:15:44.660507Z digest=sha256:ddd6ff9ddc2b6c234d7175ac55680a0b625ed83766a2d633e636ef528ae5f995

Observation 6fcb612f-1c20-47a5-96de-8b4e4f5eda65 · inbound

MixReasoning: Switching Modes to Think cites this paper.

MixReasoning: Switching Modes to Think Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T11:16:36.269089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:16:36.269089Z digest=sha256:5c9f6c079e2256fc273c1417ea8b04429aa2e47780883ada44bfa5f62b4977da

Observation 8a082f96-ef59-4865-b01f-3914f9c8a059 · inbound

Policy-Guided Stepwise Model Routing for Cost-Effective Reasoning cites this paper.

Policy-Guided Stepwise Model Routing for Cost-Effective Reasoning Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:01:11.844961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T10:29:23.538806Z digest=sha256:1d3d9125119f1e1f1f0926b712c6ae5bdc45bc01e52003ce9430c8173407e9ff

Observation beb4d729-6445-4488-a429-aa9471f96554 · inbound

Post Reasoning: Improving the Performance of Non-Thinking Models at No Cost cites this paper.

Post Reasoning: Improving the Performance of Non-Thinking Models at No Cost Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 115

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:06:09.780452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-08T10:19:08.451445Z digest=sha256:f9c8f98039f3ccb6882d3196133225b073606171da94e836fe4d493470b2b96d

Observation b1bc2428-7e1f-4d2d-a8f2-886b8997ceac · inbound

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration cites this paper.

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:51:30.110828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T03:51:52.375703Z digest=sha256:fca219f87fff0745315b4cdf48c333371c87b1ca0bc126d06a2fc32e209712a1

Observation b9aa2bb9-de64-4d15-9a03-810c6695c4bd · inbound

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration cites this paper.

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:15:03.414244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T05:11:32.053440Z digest=sha256:fde56c44caa4cd92d230ba16257e27b90c357b27ec21ac7edbfd81bc99570ad3

Observation 6b4b70f5-f254-4976-8fd3-020ac22a6c20 · inbound

Pause and Reflect: Conformal Aggregation for Chain-of-Thought Reasoning cites this paper.

Pause and Reflect: Conformal Aggregation for Chain-of-Thought Reasoning Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:23:31.576800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-15T02:23:22.298606Z digest=sha256:cb6d5fe8b857408ecc8f005662302c7cd8b5adeb8db17543df4879183d0a939d

Observation 6528dc9d-a31c-44d6-937c-719ad32bb994 · inbound

Performance-Driven Policy Optimization for Speculative Decoding with Adaptive Windowing cites this paper.

Performance-Driven Policy Optimization for Speculative Decoding with Adaptive Windowing Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-19T16:42:39.750479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T16:40:17.370100Z digest=sha256:cf0c28b95f71dcc5c0c81d795f45906e30f8a97ba9104a101fee56a360f2096e

Observation 8c66803d-ef90-423c-94bd-8205ca26b35d · inbound

ThoughtFold: Folding Reasoning Chains via Introspective Preference Learning cites this paper.

ThoughtFold: Folding Reasoning Chains via Introspective Preference Learning Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:26:28.717883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T10:07:16.700499Z digest=sha256:cd0c7251aaa27a13d0a0d205c3e2b433280692c7f26a6f540df791c324acbe5e

Observation 0486a096-6862-4eef-87df-9433fe9830e1 · inbound

KCSAT-ML: Probing Reasoning Models with Nationwide-Cohort Human Difficulty cites this paper.

KCSAT-ML: Probing Reasoning Models with Nationwide-Cohort Human Difficulty Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-06-27T13:10:55.947824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T13:07:59.509660Z digest=sha256:58f177b2ce0053fc29135cca55d05c46ee7ec9cc0b980da1aff4baa19ab16c36

Observation d2fb637d-0b15-42bd-9b2f-e1d868ee854d · inbound

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes cites this paper.

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 145

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:57:41.512993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T12:59:51.091008Z digest=sha256:b4cd00f404d6ee28a8b762ac5aa67c5816b026e957d455a3ab2ad0f262568092

Observation c7a64538-8131-4f38-9c2b-110847bc4108 · inbound

Efficient and Trainable Language Model Test-Time Scaling via Local Branch Routing cites this paper.

Efficient and Trainable Language Model Test-Time Scaling via Local Branch Routing Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:20:06.711514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-25T21:29:40.564852Z digest=sha256:752713872f75df0c4a817801f0c0ea4b3b3127aec9d00963f196ec7e6717427d

Observation aac4b144-e9fe-4ae3-a728-0a9b1dbe4f9d · inbound

Efficient and Trainable Language Model Test-Time Scaling via Local Branch Routing cites this paper.

Efficient and Trainable Language Model Test-Time Scaling via Local Branch Routing Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-01T06:35:29.567228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-01T06:33:45.824727Z digest=sha256:b311c85d712eb6464597aac9f7e54dde7281afe5581b63105636f41d9346943c

Observation 928758a2-7325-4ee5-a6b3-a370bab1e0a7 · inbound

Before Thinking, Learn to Decide: Proactive Routing for Efficient Visual Reasoning cites this paper.

Before Thinking, Learn to Decide: Proactive Routing for Efficient Visual Reasoning Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T13:34:41.134711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T05:55:19.083517Z digest=sha256:2b802de6bdca4913906c09629f0b85c2ecfca7d619d4188b45eb69c48f061c53

Observation 84169b7c-5669-40d0-9f03-9c4d03678213 · inbound

Multi-Head Latent Control: A Unified Interface for LLM Agent Decision Making cites this paper.

Multi-Head Latent Control: A Unified Interface for LLM Agent Decision Making Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T02:42:03.779702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T02:42:03.779702Z digest=sha256:f03e315aaf0df502fbb55f8f90e5be50bcf412b75c64d81ae88a615b74c0b4f9

Observation 8ba2f2ae-deba-4dbb-a4f7-8e2ef50281cd · inbound

When to Plan: Learning to Select Between Reactive Control and Deliberative Planning cites this paper.

When to Plan: Learning to Select Between Reactive Control and Deliberative Planning Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T21:04:09.724503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:04:09.724503Z digest=sha256:95e55c89e81b796852a4a93c6ecdd12f045dc64837c1113fdbccdbd67a3de133

Observation 14bc6ab4-22d8-4176-9126-858b46c5137c · inbound

Is Your Model Thinking or Just Stagnating? PUMA: Diagnosing Reasoning Pathology via Phase-Momentum Alignment cites this paper.

Is Your Model Thinking or Just Stagnating? PUMA: Diagnosing Reasoning Pathology via Phase-Momentum Alignment Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T18:49:28.872386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:49:28.872386Z digest=sha256:4b3862859621e417b5e1de069e78c05dfdc97a19feec1a96e2a77363aa162bfe