Pith. sign in

Paper Citation Record · LEDGER

Enhancing LLM Reasoning with Reward-guided Tree Search

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 29 inbound Pith citation observations for arXiv:2411.11694.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.11694 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 29 of 29 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T21:40:26.812624Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b747b821-6732-41e9-9d67-8e74890b5ad4 · inbound

Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems cites this paper.

Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:35:31.441347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T00:35:31.375020Z digest=sha256:045ca53532c457ea28918eba341d743f9bbc52ebf1f23891f8d3e313b0700768

Observation 6aa1724b-52d3-41c5-8e25-14eb70d34815 · inbound

Search-o1: Agentic Search-Enhanced Large Reasoning Models cites this paper.

Search-o1: Agentic Search-Enhanced Large Reasoning Models Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:36:27.714815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T17:36:27.515468Z digest=sha256:595423d9438498f277b091b436dea56d6d44cd9c6d2ec05ea08b1a27cabfa527

Observation 5ed12ada-9c94-4b3d-b128-ccc7f6903693 · inbound

Holistically Guided Monte Carlo Tree Search for Intricate Information Seeking cites this paper.

Holistically Guided Monte Carlo Tree Search for Intricate Information Seeking Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T21:40:26.812624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:40:26.812624Z digest=sha256:c85c11fa72d13cc0ada44ff6bc43f490c06b88e18eadc317adbd86c49c3f65bc

Observation 9a181384-00a6-44af-a6d7-0d6ebdc467e3 · inbound

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition cites this paper.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.387950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.387950Z digest=sha256:c27ab770c308c070b2feb0dba1da9c78b40e02846155e1c66ff5466fde2352cb

Observation 0826ecf3-81b6-4e94-976d-9fe0cd154d7f · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 228

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.527788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:610e6e08862e297c52094e7bcfcadf293095450d4f5e6315cd8eb423f8a1ce6d

Observation 4731a328-5df5-481d-91b0-31ab3a55478f · inbound

Large Language Model Agent: A Survey on Methodology, Applications and Challenges cites this paper.

Large Language Model Agent: A Survey on Methodology, Applications and Challenges Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 108

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:52:10.570222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T21:51:34.309870Z digest=sha256:15d67e752b10c63526a74f94ab9c5de392ac8f05dad143c4d38c43678b7939ae

Observation 34a2bdca-22ac-4241-bad8-498eb7451526 · inbound

ZeroSearch: Incentivize the Search Capability of LLMs without Searching cites this paper.

ZeroSearch: Incentivize the Search Capability of LLMs without Searching Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-17T17:44:13.470477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T17:44:13.310155Z digest=sha256:e1994349781ea8364eb3c39520ae4711e502188985f996f58433e3a3b06f93ce

Observation 56fe818b-eff3-4366-8fa8-a5ab4045d2de · inbound

ZeroSearch: Incentivize the Search Capability of LLMs without Searching cites this paper.

ZeroSearch: Incentivize the Search Capability of LLMs without Searching Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-22T16:06:45.978606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T16:05:04.715678Z digest=sha256:40c448527a8027850760ef8916f7e1ad6bf654c98324697353fb4391753bc49e

Observation a80ca779-edbd-4c3a-80ff-78034a3c8734 · inbound

EquivPruner: Boosting Efficiency and Quality in LLM-Based Search via Action Pruning cites this paper.

EquivPruner: Boosting Efficiency and Quality in LLM-Based Search via Action Pruning Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:06:00.892822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:06:00.892822Z digest=sha256:381b6e5f278075ab50b54ba41ea8b3744e0f78a0e8696c3484504af1c60d6d4e

Observation 9f9ffe20-ea0b-4f37-9e95-257e32a45eae · inbound

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning cites this paper.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.406362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.406362Z digest=sha256:d7cd51a5b4f04e8da223363304cb53e095c1c5c78ff744ccb525a7e90a004d01

Observation f9b003d7-32fe-4b32-ab05-8497d578fab5 · inbound

Reward Model Generalization for Compute-Aware Test-Time Reasoning cites this paper.

Reward Model Generalization for Compute-Aware Test-Time Reasoning Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:42:27.309575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:42:27.309575Z digest=sha256:b66495666ed43ca4574c545800edad9103126b7ee2063010f558a14b79ac71c0

Observation c4faad56-4c78-4937-9050-740d21861537 · inbound

Evaluation is All You Need: Strategic Overclaiming of LLM Reasoning Capabilities Through Evaluation Design cites this paper.

Evaluation is All You Need: Strategic Overclaiming of LLM Reasoning Capabilities Through Evaluation Design Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:40:21.212589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:40:21.212589Z digest=sha256:0ece0630bcd4d6ec62e8c2ab4ec6b15b7d6fd6769b51aa8dd26749c1528f3cb7

Observation fa67507f-6f74-448c-a283-ebf292b38dd5 · inbound

Com$^2$: A Causal-Guided Benchmark for Exploring Complex Commonsense Reasoning in Large Language Models cites this paper.

Com$^2$: A Causal-Guided Benchmark for Exploring Complex Commonsense Reasoning in Large Language Models Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:46:43.398130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:46:43.398130Z digest=sha256:73f8a57f6fc1cdca551f609d9bea82fb2ee933a59bec1ec677d812b0ab4c5535

Observation a4cc418c-f9fe-4f52-9f47-9608d2518715 · inbound

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism cites this paper.

VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T05:09:19.582648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:09:19.582648Z digest=sha256:6871ddadc5d37a7fb11f15f896e2166a2d3f936df09c523b8c90a4f810b726ef

Observation 33c4bf2c-0c93-4766-a506-f9a81f20da20 · inbound

Ctrl-Z Sampling: Scaling Diffusion Sampling with Controlled Random Zigzag Explorations cites this paper.

Ctrl-Z Sampling: Scaling Diffusion Sampling with Controlled Random Zigzag Explorations Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T22:59:12.842671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:59:12.842671Z digest=sha256:383265734c9fc7ec47c51ca5ab46a4b35809869f6a709dfd8bc1d77b1fa37172

Observation a1bcfbea-21e6-46dd-96b8-9aa36ae0c15b · inbound

Let's Revise Step-by-Step: A Unified Local Search Framework for Code Generation with LLMs cites this paper.

Let's Revise Step-by-Step: A Unified Local Search Framework for Code Generation with LLMs Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T22:09:46.543585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:09:46.543585Z digest=sha256:960019bd07d2a1e08aa344e8e9b3ff9f61b9a1dc167454e548cf3e854c28e2db

Observation ea005420-8fac-45fb-be45-2bb5acf2547f · inbound

From Trial-and-Error to Improvement: A Systematic Analysis of LLM Exploration Mechanisms in RLVR cites this paper.

From Trial-and-Error to Improvement: A Systematic Analysis of LLM Exploration Mechanisms in RLVR Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T22:03:28.144048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:03:28.144048Z digest=sha256:ec6c002157268b042c945dcd4d6df8259fdaaad4b9e70c6762972e987c7b9cc0

Observation 0a8439d8-5404-4d31-926f-f181fcc433c4 · inbound

Why Does Reasoning Length Converge? Unveiling the Underfitting-Overfitting Trade-off in Chain-of-Thought cites this paper.

Why Does Reasoning Length Converge? Unveiling the Underfitting-Overfitting Trade-off in Chain-of-Thought Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T10:31:37.800280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T10:31:37.800280Z digest=sha256:889f8a328854ffa3f492957b1633bd4aed458679db2522218d61b5b0aa1133de

Observation 608e598d-cbc1-4caa-a836-3e3fc350f3af · inbound

Sticker-TTS: Learn to Utilize Historical Experience with a Sticker-driven Test-Time Scaling Framework cites this paper.

Sticker-TTS: Learn to Utilize Historical Experience with a Sticker-driven Test-Time Scaling Framework Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T05:45:02.737120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T05:45:02.737120Z digest=sha256:c8747e09aee84451906542066e1f0fa29be54bf5ca658adcc5335c7e99463fe9

Observation 84507aa2-47d9-400c-a03c-d6af0fc57d8d · inbound

Retrieval-of-Thought: Efficient Reasoning via Reusing Thoughts cites this paper.

Retrieval-of-Thought: Efficient Reasoning via Reusing Thoughts Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:42:38.635096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T13:42:07.883909Z digest=sha256:7a1bb59f79690cbfab6f8a547807e97db4e2f4d1f8e46d6e14b5142bfb50599e

Observation 98ab1b38-b39d-4c71-bf79-bb57fba73fd2 · inbound

When control meets large language models: From words to dynamics cites this paper.

When control meets large language models: From words to dynamics Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-05-21T14:54:13.213179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T14:52:44.632671Z digest=sha256:58e9f94c51c4830cc55ee792d3d5b8bb009594d1ab9ac337097520d282714381

Observation 537f8259-853b-49fa-bb55-17839a0768ab · inbound

PARM: Pipeline-Adapted Reward Model cites this paper.

PARM: Pipeline-Adapted Reward Model Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:28:39.510004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T05:15:26.015817Z digest=sha256:8443191887e406740ca673a27d865678db9ff501ae9c3f5c71da93cbcf78faf1

Observation 3bf1b1a0-94ac-456c-b1eb-0ba4e401864d · inbound

A Communication-Theoretic Framework for LLM Agents: Cost-Aware Adaptive Reliability cites this paper.

A Communication-Theoretic Framework for LLM Agents: Cost-Aware Adaptive Reliability Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 71

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:26:27.522360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T03:03:05.715652Z digest=sha256:261680decaabf41de3e10c53977c13d2095a17bc0569c75ecce175e86759d145

Observation 0d972d20-b559-42bb-839a-8127afaf352d · inbound

Density Field State Space Models: 1-Bit Distillation, Efficient Inference, and Knowledge Organization in Mamba-2 cites this paper.

Density Field State Space Models: 1-Bit Distillation, Efficient Inference, and Knowledge Organization in Mamba-2 Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:05:37.131006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-01T08:55:37.926334Z digest=sha256:cf162618faa759ebad9d5ff7bb178566b1a200f0df29e6eb07003e4c21f21b68

Observation cc708544-57ba-42b6-ab7d-8b6a0db9d35b · inbound

MODE-RAG: Manifold Outlier Diagnosis and Energy-based Retrieval-Augmented Generation Evaluation cites this paper.

MODE-RAG: Manifold Outlier Diagnosis and Energy-based Retrieval-Augmented Generation Evaluation Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:08:56.384800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T01:35:46.913390Z digest=sha256:6adaf1b6b3c1b90d8203629c18cc8e50c143e8f915d3cf82868420a31626ab10

Observation 7eb9e57f-a20d-44e6-8a98-f55fef3c7839 · inbound

Beyond Fixed Budgets: Characterizing the Inelasticity and Limitations of Tree-of-Thought Reasoning Strategies cites this paper.

Beyond Fixed Budgets: Characterizing the Inelasticity and Limitations of Tree-of-Thought Reasoning Strategies Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-30T18:44:59.789825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T18:38:48.512777Z digest=sha256:feb967ab71c8fe7b247e7574a5b2650abd9d9952999b28c2d48c2d4b5fc7e77c

Observation 1ecd146f-269f-4c5c-9404-be0d44e64969 · inbound

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning cites this paper.

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 285

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:59:44.800993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T09:19:50.623741Z digest=sha256:1c2f3a6e6b0ae73786ce2ff736781076e6617a76c276e36c72f949f1796524e5

Observation e4709ce6-424a-4ecb-92f6-efdc6d7ccd9d · inbound

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning cites this paper.

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 284

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T18:55:59.426665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-29T01:18:19.195007Z digest=sha256:e858bd76ad479ffa7b6b173270f2165c2481ab3f7dbfe2cbd6d8da2a448f0ef3

Observation 40872ebf-2c4b-4d17-ad29-4a171c8d48d5 · inbound

TreeThink: A Modular Tree Search Library for Mathematical Reasoning with LLMs cites this paper.

TreeThink: A Modular Tree Search Library for Mathematical Reasoning with LLMs Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 123

Resolution
unresolved
no resolver link, observed 2026-07-14T05:57:23.399019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T05:57:23.399019Z digest=sha256:2f11266d47342b9149662d238821e51e28748f08b2fcad96627139fa609a5686