Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 75 inbound Pith citation observations for arXiv:2309.17179.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T12:02:42.924167Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
3
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation aea7076a-c96b-45b0-84c0-e62d2b2c97e4 · inbound
Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a9446a2e-c207-429b-a153-efdfed86db07 · inbound
Improve Mathematical Reasoning in Language Models by Automated Process Supervision Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation cf78593d-89e2-49b2-a804-19c8160fc570 · inbound
Natural Language Reinforcement Learning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 251325a8-35f2-4446-92ba-b6cccf010c91 · inbound
Marco-o1: Towards Open Reasoning Models for Open-Ended Solutions Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7645cd5e-ad11-44b8-8dc5-db8d4b2aa4a2 · inbound
PIANIST: Learning Partially Observable World Models with LLMs for Multi-Agent Decision Making Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45e7be58-dd37-41a3-aea0-62891708d141 · inbound
ZoomEye: Enhancing Multimodal LLMs with Human-Like Zooming Capabilities through Tree-Based Image Exploration Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59da8fea-a7fd-4404-b724-dbec5d540773 · inbound
o1-Coder: an o1 Replication for Coding Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec385732-d4b8-4987-9308-d193a39b7662 · inbound
RARE: Retrieval-Augmented Reasoning Enhancement for Large Language Models Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c41dc8e-5dc9-4f61-ad6b-dba38235787a · inbound
LLMs as Debate Partners: Utilizing Genetic Algorithms and Adversarial Search for Adaptive Arguments Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0b38c5f-97f0-465f-bf7b-bfe49f0da580 · inbound
Ensembling Large Language Models with Process Reward-Guided Tree Search for Better Complex Reasoning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e41f556e-da6e-452a-af56-9974dfee858a · inbound
Offline Reinforcement Learning for LLM Multi-Step Reasoning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fffcf75f-e6d4-474b-8407-31f54595ab11 · inbound
Efficiently Scaling LLM Reasoning with Certaindex Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7acb99e-dda2-4a48-8d3c-85826cb4ee4a · inbound
BoostStep: Boosting mathematical capability of Large Language Models via improved single-step reasoning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11b85831-dbcd-4a8f-acf3-7e651b70d0c8 · inbound
Monte Carlo Tree Search for Comprehensive Exploration in LLM-Based Automatic Heuristic Design Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e221ac50-fc44-411f-b8c5-01529f990319 · inbound
Rethinking External Slow-Thinking: From Snowball Errors to Probability of Correct Reasoning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18bb0357-4e3d-4883-8230-f7d2ff9d439a · inbound
Large Language Models to Diffusion Finetuning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 1969
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 469e47ad-84b8-460e-a008-a747f6fe8abb · inbound
Metastable Dynamics of Chain-of-Thought Reasoning: Provable Benefits of Search, RL and Distillation Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfbc782a-e2b0-4875-8f83-4673f0a6e7f6 · inbound
QLASS: Boosting Language Agent Inference via Q-Guided Stepwise Search Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdd185cc-189c-4585-a351-a8aeff4a8531 · inbound
On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ef57d7c-c201-481b-b7ab-80775d952152 · inbound
Policy Guided Tree Search for Enhanced LLM Reasoning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e92021b1-0c6e-4369-8774-ec43f7b0ccf9 · inbound
Bag of Tricks for Inference-time Computation of LLM Reasoning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb68f2ca-0621-4020-bd19-b463ba31ccd1 · inbound
Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fb995c4c-c156-4409-a4c3-748bfb0c287a · inbound
QM-ToT: A Medical Tree of Thoughts Reasoning Framework for Quantized Model Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 86b7f0f0-f364-4cc2-b6a8-e887299adba3 · inbound
Generative AI Act II: Test Time Scaling Drives Cognition Engineering Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e2c4003-3996-4161-8e02-189bb2881a20 · inbound
From Reflection to Perfection: Scaling Inference-Time Optimization for Text-to-Image Diffusion Models via Reflection Tuning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b176da2c-6234-4d68-9946-9070a536d876 · inbound
Lightweight Latent Verifiers for Efficient Meta-Generation Strategies Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98b5ea52-b3df-45db-a343-775bc98530d2 · inbound
Safety in Large Reasoning Models: A Survey Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b3048f0-41ac-45e5-baa7-7ad496a900d7 · inbound
COSMOS: Predictable and Cost-Effective Adaptation of LLMs Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1da66c8c-df16-4f25-9390-592bbbca1659 · inbound
Observe-R1: Unlocking Reasoning Abilities of MLLMs with Dynamic Progressive Reinforcement Learning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eee365db-c203-4bf0-a66f-19a157c902cb · inbound
MindOmni: Unleashing Reasoning Generation in Vision Language Models with RGPO Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23cc425d-47c5-4bf8-842b-656c2bc06366 · inbound
MM-PRM: Enhancing Multimodal Mathematical Reasoning with Scalable Step-Level Supervision Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03ff09e9-1f0e-4415-ad95-afb2a0f01966 · inbound
Generalizing Large Language Model Usability Across Resource-Constrained Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation beb2dcd8-ce09-4a45-ac89-9df398fc6e18 · inbound
MMATH: A Multilingual Benchmark for Mathematical Reasoning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db427559-1d5c-4206-92ed-026d48eec8af · inbound
Can Past Experience Accelerate LLM Reasoning? Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8769437d-e1a6-40e5-a546-f22ebb8d114b · inbound
How Much Backtracking is Enough? Exploring the Interplay of SFT and RL in Enhancing LLM Reasoning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ad7cede-ca98-4c61-8f14-ecbbdd2bc5ab · inbound
Structured Pruning for Diverse Best-of-N Reasoning Optimization Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 650cc220-635d-40a7-87eb-790fc13501dd · inbound
Kinetics: Rethinking Test-Time Scaling Laws Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4234ff1d-a7b8-47fb-b9ec-4092ced0f56f · inbound
VReST: Enhancing Reasoning in Large Vision-Language Models through Tree Search and Self-Reward Mechanism Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e67df85b-d344-4f51-8833-ae66fee42643 · inbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c54a79fa-7e5e-42a8-85a3-f7797b28d389 · inbound
TreeRL: LLM Reinforcement Learning with On-Policy Tree Search Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c47a8aa-1244-4b5e-bdc4-7b2deac56277 · inbound
KnowRL: Exploring Knowledgeable Reinforcement Learning for Factuality Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 6b7d3b68-72f8-48e2-94ec-d7f9494fb526 · inbound
Enhancing User Engagement in Socially-Driven Dialogue through Interactive LLM Alignments Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5eeb45b6-3c6d-4900-93a7-03e4147598f1 · inbound
Reasoning in machine vision by learning fast and slow thinking Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a1f5f4c-bd4b-4007-af55-873c9248df71 · inbound
Data Diversification Methods In Alignment Enhance Math Performance In LLMs Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1d90dd1-c705-42d0-baef-f00aca515744 · inbound
Enhancing Test-Time Scaling of Large Language Models with Hierarchical Retrieval-Augmented MCTS Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60d6e2b2-8357-4380-9b77-2c4b9d9fc6b8 · inbound
DeepSeek: Paradigm Shifts and Technical Evolution in Large AI Models Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a39a040-6db1-4621-9b31-ed1e1b145f3a · inbound
It's Not That Simple. An Analysis of Simple Test-Time Scaling Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bacaa48-b2da-445a-9bf2-4aa70ba0a491 · inbound
LLM Economist: Large Population Models and Mechanism Design in Multi-Agent Generative Simulacra Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84845823-462d-4b71-adb7-056d11f47575 · inbound
Let's Revise Step-by-Step: A Unified Local Search Framework for Code Generation with LLMs Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14285457-87e1-4d05-a671-243df4b6e228 · inbound
Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 398b7170-af66-43e5-b8c6-f39c12d228aa · inbound
CDE: Curiosity-Driven Exploration for Efficient Reinforcement Learning in Large Language Models Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 1997
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f246477-bd08-4ba3-8922-04f6a7bc2519 · inbound
Plan Then Action:High-Level Planning Guidance Reinforcement Learning for LLM Reasoning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 028a9c3e-e00f-4664-bf40-d95bf6581261 · inbound
DiffCoT: Diffusion-styled Chain-of-Thought Reasoning in LLMs Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 788d1018-d4fc-4a17-8e50-9b4dfe58040a · inbound
Neural Chain-of-Thought Search: Searching the Optimal Reasoning Path to Enhance Large Language Models Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0754fcb6-bf24-483b-9a0b-0deb0bafddd8 · inbound
Agentic Reasoning for Large Language Models Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 120
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a5728366-d70a-4ffb-9660-d96556437e88 · inbound
Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a2df38d-1ab1-4573-b6f5-5d192859296b · inbound
Reinforce to Learn, Elect to Reason: A Dual Paradigm for Video Reasoning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 761dd4d6-29a3-4eb1-8972-b07c89f37a19 · inbound
Understanding Performance Gap Between Parallel and Sequential Sampling in Large Reasoning Models Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 66be9b4c-024e-473f-ae01-4a23529ff515 · inbound
Evaluation-driven Scaling for Scientific Discovery Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 93f3611e-068a-4ac9-b5a0-fb844a575fd3 · inbound
Distilling Long-CoT Reasoning through Collaborative Step-wise Multi-Teacher Decoding Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 64f389a3-ece9-4f54-8a47-654524236766 · inbound
SAT: Sequential Agent Tuning for Coordinator Free Plug and Play Multi-LLM Training with Monotonic Improvement Guarantees Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5210ee3d-d234-4f0a-a335-3a0f50e898ff · inbound
Maximizing Rollout Informativeness under a Fixed Budget: A Submodular View of Tree Search for Tool-Use Agentic Reinforcement Learning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ff54ab4a-a1f1-47c5-8e7f-8ebbb47b3abf · inbound
CA-SQL: Complexity-Aware Inference Time Reasoning for Text-to-SQL via Exploration and Compute Budget Allocation Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ada3a1b9-1026-41b1-9cad-17bf54fbcee6 · inbound
V-ABS: Action-Observer Driven Beam Search for Dynamic Visual Reasoning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation cb9cf998-12a0-426d-826e-0ab8d015ecb7 · inbound
Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4f34326e-528b-4d96-9ad3-1fc96a76fa5d · inbound
Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4077066b-ce29-43c3-ad50-fc741a014a0d · inbound
Small RL Controller, Large Language Model: RL-Guided Adaptive Sampling for Test-Time Scaling Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9c889228-de92-44ce-852a-8375a79b6547 · inbound
Step-by-Step Optimization-like Reasoning in LLMs over Expanding Search Spaces Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1cc95fe5-48ec-40d7-879e-988120a46c37 · inbound
DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3c60c9d5-bffe-4fc4-937f-e71ab1969215 · inbound
TRACE: A Unified Rollout Budget Allocation Framework for Efficient Agentic Reinforcement Learning Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e84f6b6a-14d9-492c-acf4-7991eabb59ea · inbound
APPO: Agentic Procedural Policy Optimization Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 956695dc-6775-41b4-930a-08ea6cbae14b · inbound
APPO: Agentic Procedural Policy Optimization Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26e4a92d-54d2-45ef-982a-7632f69be8e3 · inbound
Beyond Fixed Budgets: Characterizing the Inelasticity and Limitations of Tree-of-Thought Reasoning Strategies Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c83a8d58-2ff5-4cae-ac64-129a2da31a2c · inbound
DecompRL: Solving Harder Problems by Learning Modular Code Generation Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9fe9a5a7-d48d-46cb-a485-730dbb33a8bf · inbound
ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 165
Source-reported events for the cited work
Unavailable: canonical work link unavailable.