Pith. sign in

Paper Citation Record · LEDGER

Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:2406.14283.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.14283 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 24 of 24 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T13:33:57.666272Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

9
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 12bd5ae0-7fdf-458a-91cb-088bb32bb134 · inbound

Reason4Rec: Deliberative User Preference Alignment of Large Language Models for Recommendation cites this paper.

Reason4Rec: Deliberative User Preference Alignment of Large Language Models for Recommendation Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-09T13:33:57.666272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T13:33:57.666272Z digest=sha256:26fc84453a66d9e5817d5a9f5aeec216b0eba76ed615df016a081a41e8dac533

Observation 18124b8a-0cbb-413f-b156-cce6ef0cd6a1 · inbound

QLASS: Boosting Language Agent Inference via Q-Guided Stepwise Search cites this paper.

QLASS: Boosting Language Agent Inference via Q-Guided Stepwise Search Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-09T11:47:31.339452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T11:47:31.339452Z digest=sha256:8c69d1063704d3b8cc6b53ecd6cdfa98a5627aad47ef376ab3b28732cc7bb5a1

Observation cca85fd6-0516-4ebf-bfc5-840f237f4d72 · inbound

Holistically Guided Monte Carlo Tree Search for Intricate Information Seeking cites this paper.

Holistically Guided Monte Carlo Tree Search for Intricate Information Seeking Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T21:40:26.935996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:40:26.935996Z digest=sha256:dd6ce648d17d6b934c5740aa56e15c2da734ccbd50bf39a6539f5021eae6bb67

Observation d3394951-336a-43db-a1bc-90b509e2403a · inbound

Holistically Guided Monte Carlo Tree Search for Intricate Information Seeking cites this paper.

Holistically Guided Monte Carlo Tree Search for Intricate Information Seeking Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T21:40:26.940923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:40:26.940923Z digest=sha256:90342d01c60683811090a067a225c954f8131c8f2adfff1ed5cce9f18db12ed5

Observation 02e6b3c6-4fce-4229-bd2a-91e10e91dc0b · inbound

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition cites this paper.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.567513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.567513Z digest=sha256:2260d7691578f7759c651220a2ad4015bbd2dd55007a887d07bbb0163c362dbe

Observation 76cd48da-cc93-4392-b2b0-fee689104cab · inbound

Policy Guided Tree Search for Enhanced LLM Reasoning cites this paper.

Policy Guided Tree Search for Enhanced LLM Reasoning Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-09T11:20:31.898539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T11:20:31.898539Z digest=sha256:b5253e012f3b456bbc485bc4eede376bf818e507f072f517ba3c8e94fc016780

Observation ee4d89ca-474f-448c-b6c4-f9fc2a16800b · inbound

Bag of Tricks for Inference-time Computation of LLM Reasoning cites this paper.

Bag of Tricks for Inference-time Computation of LLM Reasoning Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T13:35:24.562156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:35:24.562156Z digest=sha256:c13096298ebf955fd28868091b5041ba9998df171f6092c17acf3c78125594eb

Observation c8c40f3d-f2c3-4a9d-b0fd-941c90ec8c8b · inbound

Interactive Post-Training for Vision-Language-Action Models cites this paper.

Interactive Post-Training for Vision-Language-Action Models Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-21T14:25:47.236938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T14:25:47.178714Z digest=sha256:8c35d814b1abf85596f8b018e07349206d2b5f85104c7d39421ddb9f2a75f892

Observation 51023dc1-8e3c-4885-ac74-d59dbc822151 · inbound

Self-Critique Guided Iterative Reasoning for Multi-hop Question Answering cites this paper.

Self-Critique Guided Iterative Reasoning for Multi-hop Question Answering Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:23:47.820977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:23:47.820977Z digest=sha256:af76642ebefde325e40c987668e9db40e519111701590db57df0c57767b93568

Observation 9fed6fb4-c5ca-4ca0-b9e2-a3c83dae54bb · inbound

Large Language Models for Planning: A Comprehensive and Systematic Survey cites this paper.

Large Language Models for Planning: A Comprehensive and Systematic Survey Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 252

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:07.611284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:07.611284Z digest=sha256:b3601c2608b9a07799340485c8ae78e5123b94bcb1d80f761e1852275aff3ba7

Observation dbbb6368-36b0-4206-be47-b8f5c9665306 · inbound

Reasoning LLMs are Wandering Solution Explorers cites this paper.

Reasoning LLMs are Wandering Solution Explorers Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:01:13.678294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:01:13.678294Z digest=sha256:c75ed1ab08d02f33e6d3263177afdc68bc65870007a3429b1402a3534b852b5d

Observation 37be43d1-c47d-4fa3-bea9-6bf92ae3e265 · inbound

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL cites this paper.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.339030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.339030Z digest=sha256:8050898472b17de81802c9f1d6fd9f4dfdebc0e856a6cad162dca3dcca701f32

Observation e88e6dc7-a111-4a8e-8439-e4c2192d5cd3 · inbound

FreePRM: Training Process Reward Models Without Ground Truth Process Labels cites this paper.

FreePRM: Training Process Reward Models Without Ground Truth Process Labels Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:06:42.008548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:06:42.008548Z digest=sha256:e73b9df7e51375e1ded46eb9218d2d07e13e29303f10a2fb4b9412ca663adc36

Observation 22076380-86b1-435c-9ce2-0aebe7c31ed4 · inbound

Let's Revise Step-by-Step: A Unified Local Search Framework for Code Generation with LLMs cites this paper.

Let's Revise Step-by-Step: A Unified Local Search Framework for Code Generation with LLMs Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T22:09:46.348559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:09:46.348559Z digest=sha256:ef4d0e259abcee74f97d5a45516c84f6f3e19ab695258692b518c4ec9440b36f

Observation a504796f-b7c5-4b82-8f82-e852a7351783 · inbound

Beyond Linear Probes: Dynamic Safety Monitoring for Language Models cites this paper.

Beyond Linear Probes: Dynamic Safety Monitoring for Language Models Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:42:36.865021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-18T12:41:48.620040Z digest=sha256:2cc0a12d5f74b249e4beaa051979e8d254f81407ad8b886a090a3739873a30bc

Observation 329921e6-7667-47c8-bc7b-37c87a97f4d5 · inbound

Agentic Reasoning for Large Language Models cites this paper.

Agentic Reasoning for Large Language Models Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 109

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:14:26.008209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T15:14:25.558878Z digest=sha256:76f472bf37895d76db7682d361b520e064017674aea8a067184e8a09126c9300

Observation ec8c5023-e6f9-4d50-ba88-74a9af2fd8d5 · inbound

DeepResearch-9K: A Challenging Benchmark Dataset of Deep-Research Agent cites this paper.

DeepResearch-9K: A Challenging Benchmark Dataset of Deep-Research Agent Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-02T19:46:43.187057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:46:43.187057Z digest=sha256:317e842d9043c04517613bd3b481ed5336098473d0b4525890814238c8a8a59b

Observation 235fde31-3dba-4229-9f32-8fd0c75b7696 · inbound

C-TRAIL: A Commonsense World Framework for Trajectory Planning in Autonomous Driving cites this paper.

C-TRAIL: A Commonsense World Framework for Trajectory Planning in Autonomous Driving Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:28:26.090582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T23:27:31.054348Z digest=sha256:cc3a766b0a22717e277c177acbd8a77e84c2da92c44a5f51e7fc88aa3daa1881

Observation 5cf1ca7e-1e7a-4e6b-98ef-9c377b8cba9f · inbound

HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering cites this paper.

HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-09T23:54:45.454674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-09T23:51:47.724033Z digest=sha256:a156da3e73dda21b348d14efeaf3d391a1f56c2132656a1ff2466c1c216f5107

Observation cac4d1a0-e808-47e4-bbdd-f5a01cffc835 · inbound

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration cites this paper.

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:51:30.224331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T03:51:52.375703Z digest=sha256:85cf498b65ae54be737fc889c77f51e6adbe6bb91b6c24ed46bc9ed9f81089ee

Observation 6c00c4ca-adc2-43c3-96fe-516764254e4b · inbound

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration cites this paper.

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:15:03.212080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T05:11:32.053440Z digest=sha256:7766fe8271f8457f123e880a74b1e846d42408c97e080238fa101a5ed8791ce2

Observation 0e2bdf8b-6f9c-4412-baff-d254f61d7b41 · inbound

Training the Orchestrator: A Supervised Approach to End-to-End PDDL Planning with LLM Agents cites this paper.

Training the Orchestrator: A Supervised Approach to End-to-End PDDL Planning with LLM Agents Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T06:59:38.134654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T13:56:51.914966Z digest=sha256:f080d6c5fa4dcee4125c937578a365d73a339f5c7700cb51000cf1cfc41d5b19

Observation ceb18a38-89d6-42bf-ba72-1e84b1f82223 · inbound

REAR: Test-time Preference Realignment through Reward Decomposition cites this paper.

REAR: Test-time Preference Realignment through Reward Decomposition Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 146

Resolution
verified exact
arxiv_id, observed 2026-06-30T06:24:19.164064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-30T06:20:36.864229Z digest=sha256:4ee9a70c9239757c159e357912191f4472339e8843d774909488b666525aca80

Observation 690d9359-25a2-45b1-ac17-5f455d648fac · inbound

Engineering Trustworthy Agentic AI for Critical Systems cites this paper.

Engineering Trustworthy Agentic AI for Critical Systems Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:28.704221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:28.704221Z digest=sha256:4d744733dbe2ca85212be6ae7c19c906d47563062d7e27ffbbb216385eb902a1