Pith. sign in

Paper Citation Record · LEDGER

Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2406.14283.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.14283 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T21:40:26.940923Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

9
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation cca85fd6-0516-4ebf-bfc5-840f237f4d72 · inbound

Holistically Guided Monte Carlo Tree Search for Intricate Information Seeking cites this paper.

Holistically Guided Monte Carlo Tree Search for Intricate Information Seeking Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T21:40:26.935996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:40:26.935996Z digest=sha256:a59d52eca0a68ecd1e325e518bdc66403a739383491c46b672803e631e46b5cc

Observation d3394951-336a-43db-a1bc-90b509e2403a · inbound

Holistically Guided Monte Carlo Tree Search for Intricate Information Seeking cites this paper.

Holistically Guided Monte Carlo Tree Search for Intricate Information Seeking Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T21:40:26.940923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:40:26.940923Z digest=sha256:0c6b051f56d16c17af74f624fc1264c73213847149e333b735e68782dd0a6358

Observation 02e6b3c6-4fce-4229-bd2a-91e10e91dc0b · inbound

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition cites this paper.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.567513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.567513Z digest=sha256:a42a8beaf8b8b69cd02e8f33f9c3761b699849ac26c2319c3a752ca5b8f11143

Observation ee4d89ca-474f-448c-b6c4-f9fc2a16800b · inbound

Bag of Tricks for Inference-time Computation of LLM Reasoning cites this paper.

Bag of Tricks for Inference-time Computation of LLM Reasoning Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T13:35:24.562156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:35:24.562156Z digest=sha256:25b85ccdf98b7ca398eb752c7a8fc73bfb262b184e031068749ab34e7d9305c2

Observation c8c40f3d-f2c3-4a9d-b0fd-941c90ec8c8b · inbound

Interactive Post-Training for Vision-Language-Action Models cites this paper.

Interactive Post-Training for Vision-Language-Action Models Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-21T14:25:47.236938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T14:25:47.178714Z digest=sha256:3d532f60aaba5aa55081ca7800155354081ab4149c68faf8e58baac9aa49c5f8

Observation 51023dc1-8e3c-4885-ac74-d59dbc822151 · inbound

Self-Critique Guided Iterative Reasoning for Multi-hop Question Answering cites this paper.

Self-Critique Guided Iterative Reasoning for Multi-hop Question Answering Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:23:47.820977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:23:47.820977Z digest=sha256:fc03c576f74074ece6f6bc1535266f7f1399da6ca51e1c0f8a19fa458da5493a

Observation 9fed6fb4-c5ca-4ca0-b9e2-a3c83dae54bb · inbound

Large Language Models for Planning: A Comprehensive and Systematic Survey cites this paper.

Large Language Models for Planning: A Comprehensive and Systematic Survey Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 252

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:07.611284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:07.611284Z digest=sha256:eb0098af726d1f9b536ecb7311018c581e7fb507cfbb38f11169f1a36b03b16d

Observation dbbb6368-36b0-4206-be47-b8f5c9665306 · inbound

Reasoning LLMs are Wandering Solution Explorers cites this paper.

Reasoning LLMs are Wandering Solution Explorers Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:01:13.678294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:01:13.678294Z digest=sha256:c0bf974622cf53b5cf0ef32c98764eeda13ca0607b98fdecdebd16f2c626ed6c

Observation 37be43d1-c47d-4fa3-bea9-6bf92ae3e265 · inbound

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL cites this paper.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.339030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.339030Z digest=sha256:30479d921f3958f35350328e08dcc6355b84d1d7fef2660302a71653c2ab8a1a

Observation e88e6dc7-a111-4a8e-8439-e4c2192d5cd3 · inbound

FreePRM: Training Process Reward Models Without Ground Truth Process Labels cites this paper.

FreePRM: Training Process Reward Models Without Ground Truth Process Labels Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:06:42.008548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:06:42.008548Z digest=sha256:5a9385f49df6ed6a8280165da650df7aa9c153155ec04e57f42c3e7da8f1bc58

Observation 22076380-86b1-435c-9ce2-0aebe7c31ed4 · inbound

Let's Revise Step-by-Step: A Unified Local Search Framework for Code Generation with LLMs cites this paper.

Let's Revise Step-by-Step: A Unified Local Search Framework for Code Generation with LLMs Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T22:09:46.348559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:09:46.348559Z digest=sha256:bd823d6caa71b29ad4e6cce23fa14c310d2bb6b79fbf0a772c8be6e3209d9d17

Observation a504796f-b7c5-4b82-8f82-e852a7351783 · inbound

Beyond Linear Probes: Dynamic Safety Monitoring for Language Models cites this paper.

Beyond Linear Probes: Dynamic Safety Monitoring for Language Models Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:42:36.865021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-18T12:41:48.620040Z digest=sha256:13544f7414827a640ac831f63c49aec59a8affe1db09d923eb6f313190083f3d

Observation 329921e6-7667-47c8-bc7b-37c87a97f4d5 · inbound

Agentic Reasoning for Large Language Models cites this paper.

Agentic Reasoning for Large Language Models Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 109

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:14:26.008209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T15:14:25.558878Z digest=sha256:128ed4c25fdf41438c07d75bcd4741679fbaecab26deec3bdd77679404619664

Observation ec8c5023-e6f9-4d50-ba88-74a9af2fd8d5 · inbound

DeepResearch-9K: A Challenging Benchmark Dataset of Deep-Research Agent cites this paper.

DeepResearch-9K: A Challenging Benchmark Dataset of Deep-Research Agent Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-02T19:46:43.187057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:46:43.187057Z digest=sha256:91b92a27eb368ca0b58e3744d8ba646f24c8734f2042ed32a788bf142a48df69

Observation 235fde31-3dba-4229-9f32-8fd0c75b7696 · inbound

C-TRAIL: A Commonsense World Framework for Trajectory Planning in Autonomous Driving cites this paper.

C-TRAIL: A Commonsense World Framework for Trajectory Planning in Autonomous Driving Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:28:26.090582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T23:27:31.054348Z digest=sha256:fff3e62d4dccaa7dc8f472e7b0e6d1bc3afa4c53e556f63e0455f54d71e7e2b6

Observation 5cf1ca7e-1e7a-4e6b-98ef-9c377b8cba9f · inbound

HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering cites this paper.

HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-09T23:54:45.454674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-09T23:51:47.724033Z digest=sha256:03f79fb249e8d1ddfcf73a8b1408c27ba34c19f31e31c6baf57ea7f991edce25

Observation cac4d1a0-e808-47e4-bbdd-f5a01cffc835 · inbound

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration cites this paper.

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:51:30.224331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T03:51:52.375703Z digest=sha256:a2d4855b2fa2cd8c36ed72b467fc341f241e3bd486a2d4c9813d25336aa105cb

Observation 6c00c4ca-adc2-43c3-96fe-516764254e4b · inbound

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration cites this paper.

Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative Exploration Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:15:03.212080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T05:11:32.053440Z digest=sha256:f1d24f02385603fbeaa767375f064aebcd750430a96e68d43745d63db3d29c56

Observation 0e2bdf8b-6f9c-4412-baff-d254f61d7b41 · inbound

Training the Orchestrator: A Supervised Approach to End-to-End PDDL Planning with LLM Agents cites this paper.

Training the Orchestrator: A Supervised Approach to End-to-End PDDL Planning with LLM Agents Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T06:59:38.134654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T13:56:51.914966Z digest=sha256:61b87e40dcbd916423a9c2c372c3ff055ca0e9e0106c08e58821a5123a48322b

Observation ceb18a38-89d6-42bf-ba72-1e84b1f82223 · inbound

REAR: Test-time Preference Realignment through Reward Decomposition cites this paper.

REAR: Test-time Preference Realignment through Reward Decomposition Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 146

Resolution
verified exact
arxiv_id, observed 2026-06-30T06:24:19.164064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-30T06:20:36.864229Z digest=sha256:a76cc94846d0c59b7342a4a04f013d340f3372bed786792b7102bc57580d503c

Observation 690d9359-25a2-45b1-ac17-5f455d648fac · inbound

Engineering Trustworthy Agentic AI for Critical Systems cites this paper.

Engineering Trustworthy Agentic AI for Critical Systems Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T15:07:28.704221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:07:28.704221Z digest=sha256:e875c435e9dacfeb663b00b84481ad7cbcd0c09e2a11579a6499e1af311e5094