Pith. sign in

Paper Citation Record · LEDGER

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively

As of 15 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2506.00396.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.00396 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:11:46.192441Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f6377368-b056-412a-96ae-e14ccad40c31 · outbound

This paper cites URL: " 'urlintro :=.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively URL: " 'urlintro :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:41.751456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:41.751456Z digest=sha256:2f588be3d10d7ebc55eb6580e23ef087c066506bbbaa810c56b28cf5483e5baa

Observation 33a3c908-e5a4-4da1-9430-d564e72bd382 · outbound

This paper cites write newline.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:41.847455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:41.847455Z digest=sha256:316c5d7532e23548d02d5faed38299950896455998feda96c190aa4343f45c65

Observation f4499d80-d173-495d-aa62-b5e11faf4df3 · outbound

This paper cites Graph of Thoughts: Solving Elaborate Problems with Large Language Models.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Graph of Thoughts: Solving Elaborate Problems with Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:42.045200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:42.045200Z digest=sha256:254165cc79bc971ac62844b03bf38e597c37d901902a90c5ffdb41474f81be2c

Observation a774c6f8-9cbb-4f70-a7eb-df45b99c2156 · outbound

This paper cites Accelerating Large Language Model Decoding with Speculative Sampling.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Accelerating Large Language Model Decoding with Speculative Sampling

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:42.178654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:42.178654Z digest=sha256:26e8e6dc5714e541d74a6c25896d2084c18d01b425a3b497107c031d93236a2b

Observation cc94f70a-b3fa-48af-abc8-c7a354fcb0f0 · outbound

This paper cites FinQA: A Dataset of Numerical Reasoning over Financial Data.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively FinQA: A Dataset of Numerical Reasoning over Financial Data

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:42.331835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:42.331835Z digest=sha256:f05bcb9bbabc24c2dcc8ea65b5fb06d4246475c613f8b993d00c2fc20ca66253

Observation 7a086ed3-0756-41da-896f-065a4a73b3f6 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Training Verifiers to Solve Math Word Problems

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:42.497773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:42.497773Z digest=sha256:289f30879b6ae3365e1a6227c23c6ff758f35f3110694f174fd7644a1ad01355

Observation 95746e22-c505-4434-a73e-f0f025859af1 · outbound

This paper cites an unresolved cited work.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:11:48.912689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T12:11:42.637601Z digest=sha256:fcc562c21c1ec3757d4f9e88e12ca6af8da67d8e863c6a78add04e2c87ffa0eb

Observation 4bb8b3b9-fc0f-4f72-a8fd-e7e244cb4162 · outbound

This paper cites Everything of Thoughts: Defying the Law of Penrose Triangle for Thought Generation.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Everything of Thoughts: Defying the Law of Penrose Triangle for Thought Generation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:42.761281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:42.761281Z digest=sha256:70eec91fda5881af44fb7440032099a1926547f1b875dad0ed472d0021e304cb

Observation 4e2df33f-7c43-4e80-bbd8-fd38cc1f67d5 · outbound

This paper cites The Llama 3 Herd of Models.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively The Llama 3 Herd of Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:42.881590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:42.881590Z digest=sha256:db7a28dfc5316c2542cc0e1f9915a4a6a1d837ecb5bf8adfb60b3077d098feb2

Observation 188d4c1d-6f1a-4d79-a559-54be10710330 · outbound

This paper cites Reasoning with Language Model is Planning with World Model.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Reasoning with Language Model is Planning with World Model

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:43.039389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:43.039389Z digest=sha256:1b4ab94adbf8c21ac14c83f56c06a225e3ce18016b65642cfa0cca34f163de6e

Observation 75a67c94-5ce5-4074-bbb4-7174c7500b37 · outbound

This paper cites Large Language Models Cannot Self-Correct Reasoning Yet.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Large Language Models Cannot Self-Correct Reasoning Yet

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:43.197324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:43.197324Z digest=sha256:fc7850680956ab13848cca9e7aab8f08be09ea8e2159859f1eaf83ba78390d87

Observation bb996b5b-de16-4817-a101-d0dafa3b1dfe · outbound

This paper cites an unresolved cited work.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:11:48.581873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T12:11:43.343195Z digest=sha256:ce6d21d2c92ea2cc95d414366cf1f9eefa08988fc4d5838bec66372f8fb5323e

Observation cd972443-e59a-40ce-b06d-3303f40eea88 · outbound

This paper cites Reward Design with Language Models.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Reward Design with Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:43.508406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:43.508406Z digest=sha256:29b9ba913588fbaf8d3d8baf9d43dea6435f311d16b5dd6a705cda2e81318419

Observation ac6ecb87-4432-4707-b893-a0045ec5f7c2 · outbound

This paper cites an unresolved cited work.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:11:48.348430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T12:11:43.663834Z digest=sha256:b8545a94c5264b60f9176143dd2ba35f895afe2b37390ebda3d264baa382b5db

Observation 62239cb0-9347-4121-8d7d-05f951a4c1c9 · outbound

This paper cites an unresolved cited work.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:11:48.144129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T12:11:43.827341Z digest=sha256:914596e66fdcdb5c973300aaa1e51cd976062ed9374baeda15cf9f1d45ad37a2

Observation 73869741-03f6-4365-8c9b-9dc415079e0e · outbound

This paper cites GPT-4 Technical Report.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively GPT-4 Technical Report

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:43.975924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:43.975924Z digest=sha256:fbe37db2d523caeef09eba00760b8aee268dc86f54ff9a114688318b14ee265e

Observation 4055b49f-ec1b-4352-84a0-f9ce9dd6a23b · outbound

This paper cites an unresolved cited work.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:11:47.926136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T12:11:44.122879Z digest=sha256:2a53f11fe0be4141e18062a3d560b9ef2a6d424ce9d26ab16b193e80e64e7d87

Observation 49912d8b-c8cf-4072-bdc1-e0727f766815 · outbound

This paper cites an unresolved cited work.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:44.221997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:44.221997Z digest=sha256:7b9530c33f8d0af54b398250c026ac02b76bbc073d9105901ccbfcf31e491a45

Observation 65b26389-f98b-4547-aab2-6754cc213b0b · outbound

This paper cites an unresolved cited work.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:44.370193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:44.370193Z digest=sha256:4b1d890c93fb038114bc1a1d59c5cf8d544a3a8294ccce683b0df2d7b3bea159

Observation a4790b78-16e9-45ec-bb8b-b429c0fc16df · outbound

This paper cites an unresolved cited work.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:11:47.759436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T12:11:44.457847Z digest=sha256:3035e2495154bc399f233920599e61a0dc169451ae4bba14a970a45f67c030bf

Observation b9aaa710-1c47-41ed-9dd1-89d68d6435ff · outbound

This paper cites Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:44.548976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:44.548976Z digest=sha256:f9de6424f4f76347fbde662c5ea51a6424cb98358394a4ba8a590fd4faefcc02

Observation fdbd1863-aabd-4658-bf01-a14a595989a2 · outbound

This paper cites an unresolved cited work.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:44.662344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:44.662344Z digest=sha256:9abe8eceae0072bbef950b2d7ffa6049ab90fc616cc0459c1e2768479d5a98ac

Observation b52a0d0b-9606-428e-8a7f-0f4c3a020837 · outbound

This paper cites an unresolved cited work.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:11:47.595645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T12:11:44.725755Z digest=sha256:ba82d7afd269b0ac900561b8f17334213648a023a83062acf9420ea58a2ada62

Observation 25ecfa1a-a068-49e0-beea-9488145e4ca5 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:44.810602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:44.810602Z digest=sha256:9e7f734996cee752f32e5b3daa95e8ca853d1cc5ce7588d1c34937bc64ac02a9

Observation 54b69683-c3c1-46c1-919f-080e91e0a93d · outbound

This paper cites an unresolved cited work.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:44.904533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:44.904533Z digest=sha256:76f0c130ff92a12dfa7381cfcd6dc32c66e87d0c80151a64b4a25f526534ccb4

Observation 20b8ee25-6ca8-437f-a48c-90b04c7aca29 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:44.977819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:44.977819Z digest=sha256:490ca24aa64b7d86a258a00a42452db436bd71ca50f0996694934e6a3c815871

Observation 22ae8d7f-691c-417f-975f-b2e50f5f1b5c · outbound

This paper cites PRMBench: A Fine-grained and Challenging Benchmark for Process-Level Reward Models.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively PRMBench: A Fine-grained and Challenging Benchmark for Process-Level Reward Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:45.058024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:45.058024Z digest=sha256:93095e84e0d70b3cfc5defa857ea9dfa4a95350189d95385bc206bb2d9c898c6

Observation 06ebea2b-a996-450d-90de-d4a4cf82c051 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:45.120617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:45.120617Z digest=sha256:726f5b45bc1b6ef9543ef710fd0e23ae6d7f2a9ea7dbbc4092a0c7cdeec1dfd3

Observation 4dce3f3e-4cef-4567-9e38-923e7a5aaa63 · outbound

This paper cites an unresolved cited work.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:11:47.282226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T12:11:45.183971Z digest=sha256:c5c0dce4c09479f0d1a864905f8c39c3385fe3c146465ef9f2eb3dccc4f0a51d

Observation 58fb2e04-6789-4528-ab20-838d42e59af2 · outbound

This paper cites an unresolved cited work.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:45.260583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:45.260583Z digest=sha256:967834dc75b6708eaec376506ac4a869717770b5a1a875fc905459ffa097fc8d

Observation 00d41d8c-0668-40df-9829-13e3d9fee315 · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:45.365323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:45.365323Z digest=sha256:b2e3b6ddcb704817a2810f5f7ddd1de3dfcfe0126a63b670ca6a295721f1f007

Observation 22c6d246-76e5-403c-be61-4cca2e589cfc · outbound

This paper cites Chain-of-Table: Evolving Tables in the Reasoning Chain for Table Understanding.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Chain-of-Table: Evolving Tables in the Reasoning Chain for Table Understanding

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:45.444335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:45.444335Z digest=sha256:11d01e752bf3824be23169874493cabaa46db5f35e6730223bca36b8a8f79f10

Observation c97218ba-3f3b-4c83-9e10-22a1c2722c86 · outbound

This paper cites an unresolved cited work.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:45.475544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:45.475544Z digest=sha256:b1d3444cb90e6d939f4b9cd54351c5d2cb90f162589fd712aec047fc9881e9f6

Observation 02aadac5-3a10-41c4-ab28-8da1516f555c · outbound

This paper cites an unresolved cited work.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:45.568213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:45.568213Z digest=sha256:fe42c370b9417ccdad43c68991b35e5ffb347365954c3c373b60d5c598d4b9e8

Observation 26fa552f-2b33-4684-9e9a-661ef3c1d765 · outbound

This paper cites an unresolved cited work.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:11:47.096342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T12:11:45.666401Z digest=sha256:fb567baea3d0d4c8c05192d10bc50eeac98d5a393e27cee763ea525c5bded5e4

Observation 7e877f1d-7658-4cd9-8983-0f96efd48657 · outbound

This paper cites Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:45.751564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:45.751564Z digest=sha256:4ab9c9a585eeeb3f4cc327e78dcdc540a865140b2652795afd951e771c0bf020

Observation 77ab271b-cb31-40fb-a81d-147fbdc2ca4a · outbound

This paper cites an unresolved cited work.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:11:46.962735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T12:11:45.857180Z digest=sha256:f271b9ef28328e0c072cb69a7eaf9a5c1588a5e74fcab8f6d239aa121390869e

Observation 5d01445e-2841-406d-808a-45f48acc6633 · outbound

This paper cites Tree of Thoughts: Deliberate Problem Solving with Large Language Models.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Tree of Thoughts: Deliberate Problem Solving with Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:45.942978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:45.942978Z digest=sha256:b18b96992964ee991d0bfce293838800343a31d89414f65ba86dde77def5b98c

Observation fb7fb42d-6cd0-4eca-a422-d9b0e69eddbf · outbound

This paper cites an unresolved cited work.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:46.019554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:46.019554Z digest=sha256:c9694d8612f82c62a7189cddf2c07d664f69e0946da8f32be669018c78ecb6db

Observation 23584c56-45e1-40e6-a1e6-fb8c3a2c20d5 · outbound

This paper cites ToolChain*: Efficient Action Space Navigation in Large Language Models with A* Search.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively ToolChain*: Efficient Action Space Navigation in Large Language Models with A* Search

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:46.100639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:46.100639Z digest=sha256:210ac296704e5dcb35d3a052058db7f9bccb3fcdc5363b3cb8c075c79649dcae

Observation cb77b58f-2502-4520-af81-f354e4335d57 · outbound

This paper cites an unresolved cited work.

Speculative Reward Model Boosts Decision Making Ability of LLMs Cost-Effectively Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:11:46.822688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-07T12:11:46.192441Z digest=sha256:aff2e82f380b6bbfa5d865e5694f4fd894af5b0ca677283e71d50c6f26410226

Pith citing papers

No inbound Pith citation observations are available.