Pith. sign in

Paper Citation Record · LEDGER

Logical Reasoning with Outcome Reward Models for Test-Time Scaling

As of 8 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 0 inbound Pith citation observations for arXiv:2508.19903.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.19903 v1

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T15:26:52.816300Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

23 of 23 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b7087924-0729-49ef-a3bf-8212c0530578 · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.728328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.728328Z digest=sha256:151d0cd340af3062c8e3b6bdd588f739955464721dce165ff109d32a1a259537

Observation c3056f3b-8e04-49fa-acf3-cd2deee7a1bb · outbound

This paper cites JustLogic: A Comprehensive Benchmark for Evaluating Deductive Reasoning in Large Language Models.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling JustLogic: A Comprehensive Benchmark for Evaluating Deductive Reasoning in Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.732798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.732798Z digest=sha256:b5a4015d43a582d838e4d394e2343ad70a611496a4eb3b2cb60dd9f166d87a87

Observation 40436993-5ee3-41b0-853b-afd8acce21bd · outbound

This paper cites The Llama 3 Herd of Models.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling The Llama 3 Herd of Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.736892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.736892Z digest=sha256:adc1ce20701fe38ad3ea57b00f68be00da5291dbb204c6e925ccf3c1a931bac7

Observation 053b84d0-a2eb-4912-94f8-93da3eb77f65 · outbound

This paper cites Fabbri, Wojciech Kryscinski, Semih Yavuz, Ye Liu, Xi Victoria Lin, Shafiq Joty, Yingbo Zhou, Caiming Xiong, Rex Ying, Arman Cohan, and Dragomir Radev.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Fabbri, Wojciech Kryscinski, Semih Yavuz, Ye Liu, Xi Victoria Lin, Shafiq Joty, Yingbo Zhou, Caiming Xiong, Rex Ying, Arman Cohan, and Dragomir Radev

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:26:53.093107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T15:26:52.741046Z digest=sha256:2dc29fc0d865a8ecc40ea5a98c8c762efe1e1918f03605cb7912fd423eb88df7

Observation 1bacd92f-c50c-450c-9dc0-4e92fdbacd45 · outbound

This paper cites an unresolved cited work.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.744805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.744805Z digest=sha256:3f4f84cde88a429a000453ddea297cc8472f659e4ad5068780479b7eb35296b2

Observation 67afe112-f1a2-483e-a564-6376c518eee4 · outbound

This paper cites Qwen2.5-Coder Technical Report.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Qwen2.5-Coder Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.748712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.748712Z digest=sha256:fe09890adfbeaa77a16a73cf098224e478c7f03fc307d1705e8a935d5db4309b

Observation 39104c15-7d71-463c-9d02-66be191b2844 · outbound

This paper cites Gemma 3 Technical Report.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Gemma 3 Technical Report

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.752965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.752965Z digest=sha256:e201e066a42d38cf7692a92b9587e8f42aed37e92cd183597b39ea3a3ca1d12d

Observation 37c1c938-6b4f-4244-8901-a6f83db59867 · outbound

This paper cites A Closer Look at Logical Reasoning with LLMs: The Choice of Tool Matters.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling A Closer Look at Logical Reasoning with LLMs: The Choice of Tool Matters

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.757698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.757698Z digest=sha256:54d1b778f5fbde3fffa028bc21ba9d24e025facb01ea042a9804fad7669b005f

Observation b0b8ccb1-fad0-4035-9e9e-14ad90696a83 · outbound

This paper cites START: Self-taught Reasoner with Tools.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling START: Self-taught Reasoner with Tools

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.761394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.761394Z digest=sha256:040e10595e8aa3ea23e544ddf541fcabbce362326a01746b7acebc2a20639258

Observation f402c349-5e83-409c-ac41-8e1fd96a9448 · outbound

This paper cites an unresolved cited work.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.765584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.765584Z digest=sha256:ddb79253e74836f8f295d43290ac850362f91e613290d4ae1ffb730a722bc639

Observation cdac56e0-1ff4-4ba7-8015-06b60a9d04a6 · outbound

This paper cites Zhang, Armando Solar - Lezama, Joshua B.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Zhang, Armando Solar - Lezama, Joshua B

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.768976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.768976Z digest=sha256:2d2236d324d82285b2f70d5fbef9b7e313fa182a70234fd4d79cb9ca30634231

Observation e003ffb1-1682-48dd-8b3c-14cc9c6be5f2 · outbound

This paper cites an unresolved cited work.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.773005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.773005Z digest=sha256:8c37776612ec3a32531663385b9e9aa369a4f2a4878d3d10adf0e35517552972

Observation ce3e289a-d08b-4d3d-8d3f-f37884a6f23d · outbound

This paper cites Large Language Models Meet Symbolic Provers for Logical Reasoning Evaluation.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Large Language Models Meet Symbolic Provers for Logical Reasoning Evaluation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.776686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.776686Z digest=sha256:44a2791f93525e2ab54bc0127349f1934e03b50143baa4d86928c2463b20519b

Observation 1031d570-310b-4458-acf8-86cdf3ff0bc1 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.780303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.780303Z digest=sha256:cfb33652a41e3cd58e7ef5836d66b5b1e7613793adf6cf0b3c2787460c10e86b

Observation 364cffa8-0ada-4631-844e-ad15dfe307d4 · outbound

This paper cites Strategies for Improving NL-to-FOL Translation with LLMs: Data Generation, Incremental Fine-Tuning, and Verification.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Strategies for Improving NL-to-FOL Translation with LLMs: Data Generation, Incremental Fine-Tuning, and Verification

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.784006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.784006Z digest=sha256:74931291b51c88a41a9f0aa819cfce0f1bf731e136ab1ed2bc155eb61c576ee2

Observation b98684d4-f6b6-4466-a858-9772f8c2be91 · outbound

This paper cites Solving math word problems with process- and outcome-based feedback.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Solving math word problems with process- and outcome-based feedback

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.787792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.787792Z digest=sha256:e16c43fd0842a7d3c97fe9a5b1632c540789cc750cb7aba97f3c868bd3d22b8e

Observation 353a5442-2b9c-4ae6-824e-3574bcc1efc0 · outbound

This paper cites an unresolved cited work.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.791295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.791295Z digest=sha256:557925f3b3505276970f6cc682eb0865fa4658ac546ec32cf9912fb0e616af9b

Observation 3920aab4-2a9c-41be-80ec-c2477fbf92bd · outbound

This paper cites an unresolved cited work.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.794880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.794880Z digest=sha256:8eaf9297f323c8d1be6676f2379383037dee1bcdfacda7e80ab485af1ca15020

Observation 972dbe22-2d59-4f61-8e6b-b249ffb6da92 · outbound

This paper cites Qwen3 Technical Report.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Qwen3 Technical Report

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.798870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.798870Z digest=sha256:d48d06cabf125cdccbc5c92c3758be30916ef841e2c77c4abb8fa9cd5d035b42

Observation ea7e3f4b-556a-409a-88cd-7619fc670b71 · outbound

This paper cites an unresolved cited work.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:26:53.054868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-05T15:26:52.802869Z digest=sha256:58e3417c7492acfbf161ccb668f2b75a9939fc4b4f1084ff9f45c04a0c60ea38

Observation eb39a6cd-e778-4ec0-b42a-f65491dbd8f1 · outbound

This paper cites The Lessons of Developing Process Reward Models in Mathematical Reasoning.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling The Lessons of Developing Process Reward Models in Mathematical Reasoning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.807424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.807424Z digest=sha256:08ce1ff880863700f1d1b858462bf7b9622a5e559545dbcc78b8c1bbf8f74379

Observation 6bc34e11-3105-44b6-a9cb-79f84048d237 · outbound

This paper cites online" 'onlinestring :=.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling online" 'onlinestring :=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.811818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.811818Z digest=sha256:cff783a42cbbca79cfbd95ba77555d44ca4a4a7170fa16df876b4c669fc79d40

Observation ba7e1584-de20-4324-abec-2241b515dea5 · outbound

This paper cites write newline.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling write newline

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.816300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.816300Z digest=sha256:d8b74a96fc9db69b5386aa5e2c8e22a13cd86a9a8754da12d1ce721f58abb1a5

Pith citing papers

No inbound Pith citation observations are available.