Pith. sign in

Paper Citation Record · LEDGER

Logical Reasoning with Outcome Reward Models for Test-Time Scaling

As of 17 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 0 inbound Pith citation observations for arXiv:2508.19903.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.19903 v1

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T15:26:52.816300Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

23 of 23 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b7087924-0729-49ef-a3bf-8212c0530578 · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.728328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.728328Z digest=sha256:49042194b9e742dedb723b5165cde6c09178101397c23fa35aa4881025a280b4

Observation c3056f3b-8e04-49fa-acf3-cd2deee7a1bb · outbound

This paper cites JustLogic: A Comprehensive Benchmark for Evaluating Deductive Reasoning in Large Language Models.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling JustLogic: A Comprehensive Benchmark for Evaluating Deductive Reasoning in Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.732798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.732798Z digest=sha256:e256b888f40873e350eeca9ebece32ee191fbc23c775d917533ab492002b88fd

Observation 40436993-5ee3-41b0-853b-afd8acce21bd · outbound

This paper cites The Llama 3 Herd of Models.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling The Llama 3 Herd of Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.736892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.736892Z digest=sha256:ee07afbe82c7e4c54fb1566f0906537fdfb60b06a1f62cf58c3a7ac427e7bc75

Observation 053b84d0-a2eb-4912-94f8-93da3eb77f65 · outbound

This paper cites Fabbri, Wojciech Kryscinski, Semih Yavuz, Ye Liu, Xi Victoria Lin, Shafiq Joty, Yingbo Zhou, Caiming Xiong, Rex Ying, Arman Cohan, and Dragomir Radev.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Fabbri, Wojciech Kryscinski, Semih Yavuz, Ye Liu, Xi Victoria Lin, Shafiq Joty, Yingbo Zhou, Caiming Xiong, Rex Ying, Arman Cohan, and Dragomir Radev

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:26:53.093107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-05T15:26:52.741046Z digest=sha256:37fe61f4411ef402a8580095920c482625c1d55c436f479d7c1b9049bfae1446

Observation 1bacd92f-c50c-450c-9dc0-4e92fdbacd45 · outbound

This paper cites an unresolved cited work.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.744805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.744805Z digest=sha256:3f4f84cde88a429a000453ddea297cc8472f659e4ad5068780479b7eb35296b2

Observation 67afe112-f1a2-483e-a564-6376c518eee4 · outbound

This paper cites Qwen2.5-Coder Technical Report.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Qwen2.5-Coder Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.748712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.748712Z digest=sha256:fe09890adfbeaa77a16a73cf098224e478c7f03fc307d1705e8a935d5db4309b

Observation 39104c15-7d71-463c-9d02-66be191b2844 · outbound

This paper cites Gemma 3 Technical Report.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Gemma 3 Technical Report

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.752965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.752965Z digest=sha256:e201e066a42d38cf7692a92b9587e8f42aed37e92cd183597b39ea3a3ca1d12d

Observation 37c1c938-6b4f-4244-8901-a6f83db59867 · outbound

This paper cites A Closer Look at Logical Reasoning with LLMs: The Choice of Tool Matters.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling A Closer Look at Logical Reasoning with LLMs: The Choice of Tool Matters

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.757698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.757698Z digest=sha256:34733ca48f2c428394be06d85ce4b3bd672b8088f15fcdcb26947c326a9bb428

Observation b0b8ccb1-fad0-4035-9e9e-14ad90696a83 · outbound

This paper cites START: Self-taught Reasoner with Tools.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling START: Self-taught Reasoner with Tools

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.761394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.761394Z digest=sha256:16357bf7c81f44099fbc1cc06cc614763a9c39afde4e921d65af652aba679182

Observation f402c349-5e83-409c-ac41-8e1fd96a9448 · outbound

This paper cites an unresolved cited work.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.765584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.765584Z digest=sha256:ddb79253e74836f8f295d43290ac850362f91e613290d4ae1ffb730a722bc639

Observation cdac56e0-1ff4-4ba7-8015-06b60a9d04a6 · outbound

This paper cites Zhang, Armando Solar - Lezama, Joshua B.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Zhang, Armando Solar - Lezama, Joshua B

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.768976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.768976Z digest=sha256:2d2236d324d82285b2f70d5fbef9b7e313fa182a70234fd4d79cb9ca30634231

Observation e003ffb1-1682-48dd-8b3c-14cc9c6be5f2 · outbound

This paper cites an unresolved cited work.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.773005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.773005Z digest=sha256:8c37776612ec3a32531663385b9e9aa369a4f2a4878d3d10adf0e35517552972

Observation ce3e289a-d08b-4d3d-8d3f-f37884a6f23d · outbound

This paper cites Large Language Models Meet Symbolic Provers for Logical Reasoning Evaluation.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Large Language Models Meet Symbolic Provers for Logical Reasoning Evaluation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.776686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.776686Z digest=sha256:b65948c9853f2438d2b184d4ac42576549eb6cf19e41ce73da69dfb459eac729

Observation 1031d570-310b-4458-acf8-86cdf3ff0bc1 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.780303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.780303Z digest=sha256:cfb33652a41e3cd58e7ef5836d66b5b1e7613793adf6cf0b3c2787460c10e86b

Observation 364cffa8-0ada-4631-844e-ad15dfe307d4 · outbound

This paper cites Strategies for Improving NL-to-FOL Translation with LLMs: Data Generation, Incremental Fine-Tuning, and Verification.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Strategies for Improving NL-to-FOL Translation with LLMs: Data Generation, Incremental Fine-Tuning, and Verification

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.784006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.784006Z digest=sha256:354362e29dd470a98d24200e2099aec054c05580907d202d16d58b29437c1d60

Observation b98684d4-f6b6-4466-a858-9772f8c2be91 · outbound

This paper cites Solving math word problems with process- and outcome-based feedback.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Solving math word problems with process- and outcome-based feedback

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.787792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.787792Z digest=sha256:e16c43fd0842a7d3c97fe9a5b1632c540789cc750cb7aba97f3c868bd3d22b8e

Observation 353a5442-2b9c-4ae6-824e-3574bcc1efc0 · outbound

This paper cites an unresolved cited work.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.791295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.791295Z digest=sha256:557925f3b3505276970f6cc682eb0865fa4658ac546ec32cf9912fb0e616af9b

Observation 3920aab4-2a9c-41be-80ec-c2477fbf92bd · outbound

This paper cites an unresolved cited work.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.794880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.794880Z digest=sha256:8eaf9297f323c8d1be6676f2379383037dee1bcdfacda7e80ab485af1ca15020

Observation 972dbe22-2d59-4f61-8e6b-b249ffb6da92 · outbound

This paper cites Qwen3 Technical Report.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Qwen3 Technical Report

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.798870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.798870Z digest=sha256:d48d06cabf125cdccbc5c92c3758be30916ef841e2c77c4abb8fa9cd5d035b42

Observation ea7e3f4b-556a-409a-88cd-7619fc670b71 · outbound

This paper cites an unresolved cited work.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:26:53.054868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-05T15:26:52.802869Z digest=sha256:1835235b3e721e82ad095ccb2cb01c7ff44e4a7d310e69bcebc2645c88e679ec

Observation eb39a6cd-e778-4ec0-b42a-f65491dbd8f1 · outbound

This paper cites The Lessons of Developing Process Reward Models in Mathematical Reasoning.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling The Lessons of Developing Process Reward Models in Mathematical Reasoning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.807424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.807424Z digest=sha256:08ce1ff880863700f1d1b858462bf7b9622a5e559545dbcc78b8c1bbf8f74379

Observation 6bc34e11-3105-44b6-a9cb-79f84048d237 · outbound

This paper cites online" 'onlinestring :=.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling online" 'onlinestring :=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.811818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.811818Z digest=sha256:cff783a42cbbca79cfbd95ba77555d44ca4a4a7170fa16df876b4c669fc79d40

Observation ba7e1584-de20-4324-abec-2241b515dea5 · outbound

This paper cites write newline.

Logical Reasoning with Outcome Reward Models for Test-Time Scaling write newline

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T15:26:52.816300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:26:52.816300Z digest=sha256:d8b74a96fc9db69b5386aa5e2c8e22a13cd86a9a8754da12d1ce721f58abb1a5

Pith citing papers

No inbound Pith citation observations are available.