Pith. sign in

Paper Citation Record · LEDGER

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence

As of 13 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 3 inbound Pith citation observations for arXiv:2505.20325.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.20325 v1

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:38:09.669496Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T15:32:34.050519Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T08:16:47.988122Z

Reference resolution

51 of 51 outbound references displayed

  • verified exact0
  • verified fuzzy10
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 620d9cf2-c30c-4480-b918-1b33ce129e91 · outbound

This paper cites GPT-4 Technical Report.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:05.873523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:05.873523Z digest=sha256:2687411b76a358d13ecfee1050d3050f79c7685c9f511960f7e1dc5bd7dad6cd

Observation 1776254f-df8f-4d91-88df-5b91a25e7582 · outbound

This paper cites Aime 2024, 2024.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Aime 2024, 2024

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:38:12.211214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T14:38:05.925251Z digest=sha256:b4ca32a44b22dd7257a615b2957e5062e728a267c3bdfe433ef5ea657b612e12

Observation 43978ab9-3435-4f45-b887-e6ecaee39293 · outbound

This paper cites Amc 2023, 2024.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Amc 2023, 2024

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:38:12.073545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T14:38:05.977590Z digest=sha256:a2633b681f371f5de35d45ff54bf35dc85528aca609b79ba8706d65bc322a5be

Observation af9197ef-2360-4ca2-9f92-a8fb13d9563d · outbound

This paper cites Qwen2.5-VL Technical Report.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Qwen2.5-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:06.056917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:06.056917Z digest=sha256:51dea205d6431cade3e8a07e8715a1e934483cc71b6f8723a257a0f406d4c78d

Observation ede02f6e-2793-42b7-8f93-0f11b4156857 · outbound

This paper cites Beeching, L.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Beeching, L

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:38:11.880781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T14:38:06.142439Z digest=sha256:7f7c030ef39d604b783302ffaf13862c484173c3b104b6559a52a8d629380415

Observation b6de529b-cb61-4b4d-ad25-7120d28bc587 · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:06.243955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:06.243955Z digest=sha256:96df8269f0718cfce0684bb914540b664de3c740ada7ee8ce5b587382e8718f4

Observation a2f16126-dfd9-4e77-b446-51c3ca062809 · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:06.368130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:06.368130Z digest=sha256:428d89a423c2875dfd496aa6f3db95128e42eb9148fabec0bcee6a23d09dc7ba

Observation b3b13825-21a9-43d5-b2ea-ec3e2a514f09 · outbound

This paper cites an unresolved cited work.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:38:11.753855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T14:38:06.447995Z digest=sha256:b00e2263dc9a71b445f887e64d7bae43c876513d900d116b86fe3e71a6d27d0f

Observation 9a803f04-33f8-40ea-92ce-30a8913881d5 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:06.546725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:06.546725Z digest=sha256:9dd819312c87844f0530941aeb40a2f2c09131c54005cb4d9e2e35c3c9d8df74

Observation fd848a76-2d71-49c5-90a7-4add7dba8010 · outbound

This paper cites Hendrycks, C.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Hendrycks, C

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:38:11.591045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T14:38:06.663440Z digest=sha256:cec834adff9c1537a9869425d66cd06b72210a3a5e9979766e0ab1accdf2642f

Observation 091945ec-7752-400b-9747-496389e0deb4 · outbound

This paper cites an unresolved cited work.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:06.744841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:06.744841Z digest=sha256:4795bdd3c34331a6f1e2eb55b6970cc1451e5d7df5f62d51d2f3810640702f67

Observation c1c4ea3d-8a90-427e-90a0-41bd95cf66ad · outbound

This paper cites Fewer is More: Boosting LLM Reasoning with Reinforced Context Pruning.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Fewer is More: Boosting LLM Reasoning with Reinforced Context Pruning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:06.836586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:06.836586Z digest=sha256:45d46b88017ab7b54f22bd7d0ab748bd3c89add0919d5c3fe4b1bf5e4bc378df

Observation 68e15420-fe97-4a1a-a2a9-2b751259aa77 · outbound

This paper cites A Survey of Test-Time Compute: From Intuitive Inference to Deliberate Reasoning.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence A Survey of Test-Time Compute: From Intuitive Inference to Deliberate Reasoning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:06.844783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:06.844783Z digest=sha256:fe61e1d49b1ec1d974f7dd449dbb160bee502923070e003b3c4c696356f4fa11

Observation f834cc4c-2f21-4129-8a55-66cebada5555 · outbound

This paper cites Lightman, V.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Lightman, V

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:38:11.436050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T14:38:06.936135Z digest=sha256:f67e1b84bffbab89717c19b2624e91ab124cbd42fa49dd58d8bb6b47528dd127

Observation 0c218e16-140b-455b-92f5-e2e6dc5776f2 · outbound

This paper cites Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:06.969200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:06.969200Z digest=sha256:dfcc10c6dfacde894e8a5d40fa743ec9fe275d23db74d58a505c7c5786638693

Observation 7e0d1284-5ed0-4a2e-b3bf-a284df2d6aa8 · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Understanding R1-Zero-Like Training: A Critical Perspective

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.116566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.116566Z digest=sha256:4055acd8e9c1f4010faddb0bb30def02a26ee3a42d105b2a34229a508c3bf5d0

Observation 130743de-7a1f-4339-9ac6-331caf5fedfe · outbound

This paper cites Loshchilov and F.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Loshchilov and F

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:38:11.287185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T14:38:07.291431Z digest=sha256:e88fcd71d3ab9e419351d6a0353c64c4bf58bd8e1bc842e633bf051e2fcaf111

Observation 04ad16e2-6ad2-40bc-bd9e-e6e95281c95a · outbound

This paper cites an unresolved cited work.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:38:11.112908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T14:38:07.331385Z digest=sha256:5c4429f4a6627e52b5f27905bd18b5584296163e7e548ec2fdeddb683e809778

Observation b77e8eb1-cffd-4a6e-8c4c-dc90b07f2190 · outbound

This paper cites ReFT: Reasoning with Reinforced Fine-Tuning.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence ReFT: Reasoning with Reinforced Fine-Tuning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.369948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.369948Z digest=sha256:3168df9533d22285beb20fc462c628eff5998cbffcadae7c0f4c6b967dd730df

Observation 7b21f4d5-fec1-46d4-9fe4-59f8646c9a14 · outbound

This paper cites s1: Simple test-time scaling.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence s1: Simple test-time scaling

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.442769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.442769Z digest=sha256:a42017f57922146e225bb878a706f925917a1ad7a824119ceb8c8332e12ffae8

Observation b6bcb2ce-1559-4df0-b147-bd3b38466f04 · outbound

This paper cites Mukherjee, A.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Mukherjee, A

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.486642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.486642Z digest=sha256:3d455cc933ac18404414e2b98158a0dcb076a0e1490d89f22892a88a1b7e0bcf

Observation 152f48db-b446-4d9c-a8a7-b8d530ba88a4 · outbound

This paper cites Learning to reason with llms, 2024.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Learning to reason with llms, 2024

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.552577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.552577Z digest=sha256:acadcc8b4f7fbcf4d24574059dd395be8c336c485a53dc67f8705c7eadb3900f

Observation 84830fb7-ea87-4631-8a6b-f932f90731e4 · outbound

This paper cites Ouyang, J.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Ouyang, J

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.599879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.599879Z digest=sha256:420065a068fce02499637accd58a79f3a557e676f115ceedd057bd49cbbb3731

Observation 54476c1a-1a20-4f66-be78-1d2390021cb7 · outbound

This paper cites CER: Confidence Enhanced Reasoning in LLMs.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence CER: Confidence Enhanced Reasoning in LLMs

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.670992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.670992Z digest=sha256:79c8e810b92459d58eaec286a2abc9dc7960f21a56ae929f3af619291a5b1ef4

Observation ec5afd97-7704-4d95-b9e8-6994af35ee7c · outbound

This paper cites Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.712548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.712548Z digest=sha256:339c2894b17723086e76bedc741fb055ca07a3bd0851969306f186b65a5782c3

Observation 146b9d54-b8b3-4aed-ac1b-126f16c69c34 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Proximal Policy Optimization Algorithms

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.767361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.767361Z digest=sha256:9604acff7b13fb08d6d176418dadcd6ccaebd2c1520be53639d20e8c3f3dbce5

Observation d421a740-bcc8-454f-bcf0-264db62b192e · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.838485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.838485Z digest=sha256:136e9bc3ef6a800d77905cd1f0d5989a6e0b46b1befc7068f8f3eb8b48f6e9f3

Observation 8ece51a7-daa5-4db2-ae73-23d767fa5cc0 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.878929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.878929Z digest=sha256:95008f2479465228f6295f99d31a1d79ad6f35d419256b63642bf99a15700e61

Observation 676b6f88-109d-4035-9a73-331fb183a32d · outbound

This paper cites Taubenfeld, T.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Taubenfeld, T

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.924312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.924312Z digest=sha256:313cd48fa5b7e0c2904a2d0da27a5b76bb1ed77c1f1fb2372f1800324a34d66f

Observation 90d4aa54-fef9-4c26-875d-7faf4cda9a01 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.958180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.958180Z digest=sha256:462bb3a3c9237a998132a9ae58401a5bbefd6441c9c7503226e44e1134e0e593

Observation fbcb6a76-e838-4dc0-ad3a-c6b7899f51fc · outbound

This paper cites Will we run out of data? Limits of LLM scaling based on human-generated data.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Will we run out of data? Limits of LLM scaling based on human-generated data

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:08.046459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:08.046459Z digest=sha256:cf455ea41fa3ca90c30fb6b2eb07a3dc4d35a7811eed7ef8a5f58dacdbc155fd

Observation 2bd75d16-8eb9-4e98-853f-418f13e659bc · outbound

This paper cites von Werra, Y.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence von Werra, Y

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:38:11.000450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T14:38:08.179759Z digest=sha256:4dde0e873e35fd70b3f28b357c34166840ff5c1fcb8634c7a564f9aab0ea2354

Observation 9265c321-11c5-4633-99d7-8e7d6836272b · outbound

This paper cites Reasoning Aware Self-Consistency: Leveraging Reasoning Paths for Efficient LLM Sampling.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Reasoning Aware Self-Consistency: Leveraging Reasoning Paths for Efficient LLM Sampling

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:08.271853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:08.271853Z digest=sha256:0d93a9f4fe33945f60cdbc73743d1058f391e8419d94ea7a64ef98b9deaf175a

Observation 17a1d9d0-06be-4a7e-92ef-e9794a1bc5e6 · outbound

This paper cites OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:08.302465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:08.302465Z digest=sha256:30dc4c34c263a24b2ed044cd062a88a493ed161f8770c2e91f205ca4eb296626

Observation 9a200362-c2cc-46b8-a892-b9921d212d6c · outbound

This paper cites Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:08.379040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:08.379040Z digest=sha256:5b61b98e6254a08689e6b89c0261de2db2acb41b1dabd0a396cf67b199726713

Observation 06520007-f52c-4fba-a7fc-bbb0050358ae · outbound

This paper cites an unresolved cited work.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:08.450860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:08.450860Z digest=sha256:2249f6a268d6d1997448f9e85b105da3624debeda9046e6c609f0b2b2ea1fa25

Observation 303e843d-b984-4aa7-952c-3feec4d5ad60 · outbound

This paper cites When More is Less: Understanding Chain-of-Thought Length in LLMs.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence When More is Less: Understanding Chain-of-Thought Length in LLMs

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:08.500777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:08.500777Z digest=sha256:d661295c426f8fb0812723d92ac00d6ce8cc07941e33cb2fbacd50ee15b32fed

Observation f10ada30-92b5-401a-9376-3f3ec6a61296 · outbound

This paper cites an unresolved cited work.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:38:10.832963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T14:38:08.572256Z digest=sha256:da522f68a03a76e131ea6e171f6bc4109b4b17bec3db11c507ac092c26560e38

Observation 9bf720c4-d531-484a-9cea-ec5a7b6a8677 · outbound

This paper cites Monte Carlo Tree Search Boosts Reasoning via Iterative Preference Learning.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Monte Carlo Tree Search Boosts Reasoning via Iterative Preference Learning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:08.638944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:08.638944Z digest=sha256:21fbfc3c767bbcfaff1651b4d672b9b9ddf9c8ae31e23519939ce8315395038d

Observation b8c10bea-96b1-464f-8913-f695876c3ad7 · outbound

This paper cites Xiong, H.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Xiong, H

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:38:10.708187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T14:38:08.738098Z digest=sha256:96038ecaee33595c21f6ff6c7e804f5a45d0a6cded9b8da08aa26b6511078791

Observation 5943cc40-aedc-41fc-b156-8170db198f12 · outbound

This paper cites Qwen2.5 Technical Report.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Qwen2.5 Technical Report

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:08.830653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:08.830653Z digest=sha256:43c56849c4d5da1a0804d4df0103e7dc6ff93284314b4df200b61160d13b935f

Observation a236f04d-547f-45fb-a062-fec848137d57 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:08.920330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:08.920330Z digest=sha256:86519f71ee6f74ad3c8416dca93d4103798db94d3c22cd94c23dbe0fbd623306

Observation 39d684a3-7e61-4dd6-ab8d-67099c7dd967 · outbound

This paper cites LIMO: Less is More for Reasoning.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence LIMO: Less is More for Reasoning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:08.997900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:08.997900Z digest=sha256:55e3f1997c19be70925fd996e9fa5c4e677be031fab69f695de7b4761c692c67

Observation fde5c451-ff86-4a5d-b62c-d125c5e5dedd · outbound

This paper cites Aime 2025, 2025.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Aime 2025, 2025

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:38:10.550902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T14:38:09.095172Z digest=sha256:dd522026731865dbfc941a30748460513b9d4801b34dac75e6cbc9a10771af23

Observation 1e79e282-ae10-462a-b77d-59ed721a99a5 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:09.183758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:09.183758Z digest=sha256:5933e6c41a6b5eae26f22e9063055a61c4042cd95890f5ff67dad487fc403ce5

Observation 8e266061-ea17-44d7-83b7-efbca2fd6506 · outbound

This paper cites SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:09.260859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:09.260859Z digest=sha256:1fb053fe3360b3253509ab9ef7f217b76eefbe180ed08c1baafdd0c0ac2800b4

Observation 09c651a5-2afd-4e38-9050-cf023f822fa7 · outbound

This paper cites The Lessons of Developing Process Reward Models in Mathematical Reasoning.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence The Lessons of Developing Process Reward Models in Mathematical Reasoning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:09.339406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:09.339406Z digest=sha256:c565b4f255013e4e2ccbb5189f7632bea4bf57c4e2ce3e15fd757ef277f723eb

Observation 389238e8-7248-4a42-9359-e6cc8a6a1af8 · outbound

This paper cites ProcessBench: Identifying Process Errors in Mathematical Reasoning.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence ProcessBench: Identifying Process Errors in Mathematical Reasoning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:09.415393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:09.415393Z digest=sha256:6107b6f3f292c4de84df834514b4c2a72a4a342cf4b5e8325f1b26f0f3b4ec65

Observation 7647961d-c5e5-4c55-a8a5-d33959af433d · outbound

This paper cites A Comprehensive Survey of Reward Models: Taxonomy, Applications, Challenges, and Future.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence A Comprehensive Survey of Reward Models: Taxonomy, Applications, Challenges, and Future

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:09.508714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:09.508714Z digest=sha256:8c5c997a8e731b1df22df45294b7b3472e2244a88f01b5320bb79e3016eb9cad

Observation d1af8ecc-9c5f-4d4c-9b6a-7cede6644f4c · outbound

This paper cites A Survey on Efficient Inference for Large Language Models.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence A Survey on Efficient Inference for Large Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:09.583950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:09.583950Z digest=sha256:c4a9fecd2fb64b38869e66d1108f59baa6c5847957866c46150cbb71cbe36e48

Observation 206bae87-fb7d-44bf-9b70-6bb74186a3fd · outbound

This paper cites Final Answer.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Final Answer

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:38:10.478152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T14:38:09.669496Z digest=sha256:3f40ba4862d532b7e280bcd81b279d381371c0aa9104790c443c71b66e911e09

Pith citing papers

Observation 15ec245d-9310-4ad4-8c1e-9028e1dee8e9 · inbound

Charting the Future of Scholarly Knowledge with AI: A Community Perspective cites this paper.

Charting the Future of Scholarly Knowledge with AI: A Community Perspective Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T15:32:34.050519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:32:34.050519Z digest=sha256:2b5b4a5465cb84e655c2cc7ac56857d764ca94717efc13043b1da974f1252469

Observation 844d56b7-b373-48ef-a857-ded12ee0f99f · inbound

Probing the Difficulty Perception Mechanism of Large Language Models cites this paper.

Probing the Difficulty Perception Mechanism of Large Language Models Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T11:17:47.380695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:17:47.380695Z digest=sha256:43f1a43ffa3d66d8c660ad14787698825962e4e2581cfbf09cdeaa67a435e597

Observation bc13135e-6691-44a3-9540-2bc9e08097c0 · inbound

Can We Predict The Human Preference For Text-to-Image Content Prior To Generation And Is It Even Useful To Do So? cites this paper.

Can We Predict The Human Preference For Text-to-Image Content Prior To Generation And Is It Even Useful To Do So? Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-02T08:16:47.989711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-28T06:14:02.804743Z digest=sha256:ea81cf321d77c801b28ce9ea82616a5936a7ba8c6198e5508e213e1fa77a23dc