Pith. sign in

Paper Citation Record · LEDGER

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence

As of 18 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 3 inbound Pith citation observations for arXiv:2505.20325.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.20325 v1

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:38:09.669496Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T15:32:34.050519Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T08:16:47.988122Z

Reference resolution

51 of 51 outbound references displayed

  • verified exact0
  • verified fuzzy10
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 620d9cf2-c30c-4480-b918-1b33ce129e91 · outbound

This paper cites GPT-4 Technical Report.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:05.873523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:05.873523Z digest=sha256:e6015fad87cc9b9adc48814f499a9177b71b59099d40dd72f0450fc02e2a3115

Observation 1776254f-df8f-4d91-88df-5b91a25e7582 · outbound

This paper cites Aime 2024, 2024.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Aime 2024, 2024

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:38:12.211214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:38:05.925251Z digest=sha256:f468f65d905675d41e0eef24b12054aef2812b8a1b9f395f6e7e751bcf3ecbdb

Observation 43978ab9-3435-4f45-b887-e6ecaee39293 · outbound

This paper cites Amc 2023, 2024.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Amc 2023, 2024

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:38:12.073545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:38:05.977590Z digest=sha256:6ae02c1a00f681f3af6def5e1eaf72202ba790ac3c183948680ed62ba56146f7

Observation af9197ef-2360-4ca2-9f92-a8fb13d9563d · outbound

This paper cites Qwen2.5-VL Technical Report.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Qwen2.5-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:06.056917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:06.056917Z digest=sha256:212e5ca2f1892e639e86d22269b40eb1cb6245a233c5cc10c41e9b6f0ab8f358

Observation ede02f6e-2793-42b7-8f93-0f11b4156857 · outbound

This paper cites Beeching, L.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Beeching, L

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:38:11.880781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:38:06.142439Z digest=sha256:31955927a31dbd7bb6afb1d1cf4810ed15593b21ae94687d8c2ef1aa7cd3a7ed

Observation b6de529b-cb61-4b4d-ad25-7120d28bc587 · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:06.243955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:06.243955Z digest=sha256:28dd4aa6fbd79eacc9f0ef16b298843473f77197aeebe3ea9dc495f61fd39301

Observation a2f16126-dfd9-4e77-b446-51c3ca062809 · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:06.368130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:06.368130Z digest=sha256:1bf576b1a89294d5263f50050ff59d0a1c96f10597f5739797d27e86fb2166d6

Observation b3b13825-21a9-43d5-b2ea-ec3e2a514f09 · outbound

This paper cites an unresolved cited work.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:38:11.753855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:38:06.447995Z digest=sha256:21132064d1937099f589c8c2b246f1f67f0b6522d7b69efa11f4a5d2f844f76d

Observation 9a803f04-33f8-40ea-92ce-30a8913881d5 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:06.546725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:06.546725Z digest=sha256:ab13d29942fc2a7a06357467e8e545bae02032ba04e751f0037e4aad74210e4d

Observation fd848a76-2d71-49c5-90a7-4add7dba8010 · outbound

This paper cites Hendrycks, C.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Hendrycks, C

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:38:11.591045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:38:06.663440Z digest=sha256:4c04b0885ced6ef9237307c72e43b718e8956dc118c279b0d3e14efc7a5737c2

Observation 091945ec-7752-400b-9747-496389e0deb4 · outbound

This paper cites an unresolved cited work.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:06.744841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:06.744841Z digest=sha256:97f395b0c7815dc44eb492f13b7ef536dfffe0dd5f64770fdb37c9c4a6964c64

Observation c1c4ea3d-8a90-427e-90a0-41bd95cf66ad · outbound

This paper cites Fewer is More: Boosting LLM Reasoning with Reinforced Context Pruning.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Fewer is More: Boosting LLM Reasoning with Reinforced Context Pruning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:06.836586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:06.836586Z digest=sha256:df0656835832b27eeacb932f3c844192174eae694faf53a1d1911c68d19ac3c3

Observation 68e15420-fe97-4a1a-a2a9-2b751259aa77 · outbound

This paper cites A Survey of Test-Time Compute: From Intuitive Inference to Deliberate Reasoning.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence A Survey of Test-Time Compute: From Intuitive Inference to Deliberate Reasoning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:06.844783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:06.844783Z digest=sha256:f0296921907003ffeb29ba9a9dbe4ee677697e6d9b810749fe1c38207ae428e7

Observation f834cc4c-2f21-4129-8a55-66cebada5555 · outbound

This paper cites Lightman, V.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Lightman, V

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:38:11.436050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:38:06.936135Z digest=sha256:88be6b55aabd8d817fb6fac77ddae959a0606ff2789717a462680935101d8664

Observation 0c218e16-140b-455b-92f5-e2e6dc5776f2 · outbound

This paper cites Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:06.969200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:06.969200Z digest=sha256:b686bbd27ec07e16594246ba700e2a7bd5d86df3a8fb8c1673d66270f1bd9df1

Observation 7e0d1284-5ed0-4a2e-b3bf-a284df2d6aa8 · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Understanding R1-Zero-Like Training: A Critical Perspective

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.116566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.116566Z digest=sha256:dd8ba8a14dca5f6fb5fc88cd8b0e79b5d536f21c2b0ec5b33bf361b1d98c5c1c

Observation 130743de-7a1f-4339-9ac6-331caf5fedfe · outbound

This paper cites Loshchilov and F.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Loshchilov and F

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:38:11.287185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:38:07.291431Z digest=sha256:15c50ae35aaf86b592ca3d885756b0b4b76661b4d60b1bc42c3b3a3b18cdde2c

Observation 04ad16e2-6ad2-40bc-bd9e-e6e95281c95a · outbound

This paper cites an unresolved cited work.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:38:11.112908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:38:07.331385Z digest=sha256:e2fbcabb9649d2116bbd8c152701459de5b30a7cd690c416488c345fa523f790

Observation b77e8eb1-cffd-4a6e-8c4c-dc90b07f2190 · outbound

This paper cites ReFT: Reasoning with Reinforced Fine-Tuning.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence ReFT: Reasoning with Reinforced Fine-Tuning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.369948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.369948Z digest=sha256:ac96ab941464800146f0df9f23a10e34192a872281c89471808d46c10a201855

Observation 7b21f4d5-fec1-46d4-9fe4-59f8646c9a14 · outbound

This paper cites s1: Simple test-time scaling.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence s1: Simple test-time scaling

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.442769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.442769Z digest=sha256:d80da5a7c96b40110f647651045865e4dfbad72385debc86895aab00bd000b54

Observation b6bcb2ce-1559-4df0-b147-bd3b38466f04 · outbound

This paper cites Mukherjee, A.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Mukherjee, A

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.486642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.486642Z digest=sha256:212a3278a1f25451996b9371ce9569878b3321289f81714272b9358f13632cbb

Observation 152f48db-b446-4d9c-a8a7-b8d530ba88a4 · outbound

This paper cites Learning to reason with llms, 2024.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Learning to reason with llms, 2024

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.552577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.552577Z digest=sha256:ad2af03acf2069a566dedcbbfff01aeedf91d2ab1b928a4daacbeb9df5afa5ed

Observation 84830fb7-ea87-4631-8a6b-f932f90731e4 · outbound

This paper cites Ouyang, J.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Ouyang, J

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.599879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.599879Z digest=sha256:01e3bebdfda8fed090a91e45df1316d41bfb9030130ae6878f3bb3fc5b5f07e4

Observation 54476c1a-1a20-4f66-be78-1d2390021cb7 · outbound

This paper cites CER: Confidence Enhanced Reasoning in LLMs.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence CER: Confidence Enhanced Reasoning in LLMs

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.670992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.670992Z digest=sha256:2921203e7ae5f9f2a6b57ac86bcbca7544cd86d74a71019a95e3f139d01dbaf4

Observation ec5afd97-7704-4d95-b9e8-6994af35ee7c · outbound

This paper cites Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.712548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.712548Z digest=sha256:24716350bfa6443c4e9e50e83d351ef8976a5c77aed2d8692be4daa2a4b904f8

Observation 146b9d54-b8b3-4aed-ac1b-126f16c69c34 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Proximal Policy Optimization Algorithms

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.767361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.767361Z digest=sha256:062f2d72e43bd7f2b6fd13f4b1137d52afa709e35496d53a1964b06859e2fe37

Observation d421a740-bcc8-454f-bcf0-264db62b192e · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.838485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.838485Z digest=sha256:f37739fab41b5e36f9a153c1cd230bc7e16e03c8cb6bc624a235209db2682010

Observation 8ece51a7-daa5-4db2-ae73-23d767fa5cc0 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.878929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.878929Z digest=sha256:2206a9468b7ba4a46b8417079cb60966f29afb336619ff0e0f89ed606336372c

Observation 676b6f88-109d-4035-9a73-331fb183a32d · outbound

This paper cites Taubenfeld, T.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Taubenfeld, T

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.924312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.924312Z digest=sha256:22c2739f2cba28252a5ab9c80ef13a798c7753958919d1fb32001b08e119ab1c

Observation 90d4aa54-fef9-4c26-875d-7faf4cda9a01 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:07.958180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:07.958180Z digest=sha256:e8b8ed573e8d60135f4a2197a9bedc61af51c2fd86916f797f3408ad8d6dc323

Observation fbcb6a76-e838-4dc0-ad3a-c6b7899f51fc · outbound

This paper cites Will we run out of data? Limits of LLM scaling based on human-generated data.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Will we run out of data? Limits of LLM scaling based on human-generated data

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:08.046459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:08.046459Z digest=sha256:cc4188ba6dd450b0dd0b25141214d84c0f7fc50283a9164443ff1276041fe2ac

Observation 2bd75d16-8eb9-4e98-853f-418f13e659bc · outbound

This paper cites von Werra, Y.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence von Werra, Y

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:38:11.000450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:38:08.179759Z digest=sha256:8fa91edab624dcfac1bead5fe3a83141e16623ff01b340e8819ae32c9986bb21

Observation 9265c321-11c5-4633-99d7-8e7d6836272b · outbound

This paper cites Reasoning Aware Self-Consistency: Leveraging Reasoning Paths for Efficient LLM Sampling.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Reasoning Aware Self-Consistency: Leveraging Reasoning Paths for Efficient LLM Sampling

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:08.271853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:08.271853Z digest=sha256:26a36b827b499e0beffc46ffff2ef2992fa99a199f20a573446a35146b1b627b

Observation 17a1d9d0-06be-4a7e-92ef-e9794a1bc5e6 · outbound

This paper cites OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:08.302465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:08.302465Z digest=sha256:3cfe187091e6ccd639084feaaf3922dcabe1b468b7e0a54db0c1a5a03b54bc1c

Observation 9a200362-c2cc-46b8-a892-b9921d212d6c · outbound

This paper cites Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:08.379040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:08.379040Z digest=sha256:c0bd56c673aa65764dcfc17cda71fa6654e3a17860d4d04744bc9f6cf425cfa5

Observation 06520007-f52c-4fba-a7fc-bbb0050358ae · outbound

This paper cites an unresolved cited work.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:08.450860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:08.450860Z digest=sha256:3a050b4f31f8fa02c8fe0ce1bb703029177261d3b49c36a16423e9a70fbeb6a6

Observation 303e843d-b984-4aa7-952c-3feec4d5ad60 · outbound

This paper cites When More is Less: Understanding Chain-of-Thought Length in LLMs.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence When More is Less: Understanding Chain-of-Thought Length in LLMs

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:08.500777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:08.500777Z digest=sha256:bc1f5deaf8763779b1f0dccbc26ba41bee36455939b4833ad3399f8f2a31068d

Observation f10ada30-92b5-401a-9376-3f3ec6a61296 · outbound

This paper cites an unresolved cited work.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:38:10.832963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:38:08.572256Z digest=sha256:9f2bf17cf1ff55306caa33b8909f70e582acf262d601f56da3cd0d5014c01930

Observation 9bf720c4-d531-484a-9cea-ec5a7b6a8677 · outbound

This paper cites Monte Carlo Tree Search Boosts Reasoning via Iterative Preference Learning.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Monte Carlo Tree Search Boosts Reasoning via Iterative Preference Learning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:08.638944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:08.638944Z digest=sha256:aa3bbe58808be51ee9b92ead4f4f7f8521c5f4a7dbb51e33a479e68bdbe2f32f

Observation b8c10bea-96b1-464f-8913-f695876c3ad7 · outbound

This paper cites Xiong, H.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Xiong, H

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:38:10.708187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:38:08.738098Z digest=sha256:fca4cd23c46b41dacf0301808a91604628fd1cad4507c4ea99898dad76c6ce1b

Observation 5943cc40-aedc-41fc-b156-8170db198f12 · outbound

This paper cites Qwen2.5 Technical Report.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Qwen2.5 Technical Report

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:08.830653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:08.830653Z digest=sha256:1b5be3e61b7d02d1cd4fb4d67dfd85b110fbed9bae7357521c25e1d592321227

Observation a236f04d-547f-45fb-a062-fec848137d57 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:08.920330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:08.920330Z digest=sha256:c095f87cb9ade9089a7717feab7cfd0e87137326b5ff15fc911b0f3242427eda

Observation 39d684a3-7e61-4dd6-ab8d-67099c7dd967 · outbound

This paper cites LIMO: Less is More for Reasoning.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence LIMO: Less is More for Reasoning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:08.997900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:08.997900Z digest=sha256:4f5c288d1298af2be778ffb7984a82a38184a15ca269230853cb498f07aa8715

Observation fde5c451-ff86-4a5d-b62c-d125c5e5dedd · outbound

This paper cites Aime 2025, 2025.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Aime 2025, 2025

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:38:10.550902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:38:09.095172Z digest=sha256:fad2296db2570d601ebcba85ed6aa9ca773ef785077025e6d35f10f202a105e3

Observation 1e79e282-ae10-462a-b77d-59ed721a99a5 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:09.183758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:09.183758Z digest=sha256:2c28df272dfa01f3e5ebab9f27b7886948e591de2a48ba24f54bc241a1e67861

Observation 8e266061-ea17-44d7-83b7-efbca2fd6506 · outbound

This paper cites SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:09.260859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:09.260859Z digest=sha256:895163ffc9cca0ed4e70069ff51a2c5a2803977e74e7f3edf1c2ee46adbc4ee8

Observation 09c651a5-2afd-4e38-9050-cf023f822fa7 · outbound

This paper cites The Lessons of Developing Process Reward Models in Mathematical Reasoning.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence The Lessons of Developing Process Reward Models in Mathematical Reasoning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:09.339406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:09.339406Z digest=sha256:379fdd126768bb0aef516a88c15c052157c808a8eb91f8fdc256314cce2998d7

Observation 389238e8-7248-4a42-9359-e6cc8a6a1af8 · outbound

This paper cites ProcessBench: Identifying Process Errors in Mathematical Reasoning.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence ProcessBench: Identifying Process Errors in Mathematical Reasoning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:09.415393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:09.415393Z digest=sha256:116cee68d0392ba8ee61c717dda53756ad8950cf5d35108d223acc58ba95de81

Observation 7647961d-c5e5-4c55-a8a5-d33959af433d · outbound

This paper cites A Comprehensive Survey of Reward Models: Taxonomy, Applications, Challenges, and Future.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence A Comprehensive Survey of Reward Models: Taxonomy, Applications, Challenges, and Future

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:09.508714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:09.508714Z digest=sha256:99dd93050154fae4ecc8a82bd53e090b9390d69af17c6c1140c555a26f62af39

Observation d1af8ecc-9c5f-4d4c-9b6a-7cede6644f4c · outbound

This paper cites A Survey on Efficient Inference for Large Language Models.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence A Survey on Efficient Inference for Large Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:38:09.583950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:38:09.583950Z digest=sha256:37dc63a8021cdc11ecc5fb4c69f2d1749eb9b570104243b4f4756fced361cac1

Observation 206bae87-fb7d-44bf-9b70-6bb74186a3fd · outbound

This paper cites Final Answer.

Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence Final Answer

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:38:10.478152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:38:09.669496Z digest=sha256:2a523b3154a8c781d4a1b3db6a055d3834ebadac2e900a36c867bcc35c27f003

Pith citing papers

Observation 15ec245d-9310-4ad4-8c1e-9028e1dee8e9 · inbound

Charting the Future of Scholarly Knowledge with AI: A Community Perspective cites this paper.

Charting the Future of Scholarly Knowledge with AI: A Community Perspective Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T15:32:34.050519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:32:34.050519Z digest=sha256:d551c4ebc9931cffca91b32b3b5d1480b4e9335773cd1e64c6c50020ea4296df

Observation 844d56b7-b373-48ef-a857-ded12ee0f99f · inbound

Probing the Difficulty Perception Mechanism of Large Language Models cites this paper.

Probing the Difficulty Perception Mechanism of Large Language Models Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T11:17:47.380695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:17:47.380695Z digest=sha256:32abe560a75e6e2ee0d0f6bcf89457f371b55e81eb9879425b4073efbe0f5e92

Observation bc13135e-6691-44a3-9540-2bc9e08097c0 · inbound

Can We Predict The Human Preference For Text-to-Image Content Prior To Generation And Is It Even Useful To Do So? cites this paper.

Can We Predict The Human Preference For Text-to-Image Content Prior To Generation And Is It Even Useful To Do So? Guided by Gut: Efficient Test-Time Scaling with Reinforced Intrinsic Confidence

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-02T08:16:47.989711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T06:14:02.804743Z digest=sha256:6764757b1ae369bf4196c4dfa308128de94915a8bf8adf3368ba1ffe2cde571a