Pith. sign in

Paper Citation Record · LEDGER

Efficient Test-Time Scaling via Self-Calibration

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 27 inbound Pith citation observations for arXiv:2503.00031.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.00031 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 27 of 27 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:34:53.349405Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T14:28:31.655268Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 96a6fbb1-fe64-4705-8fbe-c08fa99344ae · inbound

Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models cites this paper.

Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models Efficient Test-Time Scaling via Self-Calibration

Reference 283

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:40:41.663278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-12T08:40:40.910461Z digest=sha256:4b8e8376646513e734ad279d55f90b255360bc0b531809c72a9d59dd6e160231

Observation be9a2fb3-d48e-407b-beb2-bc0a6a27c833 · inbound

Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models cites this paper.

Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models Efficient Test-Time Scaling via Self-Calibration

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-05-14T01:29:56.656462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-14T01:29:56.480020Z digest=sha256:7fee667ddfb951531ce06ec66ddc6a5ec17082c166c8bb2b036cb3d8eb2c9c35

Observation 20d67b80-af51-4d82-8d60-d8e885ce2f81 · inbound

General-Reasoner: Advancing LLM Reasoning Across All Domains cites this paper.

General-Reasoner: Advancing LLM Reasoning Across All Domains Efficient Test-Time Scaling via Self-Calibration

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:34:08.320486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:34:08.320486Z digest=sha256:df56ccbff7e8aa8b2b5bd58b4804a932736c1b0fbb06532d16979b65a483936d

Observation 02d160ea-3ef4-4150-b529-792935c7a446 · inbound

UniGen: Enhanced Training & Test-Time Strategies for Unified Multimodal Understanding and Generation cites this paper.

UniGen: Enhanced Training & Test-Time Strategies for Unified Multimodal Understanding and Generation Efficient Test-Time Scaling via Self-Calibration

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:34:53.349405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:34:53.349405Z digest=sha256:e65c4dd93d677b58903181ad1ffbc3018bbb6924978dc83957c8770714b6a2ff

Observation 895124a2-d20a-4fc3-8433-9edc9e6e7bd9 · inbound

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning cites this paper.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Efficient Test-Time Scaling via Self-Calibration

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.220177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.220177Z digest=sha256:5cff44b753a09be772559780f519a53077559b716002fc97ae36aa9756f98ddf

Observation a5f3239e-5b48-4445-a825-0b4914e56a75 · inbound

Generalizable LLM Learning of Graph Synthetic Data with Post-training Alignment cites this paper.

Generalizable LLM Learning of Graph Synthetic Data with Post-training Alignment Efficient Test-Time Scaling via Self-Calibration

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:02:42.612503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:02:42.612503Z digest=sha256:ef31ffd0e662a5b49719a112fe7071df8bf4f1377cc1e4cc0378824520dab4a3

Observation 63d4ab6a-717c-4624-ac71-f853ca80a07d · inbound

POSS: Position Specialist Generates Better Draft for Speculative Decoding cites this paper.

POSS: Position Specialist Generates Better Draft for Speculative Decoding Efficient Test-Time Scaling via Self-Calibration

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:04:15.687034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:04:15.687034Z digest=sha256:a38b5c9b43a6f6d29a1d759f566c74d1b65c27d63ccf0edbd663152886d24e81

Observation 92faf0a8-b416-4acc-8d08-63636b1f2689 · inbound

AdaDecode: Accelerating LLM Decoding with Adaptive Layer Parallelism cites this paper.

AdaDecode: Accelerating LLM Decoding with Adaptive Layer Parallelism Efficient Test-Time Scaling via Self-Calibration

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:02:39.251956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:02:39.251956Z digest=sha256:85884830f56cbec86f21d9b242d08e5fed28407df4cfce25824ad0ac6c95cfba

Observation 8dc60cc1-3c86-4670-a257-916140911fbe · inbound

Revisiting Test-Time Scaling: A Survey and a Diversity-Aware Method for Efficient Reasoning cites this paper.

Revisiting Test-Time Scaling: A Survey and a Diversity-Aware Method for Efficient Reasoning Efficient Test-Time Scaling via Self-Calibration

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T10:42:39.758921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:42:39.758921Z digest=sha256:7b8e4394ed3a26b63246599e6f75ba2c776c140ce081afae30479ac7c4729361

Observation 96c1265e-c539-45b8-b150-55341afd6eac · inbound

How Far Are We from Optimal Reasoning Efficiency? cites this paper.

How Far Are We from Optimal Reasoning Efficiency? Efficient Test-Time Scaling via Self-Calibration

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:49:36.050461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:49:36.050461Z digest=sha256:b2600f593450c41f62eee8f157cf50a3344a50cb8e7b17266940eecb19f2df38

Observation f612e9d5-3c2c-4b20-9423-4174e9aa716c · inbound

AALC: Large Language Model Efficient Reasoning via Adaptive Accuracy-Length Control cites this paper.

AALC: Large Language Model Efficient Reasoning via Adaptive Accuracy-Length Control Efficient Test-Time Scaling via Self-Calibration

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:00:45.940404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:00:45.940404Z digest=sha256:bf14548ea6c392333f73611a2ecb36010c207a1dfa6316ff8cd8f9c7228b2756

Observation 12a18604-f904-4086-80b3-652a7e821b24 · inbound

Reasoning on a Budget: A Survey of Adaptive and Controllable Test-Time Compute in LLMs cites this paper.

Reasoning on a Budget: A Survey of Adaptive and Controllable Test-Time Compute in LLMs Efficient Test-Time Scaling via Self-Calibration

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T20:43:10.487393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:43:10.487393Z digest=sha256:2c41d374fd11d424985a9b035e67e0551849af87efec0761f1116ca6cd72ff94

Observation e56fe328-acc4-439b-a8f9-3d99fa800486 · inbound

Confidence-Weighted Token Set Cover for Early Hypothesis Pruning in Self-Consistency cites this paper.

Confidence-Weighted Token Set Cover for Early Hypothesis Pruning in Self-Consistency Efficient Test-Time Scaling via Self-Calibration

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T01:02:45.824588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T01:02:45.824588Z digest=sha256:354ab91b78b19ca02425754314d6bf8187c1c9fede43161dd048be2d50f84150

Observation d60036a1-5e84-4f5c-a3d6-b9a9c62b5162 · inbound

Parallel-R1: Towards Parallel Thinking via Reinforcement Learning cites this paper.

Parallel-R1: Towards Parallel Thinking via Reinforcement Learning Efficient Test-Time Scaling via Self-Calibration

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T21:28:49.143122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:28:49.143122Z digest=sha256:527eaaeeaeac1c459c9cfd9847dbc781a09f99763a3d5a11a7e7fc8223462351

Observation 1558b01b-1e7e-4872-b6a6-17c016193889 · inbound

GrACE: A Generative Approach to Better Confidence Elicitation and Efficient Test-Time Scaling in Large Language Models cites this paper.

GrACE: A Generative Approach to Better Confidence Elicitation and Efficient Test-Time Scaling in Large Language Models Efficient Test-Time Scaling via Self-Calibration

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-18T17:56:42.072706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-18T17:54:18.515174Z digest=sha256:2339e22438de6b54975697d2d687087e73cfdf84f4ac9d02f2d75b5563e7ee88

Observation 422fc5c0-9399-48f9-991c-e8a99332e346 · inbound

CGES: Confidence-Guided Early Stopping for Efficient and Accurate Self-Consistency cites this paper.

CGES: Confidence-Guided Early Stopping for Efficient and Accurate Self-Consistency Efficient Test-Time Scaling via Self-Calibration

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T00:12:31.820038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T00:12:31.820038Z digest=sha256:3801f6169d184459c2b605275ee1231276c93b3ae079a7108878595b545de75f

Observation 23fb203e-c842-48a3-bbe3-dc713987d2b4 · inbound

Process Supervision of Confidence Margin for Calibrated LLM Reasoning cites this paper.

Process Supervision of Confidence Margin for Calibrated LLM Reasoning Efficient Test-Time Scaling via Self-Calibration

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:41:12.347697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-08T08:19:09.437464Z digest=sha256:bcd7185b59077cfe4cdc9c08518c050de8ec8f2857eecd631a65478388bf7a1b

Observation b41f84ed-1987-41e2-84f0-8b6ea6c0a639 · inbound

VLA-ATTC: Adaptive Test-Time Compute for VLA Models with Relative Action Critic Model cites this paper.

VLA-ATTC: Adaptive Test-Time Compute for VLA Models with Relative Action Critic Model Efficient Test-Time Scaling via Self-Calibration

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:46:06.462686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-09T15:10:16.533927Z digest=sha256:d8b6da454cb4834dc3634851d997a2537a17aa7718b24057d24dc3223e061d94

Observation 66fd0a2e-5571-4f53-b377-56a66b7add14 · inbound

Distilling Long-CoT Reasoning through Collaborative Step-wise Multi-Teacher Decoding cites this paper.

Distilling Long-CoT Reasoning through Collaborative Step-wise Multi-Teacher Decoding Efficient Test-Time Scaling via Self-Calibration

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:31:08.833359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-09T16:29:05.186607Z digest=sha256:3283409caef16e2816073954b20a5f7ee9893e8511dbfc78898c65b8fbf569c8

Observation 42b55634-9bf9-4074-9845-4d76bfe19f7d · inbound

Nonsense Helps: Prompt Space Perturbation Broadens Reasoning Exploration cites this paper.

Nonsense Helps: Prompt Space Perturbation Broadens Reasoning Exploration Efficient Test-Time Scaling via Self-Calibration

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T19:21:09.011587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-08T12:08:48.212651Z digest=sha256:b98a9940eb9b3f99287d29d6bbbef95711e73b1ecd620f0759c3c716ba80219f

Observation 83933e4c-2d3f-4935-a735-c353361ea082 · inbound

Post Reasoning: Improving the Performance of Non-Thinking Models at No Cost cites this paper.

Post Reasoning: Improving the Performance of Non-Thinking Models at No Cost Efficient Test-Time Scaling via Self-Calibration

Reference 179

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:06:09.556971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-08T10:19:08.451445Z digest=sha256:21f47d4ed2e0f670640d09d8c5afcc73844d64699447d237cafe942ea2a69977

Observation d2deb387-d908-4b11-94c4-f2656c1a8dc6 · inbound

LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling cites this paper.

LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling Efficient Test-Time Scaling via Self-Calibration

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:10:58.250352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-11T01:54:34.000216Z digest=sha256:da5fdd003f51bd790ba0aa8c580b476cc96640d493bcdabe3371f8649fa81a39

Observation c71ed7e6-6bff-4d6d-b46f-e78ac34508fe · inbound

LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling cites this paper.

LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling Efficient Test-Time Scaling via Self-Calibration

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:12:28.402145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-13T07:09:02.672233Z digest=sha256:c8d5196c64065775af3b9ee7526d2a4eab46cc8d2ccfc64006843856848a9d53

Observation 7698f45e-2618-407d-bd2e-874218dadc7d · inbound

Process Rewards with Learned Reliability cites this paper.

Process Rewards with Learned Reliability Efficient Test-Time Scaling via Self-Calibration

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-19T14:53:06.951817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-19T14:51:13.966538Z digest=sha256:30d376c2e8c014f2f8eb542b575382266862f360425b6c941f08a891320676e2

Observation 3d1a91bb-1c87-455b-a51b-fefc131f1ea3 · inbound

Small RL Controller, Large Language Model: RL-Guided Adaptive Sampling for Test-Time Scaling cites this paper.

Small RL Controller, Large Language Model: RL-Guided Adaptive Sampling for Test-Time Scaling Efficient Test-Time Scaling via Self-Calibration

Reference 80

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:56:29.829947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-28T10:25:10.559953Z digest=sha256:6388a3713c2761c98f49411d78e290345f685dcee977223fa03884521dc673ca

Observation 45db5ef9-c6be-45eb-88f0-933850611de1 · inbound

MARS: Margin-Adversarial Risk-controlled Stopping for Parallel LLM Test-time Scaling cites this paper.

MARS: Margin-Adversarial Risk-controlled Stopping for Parallel LLM Test-time Scaling Efficient Test-Time Scaling via Self-Calibration

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:28:31.656816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-27T07:03:01.918769Z digest=sha256:324e0803cf7159dd0fe35f9639d1c4ad773c4fceca8e9493574b7e02811267dc

Observation a02354e8-4d9f-4555-8fd7-b02ee024a2d4 · inbound

Interpretable Adaptive Sampling for LLM Test-Time Scaling cites this paper.

Interpretable Adaptive Sampling for LLM Test-Time Scaling Efficient Test-Time Scaling via Self-Calibration

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T05:02:10.749359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:02:10.749359Z digest=sha256:d398e9985b79a9d72f6fb6d03e8b6c7034e47597dec617a18053d0413ad79f14