Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:37:23.518356Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 5 inbound Pith citation observations for arXiv:2505.18404.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:37:23.518356Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T19:10:15.553703Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T20:38:55.947620Z
12 of 12 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 0258f757-10db-4251-afb9-63c4cda29bdb · outbound
Thought calibration: Efficient and confident test-time scaling The Llama 3 Herd of Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bb155b5-1a51-4ea5-a88f-1d46bbf4f350 · outbound
Thought calibration: Efficient and confident test-time scaling In The Twelfth International Conference on Learning Representa- tions
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 53af5fe8-d11c-409e-a8dc-794ceffe5efe · outbound
Thought calibration: Efficient and confident test-time scaling Reasoning Models Can Be Effective Without Thinking
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8df164f5-697f-43d6-b5ff-144f8a4bbc79 · outbound
Thought calibration: Efficient and confident test-time scaling Calibrating Reasoning in Language Models with Internal Consistency
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d164c14a-5f06-4ae7-9b61-5609356657d7 · outbound
Thought calibration: Efficient and confident test-time scaling Qwen2.5 Technical Report
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4a2d088-2575-4444-8f3d-fda0fc9cc30f · outbound
Thought calibration: Efficient and confident test-time scaling lmdeploy natively supports the saving of last layer representations, so it was used for almost all experiments
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4f531f36-c6ba-4abe-8755-b4012a6a8f52 · outbound
Thought calibration: Efficient and confident test-time scaling Due to computational constraints, we report the mean over a single run
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f76a5ef6-a1d7-49d5-83ee-23c09141ce55 · outbound
Thought calibration: Efficient and confident test-time scaling Learn then Test: Calibrating Predictive Algorithms to Achieve Risk Control
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0b6a32a-9223-464a-bf6b-cb6ed13451a9 · outbound
Thought calibration: Efficient and confident test-time scaling Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9f1127d-cd47-4510-9e02-350d7df22f95 · outbound
Thought calibration: Efficient and confident test-time scaling In International Conference on Machine Learning, pages 19274–19286
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6cbbd8c0-6c99-4e97-a9e2-67089320a0a4 · outbound
Thought calibration: Efficient and confident test-time scaling Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fde1f4ae-9b61-40ed-9d38-0bf9b2fe0258 · outbound
Thought calibration: Efficient and confident test-time scaling in-progress
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bb434ad-6b19-4129-a9c9-0794971002ac · inbound
E-valuator: Reliable Agent Verifiers with Sequential Hypothesis Testing Thought calibration: Efficient and confident test-time scaling
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e3cfcf8-3ebb-458c-b49d-836e39674a95 · inbound
Reliable LLM-Based Edge-Cloud-Expert Cascades for Telecom Knowledge Systems Thought calibration: Efficient and confident test-time scaling
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6c06d2d7-4b16-4f51-88ef-f9fff249c9ff · inbound
When Should an AI Workflow Release? Always-Valid Inference for Black-Box Generate-Verify Systems Thought calibration: Efficient and confident test-time scaling
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d9788a92-8813-4e2e-bb27-ca962f2faa01 · inbound
From Reasoning Traces to Reusable Modules: Understanding Compositional Generalization in Language Model Reasoning Thought calibration: Efficient and confident test-time scaling
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 389dc7f3-29dc-46cf-aa59-77823004f54a · inbound
Think Short, Defer Smart, Act, and Repeat: Calibrated Reasoning and Uncertainty-Aware Deferral for Edge LLM Agents Thought calibration: Efficient and confident test-time scaling
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.