Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-13T13:23:25.998584Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 1 inbound Pith citation observation for arXiv:2605.23909.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-13T13:23:25.998584Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T22:46:02.023365Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
26 of 26 outbound references displayed
External citation measurements
0
pith, observed 2026-08-05T02:28:24.338817Z
Observation 3bb1368b-8637-42bd-8fea-4bd2c46109f8 · outbound
Confidence Calibration in Large Language Models A Review of Uncertainty Quantification in Deep Learning: Techniques, Applications and Challenges
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f17d7937-dc3a-4ff7-acbc-1a3d5286361f · outbound
Confidence Calibration in Large Language Models Christopher Clark, Kenton Lee, Ming-Wei Chang, Tom Kwiatkowski, Michael Collins, and Kristina Toutanova
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed143514-84d3-438f-9bbe-90b5e7a1ef07 · outbound
Confidence Calibration in Large Language Models BoolQ: Exploring the Surprising Difficulty of Natural Yes/No Questions
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a1ef2c0-ac61-409b-a64e-69803af5b3c1 · outbound
Confidence Calibration in Large Language Models Do LLMs Implicitly Determine the Suitable Text Difficulty for Users?
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e07311e7-fc2b-46db-b7b9-d8fe1c409e19 · outbound
Confidence Calibration in Large Language Models On Calibration of Modern Neural Networks
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc9285a3-efde-4c96-b9c2-73009a49c621 · outbound
Confidence Calibration in Large Language Models InProceedings of the 56th ACM Technical Symposium on Computer Science Education V
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bad2fab0-8ad3-4381-b40a-30cfbb823d6d · outbound
Confidence Calibration in Large Language Models Can LLMs Estimate Cognitive Complexity of Reading Comprehension Items?
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cf8fc05-6147-4117-842d-0113394201a8 · outbound
Confidence Calibration in Large Language Models Language Models (Mostly) Know What They Know
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39db4301-d476-4c17-9ad8-11ad53e2a37c · outbound
Confidence Calibration in Large Language Models Why Language Models Hallucinate
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c77dc202-54d1-4fe5-b38e-618c19b8e9aa · outbound
Confidence Calibration in Large Language Models Taming Overconfidence in LLMs: Reward Calibration in RLHF
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aaed1cd9-d46e-486d-a891-31a5b4da21b8 · outbound
Confidence Calibration in Large Language Models Sarah Lichtenstein and Baruch Fischhoff
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8757ac57-c559-4823-82e4-01723cbc4ee2 · outbound
Confidence Calibration in Large Language Models Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b4ebb65-9591-4c6a-81d9-b03d351a3933 · outbound
Confidence Calibration in Large Language Models When are Bayesian model probabilities overconfident?
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e9e26c0-1259-44d9-8d52-67e7696efd72 · outbound
Confidence Calibration in Large Language Models Understanding Model Calibration -- A gentle introduction and visual exploration of calibration and the expected calibration error (ECE)
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 377f3949-5341-42fb-a312-9654d38ea0d8 · outbound
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78c615ca-479a-4ca0-8ba3-5a52ab9fc33c · outbound
Confidence Calibration in Large Language Models GRACE: A Granular Benchmark for Evaluating Model Calibration against Human Calibration
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d31f1d48-005f-4085-8e42-7ad2cafb3e90 · outbound
Confidence Calibration in Large Language Models Katherine Tian, Eric Mitchell, Allan Zhou, Archit Sharma, Rafael Rafailov, Huaxiu Yao, Chelsea Finn, and Christo- pher D
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4f7439e-d337-448d-8cf7-f44492eb1078 · outbound
Confidence Calibration in Large Language Models Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef494730-f362-44a7-bf5e-2940ada56aa8 · outbound
Confidence Calibration in Large Language Models ArXiv:2506.23464 [cs]
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77190e65-a272-4f00-8c38-62441e58fe05 · outbound
Confidence Calibration in Large Language Models Crowdsourcing Multiple Choice Science Questions
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65038dd5-d727-406b-8cf7-2ea4e468fde1 · outbound
Confidence Calibration in Large Language Models ArXiv:2505.01997 [cs]
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 212fface-ed7d-47e1-bc23-d93c823bcc6d · outbound
Confidence Calibration in Large Language Models Can LLMs Express Their Uncertainty? An Empirical Evaluation of Confidence Elicitation in LLMs
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e44e662-0984-45c6-9ca1-bdd9020e8207 · outbound
Confidence Calibration in Large Language Models Do Language Models Mirror Human Confidence? Exploring Psychological Insights to Address Overconfidence in LLMs
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 572da502-e57c-4723-8c76-364e78efbe62 · outbound
Confidence Calibration in Large Language Models AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de24ed61-024e-45bf-b294-624b5215f3ca · outbound
Confidence Calibration in Large Language Models AR-LSAT: Investigating Analytical Reasoning of Text
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e5dabe9-470e-4ffd-b9fa-fd600f0cc18e · outbound
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c08ee2c7-56df-4e05-8858-90b412ecf3b2 · inbound
Neuro-Symbolic AI for LEED compliance: Document-Centric Benchmarking, Deterministic Numeric Checking, and When Multimodal Hurts Confidence Calibration in Large Language Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.