Pith. sign in

Paper Citation Record · LEDGER

Thermometer: Towards Universal Calibration for Large Language Models

As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2403.08819.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.08819 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T14:07:14.385726Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6ca28ebb-515d-47c6-a0af-2556f5a1df72 · inbound

Knowledge Boundary of Large Language Models: A Survey cites this paper.

Knowledge Boundary of Large Language Models: A Survey Thermometer: Towards Universal Calibration for Large Language Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-11T14:07:14.385726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:07:14.385726Z digest=sha256:e7104d6d68b423948445204fc8798983afd8b96676bacb1f056ba2274fb3188c

Observation 63a1bb91-648e-4bfd-8ae3-7a60e6c33297 · inbound

Towards Harmonized Uncertainty Estimation for Large Language Models cites this paper.

Towards Harmonized Uncertainty Estimation for Large Language Models Thermometer: Towards Universal Calibration for Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:22.485623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:25:22.485623Z digest=sha256:c6df6e6887aab1fd33b3b0fd58d38e239c1f8d3ab919c4c393e1a2704ba2461d

Observation 12576635-f91a-4a4f-9475-6f495a25b7b2 · inbound

Towards Objective Fine-tuning: How LLMs' Prior Knowledge Causes Potential Poor Calibration? cites this paper.

Towards Objective Fine-tuning: How LLMs' Prior Knowledge Causes Potential Poor Calibration? Thermometer: Towards Universal Calibration for Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T13:51:12.698057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:51:12.698057Z digest=sha256:add77a6577808418128c6bc1a96b2242aa2b6c244563bfff6591e84280df4958

Observation 13123c1d-e4be-4616-b5c4-c48acce1a9aa · inbound

Overconfidence in LLM-as-a-Judge: Diagnosis and Confidence-Driven Solution cites this paper.

Overconfidence in LLM-as-a-Judge: Diagnosis and Confidence-Driven Solution Thermometer: Towards Universal Calibration for Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T22:54:25.079218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:54:25.079218Z digest=sha256:65a3c22fffab1fa27e7d84ba6b9e56f38bb7454ccfcf1123d5291630f9e44f16

Observation 5b9dc4ac-7e07-47d2-aa6b-df7f43671200 · inbound

Insights into User Interface Innovations from a Design Thinking Workshop at deRSE25 cites this paper.

Insights into User Interface Innovations from a Design Thinking Workshop at deRSE25 Thermometer: Towards Universal Calibration for Large Language Models

Reference 1987

Resolution
unresolved
no resolver link, observed 2026-08-05T16:15:38.387545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:15:38.387545Z digest=sha256:db7a6bc6dc37dc922ad0387d8809b6f8ecbbee14eac4d02ee813da51153ef702

Observation 5f4aa9d9-897b-4962-9273-2271b2e68222 · inbound

CDE: Curiosity-Driven Exploration for Efficient Reinforcement Learning in Large Language Models cites this paper.

CDE: Curiosity-Driven Exploration for Efficient Reinforcement Learning in Large Language Models Thermometer: Towards Universal Calibration for Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T18:53:02.742636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:53:02.742636Z digest=sha256:6690d94b0d967eaf580695ebfcb33a10380b49f43167eea6d64d51c8ee8f3300

Observation f4904710-1229-4c68-8d01-5b32f88ff9f0 · inbound

The Well-Tempered Classifier: Some Elementary Properties of Temperature Scaling cites this paper.

The Well-Tempered Classifier: Some Elementary Properties of Temperature Scaling Thermometer: Towards Universal Calibration for Large Language Models

Reference 1978

Resolution
unresolved
no resolver link, observed 2026-08-02T23:14:39.507295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:14:39.507295Z digest=sha256:ce4a42e6830df0e8feb180eaa32bcd15e49367d0617bbf03cea06b2c49b68b8b

Observation 45844de7-df57-4c1c-9478-0cdc85e30134 · inbound

Beyond Medical Diagnostics: How Medical Multimodal Large Language Models Think in Space cites this paper.

Beyond Medical Diagnostics: How Medical Multimodal Large Language Models Think in Space Thermometer: Towards Universal Calibration for Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T18:16:33.382429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:16:33.382429Z digest=sha256:c0d376cec4fa74b8bc598118ab52888ac6c1efa27e4c0d9ee38a0cd57505267f

Observation 4200b305-bba2-4c08-a1c7-30b03acd1c6b · inbound

Overconfidence and Calibration in Medical VQA: Empirical Findings and Hallucination-Aware Mitigation cites this paper.

Overconfidence and Calibration in Medical VQA: Empirical Findings and Hallucination-Aware Mitigation Thermometer: Towards Universal Calibration for Large Language Models

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T21:28:17.720506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-13T21:24:55.411001Z digest=sha256:09b027b96ccd465a8d250bc3ccc89a6a9e25459a53b65b162af87ede14e683f7

Observation 080bbfe9-ff9e-4bed-92d7-1147d9542dfb · inbound

Learning When Not to Decide: A Framework for Overcoming Factual Presumptuousness in AI Adjudication cites this paper.

Learning When Not to Decide: A Framework for Overcoming Factual Presumptuousness in AI Adjudication Thermometer: Towards Universal Calibration for Large Language Models

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-10T02:11:57.293912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T02:08:24.770003Z digest=sha256:7e7e67c2ad685e08979908902d001940f36411984b4c087f91e562dcc815931d

Observation 8bcaf9d2-095f-421b-ae27-92c868321fe7 · inbound

Inducing Artificial Uncertainty in Language Models cites this paper.

Inducing Artificial Uncertainty in Language Models Thermometer: Towards Universal Calibration for Large Language Models

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:12:55.364521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-14T20:11:44.878211Z digest=sha256:cf523554006d9932444b4ee1892e01697824db0bd8e543a9776dab7de0e5ec74

Observation b56c1917-6c4e-41e6-bc49-cd7948e331a5 · inbound

Can LLMs Use Linguistic Uncertainty Markers to Reliably Reflect Intrinsic Confidence? cites this paper.

Can LLMs Use Linguistic Uncertainty Markers to Reliably Reflect Intrinsic Confidence? Thermometer: Towards Universal Calibration for Large Language Models

Reference 65

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T12:23:24.418967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-29T12:18:36.854164Z digest=sha256:5f15eae62b4ea47fff1a8674e7f940334e0ba1eaa34af2b76e31360a98b7e290

Observation c43a8b4a-a95e-4633-b093-607f86652094 · inbound

Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs cites this paper.

Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs Thermometer: Towards Universal Calibration for Large Language Models

Reference 83

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:35:42.097153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-01T05:22:38.232552Z digest=sha256:4c075011147ec7e5525e159fcb6c502cb8d41c7771efeeb30401a3a7a12a4ffd