Pith. sign in

Paper Citation Record · LEDGER

Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 35 inbound Pith citation observations for arXiv:2307.10236.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.10236 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 35 of 35 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T18:53:44.196975Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T18:08:46.635429Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d61e38a5-dee2-4724-b913-3c5a57d5f793 · inbound

Bridging Internal Probability and Self-Consistency for Effective and Efficient LLM Reasoning cites this paper.

Bridging Internal Probability and Self-Consistency for Effective and Efficient LLM Reasoning Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T18:53:44.196975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:53:44.196975Z digest=sha256:8bd26ffca01e7fae1c62c5c0fc2879f52bc4c375f07b8f52785e55ecc23c177d

Observation 8c9985b1-0ac7-42d2-844b-922eff5d1786 · inbound

Confident or Seek Stronger: Exploring Uncertainty-Based On-device LLM Routing From Benchmarking to Generalization cites this paper.

Confident or Seek Stronger: Exploring Uncertainty-Based On-device LLM Routing From Benchmarking to Generalization Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T22:49:51.741605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:49:51.741605Z digest=sha256:3cd0a58bf956cc2151cc1172630f80c6935ed66c8bbb073c375093c142f061a9

Observation 04ce6625-c5ca-406c-8ae8-a880fa4b2be6 · inbound

Consensus Entropy: Harnessing Multi-VLM Agreement for Self-Verifying and Self-Improving OCR cites this paper.

Consensus Entropy: Harnessing Multi-VLM Agreement for Self-Verifying and Self-Improving OCR Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-22T20:32:04.647581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T20:31:34.074705Z digest=sha256:e3f0c458d5ebbc74118641863f9b9bc3e42488f067454a2db7918886411173f1

Observation 912e6fa5-4b63-4d10-a79b-15e3f1ec704c · inbound

Divide-Then-Align: Honest Alignment based on the Knowledge Boundary of RAG cites this paper.

Divide-Then-Align: Honest Alignment based on the Knowledge Boundary of RAG Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:50:39.165790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:50:39.165790Z digest=sha256:218744748506c5ffd7ae5d070cd5afe85909bf73bac7252ad5ac863bbe815729

Observation 17d49396-5d55-4116-ab29-06355a8eb962 · inbound

Is Your Model Fairly Certain? Uncertainty-Aware Fairness Evaluation for LLMs cites this paper.

Is Your Model Fairly Certain? Uncertainty-Aware Fairness Evaluation for LLMs Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:44:05.183520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:44:05.183520Z digest=sha256:8dc4b7e38c6d2689288cde5e7c2dcf1a5a12a8eadef66cc32048bda8b47fa112

Observation d542e00c-9b25-4a32-b4d7-fc025e8ed3db · inbound

HD-NDEs: Neural Differential Equations for Hallucination Detection in LLMs cites this paper.

HD-NDEs: Neural Differential Equations for Hallucination Detection in LLMs Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:34:04.092510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:34:04.092510Z digest=sha256:03c01c08743f8b34b3e5a763bc626c32f0641445941a7e174f31aac6f9a1fa79

Observation 28f84359-f68c-4e41-9021-fb91d2fd62d0 · inbound

Improving the Calibration of Confidence Scores in Text Generation Using the Output Distribution's Characteristics cites this paper.

Improving the Calibration of Confidence Scores in Text Generation Using the Output Distribution's Characteristics Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:07:28.528388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:07:28.528388Z digest=sha256:1f0ad10d6445f3da60083ebbee2ca8daaa7e5deef11e8c0814be0e2e410fb149

Observation c4ad12f7-8bf7-4960-95b9-1f84ea729238 · inbound

ChemAU: Harness the Reasoning of LLMs in Chemical Research with Adaptive Uncertainty Estimation cites this paper.

ChemAU: Harness the Reasoning of LLMs in Chemical Research with Adaptive Uncertainty Estimation Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:44.964444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:44.964444Z digest=sha256:d08cea89438b37899462ca2572fc5243f28d8cd35e2b8d7060d33797cf57f8a4

Observation 94a0b584-6497-4576-a02a-ad089982c8f3 · inbound

Representations of Fact, Fiction and Forecast in Large Language Models: Epistemics and Attitudes cites this paper.

Representations of Fact, Fiction and Forecast in Large Language Models: Epistemics and Attitudes Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:44:14.313733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:44:14.313733Z digest=sha256:af03b2a59420ad1aef3bba85ea232357f1dbabecaf77f080ff59ae30128d0f1c

Observation 4e92d247-b511-49e5-8697-b48d91bde039 · inbound

Shaking to Reveal: Perturbation-Based Detection of LLM Hallucinations cites this paper.

Shaking to Reveal: Perturbation-Based Detection of LLM Hallucinations Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T11:23:45.835103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:23:45.835103Z digest=sha256:1a56c7c7d7c2bf172918d255e53ec92392c716cd48331a6ae4bc244764a5258f

Observation 8509ddc4-7e3f-46b0-81f3-93994956695f · inbound

TRUST: Test-time Resource Utilization for Superior Trustworthiness cites this paper.

TRUST: Test-time Resource Utilization for Superior Trustworthiness Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:08.144917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:07:08.144917Z digest=sha256:df25168b558534018ad762f4ac0857caa989682c124eed99d369a977b66965c5

Observation 2de1bfea-dff1-45d1-be5d-ec1373a574f2 · inbound

WebSailor: Navigating Super-human Reasoning for Web Agent cites this paper.

WebSailor: Navigating Super-human Reasoning for Web Agent Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-17T15:37:09.646082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T15:37:09.572241Z digest=sha256:d14d65e19792f8aa3db47680ed1faad086a0a631bb5b313e962130b40ca8e9eb

Observation e8f4b193-3d06-4809-833f-c1f35ce37c68 · inbound

Toward Better Generalisation in Uncertainty Estimators: Leveraging Data-Agnostic Features cites this paper.

Toward Better Generalisation in Uncertainty Estimators: Leveraging Data-Agnostic Features Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T20:02:02.991407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:02:02.991407Z digest=sha256:c22bfa83467db2e9fa866dc841c242110c718fb3ab445cb5a1ae994b547b2a6c

Observation 4627834b-dcba-4186-bdec-8c6f045205fe · inbound

Evaluating Uncertainty and Quality of Visual Language Action-enabled Robots cites this paper.

Evaluating Uncertainty and Quality of Visual Language Action-enabled Robots Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:25.438745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:02:25.438745Z digest=sha256:b7fc22b3ad499207f73ee1006b4df7cf880e8b315bca3ea77113e5a283c18562

Observation c10d5dda-f7ac-43e2-a2b4-7efd07fc44ed · inbound

Confidence Estimation for Text-to-SQL in Large Language Models cites this paper.

Confidence Estimation for Text-to-SQL in Large Language Models Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T22:36:21.614672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:36:21.614672Z digest=sha256:ceb270f0544e60c337929eb2036271923a1378cfc8f81edd6d39c73175906c2f

Observation c4be402b-50d1-4e14-a95d-45d5e84071d3 · inbound

Can Multiple Responses from an LLM Reveal the Sources of Its Uncertainty? cites this paper.

Can Multiple Responses from an LLM Reveal the Sources of Its Uncertainty? Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-05T14:34:33.108971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:34:33.108971Z digest=sha256:6ad94cac0be18afc8b648cbb94881c8e2adc4ed55a7db36470ee02be9a59de6a

Observation 7d51c7d5-ae77-44e2-81ab-67c99bb17bdb · inbound

Neural Message-Passing on Attention Graphs for Hallucination Detection cites this paper.

Neural Message-Passing on Attention Graphs for Hallucination Detection Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T13:52:09.125465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:52:09.125465Z digest=sha256:43144dba5605fa9817145b9346f08ae8ccf4d5a3ba7759d87ed40eb72bfe58f2

Observation d5e5bab8-0f45-4cbc-b79b-b6f0d9d0cdaa · inbound

Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons cites this paper.

Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 135

Resolution
verified exact
arxiv_id, observed 2026-05-15T18:00:12.654671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T17:59:52.365630Z digest=sha256:759c7a7e87db0e12c7fd7152924e6bd1f83b044a436b33fdb041aa31490a14c7

Observation 2f02c018-2875-4c24-a5e5-4ac2f5420c11 · inbound

Context Collapse: Barriers to Adoption for Generative AI in Workplace Settings cites this paper.

Context Collapse: Barriers to Adoption for Generative AI in Workplace Settings Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:35:52.626262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T18:58:18.436603Z digest=sha256:47601cc845d017d2be1fd691699eda4ba6b7e8687a4a9128e0adfe843f660406

Observation 059fdd2b-d15a-48aa-8081-7743fdbbdb19 · inbound

Confident in a Confidence Score: Investigating the Sensitivity of Confidence Scores to Supervised Fine-Tuning cites this paper.

Confident in a Confidence Score: Investigating the Sensitivity of Confidence Scores to Supervised Fine-Tuning Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:25:58.665662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T17:38:53.434978Z digest=sha256:bf250cb05df94a9f601d93ca69ffcdc8436bd4fa3b1fe9385f87a0afda82f8a4

Observation af56679e-e846-44ab-b471-5b935ea7a7e7 · inbound

Learning Uncertainty from Sequential Internal Dispersion in Large Language Models cites this paper.

Learning Uncertainty from Sequential Internal Dispersion in Large Language Models Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:48:02.443497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T08:36:39.242766Z digest=sha256:f86c07b9047a4c164d7aa2d2e8c00075a6cafafcb59ac0b29d0284fd679be6fd

Observation 5af0c885-79f7-496c-8cda-771cc62f48a1 · inbound

Temporal Difference Calibration in Sequential Tasks: Application to Vision-Language-Action Models cites this paper.

Temporal Difference Calibration in Sequential Tasks: Application to Vision-Language-Action Models Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 54

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T13:51:03.364892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-09T23:59:32.295839Z digest=sha256:112e201efda02bcde9cae9bf4d0bdff24a19db0673c5db35484b4efec4254ab9

Observation 55461839-1c00-4185-a985-8a35556bab5b · inbound

LLMs Uncertainty Quantification via Adaptive Conformal Semantic Entropy cites this paper.

LLMs Uncertainty Quantification via Adaptive Conformal Semantic Entropy Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:11:09.088849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T17:51:53.511503Z digest=sha256:009155af040e75f8bdac09f859b4e88ac7ee3229fefadbe6a9709bb536c113b9

Observation 421a65a4-8441-4e51-9d2d-4ca75d8643c0 · inbound

LLMs Uncertainty Quantification via Adaptive Conformal Semantic Entropy cites this paper.

LLMs Uncertainty Quantification via Adaptive Conformal Semantic Entropy Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-01T13:15:46.007360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T23:47:33.632049Z digest=sha256:fc1cd994e71b78d5a5f9630cdcaed0b9e710a1676c251dc544ca71a16019f801

Observation 1c2e9598-9e4a-4827-9ff7-6154ca08132c · inbound

BackFlush: Knowledge-Free Backdoor Detection and Elimination with Watermark Preservation in Large Language Models cites this paper.

BackFlush: Knowledge-Free Backdoor Detection and Elimination with Watermark Preservation in Large Language Models Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:02:58.950282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T21:01:10.756844Z digest=sha256:f871a2cb95f42e20e0847298349b83650510ee94c361be86695a1e83b9083271

Observation b33281b2-b7f2-4022-a514-287aa102cfde · inbound

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming cites this paper.

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T09:14:45.421169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T09:13:45.830192Z digest=sha256:d8ccc482c2076af828e361d5e6ac642c5ff18f240c44b185ca9825e1926571dd

Observation 88a04825-f6ac-45a1-8a0f-92b8eb30babd · inbound

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming cites this paper.

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T16:54:58.773116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T16:48:55.898188Z digest=sha256:108d325a00d1df359cf0148cdf2a8da6592d74d9c8051e7206f6d9c1de4b35a3

Observation 9175b8aa-0484-45f1-8a12-dd820e23d6d1 · inbound

MultiHaluDet: Multilingual Hallucination Detection via LLM Hidden State Probing cites this paper.

MultiHaluDet: Multilingual Hallucination Detection via LLM Hidden State Probing Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T12:34:38.667568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T12:29:59.165791Z digest=sha256:116f596ee70347abd8e90b5a80cc25fc82c54e7f9016a6af4afacf0f338fe526

Observation 81b8b31d-b857-49d0-b3eb-435b291e0f35 · inbound

Entropy Distribution as a Fingerprint for Hallucinations in Generative Models cites this paper.

Entropy Distribution as a Fingerprint for Hallucinations in Generative Models Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:13:26.691322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T12:10:43.626846Z digest=sha256:d948476e0b3f865cc77574b670ff502ecad8bf0f0f2030d63db056c209b12e2f

Observation 83e58d2a-d1e9-492f-84ad-6cc08aeaebf8 · inbound

Smart Picks in the Dark: Towards Efficient RLVR for Reasoning via Tracing Metacognitive Pivots cites this paper.

Smart Picks in the Dark: Towards Efficient RLVR for Reasoning via Tracing Metacognitive Pivots Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:26:46.219869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T06:55:09.927034Z digest=sha256:6d184e05c6c71075a9139bf43fc049d729a94cd4c5ad3067adb6fe79d466ec22

Observation caa66479-c02b-41e8-92a3-5e8caa048f64 · inbound

ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning cites this paper.

ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T21:07:23.911497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T19:56:09.820812Z digest=sha256:503c813d9524b5b4b0a1e39eee2ca7f3fb33e8ea7e1a72a9d9da884fe1f6ee8a

Observation 7e6d11a9-b4e3-459f-876b-d040a43650e3 · inbound

Quantifying Consistency in LLM Logical Reasoning via Structural Uncertainty cites this paper.

Quantifying Consistency in LLM Logical Reasoning via Structural Uncertainty Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-03T18:08:46.637567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T03:18:22.813002Z digest=sha256:8120207aa02263cd9503427b617121c454a93bf9de75892db588561c289c7f5f

Observation 68dbbede-f71e-4aa9-89a7-9fe7717fcaf9 · inbound

The "I Don't Know" Filter: Enhancing Agentic Reliability in Function Calling cites this paper.

The "I Don't Know" Filter: Enhancing Agentic Reliability in Function Calling Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-11T22:10:22.729188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T22:10:22.729188Z digest=sha256:a31d04404955592549326a5ac540b18ad3f2a59a214e448e9ac21c01f7e741ab

Observation e2804edd-e922-46dc-b696-e9411f01a820 · inbound

Hallucination Detection in Large Language Models Using Diversion Decoding cites this paper.

Hallucination Detection in Large Language Models Using Diversion Decoding Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-14T11:26:23.545785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T11:26:23.545785Z digest=sha256:4fb402c6674e58796d8c66d6031b088842c782ceae413aebbe513a3028af4a2c

Observation 31f5a8fb-221e-4974-af7b-cc7fbc308b06 · inbound

SAFECAST: Robust Failure Detection for VLA Policies with Contrast-Set Training and Calibration cites this paper.

SAFECAST: Robust Failure Detection for VLA Policies with Contrast-Set Training and Calibration Look Before You Leap: An Exploratory Study of Uncertainty Measurement for Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T00:10:44.884480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:10:44.884480Z digest=sha256:27fc30480d41e17ca72725645086ae250c141f172eb98aeb4d91ede9d373a285