Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T16:34:01.216310Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 6 inbound Pith citation observations for arXiv:2502.01126.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T16:34:01.216310Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T17:12:11.740039Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T18:08:46.660526Z
55 of 55 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 038a8a8a-2c88-40a9-a9ea-9749e2ffe95e · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Claude 3.5 Sonnet
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b62ee666-6b2d-4217-ae83-ad0abd695d9f · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Unresolved cited work
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e509e91f-eca9-494a-ab27-11b4242163d4 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Constitutional AI: Harmlessness from AI Feedback
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d4c0e2e-7db0-4283-abd4-d9c69e4036ff · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd974ade-36e1-4551-b740-d9bc4a07e3e4 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Deep reinforcement learning from human preferences
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08ee4562-d1ac-4077-96c6-ba8fcb222618 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences DeGroot and Stephen E
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c20c354f-7fac-4b5d-bcd9-8fce50d63381 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences The Llama 3 Herd of Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bab9d77-0b27-4b7b-b112-dc3b5792b670 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Sivakumar
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 25bb767f-3e02-4c55-a6fd-94ae2ecbea01 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences On the foundations of noise-free selective classification
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 96c38c8d-dad7-473d-8026-6f82bf3b170c · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences The Rating of Chessplayers, Past and Present
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ae8e036-984d-401d-bdb6-f64009ff1e13 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences KTO: Model Alignment as Prospect Theoretic Optimization
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9aa76400-b106-48be-b37c-9b7786440684 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Selective prediction-set models with coverage guarantees
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 84c0728d-b3a4-4b59-9740-b39a87669bfe · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Gptscore: Evaluate as you desire
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 3feebc79-0d96-458c-95dd-9bf0520e9029 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences On Calibration of Modern Neural Networks
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eec04e32-23a9-4e57-bff1-806039ae5465 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Weinberger
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1b1f8248-2013-4fe3-b151-4b20f7fdf2c7 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences A baseline for detecting misclassified and out-of-distribution examples in neural networks
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24e92e21-8248-4067-b93c-ffb1897ae9a2 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Measuring massive multitask language understanding
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8f8a93b4-2247-4cdd-9bfd-ceab9fcf0f74 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Measuring Mathematical Problem Solving With the MATH Dataset
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7598ce1d-e0dd-466c-8826-c37ce81c98b7 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Minka, and Thore Graepel
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ff0b0cc7-f982-4dc0-999b-8b60016070d5 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences What Disease does this Patient Have? A Large-scale Open Domain Question Answering Dataset from Medical Exams
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b5e1787-7bf2-43de-94cf-8f1335f3d6f1 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Selective classification can magnify disparities across groups
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6c3ee263-67cd-451d-8919-8526caca3851 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Language models (mostly) know what they know, 2022
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 214b97c4-088e-4703-9659-0bf53bcb0b76 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Kemeny and J
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 835eca9d-2993-4d03-97c4-b2cf7246dd56 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Unanimous prediction for 100\ In Association for Computational Linguistics (ACL), 2016
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 06999197-03dc-4598-b2eb-405ce3fdcf68 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Holistic Evaluation of Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d5cf93c-90ee-401f-8f68-ed55cd51c49c · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences TruthfulQA: Measuring How Models Mimic Human Falsehoods
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03507724-f635-4e18-9ef5-b3ed1d0d7dcf · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Teaching models to express their uncertainty in words
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6e9c625-e1fd-46bd-b41a-cfe9d4556419 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Self-Refine: Iterative Refinement with Self-Feedback
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc686808-1158-4955-a2a4-f7374a987631 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Can a suit of armor conduct electricity? a new dataset for open book question answering
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 61d12a34-b748-494d-b47d-10e35f762cbe · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Murphy and Robert L
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4bf1fa61-233e-4c5e-9403-b6c7626adde1 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Cooper, and Milos Hauskrecht
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 809fcde4-914b-4aae-a4e7-5740952b6dc1 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Cooper, and Milos Hauskrecht
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation fb0ad1e6-d1d5-4f90-a3a8-d037c20bf836 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Negahban, Sewoong Oh, and Devavrat Shah
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c8371c9a-25c4-4a57-9b5f-6d07069b7ca9 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Introducing chatgpt
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 1eba87ca-a0ff-4536-913c-2e85d49ef04f · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b6fe4d77-bb6e-438c-9d33-5bfba0793266 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c2980b56-f760-43b1-8d41-0216320d50f9 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Training language models to follow instructions with human feedback
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14675387-9817-4f41-8561-634655389597 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences The pagerank citation ranking : Bringing order to the web
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ccdef85f-4f49-4f33-99bc-d43929110c23 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Direct Preference Optimization: Your Language Model is Secretly a Reward Model
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cc7a1ab-0bfa-456e-84d3-9794834aace5 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences GPQA: A Graduate-Level Google-Proof Q&A Benchmark
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61db9865-fe90-4a1d-92a6-3fc53e1f211a · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences SocialIQA: Commonsense Reasoning about Social Interactions
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b190f9c-eaa6-47a0-9b53-f2eb4ebbd08e · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Llamas Know What GPTs Don't Show: Surrogate Models for Confidence Estimation
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f533e8a-e0eb-49b3-b257-28eb7152f05d · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Commonsenseqa: A question answering challenge targeting commonsense knowledge
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 2124298d-b595-4338-a9f5-e0e5259c4e55 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65bf4de6-e9e0-4442-8bd8-c5423652d08f · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Fine-tuning Language Models for Factuality
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6f56a58-80d5-4552-824c-d62dbe82c527 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Nicolaus Tideman
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 395bdfd9-87ea-469d-be88-b663e9bbf964 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Minimum weighted feedback arc sets for ranking from pairwise comparisons
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 675fb41e-9d80-45c8-8a65-7ea8e7942630 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Self-Consistency Improves Chain of Thought Reasoning in Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa845f90-2b3f-44b0-979b-8234e270609e · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Can LLMs Express Their Uncertainty? An Empirical Evaluation of Confidence Elicitation in LLMs
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b626cbf-2f47-49aa-982f-bf923d01ef5c · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Tree of Thoughts: Deliberate Problem Solving with Large Language Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56b68cfa-eb5d-43fe-8c34-ae891a55b512 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Self-Rewarding Language Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2ac2982-bf43-48a5-bc46-a9656d66f181 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70090b81-bcbc-4764-890a-f84d4aae0c8f · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Navigating the grey area: Expressions of overconfidence and uncertainty in language models, 2023
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 425ab6e1-d2b3-437c-a0e1-2b9c2380ce30 · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences Fine-Tuning Language Models from Human Preferences
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 360b2d10-5592-47bc-8985-5d0c1396371b · outbound
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences write newline
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb3ea3df-bf2f-4af3-b0de-ea09f7ab73b9 · inbound
Inertia in Moral and Value Judgments of Large Language Models Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 43d71121-9bf3-4a26-abc6-13a748106daf · inbound
Reasoning Strategies in Large Language Models: Can They Follow, Prefer, and Optimize? Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28674834-87fb-4268-b088-ed8ee9b67625 · inbound
Generalization of Fine-Tuned Uncertainty Communication and Metacognition in Large Language Models Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences
Reference 701
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fd14fac-9697-41a6-aa7a-98d5603db9b9 · inbound
CLSGen: A Dual-Head Fine-Tuning Framework for Joint Probabilistic Classification and Verbalized Explanation Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e72eaddd-d3d4-4942-a0af-23b64811a8ff · inbound
Quantifying Consistency in LLM Logical Reasoning via Structural Uncertainty Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 7bf6f57f-3570-48dd-bb75-b70e9ef523d9 · inbound
MARS-RA: Rank Aggregation for Credit Assignment via Multimodal Comparisons in Embodied Multi-Agent Cooperation Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences
Reference 124
Source-reported events for the cited work
Unavailable: canonical work link unavailable.