Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T05:42:36.727558Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 2 inbound Pith citation observations for arXiv:2604.17293.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T05:42:36.727558Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-09T04:49:14.556768Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-09T04:55:58.814806Z
53 of 53 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 85e44662-b91e-4ba8-b48b-a28b511c749d · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty The Twelfth International Conference on Learning Representations
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 71837a24-9a23-4b87-b6e2-88b6556a5cc3 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Transactions of the Association for Computational Linguistics , pages =
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 23f3c190-ce6f-495b-8268-bc4b6ad9b734 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Do Large Language Models Know What They Don ' t Know?
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9c414fc0-6ebf-4a5b-9fe9-16019a6ce6e9 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty O lympiad B ench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 49e9ab11-7c1e-435f-90ec-d9b02f03e668 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Missing Premise exacerbates Overthinking: Are Reasoning Models losing Critical Thinking Skill? , url =
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 33c038ac-2933-4037-b207-7c7cbf8346ef · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Training verifiers to solve math word problems , url =
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2319da1b-874e-4832-ad6e-0f34e2ee59de · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Measuring mathematical problem solving with the math dataset , url =
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b69d0bb4-3b19-4db5-9903-d5be98b25678 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Dapo: An open-source llm reinforcement learning system at scale , url =
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 7c5f3be1-b4cc-43c5-b77b-6edc45cc0b03 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Hybridflow: A flexible and efficient rlhf framework , url =
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 22d20a86-8864-4c30-ab77-34ebf310d9e3 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Metacognition: Answered and unanswered questions , volume =
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 23a6d4e6-37da-4fe8-800a-ed780d93486d · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Benchmarking Uncertainty Quantification Methods for Large Language Models with LM -Polygraph
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5396b990-1f2f-40d1-8e57-3ed427fd8196 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Deliberative alignment: Reasoning enables safer language models , url =
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation be5fcd18-f7c2-422a-a90d-c5d1ed4de737 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Does Biomedical Training Lead to Better Medical Performance?
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ca6eea5d-877d-49ca-be8d-a06d510cee44 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty A Survey on Proactive Dialogue Systems: Problems, Methods, and Prospects , url =
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 13d814fb-40ba-4f65-9371-3c465dd9bb5c · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty AbstentionBench: Reasoning LLMs Fail on Unanswerable Questions , url =
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b448555b-86b6-4d80-948c-a62eb834a9bf · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Proceedings of the 31st ACM SIGKDD Conference on Knowledge Discovery and Data Mining V
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation bd54f8a1-2c3e-4a55-b976-afab52b33446 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Don ' t Just Say `` I don ' t know''! Self-aligning Large Language Models for Responding to Unknown Questions with Explanations
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5a750a5d-f4f2-4aed-a24d-dc633e63a640 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty The Dialogue That Heals: A Comprehensive Evaluation of Doctor Agents' Inquiry Capability , url =
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b4954075-0fa4-4195-8b74-2781713d46d9 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Doctor-R1: Mastering Clinical Inquiry with Experiential Agentic Reinforcement Learning , url =
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation d8405eff-2bbc-42f3-a416-bd04eb246079 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Search-r1: Training llms to reason and leverage search engines with reinforcement learning , url =
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f85548d7-3ff2-47aa-b894-c1eb356f4af8 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty The Twelfth International Conference on Learning Representations
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c96fccb1-66d7-4802-bcd0-97696597f03f · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Transparent and Robust RAG: Adaptive-Reward Reinforcement Learning for Decision Traceability , url =
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 49730b9f-2186-4986-b120-066b89353d7f · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Adaptive Tool Use in Large Language Models with Meta-Cognition Trigger
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a6e2bdfe-6630-400f-ae44-e92f367c0a28 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Edelman , bibsource =
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 60f2440a-ffbe-4df5-889d-776a23178a03 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty W i C ke D : A Simple Method to Make Multiple Choice Benchmarks More Challenging
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a74d959a-a3f6-4784-85b4-0455d694b207 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty None of the Above, Less of the Right Parallel Patterns in Human and LLM Performance on Multi-Choice Questions Answering
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a6e4938c-637d-4b15-bf23-1ebb5a3c3f68 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Asking clarification questions to handle ambiguity in open-domain qa
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 812ff6ba-319d-49d3-94b5-87f807d9ac35 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty CLAMBER: A benchmark of identifying and clarifying ambiguous information needs in large language models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a2b90ad5-642d-4103-97ba-b037972b253f · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Benchmarking Hallucination in Large Language Models Based on Unanswerable Math Word Problem , url =
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 63ed5d99-4c30-41ec-a831-9c1aa6a3ab55 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty ArXiv preprint , title =
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 295a598d-5844-449c-8b66-4bce6a1be88d · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c5122af2-06d6-41ec-acf3-0a71e202ff90 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Knowledge of Knowledge: Exploring Known-Unknowns Uncertainty with Large Language Models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 8ea51281-aa31-4d02-b321-dc8909ad625c · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Wong and Emine Yilmaz and Shuming Shi and Zhaopeng Tu , bibsource =
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 83b270a9-b44c-4b6d-a7e5-1338bd9beb2c · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty UR ^2 : Unify RAG and Reasoning through Reinforcement Learning , url =
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 68edb052-361e-4e77-9f62-067c32e5888e · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Octothinker: Mid-training incentivizes reinforcement learning scaling.arXiv preprint arXiv:2506.20512, 2025b
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ccbd8552-e1af-438a-aaf2-9b8c3d8056db · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty A Survey of Confidence Estimation and Calibration in Large Language Models , url =
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1eb9e26c-009f-4832-a329-08af901a4f66 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty The curious case of hallucinatory (un)answerability: Finding truths in the hidden states of over-confident large language models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 15fe4bde-af2a-4418-bfba-1721456d8c4b · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Let the Model Distribute Its Doubt: Confidence Estimation through Verbalized Probability Distribution , url =
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 3b3be7c2-dc74-4bb8-9556-e4d385af4da8 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Grace: A generative approach to better confidence elicitation in large language models , url =
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation d683f474-1ed4-4bad-9b63-a0e4fda7f145 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Large Language Models Must Be Taught to Know What They Don't Know , url =
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 93c4ccc0-d789-450a-bd7c-efb7c087aa36 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Beyond binary rewards: Training lms to reason about their uncertainty , url =
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 619c0904-7de1-42cd-a857-8f1c989448cf · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Knowrl: Exploring knowledgeable reinforcement learning for factuality , url =
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f7b2b212-5290-4817-89e5-429ed5460b0d · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty KnowRL: Teaching Language Models to Know What They Know , url =
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ddeeffc7-24d3-4895-8bcf-6aae3391e7ca · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Qwen3 technical report , url =
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c4b6a4b2-04dd-4e85-aca7-3b8de40fd506 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Unresolved cited work
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4a02d689-6749-4c0e-807a-9115a70636a5 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty GPT-4o System Card , year =
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation bb7fd78b-caa6-4829-9394-6e1a451dcd6c · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty GPT-5 System Card , year =
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4acb1771-73b2-4c5e-8890-cc036411b1ed · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Introducing GPT-OSS: Open Weights for Advanced Reasoning , year =
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation daeb6169-2491-4e9c-84b2-186fd9d41663 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty The Claude 4 Model Family: Opus, Sonnet, and Haiku , year =
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 7f3e52f7-392d-4ed0-99b5-63b7bebcbb50 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Unresolved cited work
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0b0d5a55-2988-46b7-bb6e-7aab3567a462 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Deepseekmath: Pushing the limits of mathematical reasoning in open language models , url =
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e88358eb-36c5-46ca-9c47-95ce89ac8414 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena , url =
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation df1ccd2b-3824-4e7d-a830-a199194013a0 · outbound
Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty Countering capability boundary collapse of llms in reinforcement learning with hybrid-policy optimization , url =
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation aa905abd-b9dc-4bb0-800e-f217a1f6b396 · inbound
Second Guess: Detecting Uncertainty Through Abstention and Answer Stability in Small Language Models Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation fd47c197-873b-46e0-95a8-7adddc535d9d · inbound
Future Confidence Distillation in Large Language Models Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model Uncertainty
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.