Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:44:36.534742Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 9 inbound Pith citation observations for arXiv:2505.13988.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:44:36.534742Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T10:52:50.299350Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-02T18:57:16.598004Z
52 of 52 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ee8a17e6-6a68-4f1f-a790-a77f4d267511 · outbound
The Hallucination Tax of Reinforcement Finetuning online" 'onlinestring :=
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bd65997-4777-4fe7-9e09-495e03bcfbc3 · outbound
The Hallucination Tax of Reinforcement Finetuning write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 995a819f-5fbc-4bb5-90c7-4cf4c8a0ac33 · outbound
The Hallucination Tax of Reinforcement Finetuning Phi-4-reasoning Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c6d2e74-fa0b-4a8b-bb16-c419cb0eefa0 · outbound
The Hallucination Tax of Reinforcement Finetuning HalluLens: LLM Hallucination Benchmark
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24326b6b-640a-4b9a-abe6-962ebb645b7e · outbound
The Hallucination Tax of Reinforcement Finetuning Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c1703b3-86d1-46ea-9da9-93516346b420 · outbound
The Hallucination Tax of Reinforcement Finetuning Mitigating Open-Vocabulary Caption Hallucinations
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0c669cd-5eb0-4592-82c1-95bfa02b08bf · outbound
The Hallucination Tax of Reinforcement Finetuning Evaluating Hallucinations in Chinese Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41939ca4-a260-4759-91f5-d40a2cfe6b20 · outbound
The Hallucination Tax of Reinforcement Finetuning Training Verifiers to Solve Math Word Problems
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a56fac1-4f09-493f-ad1f-d88b1ad02e51 · outbound
The Hallucination Tax of Reinforcement Finetuning Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71483c38-9e7a-49ef-8260-ff5294e1f711 · outbound
The Hallucination Tax of Reinforcement Finetuning Does Fine-Tuning LLMs on New Knowledge Encourage Hallucinations?
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 854c0351-2a20-4b30-92cb-d939b26418a0 · outbound
The Hallucination Tax of Reinforcement Finetuning The Llama 3 Herd of Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce49eb88-0866-4122-8767-c1fdbd6c4c21 · outbound
The Hallucination Tax of Reinforcement Finetuning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 819102f4-c5fd-4cb0-b98d-60b315192ed4 · outbound
The Hallucination Tax of Reinforcement Finetuning OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5f2b607-eb5a-428d-ac1d-fb36a23b78d3 · outbound
The Hallucination Tax of Reinforcement Finetuning Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f7481b6-dca4-4af9-aaed-449f5323eaa3 · outbound
The Hallucination Tax of Reinforcement Finetuning REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84052823-48f1-4594-8298-fe0688dfcfe8 · outbound
The Hallucination Tax of Reinforcement Finetuning Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b37d6ff-c0f1-4c99-a109-81665178032f · outbound
The Hallucination Tax of Reinforcement Finetuning Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 25972b8f-464b-49e8-94e8-f8f5994bcaff · outbound
The Hallucination Tax of Reinforcement Finetuning O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f1ffac1-a515-450a-8696-fd8c930e9a5e · outbound
The Hallucination Tax of Reinforcement Finetuning Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 911ec246-e3ca-4ef7-a54b-253350311b84 · outbound
The Hallucination Tax of Reinforcement Finetuning The Dawn After the Dark: An Empirical Study on Factuality Hallucination in Large Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27b411fc-cbc0-4d0d-9643-2eeee438d1af · outbound
The Hallucination Tax of Reinforcement Finetuning Treble Counterfactual VLMs: A Causal Approach to Hallucination
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9418ddc7-8284-4590-81cc-d8bc205438cb · outbound
The Hallucination Tax of Reinforcement Finetuning LIMR: Less is More for RL Scaling
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc76dd61-de76-4e95-b56a-b6224cd025a2 · outbound
The Hallucination Tax of Reinforcement Finetuning Let's Verify Step by Step
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 711946dc-cdd6-42db-a1c6-d3ed63c0ed2f · outbound
The Hallucination Tax of Reinforcement Finetuning Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 152c4b32-31b7-4b31-b8a7-ce1516fce7a9 · outbound
The Hallucination Tax of Reinforcement Finetuning Tang, Manan Roongta, Colin Cai, Jeffrey Luo, Li Erran Li, Raluca Ada Popa, and Ion Stoica
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 655dcc39-5a30-44d3-ab7d-6c4fd97b7cc5 · outbound
The Hallucination Tax of Reinforcement Finetuning Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10c551ec-e18d-4b78-987f-de8e7b0ceff8 · outbound
The Hallucination Tax of Reinforcement Finetuning Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2ee8666a-b3cf-42e7-b1d0-e4cff21c07ea · outbound
The Hallucination Tax of Reinforcement Finetuning Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e825bd42-c4fd-4479-9e05-a2f9c9a63638 · outbound
The Hallucination Tax of Reinforcement Finetuning Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e638485-7ee6-4076-8934-97bbd59cd17a · outbound
The Hallucination Tax of Reinforcement Finetuning Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation be302502-a318-4dc2-87ee-44644080a929 · outbound
The Hallucination Tax of Reinforcement Finetuning Proximal Policy Optimization Algorithms
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c226e53-6050-4b90-b6c2-edf442f7a5d1 · outbound
The Hallucination Tax of Reinforcement Finetuning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16abeb0e-4769-4dee-bee0-62ab852816d5 · outbound
The Hallucination Tax of Reinforcement Finetuning Unresolved cited work
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6932d08-7faf-4c4a-9d61-9ad113926933 · outbound
The Hallucination Tax of Reinforcement Finetuning Efficient Reinforcement Finetuning via Adaptive Curriculum Learning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8527608-d5de-4952-b276-8553a0ab47df · outbound
The Hallucination Tax of Reinforcement Finetuning Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed215d06-63d0-4e1a-96ed-001d416d4174 · outbound
The Hallucination Tax of Reinforcement Finetuning Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1e62e25e-fcd5-4aa5-9bbe-597e3bacba13 · outbound
The Hallucination Tax of Reinforcement Finetuning Aligning Large Multimodal Models with Factually Augmented RLHF
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e57963e-b949-4603-86b2-f160291ca989 · outbound
The Hallucination Tax of Reinforcement Finetuning Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fdd7360a-d90f-4de7-905a-48d9044a6cfb · outbound
The Hallucination Tax of Reinforcement Finetuning A Comprehensive Survey of Hallucination Mitigation Techniques in Large Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e86e968-b904-4aa8-af69-2222fd6b1be5 · outbound
The Hallucination Tax of Reinforcement Finetuning AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e7e68fb-7fad-46bd-a097-8001bb4f84c5 · outbound
The Hallucination Tax of Reinforcement Finetuning Tina: Tiny Reasoning Models via LoRA
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec9d1ffe-a12c-470d-9379-96f40be417bc · outbound
The Hallucination Tax of Reinforcement Finetuning Reinforcement Learning for Reasoning in Large Language Models with One Training Example
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 176734b8-7600-48fd-a725-72d93dc4c722 · outbound
The Hallucination Tax of Reinforcement Finetuning Simple synthetic data reduces sycophancy in large language models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9aebf6e4-c24a-47e1-9f0a-9503c9f6beb1 · outbound
The Hallucination Tax of Reinforcement Finetuning Unresolved cited work
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 007b1964-85c9-4886-b1c9-f9788c4ad213 · outbound
The Hallucination Tax of Reinforcement Finetuning A Minimalist Approach to LLM Reasoning: from Rejection Sampling to Reinforce
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7be9aab1-c81d-42c3-9abd-b00308dcd272 · outbound
The Hallucination Tax of Reinforcement Finetuning Unresolved cited work
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 16848e90-3df9-41b2-bccf-5254ad9f571f · outbound
The Hallucination Tax of Reinforcement Finetuning Do Large Language Models Know What They Don't Know?
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b778e018-1636-4f23-891f-dbaaaa36b509 · outbound
The Hallucination Tax of Reinforcement Finetuning DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a8df657-c924-449c-b6ab-22dd12a69bd1 · outbound
The Hallucination Tax of Reinforcement Finetuning Unresolved cited work
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9268f459-e4da-4b2a-901c-79f32205026e · outbound
The Hallucination Tax of Reinforcement Finetuning How Language Model Hallucinations Can Snowball
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f32edaa0-bd51-4cb7-81a7-533956c446c1 · outbound
The Hallucination Tax of Reinforcement Finetuning Nemotron-Research-Tool-N1: Exploring Tool-Using Language Models with Reinforced Reasoning
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13000c1c-0526-4a67-bb6d-f74ffe031b72 · outbound
The Hallucination Tax of Reinforcement Finetuning Absolute Zero: Reinforced Self-play Reasoning with Zero Data
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f96d3bbd-0892-49ad-9c12-ce012dde8f0d · inbound
Beyond Passive Critical Thinking: Fostering Proactive Questioning to Enhance Human-AI Collaboration The Hallucination Tax of Reinforcement Finetuning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1ed60c3-31ad-4d33-8b55-7c3267e55276 · inbound
A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models The Hallucination Tax of Reinforcement Finetuning
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d18fcaa-1364-4c4f-9724-4d6f3d7debac · inbound
MoCo: A One-Stop Shop for Model Collaboration Research The Hallucination Tax of Reinforcement Finetuning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7a4542a9-ef81-404b-97c1-25a1b7813634 · inbound
Rethinking Agentic Reinforcement Learning In Large Language Models The Hallucination Tax of Reinforcement Finetuning
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 92e46cd5-8b6e-4d71-83a9-2633117e1cc6 · inbound
Rethinking Agentic Reinforcement Learning In Large Language Models The Hallucination Tax of Reinforcement Finetuning
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1686a16f-1bbe-407e-b0bd-10c6fe6c2dc3 · inbound
Rethinking Agentic Reinforcement Learning In Large Language Models The Hallucination Tax of Reinforcement Finetuning
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d9620be2-8cc6-45fd-88ce-ff858571b6e4 · inbound
Bridging the Detection-to-Abstention Gap in Reasoning Models under Insufficient Information The Hallucination Tax of Reinforcement Finetuning
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 74983e86-3310-4efc-aab7-7357ede6e45f · inbound
Scaling Participation in Modular AI Systems The Hallucination Tax of Reinforcement Finetuning
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a0a77b3d-b3f9-4a9d-8d75-9ebe41e5c81a · inbound
Mechanistic Attention Guidance for Agent Memory Refinement The Hallucination Tax of Reinforcement Finetuning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.