Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T04:09:49.243797Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 10 of 10 outbound references and 1 inbound Pith citation observation for arXiv:2602.05897.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T04:09:49.243797Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-14T14:52:42.813309Z
A source-named dated measurement, never combined with another source.
Source: cited_works
10 of 10 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 91a99dbd-85d5-472e-b9bf-4716629a7e93 · outbound
Stop Rewarding Hallucinated Steps: Faithfulness-Aware Step-Level Reinforcement Learning for Small Reasoning Models Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation feb71d4c-4332-4e4e-94ea-3b92ad96a4f4 · outbound
Stop Rewarding Hallucinated Steps: Faithfulness-Aware Step-Level Reinforcement Learning for Small Reasoning Models Unresolved cited work
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 287c4a17-b522-4b7a-8127-d4e9a8460bf3 · outbound
Stop Rewarding Hallucinated Steps: Faithfulness-Aware Step-Level Reinforcement Learning for Small Reasoning Models Unresolved cited work
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d98b1e1-bc17-4a8e-acae-a8fc774ebfab · outbound
Stop Rewarding Hallucinated Steps: Faithfulness-Aware Step-Level Reinforcement Learning for Small Reasoning Models new_question_1
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5b6b225-781e-4335-b577-abe861a02ca9 · outbound
Stop Rewarding Hallucinated Steps: Faithfulness-Aware Step-Level Reinforcement Learning for Small Reasoning Models The FACTS Grounding Leaderboard: Benchmarking LLMs' Ability to Ground Responses to Long-Form Input
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afd2a6cd-2ad7-4aac-ac82-e01d76411e67 · outbound
Stop Rewarding Hallucinated Steps: Faithfulness-Aware Step-Level Reinforcement Learning for Small Reasoning Models Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8924ffb2-89f7-4ce6-994b-ae8d3aea172b · outbound
Stop Rewarding Hallucinated Steps: Faithfulness-Aware Step-Level Reinforcement Learning for Small Reasoning Models Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2a3ce71-11dd-4a11-9d7d-e8459ca4b964 · outbound
Stop Rewarding Hallucinated Steps: Faithfulness-Aware Step-Level Reinforcement Learning for Small Reasoning Models Why Language Models Hallucinate
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62fab9f8-67f8-495f-ae62-a0028f908a72 · outbound
Stop Rewarding Hallucinated Steps: Faithfulness-Aware Step-Level Reinforcement Learning for Small Reasoning Models Smoothing Out Hallucinations: Mitigating LLM Hallucination with Smoothed Knowledge Distillation
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1832476d-24f2-466d-8b07-a60996cc1d98 · outbound
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2befc4e1-7fe5-4a7a-8ba7-22bc161c998d · inbound
CLIR-Bench: Benchmarking Multimodal Question Answering over Irregular Clinical Time Series Stop Rewarding Hallucinated Steps: Faithfulness-Aware Step-Level Reinforcement Learning for Small Reasoning Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.