Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2504.05294.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T15:30:30.183583Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 1a5bc834-776e-46cc-bec6-5bcd15655a02 · inbound
Analysing Chain of Thought Dynamics: Active Guidance or Unfaithful Post-hoc Rationalisation? Truthful or Fabricated? Using Causal Attribution to Mitigate Reward Hacking in Explanations
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16b77f73-248e-4669-95fe-35d2d5d1f121 · inbound
CLARity: Reasoning Consistency Alone Can Teach Reinforced Experts Truthful or Fabricated? Using Causal Attribution to Mitigate Reward Hacking in Explanations
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c099ec6-1b8d-4e5a-9e3a-f36599310aae · inbound
Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models Truthful or Fabricated? Using Causal Attribution to Mitigate Reward Hacking in Explanations
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 87c675e8-1dc2-4dcd-836c-06657ed2270d · inbound
Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation Truthful or Fabricated? Using Causal Attribution to Mitigate Reward Hacking in Explanations
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c83475e-f2c6-4ae2-b0a2-8114818e292e · inbound
Reasoning Dynamics and the Limits of Monitoring Modality Reliance in Vision-Language Models Truthful or Fabricated? Using Causal Attribution to Mitigate Reward Hacking in Explanations
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b3551660-9d06-4278-afb5-58ada5ac271c · inbound
Compared to What? Baselines and Metrics for Counterfactual Prompting Truthful or Fabricated? Using Causal Attribution to Mitigate Reward Hacking in Explanations
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 80879fc9-f7b1-4296-b2b2-1492840e169c · inbound
Interpretability Can Be Actionable Truthful or Fabricated? Using Causal Attribution to Mitigate Reward Hacking in Explanations
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.