Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2405.17888.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T21:06:55.935526Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-15T01:58:29.018876Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation e56eae66-2a65-4723-84a4-b2a9ea64b870 · inbound
Beyond External Monitors: Enhancing Transparency of Large Language Models for Easier Monitoring Getting More Juice Out of the SFT Data: Reward Learning from Human Demonstration Improves SFT for LLM Alignment
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0880f39-3b93-4bdc-aa5d-90bb0e6fabbc · inbound
RoSTE: An Efficient Quantization-Aware Supervised Fine-Tuning Approach for Large Language Models Getting More Juice Out of the SFT Data: Reward Learning from Human Demonstration Improves SFT for LLM Alignment
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 457bc523-1a1f-400a-9b91-3473eb7f2ab3 · inbound
Loki's Dance of Illusions: A Comprehensive Survey of Hallucination in Large Language Models Getting More Juice Out of the SFT Data: Reward Learning from Human Demonstration Improves SFT for LLM Alignment
Reference 129
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb7706c2-0dea-40de-8ebd-6afbeaf1cdb6 · inbound
Beyond Two-Stage Training: Cooperative SFT and RL for LLM Reasoning Getting More Juice Out of the SFT Data: Reward Learning from Human Demonstration Improves SFT for LLM Alignment
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a6abd05-fefb-4533-b552-117d5eed47b8 · inbound
On the Nature of Regularity Assumptions in Bilevel Optimization with Constrained Lower-level Problem Getting More Juice Out of the SFT Data: Reward Learning from Human Demonstration Improves SFT for LLM Alignment
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.