Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2410.18640.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:45:21.939278Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-07T12:33:45.100323Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 8d112cc8-5811-4279-aba8-2e29fe42d0d8 · inbound
BIRD: Behavior Induction via Representation-structure Distillation Weak-to-Strong Preference Optimization: Stealing Reward from Weak Aligned Model
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f121dd03-c33d-4d62-b5a5-06878a59f621 · inbound
Hybrid Policy Distillation for LLMs Weak-to-Strong Preference Optimization: Stealing Reward from Weak Aligned Model
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e0cf617e-c640-4873-a526-d4642e51e358 · inbound
Weak-to-Strong Generalization via Direct On-Policy Distillation Weak-to-Strong Preference Optimization: Stealing Reward from Weak Aligned Model
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 30059df4-9f93-4e88-8475-7f9ccd21f731 · inbound
Weak-to-Strong Generalization via Direct On-Policy Distillation Weak-to-Strong Preference Optimization: Stealing Reward from Weak Aligned Model
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.