Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2503.03746.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:21:23.467812Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T20:56:13.459331Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 861185a5-46b3-467b-ac28-671f8ca9c9b8 · inbound
PATS: Process-Level Adaptive Thinking Mode Switching Process-based Self-Rewarding Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30fd01d9-4a5a-4067-8c7f-6140ea62cb82 · inbound
Consistent Paths Lead to Truth: Self-Rewarding Reinforcement Learning for LLM Reasoning Process-based Self-Rewarding Language Models
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8e0183b-dd8c-426c-b887-fb5a31b60961 · inbound
SwS: Self-aware Weakness-driven Problem Synthesis in Reinforcement Learning for LLM Reasoning Process-based Self-Rewarding Language Models
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cefe01cc-2bf6-412e-9050-1c1a2d005fcc · inbound
Temporal Self-Rewarding Language Models: Decoupling Chosen-Rejected via Past-Future Process-based Self-Rewarding Language Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb21e2a1-3c6e-49e3-8ce8-d505200f1477 · inbound
Enhancing Speech Large Language Models through Reinforced Behavior Alignment Process-based Self-Rewarding Language Models
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6a02aacb-5e6a-4cad-b865-ae18588b1717 · inbound
Confidence-Aware Alignment Makes Reasoning LLMs More Reliable Process-based Self-Rewarding Language Models
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9982bb87-d892-4eb1-a149-e18add8e1e38 · inbound
Trust Region On-Policy Distillation Process-based Self-Rewarding Language Models
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.