Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:07:36.826395Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 3 inbound Pith citation observations for arXiv:2507.06573.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:07:36.826395Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-29T14:08:40.968105Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-29T14:13:30.120402Z
13 of 13 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ef897c4f-df5c-408d-a1dd-cdceafedd80a · outbound
From Data-Centric to Sample-Centric: Enhancing LLM Reasoning via Progressive Optimization DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 847c084f-a9ab-43c3-aebb-fc2e12b2151b · outbound
From Data-Centric to Sample-Centric: Enhancing LLM Reasoning via Progressive Optimization In The Twelfth In- ternational Conference on Learning Representations, ICLR 2024, Vienna, Austria, May 7-11, 2024
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b2e1df28-db64-4825-b65a-54c5fb249698 · outbound
From Data-Centric to Sample-Centric: Enhancing LLM Reasoning via Progressive Optimization Understanding R1-Zero-Like Training: A Critical Perspective
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48c950b9-6541-4e7b-8811-f18eff5a7c80 · outbound
From Data-Centric to Sample-Centric: Enhancing LLM Reasoning via Progressive Optimization Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0489ca83-0458-43a3-a87a-bb07e55e015d · outbound
From Data-Centric to Sample-Centric: Enhancing LLM Reasoning via Progressive Optimization DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ee80aa1-e4b6-4d90-ba9b-0c678caca20c · outbound
From Data-Centric to Sample-Centric: Enhancing LLM Reasoning via Progressive Optimization SRPO: A Cross-Domain Implementation of Large-Scale Reinforcement Learning on LLM
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06880c01-66ea-45a5-bfc6-f2f41724fff8 · outbound
From Data-Centric to Sample-Centric: Enhancing LLM Reasoning via Progressive Optimization Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 35635517-f411-4aa6-8762-66ed6ee0b8c0 · outbound
From Data-Centric to Sample-Centric: Enhancing LLM Reasoning via Progressive Optimization Aitor Lewkowycz, Anders Andreassen, David Dohan, Ethan Dyer, Henryk Michalewski, Vinay V
Reference 626
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e9f16861-1523-4116-9bcc-19502952b717 · outbound
From Data-Centric to Sample-Centric: Enhancing LLM Reasoning via Progressive Optimization Figure 8 illustrates the cosine similarity between the embeddings of model-generated solutions and oracle expert solutions, comparing training with and without PG-Sampling
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fbab5568-6cdb-48f6-9369-0112cd8020ae · outbound
From Data-Centric to Sample-Centric: Enhancing LLM Reasoning via Progressive Optimization Proximal Policy Optimization Algorithms
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b894d50c-c028-4d0c-9c0d-fb4703238dab · outbound
From Data-Centric to Sample-Centric: Enhancing LLM Reasoning via Progressive Optimization LIMR: Less is More for RL Scaling
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 219b8998-dcc0-4f6c-8368-7323e81837a5 · outbound
From Data-Centric to Sample-Centric: Enhancing LLM Reasoning via Progressive Optimization Process Reinforcement through Implicit Rewards
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f738fbeb-dc33-4e5b-9782-9a22379c6465 · outbound
From Data-Centric to Sample-Centric: Enhancing LLM Reasoning via Progressive Optimization Large-Scale Data Selection for Instruction Tuning
Reference 9061
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23659746-3e51-4ba5-b8ca-8aac15e366d7 · inbound
SCALER:Synthetic Scalable Adaptive Learning Environment for Reasoning From Data-Centric to Sample-Centric: Enhancing LLM Reasoning via Progressive Optimization
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 988c29da-eaf5-464d-b870-5250a0180a72 · inbound
Generate, Filter, Control, Replay: A Comprehensive Survey of Rollout Strategies for LLM Reinforcement Learning From Data-Centric to Sample-Centric: Enhancing LLM Reasoning via Progressive Optimization
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ba1428bf-2956-452f-bdb3-71112283d2dd · inbound
Single-Rollout Hidden-State Dynamics for Training-Free RLVR Data Selection From Data-Centric to Sample-Centric: Enhancing LLM Reasoning via Progressive Optimization
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.