Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2406.14532.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T11:47:31.286943Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T05:57:41.298545Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 0317369c-a3d0-4623-969c-eaef666a455c · inbound
Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7b6fed2f-15ec-4979-9ef2-4a60d3cf925a · inbound
Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation eb91be93-fdbe-4e6a-9726-dbe0f96d824a · inbound
Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f90cc920-7429-4ba0-b076-b4969716c1e6 · inbound
QLASS: Boosting Language Agent Inference via Q-Guided Stepwise Search RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a82e362-277a-4120-910f-1cc2d62ec641 · inbound
Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0206bbf1-b54a-4626-bcb5-a65cff803769 · inbound
InSTA: Towards Internet-Scale Training For Agents RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1279313-417a-47ed-a1a8-01a313703505 · inbound
DuaShepherd: Integrating Stepwise Correctness and Potential Rewards for Mathematical Reasoning RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c36be57-9ab2-463b-a971-a828d05c21ad · inbound
SCALER:Synthetic Scalable Adaptive Learning Environment for Reasoning RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 200d5800-cc90-4686-a480-1fc7deb21f6f · inbound
Reinforcement Learning Improves LLM Accuracy and Reasoning in Disease Classification from Radiology Reports RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e14860d0-4933-4bff-86ee-ea75944d0428 · inbound
Adaptive Teacher Exposure for Self-Distillation in LLM Reasoning RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5ab74adf-c224-48a9-b97d-76eda0cb26d4 · inbound
Adaptive Teacher Exposure for Self-Distillation in LLM Reasoning RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2bc84b52-7025-4321-a992-b85bdbac54b7 · inbound
The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold
Reference 204
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 22ac0092-d162-4cae-82a9-e23c83dbb28c · inbound
Reason Before You Retrieve: Agentic Planning for Multi-modal RAG RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-Fold
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.