Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:14:53.840638Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 0 inbound Pith citation observations for arXiv:2608.02139.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:14:53.840638Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
25 of 25 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a86376f1-91ec-4e3f-a42a-ab7681d51475 · outbound
Self-Improving Large Language Models via Progressive Experience Evolution Training Verifiers to Solve Math Word Problems
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ae55e69-9a49-4504-b1b4-43e0f52c452d · outbound
Self-Improving Large Language Models via Progressive Experience Evolution Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f9e970c-930a-4876-a105-f50ebe8d9dfc · outbound
Self-Improving Large Language Models via Progressive Experience Evolution Let's Verify Step by Step
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd8b029b-c614-420d-963a-c6901beb1782 · outbound
Self-Improving Large Language Models via Progressive Experience Evolution ReFT: Reasoning with Reinforced Fine-Tuning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ab77ad3-d340-48c2-8ca2-09ad79301ced · outbound
Self-Improving Large Language Models via Progressive Experience Evolution Mexico City, Mexico: Association for Compu- tational Linguistics
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 17bae05d-6cd2-4496-a84d-ee2afd62df82 · outbound
Self-Improving Large Language Models via Progressive Experience Evolution Privileged Information Distillation for Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f931f582-dfb6-4887-89ad-1bd468f449cd · outbound
Self-Improving Large Language Models via Progressive Experience Evolution DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf57b854-9cf4-4a50-9db3-79684a2b7804 · outbound
Self-Improving Large Language Models via Progressive Experience Evolution Beyond Trajectory Imitation: Strategy-Guided Policy Optimization for LLM Reasoning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cb45d8de-3697-4783-9d55-96d7e867809c · outbound
Self-Improving Large Language Models via Progressive Experience Evolution Learning to summarize from human feedback
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fc808de-7b5f-48e0-8ca0-15ed7ed9a7d0 · outbound
Self-Improving Large Language Models via Progressive Experience Evolution Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7ee965ff-f280-4811-a330-c9a58e959cee · outbound
Self-Improving Large Language Models via Progressive Experience Evolution Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 917f942a-4e44-4f62-9862-d96d96fa09b2 · outbound
Self-Improving Large Language Models via Progressive Experience Evolution Voyager: An Open-Ended Embodied Agent with Large Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae2fa355-1d36-42e6-b2c9-856a172610a1 · outbound
Self-Improving Large Language Models via Progressive Experience Evolution Self-Consistency Improves Chain of Thought Reasoning in Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90a4fbce-1e03-4d14-8fca-69c32febfd6f · outbound
Self-Improving Large Language Models via Progressive Experience Evolution RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 552e9929-b0b7-4a2f-a4b3-f196aa85df3c · outbound
Self-Improving Large Language Models via Progressive Experience Evolution ReAct: Synergizing Reasoning and Acting in Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69d5ac64-ecc7-4360-ad19-e41cc7ad756c · outbound
Self-Improving Large Language Models via Progressive Experience Evolution Expel:Llmagentsareexperientiallearners
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6884dd0e-b56d-48ac-b1f8-b0a7828cbbea · outbound
Self-Improving Large Language Models via Progressive Experience Evolution Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f922c2e8-bb6f-4c45-8aeb-fe438a28dee5 · outbound
Self-Improving Large Language Models via Progressive Experience Evolution Fine-Tuning Language Models from Human Preferences
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d7fd242-6134-4c8c-87ed-4cc0e74efd5a · outbound
Self-Improving Large Language Models via Progressive Experience Evolution Distilling the Knowledge in a Neural Network
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 936ce180-9f81-4839-9535-6323bcc33898 · outbound
Self-Improving Large Language Models via Progressive Experience Evolution Scaling Laws for Neural Language Models
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4058ba69-5813-41c8-83de-99ec992faf60 · outbound
Self-Improving Large Language Models via Progressive Experience Evolution CodeRL: Mastering Code Generation through Pretrained Models and Deep Reinforcement Learning
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1261451e-4e58-4460-b123-6dab728803a0 · outbound
Self-Improving Large Language Models via Progressive Experience Evolution GPT-4 Technical Report
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 562d20f0-28e4-49f3-a75b-25b9f7dc0ce7 · outbound
Self-Improving Large Language Models via Progressive Experience Evolution On-Policy Distillation of Language Models: Learning from Self-Generated Mistakes
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3dc3a32-dfd4-493b-91d6-fa61925135e4 · outbound
Self-Improving Large Language Models via Progressive Experience Evolution DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c0bdd01-606b-4211-8a2b-71f6b2bf37e1 · outbound
Self-Improving Large Language Models via Progressive Experience Evolution MiniLLM: On-Policy Distillation of Large Language Models
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.