Pith. sign in

Paper Citation Record · LEDGER

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula

As of 16 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:2602.10014.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2602.10014 v3

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T02:43:38.526645Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

43 of 43 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved43
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 51242c2b-32fb-4315-9ba8-4a9e5f9a5f36 · outbound

This paper cites Intrinsic dimensionality explains the effectiveness of language model fine-tuning.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Intrinsic dimensionality explains the effectiveness of language model fine-tuning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:33.314701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:33.314701Z digest=sha256:fe0272e5f1c276c770395937bfcc42c92ae99f400d243a0e6c57abb7e747d5ef

Observation 1246c825-e4ae-4c10-a5c6-f8e7e9dd842c · outbound

This paper cites Smaller, Weaker, Yet Better: Training LLM Reasoners via Compute-Optimal Sampling.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Smaller, Weaker, Yet Better: Training LLM Reasoners via Compute-Optimal Sampling

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:33.445001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:33.445001Z digest=sha256:d67ef3579b73595e45fb738bf6358580854401e9f87c158757687630a0e777f5

Observation baaa6c08-d736-4e3d-ba72-5e8491f35f24 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Training Verifiers to Solve Math Word Problems

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:33.551277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:33.551277Z digest=sha256:4734b52669e949b190ec4ab60010f6d0e72df730a676237317bb786b11c7e817

Observation a70328de-abc2-4832-8a73-a421b67040fa · outbound

This paper cites Understanding self-distillation in the presence of label noise.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Understanding self-distillation in the presence of label noise

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:33.684760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:33.684760Z digest=sha256:c02a87973c6b10aa6d48be03845f9f5c5d4a1f4e46d36240b7e7070c75ae05d2

Observation 68b96946-0003-4499-b0d7-c3341c91fb8c · outbound

This paper cites Beyond Model Collapse: Scaling Up with Synthesized Data Requires Verification.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Beyond Model Collapse: Scaling Up with Synthesized Data Requires Verification

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:33.780269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:33.780269Z digest=sha256:3c52734741103cfd06ec428eaaed5ba8f54c139d6dc4355fb637a98d9530b88a

Observation 2cbb0070-04db-4858-aea1-80b079085422 · outbound

This paper cites Self-Consuming Generative Models with Curated Data Provably Optimize Human Preferences.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Self-Consuming Generative Models with Curated Data Provably Optimize Human Preferences

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:33.964748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:33.964748Z digest=sha256:89f275b42ae46db66f239c86c09bec7798fb0e58f70674d8d294115829746d4a

Observation 871c7695-7369-4a87-a32b-6d309d392038 · outbound

This paper cites Towards Theoretical Understandings of Self-Consuming Generative Models.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Towards Theoretical Understandings of Self-Consuming Generative Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:34.066179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:34.066179Z digest=sha256:c7b2a19c78ae9ddb496220a733b324a02a2c610520f2791569276a3426f58b30

Observation fea1b8b6-cc38-4cb4-8a42-377033b77326 · outbound

This paper cites Self-verification provably prevents model collapse in recursive synthetic training.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Self-verification provably prevents model collapse in recursive synthetic training

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:34.174752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:34.174752Z digest=sha256:03784ea190cf8408fcf57f2d4521c8151dd26cf3b08ff0c0f9b8f853d82148b9

Observation ee23ba26-3d4d-45b6-9289-8f71247c0d0f · outbound

This paper cites A Theoretical Perspective: How to Prevent Model Collapse in Self-consuming Training Loops.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula A Theoretical Perspective: How to Prevent Model Collapse in Self-consuming Training Loops

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:34.264756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:34.264756Z digest=sha256:8cd9d4f48fa6e8a63c50920e82c4f821539417132a5e4f34f04b42b3a533ac54

Observation 902ae028-2bfa-42e2-88aa-105ebf427232 · outbound

This paper cites Cambridge university press, 2000.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Cambridge university press, 2000

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:34.312806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:34.312806Z digest=sha256:9f3b13020e324f7c1de3b0486f2ea2bc7e47ac7c3c93ae8c57f4d65fc17eb67d

Observation 7e847d73-d621-438f-83fa-1154dc2ffddb · outbound

This paper cites Is Model Collapse Inevitable? Breaking the Curse of Recursion by Accumulating Real and Synthetic Data.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Is Model Collapse Inevitable? Breaking the Curse of Recursion by Accumulating Real and Synthetic Data

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:34.442154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:34.442154Z digest=sha256:cd29798963c70a67c2b5b987e859bfc0ebcf7ae4f8f4194627a8a739713dc7c4

Observation bb326200-a19a-47d3-b716-0d536753fb65 · outbound

This paper cites Self-Correcting Self-Consuming Loops for Generative Model Training.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Self-Correcting Self-Consuming Loops for Generative Model Training

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:34.544741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:34.544741Z digest=sha256:331a904b5f45a50be620e79a54ad9145bfc31d868bfc1d90b60fa2d313134a91

Observation 69fa0049-0ba1-41dc-8cca-90e6212130fd · outbound

This paper cites The Llama 3 Herd of Models.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula The Llama 3 Herd of Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:34.644029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:34.644029Z digest=sha256:da8175a66120b9850a44988c7625a23c5ec8d652abf04bccc46155e91b336180

Observation 2a65ce3f-459f-4877-9f61-32c6531bc8d9 · outbound

This paper cites rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:34.738415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:34.738415Z digest=sha256:9bbd033ab87178a1706baf9891e1d70e8f7b6ff2ca02a2ac1276ce821dafbc1a

Observation ca9bd427-7575-464f-abd9-784748578992 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:34.983533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:34.983533Z digest=sha256:6a6f1c23cdbcc53cef94ed9bc05ec915a61bff58b83448cfd0fc7a8e1ba0ca71

Observation d9e3b4c8-2b19-4b7e-8c88-eaf2aad9f39a · outbound

This paper cites Self-Improvement in Language Models: The Sharpening Mechanism.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Self-Improvement in Language Models: The Sharpening Mechanism

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:35.074111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:35.074111Z digest=sha256:b1c65a75298cf8b92c13bf8cc30e68a0705d2d58d4cf9a60cbebee5c3f82ad9e

Observation a92164a4-cfe1-4d16-8170-e066dbd66c7f · outbound

This paper cites Adastar: Adaptive data sampling for training self-taught reasoners.arXiv preprint arXiv:2505.16322, 2025.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Adastar: Adaptive data sampling for training self-taught reasoners.arXiv preprint arXiv:2505.16322, 2025

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:35.187626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:35.187626Z digest=sha256:4138530775add28a85cf03888befe350cde869c077b8501ee1e1c339a3a8916e

Observation 999b486e-0963-44e2-9b85-109a71a815a3 · outbound

This paper cites Self-Improving Transformers Overcome Easy-to-Hard and Length Generalization Challenges.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Self-Improving Transformers Overcome Easy-to-Hard and Length Generalization Challenges

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:35.308208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:35.308208Z digest=sha256:ebf0cdc69dad6550aeb55b62eaa0ff3578db877071f26166c95f37fd557e3e51

Observation 17243480-9c70-49c8-a8c3-274f04788f60 · outbound

This paper cites Goedel-Prover: A Frontier Model for Open-Source Automated Theorem Proving.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Goedel-Prover: A Frontier Model for Open-Source Automated Theorem Proving

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:35.409770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:35.409770Z digest=sha256:b60207d16819933c9f3f7e2ff8895331488bcc44b79d7e70ef4dbb2a188593f3

Observation b9d2e9c4-25a8-4196-a655-764c20f43697 · outbound

This paper cites Goedel-Prover-V2: Scaling Formal Theorem Proving with Scaffolded Data Synthesis and Self-Correction.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Goedel-Prover-V2: Scaling Formal Theorem Proving with Scaffolded Data Synthesis and Self-Correction

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:35.496157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:35.496157Z digest=sha256:2ff7ea32d83721b74bbc8c521676b9d3b5751450a2d4828de748ec5ecc7da635

Observation 5ab37e65-a95b-4d2c-b9c9-a76742e97a27 · outbound

This paper cites Self-distillation amplifies regular- ization in hilbert space.Advances in Neural Information Processing Systems, 33:3351–3361, 2020.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Self-distillation amplifies regular- ization in hilbert space.Advances in Neural Information Processing Systems, 33:3351–3361, 2020

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:36.035387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:36.035387Z digest=sha256:a2245d5a3c7929f217f047910aba16226f4ad025e2d2626c4548657fbb685b81

Observation 0af07b42-b1e5-46ec-b42b-f12b08630805 · outbound

This paper cites Coherence mechanisms for provable self- improvement.arXiv preprint arXiv:2511.08440, 2025.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Coherence mechanisms for provable self- improvement.arXiv preprint arXiv:2511.08440, 2025

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:36.185231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:36.185231Z digest=sha256:ccd540132e9fddda99b8f48ced1cc4d95ec779c4c3de50aa4592090d26fa04ce

Observation 162066db-64c5-4a1a-a34d-48d9c10abfab · outbound

This paper cites Understanding the gains from repeated self-distillation.Advances in Neural Information Processing Systems, 37:7759–7796, 2024.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Understanding the gains from repeated self-distillation.Advances in Neural Information Processing Systems, 37:7759–7796, 2024

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:36.294740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:36.294740Z digest=sha256:95df7d82b1e7292d1cf8187285a5cebbd0f73706f631db24b2720c302d6251cd

Observation e8e5813b-eb28-4e05-b35e-49d955d1df80 · outbound

This paper cites DeepSeek-Prover-V2: Advancing Formal Mathematical Reasoning via Reinforcement Learning for Subgoal Decomposition.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula DeepSeek-Prover-V2: Advancing Formal Mathematical Reasoning via Reinforcement Learning for Subgoal Decomposition

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:36.414740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:36.414740Z digest=sha256:64dd65a5a521477c420a13da630801a7c530ef6f962c56d7e63eb47cec9781ed

Observation d3c6978b-91e7-42ca-b492-24a4c81d0505 · outbound

This paper cites Analysing Mathematical Reasoning Abilities of Neural Models.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Analysing Mathematical Reasoning Abilities of Neural Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:36.501769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:36.501769Z digest=sha256:5ec95cd0e05cb0b82e3466ae17f9438fa77f90aa98ef964496644a24e6234381

Observation 33e0d071-2412-44e2-9a11-b8b011ebea37 · outbound

This paper cites A mathematical theory of communication.The Bell system technical journal, 27(3):379–423, 1948.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula A mathematical theory of communication.The Bell system technical journal, 27(3):379–423, 1948

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:36.581709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:36.581709Z digest=sha256:9aaa72041c3e735e495be5e65de05f75f669fe0cbfb737adb6c019bb884c65ed

Observation 5f340510-3d13-47fd-b388-cdd51ecdeb84 · outbound

This paper cites Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:36.662189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:36.662189Z digest=sha256:8d936e6630bdbfdb31d29b69c36fb63c6bf0615bc30ad696a21769138024e440

Observation 88fa9f2e-85a3-466f-8bbc-a2d61ed641f9 · outbound

This paper cites Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:36.764821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:36.764821Z digest=sha256:72d146a3f6c8b38713838a071257600317977ab7675081b11b75a9a4e0db2563

Observation 02bd9ddc-50c3-4902-8b80-ff068cdc6fa0 · outbound

This paper cites Theoretical modeling of llm self- improvement training dynamics through solver-verifier gap.arXiv preprint arXiv:2507.00075, 2025.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Theoretical modeling of llm self- improvement training dynamics through solver-verifier gap.arXiv preprint arXiv:2507.00075, 2025

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:36.832695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:36.832695Z digest=sha256:24593b1afcaefa83881a573b66f4f301836f9042c4027d3304a57acba4614c8c

Observation a8587e93-a18b-4602-b665-802c3ad1ecdb · outbound

This paper cites Canlanguagemodelssolvegraphproblemsinnaturallanguage?Advances in Neural Information Processing Systems, 36:30840–30861, 2023.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Canlanguagemodelssolvegraphproblemsinnaturallanguage?Advances in Neural Information Processing Systems, 36:30840–30861, 2023

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:36.903491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:36.903491Z digest=sha256:685077c8e903989bc133d7f65025905cb216252a9683f9045cb83f44d67cce63

Observation 676c08f5-ff8e-46d0-b47b-4e2cea3bb481 · outbound

This paper cites Huxley-gödel machine: Human-level coding agent development by an approximation of the optimal self-improving machine, 2025.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Huxley-gödel machine: Human-level coding agent development by an approximation of the optimal self-improving machine, 2025

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:37.006568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:37.006568Z digest=sha256:4e9835c762254a2aa9602de2682bb91bbef758b61a8e0851659978abb653a853

Observation e00b3aec-fe22-4c45-bc02-c247ea87cf25 · outbound

This paper cites Propose, solve, verify: Self-play through formal verification.arXiv preprint arXiv:2512.18160, 2025.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Propose, solve, verify: Self-play through formal verification.arXiv preprint arXiv:2512.18160, 2025

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:37.084920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:37.084920Z digest=sha256:4a6929b3094b35a04d1ca53ad541469400b4dc8e8318fc724c850ab08af48dbd

Observation 0dc1b3f4-a652-4e3b-b1d7-be5c7def22b5 · outbound

This paper cites Probability inequalities for likelihood ratios and convergence rates of sieve mles.The Annals of Statistics, pages 339–362, 1995.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Probability inequalities for likelihood ratios and convergence rates of sieve mles.The Annals of Statistics, pages 339–362, 1995

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:37.203560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:37.203560Z digest=sha256:40243a1d222a2836538d2ed2401c3cd18419a89bddbc040032d650340b37caa1

Observation 318dd6b9-3087-411e-8bd5-08c546a795ea · outbound

This paper cites DeepSeek-Prover: Advancing Theorem Proving in LLMs through Large-Scale Synthetic Data.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula DeepSeek-Prover: Advancing Theorem Proving in LLMs through Large-Scale Synthetic Data

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:37.405053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:37.405053Z digest=sha256:77ee9574ee4c4e859f8f2001b616b8fe01f4e86e3401273d6d98b4befa434a7f

Observation f95f7968-3759-442c-9096-7548cdb8b70c · outbound

This paper cites DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:37.574813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:37.574813Z digest=sha256:a5d5f893f7cc912501974ce51a4fddc4a0a3bc21fd9fff4bf2ae2ccb22eb3e7a

Observation 97c7ebbf-986f-4ad1-9881-97eef9c48816 · outbound

This paper cites Qwen2.5 Technical Report.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Qwen2.5 Technical Report

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:37.684535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:37.684535Z digest=sha256:cb86f4325cee64bbda851220eff563912aa7873610c0a7cd2092d30b86c1d16b

Observation b9071e0e-7881-4e59-8c92-4803d3695144 · outbound

This paper cites Qwen3 Technical Report.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Qwen3 Technical Report

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:37.854813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:37.854813Z digest=sha256:0cee3b634e33a0166129ee0b81cefcabd9b5f2fe7b7ea2cfb2b7b26c128eb575

Observation 5f732155-4010-4193-b5ff-ca448e3493bd · outbound

This paper cites Spendwisely: Maximizing post-training gains in iterative synthetic data bootstrapping.arXiv preprint arXiv:2501.18962, 2025.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Spendwisely: Maximizing post-training gains in iterative synthetic data bootstrapping.arXiv preprint arXiv:2501.18962, 2025

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:37.911309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:37.911309Z digest=sha256:09272b56667bda985ada13ef9dafd7b9ed1b318962a08b71840c1f2d36eff0f1

Observation c27f7bb9-1692-4506-9202-d39226132c92 · outbound

This paper cites Optimizing Chain-of-Thought Reasoners via Gradient Variance Minimization in Rejection Sampling and RL.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Optimizing Chain-of-Thought Reasoners via Gradient Variance Minimization in Rejection Sampling and RL

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:38.039514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:38.039514Z digest=sha256:cfa20256113e7e9ab8768b63318bc673c92f81e74a537b8cd30ba97468b42693

Observation 5d617e96-254c-45fe-97fc-31896e12c58e · outbound

This paper cites Star: Bootstrapping reasoning with reasoning.Advances in Neural Information Processing Systems, 35:15476–15488, 2022.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Star: Bootstrapping reasoning with reasoning.Advances in Neural Information Processing Systems, 35:15476–15488, 2022

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:38.143324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:38.143324Z digest=sha256:0c3b03cad2dd70163934d99e593c8ec3ece751d5b3b236caabd933e132a7524a

Observation 95a82f25-5c48-47a9-b428-55d1ef7f0dd4 · outbound

This paper cites B-STaR: Monitoring and Balancing Exploration and Exploitation in Self-Taught Reasoners.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula B-STaR: Monitoring and Balancing Exploration and Exploitation in Self-Taught Reasoners

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:38.300114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:38.300114Z digest=sha256:f184a429b28410467c9f757a6a0ba640e7e33b070ab131909f86aa910fc44a8d

Observation 4981f9a3-4db2-4525-89a3-bdfec3acf38e · outbound

This paper cites Leanabell-Prover: Posttraining Scaling in Formal Reasoning.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Leanabell-Prover: Posttraining Scaling in Formal Reasoning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:38.423003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:38.423003Z digest=sha256:4eac4521d70d1361ed52195dd66898d5377fb91bba63745f1900cb1f3da03639

Observation 743d4d8b-b4bd-4634-b99b-3ef94e26d084 · outbound

This paper cites question distribution.

A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula question distribution

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-03T02:43:38.526645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:43:38.526645Z digest=sha256:6dd5530a9a4d9fddeaf2c90262764fa1e83404d7c893fb373c02381110231184

Pith citing papers

No inbound Pith citation observations are available.