Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:51:30.802158Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 7 inbound Pith citation observations for arXiv:2506.06923.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:51:30.802158Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T02:12:24.953400Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
44 of 44 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 72f5906c-14e6-4017-99f2-6a8016d04b9f · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Back to Basics: Revisiting REINFORCE Style Optimization for Learning from Human Feedback in LLMs
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ffdbfdd-1f2c-4183-a0a0-8869dca2b8a3 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction AutoPRM: Automating Procedural Supervision for Multi-Step Reasoning via Controllable Question Decomposition
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 821ba7ec-47d2-4c1a-9730-cdf55bfb9a5f · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21619a22-cada-4584-ad50-89d5977f8072 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fcfeb08-6ff6-4638-acff-9a8ffcd9f872 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Measuring Mathematical Problem Solving With the MATH Dataset
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 064fd908-3486-44a2-9119-af0b37e847b5 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Training Language Models to Self-Correct via Reinforcement Learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6480e74e-4923-47ad-a159-a15f1759e899 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Encouraging Divergent Thinking in Large Language Models through Multi-Agent Debate
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3faa4b9-3db5-410e-b94e-bd522d5d2203 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Let's Verify Step by Step
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d5e7b6f-4e1c-4faf-9f00-b351615363dd · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 639f6b94-bffe-4676-8378-ca4216a1b2b4 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Deepseek-r1 thoughtology: Let’s think about llm reasoning.arXiv preprint arXiv:2504.07128,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e255617b-dfeb-4c4b-87f2-5a314ab0dbf4 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Orca-Math: Unlocking the potential of SLMs in Grade School Math
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0864680-2e7a-469e-8c9a-8d99867be80d · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Malt: Improving reasoning with multi-agent llm training
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22a7ff88-7d46-4db5-ab0c-6d65423d6042 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Recursive Introspection: Teaching Language Model Agents How to Self-Improve
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a87dd523-30df-455b-8f44-7739100a0701 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Self-critiquing models for assisting human evaluators
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d8a13ba-1438-4563-baba-31d72530b596 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Generating Sequences by Learning to Self-Correct
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf952e5b-92b0-464d-b2bf-5bc6059a3ee0 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45d72fcc-112d-4e52-9821-6762d0f869d5 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cbcdcb0-37db-4fba-b52b-61e85f8acf78 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Iterative Preference Learning from Human Feedback: Bridging Theory and Practice for RLHF under KL-Constraint
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a00c80d3-824f-4ab2-922a-582cb76fbd30 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Self-rewarding correction for mathematical reasoning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9d12446-f629-4000-8a3e-681432b58a87 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction The Perfect Blend: Redefining RLHF with Mixture of Judges
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 481e8170-2563-43b9-ac69-ffd75e58a9eb · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6724760-13bd-44c8-bab3-1b31fdef9bd7 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Physics of Language Models: Part 2.2, How to Learn From Mistakes on Grade-School Math Problems
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10ff8058-bac5-4f70-b4b2-50d5cd7b6289 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Generative Verifiers: Reward Modeling as Next-Token Prediction
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a0cdfcf-6cc6-42af-b74c-338472202853 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Critic-CoT: Boosting the reasoning abilities of large language model via Chain-of-thoughts Critic
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 521ee6e4-ca2e-4d74-be1b-ff9f04abe176 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a46726ca-a597-4e2a-8506-f48cb7332d72 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction This test set spans five difficulty levels and seven subjects, which promotes a comprehensive evaluation of reasoning capabilities
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 206d89b9-7f9a-4304-a511-5fe7b7a0cbe1 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction I think the solution is correct
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b056555e-06dd-419d-9c78-59a1b9ef497d · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 27a73b57-0698-48bc-83af-f9b116da876a · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction ## Step 4: Verify if \( a = 1 \) satisfies the conditions of the problem
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c5523e90-8edf-4613-b87e-db957766b547 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction This means \( r^2 = 2009 \) is not possible for any integer \( r \) since 2009 is not a perfect square
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9da34c37-d85e-4443-b0ac-74daaee1a0ce · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Thus, \( b = 1 \) is not possible since \( a < b \), implying \( a \) would have to be less than 1, which is not possible for positive integers
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 316997fc-3f4d-49ad-bfa6-4147c8b15705 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Unresolved cited work
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4e7bab44-6516-491d-bac4-cf78265fcb73 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Unresolved cited work
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3a97133e-bba6-480b-856a-a065388f58cb · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Thus, \( r^2 = 1 \), giving \( r = 1 \) or \( r = -1 \)
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 97e67abb-7f35-4632-a4f8-1d62ac7e2da8 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Unresolved cited work
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f155d4f3-4747-4e7c-872a-8db3d2a0e2ea · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction REFINER: Reasoning Feedback on Intermediate Representations
Reference 1994
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c712f5b-949f-4280-a6ee-3b068b91deb3 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 2008
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5819207-0b3a-4ae6-aabe-69630fdf17be · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction To find the factors of 2009, we can start by checking for its prime factorization
Reference 2009
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7290b59d-4f0d-4708-af25-53e0c5171f20 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57ab5821-d2de-4f18-a0c1-e2d79764e2e1 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction Large Language Models Cannot Self-Correct Reasoning Yet
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1377f845-2e76-4d66-a6cc-ee47c3cde664 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1949295-9ccd-4115-96c8-d4df2ae36cc4 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction The Llama 3 Herd of Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5a81be3-6384-4657-9ab5-a0f67f876325 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction RL4F: Generating Natural Language Feedback with Reinforcement Learning for Repairing Model Outputs
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b210253c-344a-4e54-93cf-2142dcf91728 · outbound
Boosting LLM Reasoning via Spontaneous Self-Correction GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b132b7c-82ef-4045-8111-aef36f124d68 · inbound
Failure Makes the Agent Stronger: Enhancing Accuracy through Structured Reflection for Reliable Tool Interactions Boosting LLM Reasoning via Spontaneous Self-Correction
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 86faf3ca-6145-4caf-afeb-ba9ece208066 · inbound
Token-Level LLM Collaboration via FusionRoute Boosting LLM Reasoning via Spontaneous Self-Correction
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ba8d26e6-7c0c-4fbe-8735-0607b64bd0e5 · inbound
The Self-Correction Illusion: Role Relabeling Gates Explicit Error Flagging in Large Language Models Boosting LLM Reasoning via Spontaneous Self-Correction
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d859d439-49c4-4f68-94c9-544189e6e1dc · inbound
ReSum: Synergizing LLM Reasoning and Summarization with Reinforcement Learning Boosting LLM Reasoning via Spontaneous Self-Correction
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8c53babe-d41d-4111-b267-a2debe949841 · inbound
ReSum: Synergizing LLM Reasoning and Summarization with Reinforcement Learning Boosting LLM Reasoning via Spontaneous Self-Correction
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c74e4ae-a7d8-4650-a512-401250934550 · inbound
Mixture of Debaters: Learn to Debate at Architectural Level in Multi-Agent Reasoning Boosting LLM Reasoning via Spontaneous Self-Correction
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 44d7a380-53a1-459e-874b-f727b8fb1aba · inbound
SVR-R1: Bootstrapping Multi-modal Reasoning with Self-verification in Reinforcement Learning Boosting LLM Reasoning via Spontaneous Self-Correction
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.