Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T23:15:28.793471Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 18 inbound Pith citation observations for arXiv:2412.02674.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T23:15:28.793471Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T14:54:29.270448Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T20:56:13.476979Z
62 of 62 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a582408f-2d25-4c20-9cc5-d3096012ef66 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1006b770-934f-41e5-9ed3-7cad476dce90 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Critique-out-Loud Reward Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e86cb86-5877-4f2d-a9bc-0c27bb0838af · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Qwen Technical Report
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e21ec20-274a-43bf-9623-e75a19af96ee · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Constitutional AI: Harmlessness from AI Feedback
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cecfc31-dcfe-4552-9b88-b57bf309ce46 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models On the Stability of Iterative Retraining of Generative Models on their own Data
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2c8edaf-45f1-49d0-9a5a-abaa4a59eb3e · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Large Language Models Suffer From Their Own Output: An Analysis of the Self-Consuming Training Loop
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a626e7b-1941-48ca-aa31-42cb9b82f236 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Large Language Monkeys: Scaling Inference Compute with Repeated Sampling
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2eabb5c7-befa-4308-9e34-e7a3f7b09e84 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Sparks of Artificial General Intelligence: Early experiments with GPT-4
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b86df85-5d4a-4114-bf8b-29e86fa1edcd · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Learning to Generate Better Than Your LLM
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d31c9d90-01d5-4efa-a92b-d52e383a60f7 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Teaching Large Language Models to Self-Debug
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49dffea3-7aed-42a6-aa7f-9ed251e93a59 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7f030f7-35dc-4570-97ee-63008b05c85a · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Can Large Language Models Be an Alternative to Human Evaluations?
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97831683-ac83-4a52-96f8-61f4818844d7 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality.See https://vicuna
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d8a2194-e9ad-46f9-9fa4-ba7564fe7149 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Model Collapse Demystified: The Case of Regression
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb192b01-f8f8-47b4-831f-9bd212ebdeaf · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Training on the Test Task Confounds Evaluation and Emergence
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28e71f36-6cbb-4fcd-998d-7ad8038de913 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models RLHF Workflow: From Reward Modeling to Online RLHF
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20974d57-4de3-4937-bda5-070ac236be73 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models The Llama 3 Herd of Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b14dd521-d354-43a2-8a13-bd9516519e42 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Length-Controlled AlpacaEval: A Simple Way to Debias Automatic Evaluators
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed77794f-ce2e-484c-8aff-76c5130c7452 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Matthias Gerstgrasser, Rylan Schaeffer, Apratim Dey, Rafael Rafailov, Henry Sleight, John Hughes, Tomasz Korbak, Rajashree Agrawal, Dhruv Pai, Andrey Gromov, et al
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95e0dace-5d23-469c-a1dd-ff4d14cc32ae · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Self-Correcting Self-Consuming Loops for Generative Model Training
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7419b19-0da9-4ea2-9723-92b4175c42b8 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Textbooks Are All You Need
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71f6d78d-5621-4854-a1ab-cd6d43502736 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Teaching Large Language Models to Reason with Reinforcement Learning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99c1ad99-4373-41cc-84e8-58185689d9f1 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Scaling laws for downstream task performance of large language models.arXiv preprint arXiv:2402.04177,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ba54d39-b856-4fbb-985f-6dd250ead3cd · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models SELF-[IN]CORRECT: LLMs Struggle with Discriminating Self-Generated Responses
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6b417b9-0794-4870-8504-6ea296da7493 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Scaling Laws for Neural Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75b51781-c29e-4c73-9188-0be6463c098a · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Synthetic Data (Almost) from Scratch: Generalized Instruction Tuning for Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b5a27ef-cb02-45dc-a395-dbd640ca7882 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Textbooks Are All You Need II: phi-1.5 technical report
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 393ae529-41ec-4015-aede-d8de3469f8a7 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models I-SHEEP: Self-Alignment of LLM from Scratch through an Iterative Self-Enhancement Paradigm
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff6abce8-9a92-4643-ae6c-6ceb759e6950 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Let's Verify Step by Step
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e461d49-9e9e-40c5-a340-1ab0792071c5 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Statistical Rejection Sampling Improves Preference Optimization
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11592053-4d9d-43a4-880a-94a3c6b9652f · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Improve Mathematical Reasoning in Language Models by Automated Process Supervision
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f23aae0-c249-476c-bafe-64b6086acdcb · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Combining Generative Artificial Intelligence (AI) and the Internet: Heading towards Evolution or Degradation?
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ff5c2a0-a54e-467e-a1bf-9bb59e954622 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Iterative Reasoning Preference Optimization
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2969d4bf-fb2b-480c-aab9-9f3cccf6ecd0 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Observational Scaling Laws and the Predictability of Language Model Performance
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd27e811-fe89-4086-94dc-46179857da5f · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2b82ff6-6639-41aa-bc6c-a542e42df54e · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e46a5ea-b307-4b1d-bed4-778114a6a13d · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models SuperCLUE: A Comprehensive Chinese Large Language Model Benchmark
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31de1885-8f17-4558-ab61-91933eaa087e · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Gemini: A Family of Highly Capable Multimodal Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2354056d-55f4-4c98-a313-2a42c94dcf98 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0d9e045-3538-42cf-9d5c-60c33672daa0 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Self-Taught Evaluators
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fb67c28-3388-46d9-b612-cde357028e3a · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models From Decoding to Meta-Generation: Inference-time Algorithms for Large Language Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98b2924e-d1d3-456d-bba8-9439cf7f9261 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b3c2369-6c8b-499e-866a-88410e6e5615 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Qwen2 Technical Report
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dc7458b-b4c8-4da7-b51e-d960c619ba71 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Genie: Achieving Human Parity in Content-Grounded Datasets Generation
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 629bf445-2d55-4cc5-83a5-7cabcce69417 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Yi: Open Foundation Models by 01.AI
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0489c865-ea9d-4798-8599-1d8488769214 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Self-Rewarding Language Models
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d00dfe03-37cc-417d-b210-f49f2e73307d · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Quiet-STaR: Language Models Can Teach Themselves to Think Before Speaking
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 143828bb-371f-43d0-8ba7-0e6518bb23a7 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e47622b5-ed4e-465e-a80b-cfeee8c31b99 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Unresolved cited work
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 3c1fe777-9ef9-46c5-9fea-6da85f5d4fbb · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Unresolved cited work
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9b431a19-324a-48eb-af09-6baf75bbfade · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models The answer is ANSWER
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4bb79ffe-ae45-4439-87d8-7c3a80a411ff · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Learning How Hard to Think: Input-Adaptive Allocation of LM Computation
Reference 2005
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1faa3f38-df3a-46fd-b810-1c80e0fae732 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models From Lists to Emojis: How Format Bias Affects Model Alignment
Reference 2006
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63d336fa-0c4c-4805-8f21-c9204165c948 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Qwen2.5-Coder Technical Report
Reference 2007
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b028eb2-10ef-4b77-ba70-6096fc73c12b · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Measuring Mathematical Problem Solving With the MATH Dataset
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 792c3e5f-31ff-46e0-944c-8e18e8bec09d · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Training Verifiers to Solve Math Word Problems
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1064790f-6c2e-435d-a6bf-1a55ab4286f9 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models RankGen: Improving Text Generation with Large Ranking Models
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 243fad25-7fbf-4e43-ad28-32c556fbd12c · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Scaling Laws for Transfer
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96760804-8764-4cbe-b70a-66b31af334d1 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Smaller, Weaker, Yet Better: Training LLM Reasoners via Compute-Optimal Sampling
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a423a56c-b440-4db2-9426-3ce9500910d7 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Nemotron-4 340B Technical Report
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3c60f7e-5b8b-448f-81b5-2d195bebe09b · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Self-Consuming Generative Models Go MAD
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1480e59d-4c59-483d-8432-dd0882773667 · outbound
Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models Large Language Models Can Self-Improve
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 574ba37f-3705-454b-b24d-4ca3b9a0819a · inbound
Self-Improving Transformers Overcome Easy-to-Hard and Length Generalization Challenges Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e61b5bbf-b799-4ccf-ab58-29688b857d9b · inbound
Truly Self-Improving Agents Require Intrinsic Metacognitive Learning Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c0905cb-9aa6-41d5-ac77-19d69cfcfcb6 · inbound
Sample Complexity and Representation Ability of Test-time Scaling Paradigms Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b760d7ee-a2d3-4a3e-a999-b4747dad6391 · inbound
e3: Learning to Explore Enables Extrapolation of Test-Time Compute for LLMs Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecfaabd5-011e-4fb5-8f41-d6591eece8e5 · inbound
PAG: Multi-Turn Reinforced LLM Self-Correction with Policy as Generative Verifier Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb7f58e0-4da1-417f-8b70-2b500455401a · inbound
Because we have LLMs, we Can and Should Pursue Agentic Interpretability Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2a5c216-4747-45e8-91ec-72ca80aed3cd · inbound
An Empirical Study of LLM-as-a-Judge: How Design Choices Impact Evaluation Reliability Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1eb8d0e-a29f-4d6d-8e30-6db51c75e167 · inbound
Learn from What We HAVE: History-Aware VErifier that Reasons about Past Interactions Online Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73d87d62-3600-436e-8f4c-b412c9fef2d6 · inbound
Outcome-based Exploration for LLM Reasoning Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88fa9f2e-85a3-466f-8bbc-a2d61ed641f9 · inbound
A Task-Centric Theory for Iterative Self-Improvement with Easy-to-Hard Curricula Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c17bca8-58cc-4883-b5af-fff3617f1794 · inbound
Beyond Verifiable Rewards: Rubric-Based GRM for Reinforced Fine-Tuning SWE Agents Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d4970fb6-29dc-4158-9ff4-388f7f142f42 · inbound
AutoOR: Scalably Post-training LLMs to Autoformalize Operations Research Problems Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation fc296c37-263e-4b37-918d-b29e828bd96c · inbound
Tree of Concepts: Interpretable Continual Learners in Non-Stationary Clinical Domains Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation caec8726-3d6a-4f00-a7cb-0274bf57bdcc · inbound
Annotations Mitigate Post-Training Mode Collapse Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f5591126-1c6c-4e21-b113-78511853ae51 · inbound
On the Generalization Gap in Self-Evolving Language Model Reasoning Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 95b3bb05-48ac-4514-b2bf-632c27e7fbed · inbound
Trust Region On-Policy Distillation Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6d689bb3-6ea7-4f52-9251-8247dbb3f52d · inbound
Falsification, Not Exposure: An Internally Preregistered Placebo-Controlled Decomposition of Self-Repair Feedback in Frozen Small Code Models Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4a151dce-1fbe-4ea6-af68-2b1378e91128 · inbound
Judging Is Not Enumerating: Silent Omissions in LLM-Authored Acceptable Sets Mind the Gap: Examining the Self-Improvement Capabilities of Large Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.