Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 48 inbound Pith citation observations for arXiv:2410.02089.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:35:50.613928Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
6
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 91e1c7b4-1331-4b0d-8ee6-10cdad3b9da5 · inbound
Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4f86d814-6f5b-4ad7-a836-828d61c88d09 · inbound
SWE-RL: Advancing LLM Reasoning via Reinforcement Learning on Open Software Evolution RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dfd6e967-1c5b-4d72-be45-fe8404b61f48 · inbound
Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 206
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 43bae483-6515-488d-813f-dedbbcfb33ce · inbound
Gemma 3 Technical Report RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c56626ec-d521-44db-b591-9947b9a4ba79 · inbound
Reinforcement Learning from Human Feedback RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 166
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 41cd47ac-0822-4c76-8982-68592cfe2bbc · inbound
Reinforcing General Reasoning without Verifiers RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 484deb15-1a48-4b08-a34d-273bc18125e1 · inbound
Training Language Models to Generate Quality Code with Program Analysis Feedback RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5badb4da-5bed-4308-8c0b-fd622e9b4717 · inbound
Afterburner: Reinforcement Learning Facilitates Self-Improving Code Efficiency Optimization RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1aff602b-b69d-462f-8792-7d66865a4e5d · inbound
LlamaRL: A Distributed Asynchronous Reinforcement Learning Framework for Efficient Large-scale LLM Training RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42e8fd50-6085-47b2-9879-b8b77ad352ed · inbound
Improving LLM-Generated Code Quality with GRPO RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43b36156-2531-46f0-b9dd-afe892cd943c · inbound
Adaptive Guidance Accelerates Reinforcement Learning of Reasoning Models RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b9a162a-60ac-446a-8ddc-30c8f0d508da · inbound
Reinforcement learning fine-tuning of language model for instruction following and math reasoning RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b939a93-5116-451a-b01f-0489329ea5b0 · inbound
CodeGrad: Integrating Multi-Step Verification with Gradient-Based LLM Refinement RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea83a352-454f-448a-810f-8d2cc060d73a · inbound
VERIRL: Boosting the LLM-based Verilog Code Generation via Reinforcement Learning RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97650b8c-9389-4fae-9456-c96ce892e39f · inbound
Short window attention enables long-term memorization RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9651214a-fb95-4701-9ae9-a09d438c2181 · inbound
KernelEvolve: Scaling Agentic Kernel Coding for Heterogeneous AI Accelerators at Meta RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba509e48-0377-4279-bbfb-229312551200 · inbound
Beyond Binary: Turning Partial Success into Dense Verifiable Rewards for Reinforcement Learning in Code Generation RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f6b0ecb-48a8-4d63-a242-e6c291cdc467 · inbound
NEURA: A Unified and Retargetable Compilation Framework for Coarse-Grained Reconfigurable Architectures RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b38a1b66-17ac-4be1-844c-5a99239250b9 · inbound
Beyond Fixed Tests: Repository-Level Issue Resolution as Coevolution of Code and Behavioral Constraints RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f5e72726-7d9c-4e30-9322-09cbe9cfdc61 · inbound
An Iterative Test-and-Repair Framework for Competitive Code Generation RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 98fd10ca-0dbe-4f9b-bc8d-ac1fcde25d4e · inbound
An Iterative Test-and-Repair Framework for Competitive Code Generation RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6d7d84f-326b-455e-8077-d709f2e0fac7 · inbound
Scientific Graphics Program Synthesis via Dual Self-Consistency Reinforcement Learning RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1027fdd9-2fe5-4d0a-8b08-6c6c3423f5e0 · inbound
Controllable and Verifiable Tool-Use Data Synthesis for Agentic Reinforcement Learning RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d4a371a1-8bf5-49fd-8ab3-d17134a5225b · inbound
Beyond Verifiable Rewards: Rubric-Based GRM for Reinforced Fine-Tuning SWE Agents RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5335f5e0-a2ef-4c60-aca3-0809b19ee7ab · inbound
CodePivot: Bootstrapping Multilingual Transpilation in LLMs via Reinforcement Learning without Parallel Corpora RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation de5616c2-4498-479a-80ef-866762313403 · inbound
PYTHALAB-MERA: Validation-Grounded Memory, Retrieval, and Acceptance Control for Frozen-LLM Coding Agents RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1e6959cd-4a2f-4088-9d8f-8fa9aee28c31 · inbound
BoostAPR: Boosting Automated Program Repair via Execution-Grounded Reinforcement Learning with Dual Reward Models RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 71e6a669-defc-4543-bd0c-571fbabf2cb5 · inbound
BoostAPR: Boosting Automated Program Repair via Execution-Grounded Reinforcement Learning with Dual Reward Models RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 61dd3dc0-56a7-474c-93e4-1a540b0357bf · inbound
Learning from Failures: Correction-Oriented Policy Optimization with Verifiable Rewards RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 706ce023-1d4b-400f-9605-f8ed0fb8259a · inbound
Self-Supervised On-Policy Distillation for Reasoning Language Models RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 107
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cbbfb1d5-2ad0-4a09-851a-d943557595ac · inbound
HydroAgent: Closing the Gap Between Frontier LLMs and Human Experts in Hydrologic Model Calibration via Simulator-Grounded RL RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4b8c36f2-35bd-48e7-b0fe-ebdaa2ae94e0 · inbound
Code as Agent Harness RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 104
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e60c5963-68fb-4953-90eb-5fc7254f05f9 · inbound
LamPO: A Lambda Style Policy Optimization for Reasoning Language Models RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 59c618b0-de54-4a45-99fb-e3b65c89d6f4 · inbound
Self-Policy Distillation via Capability-Selective Subspace Projection RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 86301b5d-5ca7-45a6-93ff-f7022016e5a4 · inbound
Learning the Error Patterns of Language Models RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cf5dea75-95b2-4735-be05-ef2454e75c9d · inbound
Learn from Your Mistakes: Tree-like Self-Play for Secure Code LLMs RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation af5d5b50-2658-4b97-b35b-aa63b36d48df · inbound
Sakana Fugu Technical Report RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 298
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0db63301-0fce-4dd1-a679-b74a2b8679f3 · inbound
When AI Reviews Its Own Code: Recursive Self-Training Collapse in Code LLMs RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 167
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 94718f18-42ee-4043-b317-adf8860d58f3 · inbound
DecompRL: Solving Harder Problems by Learning Modular Code Generation RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 16ed8299-fb9d-4606-9c31-af71c070ade7 · inbound
Beyond the Need for Speed: Energy-Aware Code Generation via Simulation-Guided Reinforcement Learning RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfd47062-a949-4b64-969c-22bb38ce2cf8 · inbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae5926b7-33cb-4f8d-93b0-fb5d51422def · inbound
Adopting Reinforcement Learning with Verifiable Rewards for Molecular Generation RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 873b42e5-cbfb-4569-9dae-a9a5bf5f35a1 · inbound
CSPF: A Constrained Shared-Private Fusion Method for Non-Verifiable Preference Evaluation RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8597521d-8618-40b5-9d1a-cfa2c439d041 · inbound
Training Large Language Models for Self-Explanation Faithfulness RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 144
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bba3d182-65cb-4f0c-a6a5-429435763b84 · inbound
Skill Self-Play: Pushing the Frontier of LLM Capability with Co-Evolving Skills RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9989336a-ed70-49c6-9097-e3ccdfc965fc · inbound
RLPF: Reinforcement Learning from Performance Feedback for Code Generation RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b0e5296-46c0-4996-9f78-40f3af77103c · inbound
LEAP: Lean Environment-Feedback via Adaptive Pruning for Code RL in GPU Kernel Generation RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b017c0d-f03d-4ccb-8932-104c8caa8a47 · inbound
PCSD: Persistent Consistency for Self-Distillation in Agentic Reinforcement Learning RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.