Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2502.01715.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:31:54.286136Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-30T12:44:40.230064Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 18e4674b-e365-46e7-8ec7-e59fadbb27af · inbound
Improving LLM-Generated Code Quality with GRPO Process-Supervised Reinforcement Learning for Code Generation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aaffbf05-6329-479d-bf10-f10098639611 · inbound
Reinforcement Learning in hyperbolic space for multi-step reasoning Process-Supervised Reinforcement Learning for Code Generation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ebf6c85-9541-44f4-ba5d-45ae9cf8dc64 · inbound
Reinforcement Learning Improves Traversal of Parametric Knowledge in LLMs Process-Supervised Reinforcement Learning for Code Generation
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 862d71d8-3201-45e2-a169-e6611ec11551 · inbound
Beyond Binary: Turning Partial Success into Dense Verifiable Rewards for Reinforcement Learning in Code Generation Process-Supervised Reinforcement Learning for Code Generation
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0374b31-8d81-4623-8e1e-6d44d37a3bfb · inbound
TestDecision: Sequential Test Suite Generation via Greedy Optimization and Reinforcement Learning Process-Supervised Reinforcement Learning for Code Generation
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6a3feea1-6dfe-49be-aa00-6b9d94158540 · inbound
Free Energy-Driven Reinforcement Learning with Adaptive Advantage Shaping for Unsupervised Reasoning in LLMs Process-Supervised Reinforcement Learning for Code Generation
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cd0ae5ae-5a2c-43c2-abbf-20c4a72c161e · inbound
Adapt to Thrive! Adaptive Power-Mean Policy Optimization for Improved LLM Reasoning Process-Supervised Reinforcement Learning for Code Generation
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1d03fe18-37d5-47bd-8db7-482b1e88f803 · inbound
Beyond Trajectory Rewards: Step-level Credit Assignment for Agentic Search via Graph Modeling Process-Supervised Reinforcement Learning for Code Generation
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fae0bdb1-6d17-4048-8adb-07a787463b99 · inbound
BV-Blend: Uncertainty-Weighted Historical Baselines for Stable Critic-Free RL with Verifiable Rewards Process-Supervised Reinforcement Learning for Code Generation
Reference 105
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.