Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 35 inbound Pith citation observations for arXiv:2504.15930.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:18:15.353219Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 88c77a74-8c21-474b-910f-2e31b5772013 · inbound
DeepCEE: Efficient Cross-Region Model Distributed Training System under Heterogeneous GPUs and Networks StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59953177-8abe-4da9-bdfe-30c73e0f5279 · inbound
AReaL: A Large-Scale Asynchronous Reinforcement Learning System for Language Reasoning StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4f11a428-3c5d-49a9-bcf4-d8d0d4410b22 · inbound
Reinforcement Learning Optimization for Large-Scale Learning: An Efficient and User-Friendly Scaling Library StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09a90acc-3fd0-4efd-8eda-2e08a99ff9bb · inbound
Infinite Sampling: Efficient and Stable Grouped RL Training for Large Language Models StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eaef4a1d-b396-4be5-aee8-b1ce2379af6e · inbound
AsyncFlow: An Asynchronous Streaming RL Framework for Efficient LLM Post-Training StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74cff2a9-d30f-490e-bb5e-b9933ca563d1 · inbound
RLBoost: Harvesting Preemptible Resources for Cost-Efficient Reinforcement Learning on LLMs StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d7209fb2-2243-47dc-b1ec-681f4017e71a · inbound
Seer: Online Context Learning for Fast Synchronous LLM Reinforcement Learning StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5f79b4f8-9d80-424e-abdd-656aedd4537d · inbound
HetRL: Efficient Reinforcement Learning for LLMs in Heterogeneous Environments StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 940e6faa-2206-42ea-b86d-21bf3ee3ce7c · inbound
StaleFlow: Staleness-Aware Data Management for Mitigating Data Skewness in Fully Disaggregated RL Post-Training StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 850b703d-a73a-4f39-aed9-60976c229ce6 · inbound
ECHO-2: A Large-Scale Distributed Rollout Framework for Cost-Efficient Reinforcement Learning StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6764a43b-da9e-48b3-9f05-f6a9823d679f · inbound
TensorHub: Scalable and Elastic Weight Transfer for LLM RL Training StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8e99c259-8bfb-473f-ac32-523e7d0678b0 · inbound
JigsawRL: Assembling RL Pipelines for Efficient LLM Post-Training StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7d8401ba-16d8-40da-a466-2bac472ba821 · inbound
DORA: A Scalable Asynchronous Reinforcement Learning System for Language Model Training StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3dac3891-cd52-429e-baf4-ac9292a3446e · inbound
ROSE: Rollout On Serving GPUs via Cooperative Elasticity for Agentic RL StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3a096d9f-ad4a-47de-855b-debc9c8ae7bf · inbound
ROSE: Rollout On Serving GPUs via Cooperative Elasticity for Agentic RL StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f13a6e96-ffb6-4569-ba9e-afb95c171d0d · inbound
Experience Sharing in Mutual Reinforcement Learning for Heterogeneous Language Models StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 121
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation cac17c85-9c38-4b1f-9928-8f512c323582 · inbound
FlashEvolve: Accelerating Agent Self-Evolution with Asynchronous Stage Orchestration StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a4915de7-acd6-45b8-b6a4-9a4118b9a47a · inbound
ReLibra: Routing-Replay-Guided Load Balancing for MoE Training in Reinforcement Learning StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2b31c08b-f0c0-41bb-b1a6-12f9e9f58af4 · inbound
BubbleSpec: Turning Long-Tail Bubbles into Speculative Rollout Drafts for Synchronous Reinforcement Learning StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d060c8be-386b-4cbb-be5f-fad2cb4be93d · inbound
AstraFlow: Dataflow-Oriented Reinforcement Learning for Agentic LLMs StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7e9154e6-835d-4752-9f43-fafcf454e4b3 · inbound
How Off-Policy Can GRPO Be? Mu-GRPO for Efficient LLM Reinforcement Learning StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation acafb593-ecf7-4218-a7bc-8153d4c3e4aa · inbound
Accelerating Long-Tail Generation in Synchronous RLHF Training via Adaptive Tensor Parallelism StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation bafa7e87-d6b2-494a-ad89-159bcb56b9fe · inbound
SiDP: Memory-Efficient Data Parallelism for Offline LLM Inference StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 313d458a-1224-4b4d-8c51-f82bee97f89a · inbound
DARTS: Distribution-Aware Active Rollout Trajectory Shaping for Accelerating LLM Reinforcement Learning StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3066e0d2-19d8-45e1-908d-b0e073ee6e79 · inbound
Libra: Efficient Resource Management for Agentic RL Post-Training StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 48418179-a97c-45e8-bc5c-a39dbcb6ede1 · inbound
Rollout-Level Advantage-Prioritized Experience Replay for GRPO StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation eebaf89b-21f8-47ef-92d7-2c8ebb41663d · inbound
AsyncWebRL: Efficient Asynchronous Reinforcement Learning for Multi-Step Visual Web Agents StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1b028528-a2ae-46a6-a595-a4b14ed67838 · inbound
Harnessing Routing Foresight for Micro-step-level MoE load balancing in RL Post-training StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d389e0db-7c88-45cd-a650-a2682c1d0d57 · inbound
AsyncOPD: How Stale Can On-Policy Distillation Be? StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation bad8377f-0b26-4e14-a931-a2b5eb54ac00 · inbound
Moebius: Serving Mixture-of-Expert Models with Seamless Runtime Parallelism Switch StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ee440a99-b36c-4fa5-8ae7-eb967396bc06 · inbound
Next-Generation Agentic Reinforcement Learning Systems Enable Self-Evolving Agents StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 082f5c95-f61d-4dd9-8c06-0462846c29b6 · inbound
Next-Generation Agentic Reinforcement Learning Systems Enable Self-Evolving Agents StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b5221918-d94c-451b-b18c-ec0212a48ec4 · inbound
Bidirectional Resource Scheduling for Disaggregated and Asynchronous RL Post-Training StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5568b77-114e-43db-aeff-8c01158870ee · inbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc987d12-9c0f-4ba4-b4f1-b067647d1df2 · inbound
QLPO: Quadrant-weighted Sampling for Length-aware Policy Optimization StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.