Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T09:49:31.823521Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2607.21653.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T09:49:31.823521Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
57 of 57 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 169d5e94-0897-4aaf-ab07-4d181591a69e · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Unresolved cited work
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57e14c27-f090-49f1-956c-2c70226cee1b · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Unresolved cited work
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2733aa50-7cd1-4214-9f4c-d5e308c9d401 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 524d3b2b-94cd-4465-b5ae-79b1b6ccea87 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43cc60cc-2c5e-4c77-b040-5df844775a04 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Laminar: A Scalable Asynchronous
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ee07cb2-abc2-4351-83b6-0b530f24a4d7 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Stabilizing
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f495a41-c1d5-4289-8078-4685b1cbb706 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e0e083e-fdb5-4295-b431-8844095c498a · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Polar: Agentic
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8f51995-d4e6-40da-b849-dd27d25edb63 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39b9d0b4-3124-48dd-ac64-8e3fbf0632ce · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a614a4e-908e-4769-a152-dc7129e941e9 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bf6ffa0-fc0e-4a64-97ac-2ced3c95de80 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Unresolved cited work
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd0b2436-ff5a-42a4-8c99-a25afa36678e · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning and Barrett, Clark and Sheng, Ying , journal =
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85bd74cd-14a9-44a5-9adf-f519f40301d5 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7dae37a-047d-41f5-a905-123bdaa42063 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 909a516f-9848-4f18-83a7-0b6d55bde502 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20da812d-47eb-4afc-8971-3c887eefa3fc · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1465e5b7-aa28-4ebe-b50e-424149d8b2e2 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning and Zhang, Hao and Stoica, Ion , booktitle =
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff47368a-5bc5-4128-8c07-6cc64261bfae · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning and Stoica, Ion , booktitle =
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd3548b2-1621-412c-8605-c5f3e1d8a8fe · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Unresolved cited work
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7058c24d-2e1a-4947-aa6d-f1628b4c1888 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3818632c-8277-4616-88f2-2f5aba1d0357 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95a1e4b1-914b-450a-a949-ef468cdb288e · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning and Yang, Yuqing , journal =
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3f02915-804a-4d72-99fa-761805eefe7b · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning When Speed Kills Stability: Demystifying
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2694c4b7-3202-4cc7-b893-f0fd11245bc9 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Understanding
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce0c8919-f800-4303-aa4a-0813e4f916a9 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning International Conference on Learning Representations (ICLR) , year =
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f74df15-cba2-4cb2-a4c1-20b56a18c131 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Back to Basics: Revisiting REINFORCE Style Optimization for Learning from Human Feedback in LLMs
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ac76f72-7b09-4b87-ad9b-7fab3beeab8e · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning OpenAI Gym
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7164737e-8b03-4d42-8053-986528121718 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning RolloutPipe: Overlapping Pipelined Rollout and Training in Disaggregated On-Policy LLM Reinforcement Learning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa3c69e7-334c-4751-850c-eb974ac2e515 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning DeepSeek-V3 Technical Report
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f2db113-7f86-4a24-8845-7519a8acf1ce · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning AReaL: A Large-Scale Asynchronous Reinforcement Learning System for Language Reasoning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65851560-8aec-47ff-a848-82a693f9d033 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning RollArt : Disaggregated multi-task agentic RL training at scale
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c8d6e8d-30b9-4e5e-916c-53c7ca0360fd · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f80a6e7-f801-412e-bda9-2a853337e2ba · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a54e877-88d7-4288-a924-63f9b6dfc584 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning REINFORCE++: Stabilizing Critic-Free Policy Optimization with Global Advantage Normalization
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 032c9881-1857-46f3-b18f-ef9548bc4cce · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning DORA: A Scalable Asynchronous Reinforcement Learning System for Language Model Training
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c2096ac-b5bd-4510-9b9f-88eeaf63604e · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Gonzalez, Hao Zhang, and Ion Stoica
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f977c3a0-f5bf-425b-ba57-c2fcc1bcfd04 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning When speed kills stability: Demystifying RL collapse from the training-inference mismatch
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32a94d8c-bfa4-461b-aca9-a886a35ff35b · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Understanding R1-Zero-Like Training: A Critical Perspective
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10f4918a-182a-47d9-bff4-716a6e3ab1ac · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Agent Lightning: Train ANY AI Agents with Reinforcement Learning
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3766b57f-6236-4697-82a1-7c89dd2c0d35 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Stabilizing MoE reinforcement learning by aligning training and inference routers
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1775cd88-1944-43b8-a6ea-9d0d19736d4c · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Jordan, and Ion Stoica
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53ee4424-189e-4cf7-9f4c-3eb8fe59f823 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning High-dimensional continuous control using generalized advantage estimation
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f052a68-1122-4f63-bf7c-c808b0a9f3dd · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53615c33-5cd4-4638-b25f-e2d606bffee1 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a18c5adc-9c31-416c-bef8-523fd5b64570 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Laminar: A scalable asynchronous RL post-training framework
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a85b83f0-2982-4164-8d08-03090b49d7d6 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning HybridFlow : A flexible and efficient RLHF framework
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67da6d4a-df23-49f6-a9a6-395678902471 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning OSWorld : Benchmarking multimodal agents for open-ended tasks in real computer environments
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b21a729-fd86-4cf5-b580-1adf8cc3ca45 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Polar: Agentic RL on Any Harness at Scale
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a869fb7-4c00-430a-973c-afe7325b5047 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Next-Generation Agentic Reinforcement Learning Systems Enable Self-Evolving Agents
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae069bff-5684-4e86-9919-8ab9ea6f4ab6 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd042aea-7619-4a94-ba34-9e89b795b488 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning ProRL agent: Rollout-as-a-service for RL training of multi-turn LLM agents
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b849eeb-c909-49a0-82db-ab62884f9c63 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Relax: An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 265e3e5a-e547-4283-a722-a98add457ae0 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning PyTorch FSDP : Experiences on scaling fully sharded data parallel
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37981a8a-1c61-49b0-ae8a-b08f35d2efda · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning Group Sequence Policy Optimization
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4c6c9c6-09b0-4fdd-b732-68d274133d64 · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning SGLang: Efficient Execution of Structured Language Model Programs
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5568b77-114e-43db-aeff-8c01158870ee · outbound
Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.