Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 32 inbound Pith citation observations for arXiv:2308.01320.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-10T18:21:37.458748Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
10
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation a8da933d-fddd-4d6c-9a12-82d3fa422edf · inbound
A Survey of Large Language Models DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 214
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a1c3eae2-2f4a-4e22-8be7-d9e6e31897b5 · inbound
OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 81152436-d2f4-4828-83bc-53f101f250ce · inbound
HybridFlow: A Flexible and Efficient RLHF Framework DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9fb83aab-dfe0-4eaa-83ea-22db674c60c9 · inbound
Curiosity-Driven Reinforcement Learning from Human Feedback DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cb093db-74b0-4c15-9e31-db362efb60fc · inbound
Leveraging Sparsity for Sample-Efficient Preference Learning: A Theoretical Perspective DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 105
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae705c90-d7b1-4065-86eb-f4a8027ba4af · inbound
Disentangling Length Bias In Preference Learning Via Response-Conditioned Modeling DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a346cb7-c05c-40d9-adb2-bbddb6a051d4 · inbound
Trustworthy AI: Safety, Bias, and Privacy -- A Survey DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97efdee9-ee8b-4ccc-bce1-73c6eaed10ff · inbound
hdl2v: A Code Translation Dataset for Enhanced LLM Verilog Generation DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31e20e14-e2d7-48c2-8f4b-62ff40c7919f · inbound
AsyncFlow: An Asynchronous Streaming RL Framework for Efficient LLM Post-Training DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c61dd39-2429-487e-a5fc-91c87cd3b8a4 · inbound
A Technical Survey of Reinforcement Learning Techniques for Large Language Models DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 125
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 925590c5-9fb1-4b36-a820-e55fa5a3b877 · inbound
Towards Hallucination-Free Music: A Reinforcement Learning Preference Optimization Framework for Reliable Song Generation DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2e9a5d5-ed78-4f19-9f40-9bf2c675bda2 · inbound
Echo: Decoupling Inference and Training for Large-Scale RL Alignment on Heterogeneous Swarms DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e334d77-5ec9-4c76-b73f-dc6bf5d9b267 · inbound
Reinforcement Learning Meets Large Language Models: A Survey of Advancements and Applications Across the LLM Lifecycle DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 225
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 854279e9-d260-4a88-9ef2-3f3bf087e8aa · inbound
Seer: Online Context Learning for Fast Synchronous LLM Reinforcement Learning DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 320429c6-cee5-4569-8a60-7b8e42716a9b · inbound
Periodic Asynchrony: An On-Policy Approach for Accelerating LLM Reinforcement Learning DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 91d20183-0493-4ff5-9269-c0bd1ad03359 · inbound
HetRL: Efficient Reinforcement Learning for LLMs in Heterogeneous Environments DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6a8f4821-a09c-4daa-a10a-b230ba7d5b2e · inbound
FP8-RL: A Practical and Stable Low-Precision Stack for LLM Reinforcement Learning DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 46372f89-f4c5-4dd3-a354-abca0291a1ad · inbound
TENT: A Declarative Slice Spraying Engine for Performant and Resilient Data Movement in Disaggregated LLM Serving DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0025aedf-01b9-4e91-a04c-311a574f183e · inbound
Reinforcement Learning from Human Feedback: A Statistical Perspective DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 49c70956-40a4-4cb5-9270-5de56d9f3a92 · inbound
JigsawRL: Assembling RL Pipelines for Efficient LLM Post-Training DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9a61df6f-04b7-41ff-b383-ab7013fb6cdf · inbound
DORA: A Scalable Asynchronous Reinforcement Learning System for Language Model Training DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7c63bc60-5e50-4d1b-a984-356098c4ecc9 · inbound
Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6c0aa9ae-c00e-41e6-97bc-eaca23fed6e4 · inbound
Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ec74e6e1-88e0-407e-97f9-b0b76e09f3b6 · inbound
ROSE: Rollout On Serving GPUs via Cooperative Elasticity for Agentic RL DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 81aa4c7c-fa99-4278-bbf4-f768a620a14b · inbound
ROSE: Rollout On Serving GPUs via Cooperative Elasticity for Agentic RL DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c8cdc8e8-cd28-44a5-9c50-c99de0989ddb · inbound
PlexRL: Cluster-Level Orchestration of Serviceized LLM Execution for RLVR DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 36f0d166-0b90-4c03-9c82-f3d3b7ca186f · inbound
Spend Your Rollouts Where It Counts: Rollout Allocation for Group-Based RL Post-Training DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5fd2bbd8-573f-4685-8264-38363a5652fc · inbound
Libra: Efficient Resource Management for Agentic RL Post-Training DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b60fc586-0338-4925-bad5-7fe6ee500f3a · inbound
The Hitchhiker's Guide to Agentic AI: From Foundations to Systems DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 230
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation cb13f38a-7fff-4397-8a91-7c58abd0a5da · inbound
The Hitchhiker's Guide to Agentic AI: From Foundations to Systems DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 216
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 907b9312-f955-48fd-917a-976172431ce0 · inbound
Bidirectional Resource Scheduling for Disaggregated and Asynchronous RL Post-Training DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5b7eb1d-f9a0-48cb-9da4-c1c052ed7136 · inbound
JoyNexus: Service-Oriented Multi-Tenant Post-Training for VLA Models DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.