Pith. sign in

Paper Citation Record · LEDGER

DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 32 inbound Pith citation observations for arXiv:2308.01320.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.01320 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 32 of 32 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T18:21:37.458748Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

10
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a8da933d-fddd-4d6c-9a12-82d3fa422edf · inbound

A Survey of Large Language Models cites this paper.

A Survey of Large Language Models DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 214

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:46:40.026523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T22:46:39.268353Z digest=sha256:f4a7c9786d922a52a34ed2adeb605dc108dc6a41386d6fe3393949a8c023562c

Observation a1c3eae2-2f4a-4e22-8be7-d9e6e31897b5 · inbound

OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework cites this paper.

OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-15T03:28:57.132277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T03:28:57.008431Z digest=sha256:3bc999c652515195765ba28d2b34954f0cfbc2279b875a29566f4e2d8fe728c3

Observation 81152436-d2f4-4828-83bc-53f101f250ce · inbound

HybridFlow: A Flexible and Efficient RLHF Framework cites this paper.

HybridFlow: A Flexible and Efficient RLHF Framework DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 94

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:53:38.949833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T07:53:38.715353Z digest=sha256:66d47f66b5202db16170947a0f97fa92e47f60de0f7504df2c1f0586ed83b37b

Observation 9fb83aab-dfe0-4eaa-83ea-22db674c60c9 · inbound

Curiosity-Driven Reinforcement Learning from Human Feedback cites this paper.

Curiosity-Driven Reinforcement Learning from Human Feedback DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T18:21:37.458748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T18:21:37.458748Z digest=sha256:669b04eaeaadf4d4c00206e61510134965303a5dc1a86f3d3ae5f1eeef78735b

Observation 2cb093db-74b0-4c15-9e31-db362efb60fc · inbound

Leveraging Sparsity for Sample-Efficient Preference Learning: A Theoretical Perspective cites this paper.

Leveraging Sparsity for Sample-Efficient Preference Learning: A Theoretical Perspective DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 105

Resolution
unresolved
no resolver link, observed 2026-08-10T00:11:59.487147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T00:11:59.487147Z digest=sha256:1cba3aea5d2cdd1c2e360129d19ce0b18c017bdd188a70e62901ac51b387bfc1

Observation ae705c90-d7b1-4065-86eb-f4a8027ba4af · inbound

Disentangling Length Bias In Preference Learning Via Response-Conditioned Modeling cites this paper.

Disentangling Length Bias In Preference Learning Via Response-Conditioned Modeling DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-09T17:46:29.299748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T17:46:29.299748Z digest=sha256:998f35ee0b9786810db681b0bcbef52447d8c0e3aaf18511c6d87e132d934220

Observation 8a346cb7-c05c-40d9-adb2-bbddb6a051d4 · inbound

Trustworthy AI: Safety, Bias, and Privacy -- A Survey cites this paper.

Trustworthy AI: Safety, Bias, and Privacy -- A Survey DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T11:26:27.871302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:26:27.871302Z digest=sha256:7baf08198893393c31b9b9ec42817385f5ef9bdfd6c07a32bd2356178bf79918

Observation 97efdee9-ee8b-4ccc-bce1-73c6eaed10ff · inbound

hdl2v: A Code Translation Dataset for Enhanced LLM Verilog Generation cites this paper.

hdl2v: A Code Translation Dataset for Enhanced LLM Verilog Generation DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T10:45:07.106711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:45:07.106711Z digest=sha256:e4177dae8484dd0986e81b974f42068c8cf4581ac486088f37eb7f8cf467a8f6

Observation 31e20e14-e2d7-48c2-8f4b-62ff40c7919f · inbound

AsyncFlow: An Asynchronous Streaming RL Framework for Efficient LLM Post-Training cites this paper.

AsyncFlow: An Asynchronous Streaming RL Framework for Efficient LLM Post-Training DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T20:50:26.921368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:50:26.921368Z digest=sha256:2d45d28261c00955abc127d1848bcd69563de6f3099dbf4a3a4e4d119297751d

Observation 1c61dd39-2429-487e-a5fc-91c87cd3b8a4 · inbound

A Technical Survey of Reinforcement Learning Techniques for Large Language Models cites this paper.

A Technical Survey of Reinforcement Learning Techniques for Large Language Models DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 125

Resolution
unresolved
no resolver link, observed 2026-08-06T19:59:36.363859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:59:36.363859Z digest=sha256:4775dc0d90ffc057e5ce8f408bec32e539c65a7baa35489e333628b8237d3a2e

Observation 925590c5-9fb1-4b36-a820-e55fa5a3b877 · inbound

Towards Hallucination-Free Music: A Reinforcement Learning Preference Optimization Framework for Reliable Song Generation cites this paper.

Towards Hallucination-Free Music: A Reinforcement Learning Preference Optimization Framework for Reliable Song Generation DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T23:43:05.560349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:43:05.560349Z digest=sha256:97eab3da4d3996344e07a368fb71f4f59b303c3baa8fd0a149d9828c430b9f77

Observation c2e9a5d5-ed78-4f19-9f40-9bf2c675bda2 · inbound

Echo: Decoupling Inference and Training for Large-Scale RL Alignment on Heterogeneous Swarms cites this paper.

Echo: Decoupling Inference and Training for Large-Scale RL Alignment on Heterogeneous Swarms DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T23:26:57.319056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:26:57.319056Z digest=sha256:58e3b3b03f3dda0fdeb53142feb5e0710a14727350bc05577170bdb5b5e3d77e

Observation 9e334d77-5ec9-4c76-b73f-dc6bf5d9b267 · inbound

Reinforcement Learning Meets Large Language Models: A Survey of Advancements and Applications Across the LLM Lifecycle cites this paper.

Reinforcement Learning Meets Large Language Models: A Survey of Advancements and Applications Across the LLM Lifecycle DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 225

Resolution
unresolved
no resolver link, observed 2026-08-04T16:07:45.810430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:07:45.810430Z digest=sha256:450fb02e9d36607d74083596c8f7a3206f0927faeb2ec754873b72997d373cff

Observation 854279e9-d260-4a88-9ef2-3f3bf087e8aa · inbound

Seer: Online Context Learning for Fast Synchronous LLM Reinforcement Learning cites this paper.

Seer: Online Context Learning for Fast Synchronous LLM Reinforcement Learning DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:40:14.512467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T20:38:30.169363Z digest=sha256:8608917c79137757bedd6ac19d611bfc66759720d143eab4e34329dccec67acb

Observation 320429c6-cee5-4569-8a60-7b8e42716a9b · inbound

Periodic Asynchrony: An On-Policy Approach for Accelerating LLM Reinforcement Learning cites this paper.

Periodic Asynchrony: An On-Policy Approach for Accelerating LLM Reinforcement Learning DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:34:04.919439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T05:33:46.817183Z digest=sha256:cbc6d81ace3b4a7616eb1771189a74d8e3c8be2839781e67d35e7a33d83399e1

Observation 91d20183-0493-4ff5-9269-c0bd1ad03359 · inbound

HetRL: Efficient Reinforcement Learning for LLMs in Heterogeneous Environments cites this paper.

HetRL: Efficient Reinforcement Learning for LLMs in Heterogeneous Environments DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T22:23:36.659585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-16T22:21:26.271796Z digest=sha256:e79c9acafd38d17676dd662d1a49e16078bb70dd28b2d743c0e50aca507725ac

Observation 6a8f4821-a09c-4daa-a10a-b230ba7d5b2e · inbound

FP8-RL: A Practical and Stable Low-Precision Stack for LLM Reinforcement Learning cites this paper.

FP8-RL: A Practical and Stable Low-Precision Stack for LLM Reinforcement Learning DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:07:47.104402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T11:07:18.839899Z digest=sha256:69781b54be8c3d9e769c608fc3bca240ae5e385b893bd50662f5959d8a3d904c

Observation 46372f89-f4c5-4dd3-a354-abca0291a1ad · inbound

TENT: A Declarative Slice Spraying Engine for Performant and Resilient Data Movement in Disaggregated LLM Serving cites this paper.

TENT: A Declarative Slice Spraying Engine for Performant and Resilient Data Movement in Disaggregated LLM Serving DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T17:01:43.630573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T17:01:43.630573Z digest=sha256:08771c08b9267e9dd5cf59983a7067afa3c1bbd940394d2ebca145385091378b

Observation 0025aedf-01b9-4e91-a04c-311a574f183e · inbound

Reinforcement Learning from Human Feedback: A Statistical Perspective cites this paper.

Reinforcement Learning from Human Feedback: A Statistical Perspective DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 87

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:13:13.726538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-13T20:10:43.578904Z digest=sha256:5f1601c5445803048287887ad9466b9cd52b866ce0b3a7908f98bf0cf2979a8c

Observation 49c70956-40a4-4cb5-9270-5de56d9f3a92 · inbound

JigsawRL: Assembling RL Pipelines for Efficient LLM Post-Training cites this paper.

JigsawRL: Assembling RL Pipelines for Efficient LLM Post-Training DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:11:11.055339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T06:36:33.914193Z digest=sha256:ffc5c5d58af55f44e5e56331ad03a69783331e3b80b801225bb1e93fc807ff27

Observation 9a61df6f-04b7-41ff-b383-ab7013fb6cdf · inbound

DORA: A Scalable Asynchronous Reinforcement Learning System for Language Model Training cites this paper.

DORA: A Scalable Asynchronous Reinforcement Learning System for Language Model Training DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:46:26.338424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-07T13:49:26.560459Z digest=sha256:72f63331367a9602e60329c594c5da34013f33ec27e9b6424c6038cd3a0b7c29

Observation 7c63bc60-5e50-4d1b-a984-356098c4ecc9 · inbound

Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence cites this paper.

Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 94

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:06:18.446759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T10:12:58.421050Z digest=sha256:bc3d27961886cde255b26cbd214d3d8c3c0f6a8b1df88deb36146f93d88daf8f

Observation 6c0aa9ae-c00e-41e6-97bc-eaca23fed6e4 · inbound

Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence cites this paper.

Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 94

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:00:54.475637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T00:56:48.838028Z digest=sha256:551c7c071ae3697d88bc8c4ee99342d4b901a41f0c26d08f4c41022a01d42558

Observation ec74e6e1-88e0-407e-97f9-b0b76e09f3b6 · inbound

ROSE: Rollout On Serving GPUs via Cooperative Elasticity for Agentic RL cites this paper.

ROSE: Rollout On Serving GPUs via Cooperative Elasticity for Agentic RL DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 87

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:31:16.480776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T05:14:14.168753Z digest=sha256:dcfd94d201bcc6dbd3dcea629e660f985305aa350fcd7c0b50ce59532e356efd

Observation 81aa4c7c-fa99-4278-bbf4-f768a620a14b · inbound

ROSE: Rollout On Serving GPUs via Cooperative Elasticity for Agentic RL cites this paper.

ROSE: Rollout On Serving GPUs via Cooperative Elasticity for Agentic RL DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:39:53.327060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T08:39:31.911497Z digest=sha256:8a7565d757b4902392d5f7ae3fc5a70c92822b78f7a5a614a8552e1d78f32e24

Observation c8cdc8e8-cd28-44a5-9c50-c99de0989ddb · inbound

PlexRL: Cluster-Level Orchestration of Serviceized LLM Execution for RLVR cites this paper.

PlexRL: Cluster-Level Orchestration of Serviceized LLM Execution for RLVR DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-21T02:29:25.369999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T02:24:48.872065Z digest=sha256:1ba6a4037d0d3d39d86de792721d30060f861e85906b58cb1df68cccca532d15

Observation 36f0d166-0b90-4c03-9c82-f3d3b7ca186f · inbound

Spend Your Rollouts Where It Counts: Rollout Allocation for Group-Based RL Post-Training cites this paper.

Spend Your Rollouts Where It Counts: Rollout Allocation for Group-Based RL Post-Training DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-06-29T19:53:55.785979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T19:47:09.817243Z digest=sha256:8b685a0adf71e360f81f84499e5b11fe64960dfdc9ec9e144fb4ec74ad2eaae6

Observation 5fd2bbd8-573f-4685-8264-38363a5652fc · inbound

Libra: Efficient Resource Management for Agentic RL Post-Training cites this paper.

Libra: Efficient Resource Management for Agentic RL Post-Training DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:16:27.143876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-28T11:02:00.385932Z digest=sha256:72a4b0f83d94847a5acfd0c9cef4ec3215eb79173c2c8362d1f9b78ebbcba9ed

Observation b60fc586-0338-4925-bad5-7fe6ee500f3a · inbound

The Hitchhiker's Guide to Agentic AI: From Foundations to Systems cites this paper.

The Hitchhiker's Guide to Agentic AI: From Foundations to Systems DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 230

Resolution
verified exact
arxiv_id, observed 2026-07-04T11:09:46.498426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T08:09:57.542558Z digest=sha256:3d606740784e1eeaef35ace22a9b832d0885b17cebdd786442dd67b294389d05

Observation cb13f38a-7fff-4397-8a91-7c58abd0a5da · inbound

The Hitchhiker's Guide to Agentic AI: From Foundations to Systems cites this paper.

The Hitchhiker's Guide to Agentic AI: From Foundations to Systems DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 216

Resolution
unresolved
no resolver link, observed 2026-08-02T10:27:18.522221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:27:18.522221Z digest=sha256:0f2745d8fb37a24746e14bd901ff5e8c4beddb07de71c939058a30c493c7ff1a

Observation 907b9312-f955-48fd-917a-976172431ce0 · inbound

Bidirectional Resource Scheduling for Disaggregated and Asynchronous RL Post-Training cites this paper.

Bidirectional Resource Scheduling for Disaggregated and Asynchronous RL Post-Training DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-13T04:42:35.589143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T04:42:35.589143Z digest=sha256:66c0b4e82900ac42e6429960f714d643845d0bf01efe8587e0240ff81ca963cf

Observation b5b7eb1d-f9a0-48cb-9da4-c1c052ed7136 · inbound

JoyNexus: Service-Oriented Multi-Tenant Post-Training for VLA Models cites this paper.

JoyNexus: Service-Oriented Multi-Tenant Post-Training for VLA Models DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T21:32:07.773505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T21:32:07.773505Z digest=sha256:011cd791156965d1cb435e305f73db3da6e42d808fc2ffa65461daff903c0101