Pith. sign in

Paper Citation Record · LEDGER

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster

As of 20 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 2 inbound Pith citation observations for arXiv:2507.19017.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.19017 v1

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:07:35.118320Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T14:26:49.704505Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T20:26:10.436046Z

Reference resolution

55 of 55 outbound references displayed

  • verified exact2
  • verified fuzzy4
  • unresolved49
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 94f1291f-620c-4117-9c39-f664a2d054f0 · outbound

This paper cites an unresolved cited work.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:07:35.466129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:34.951530Z digest=sha256:605d22fd2db3a2514f82a6ffb5595ac9338d7e9d4703170785dc344407163b20

Observation e7b00ddf-2bdc-419d-b60d-141835a96aaf · outbound

This paper cites an unresolved cited work.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:07:35.458483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:34.954797Z digest=sha256:90beba7e0f016092d8b91c97bddd6b0c0ffe9bf453074ed9bec68777791d8f36

Observation 840a01f6-312f-4f4e-abd8-9b024ffb70bc · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:34.958676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:34.958676Z digest=sha256:c09ca0702eed327f6d91c28679fac58594ad2132b70d3aabfcba453d7c8e257b

Observation ebbcfd94-bb66-4e87-92d0-62649c8bea83 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:34.961985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:34.961985Z digest=sha256:766ac4cef43900b2202ae7105d05498d527325fe7d2116e8a3f170888bd479d1

Observation 89ad4549-b54c-490c-b514-75a11640561b · outbound

This paper cites AReaL: A Large-Scale Asynchronous Reinforcement Learning System for Language Reasoning.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster AReaL: A Large-Scale Asynchronous Reinforcement Learning System for Language Reasoning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:34.964925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:34.964925Z digest=sha256:48e9fd23182884c26c6cb5c13bb1376100c89ab2865047121ee3157b894d0805

Observation 753cb2c1-a66a-4c6d-89b9-a5cfaf212dd5 · outbound

This paper cites FlashNorm: Fast Normalization for Transformers.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster FlashNorm: Fast Normalization for Transformers

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-15T18:07:35.295439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:34.968888Z digest=sha256:7b506869693e6637c57b295bde3198ab456bb2df9dcfe030d99c82f565ce7f91

Observation ef0dfd9a-1b6a-4f4c-9052-d6bda61c6aae · outbound

This paper cites an unresolved cited work.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:07:35.451899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:34.973000Z digest=sha256:580caf1c807d2f7bd81c6f06ac410c7992db6581fa9a229d4d9c80100ddf4387

Observation 8619193d-a692-4691-a714-cdea857271be · outbound

This paper cites an unresolved cited work.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:07:35.445264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:34.975646Z digest=sha256:a27d8613e9ada270d493251ee64bcd68694998e2fbca8220795a8e9147bf6454

Observation 56f1cc2c-0b98-4d23-a43e-d1ff3c582b8b · outbound

This paper cites OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:34.979251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:34.979251Z digest=sha256:3dea099b12a6983c06622d353d91cc8c1c197008d5b5867edb45e1ba959da9da

Observation 74a59b31-9d2a-4496-8778-49ae4baef79e · outbound

This paper cites X.; Chen, D.; Lee, H.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster X.; Chen, D.; Lee, H

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:07:35.438334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:34.982745Z digest=sha256:e4aaa0454ffe03e8457bbf98074845141a4ff31966dded9898aafddcb05da762

Observation 27378040-ba78-45dd-a90f-3ab75a6b130c · outbound

This paper cites A.; Tanaka, M.; Zhang, C.; Zhang, M.; Aminabadi, R.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster A.; Tanaka, M.; Zhang, C.; Zhang, M.; Aminabadi, R

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:07:35.432069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:34.986302Z digest=sha256:10e953679fe5c94443ca561c6fcb77fbe570335d34d78812f7a425b2992b4e9f

Observation 0c3921ca-5b03-4c57-8e65-a454473a4c44 · outbound

This paper cites an unresolved cited work.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:07:35.425296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:34.989539Z digest=sha256:36ef3d39c0c6621691248501a9189503ecd5b4a7ad06f0eace3b907be2f42700

Observation 7424fa68-1741-4a3a-b37c-cd5759b34ede · outbound

This paper cites RLlib Flow: Distributed Reinforcement Learning is a Dataflow Problem.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster RLlib Flow: Distributed Reinforcement Learning is a Dataflow Problem

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:34.992842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:34.992842Z digest=sha256:b44c69cfdc9a4e1737ca6d3de5cfeb5de69ab8b7f069e28d6fabe2efec937130

Observation 507d8c7c-ac35-46ff-8dff-be1967d58240 · outbound

This paper cites Ring Attention with Blockwise Transformers for Near-Infinite Context.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Ring Attention with Blockwise Transformers for Near-Infinite Context

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:34.996287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:34.996287Z digest=sha256:ddc6e5bd2bc7493a58ec364951d6c88c826de6a04d49bbcf36e4c2cd2395e303

Observation 309d336c-11dc-424b-b304-cce559117312 · outbound

This paper cites Y.; Roongta, M.; Cai, C.; Luo, J.; Li, L.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Y.; Roongta, M.; Cai, C.; Luo, J.; Li, L

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:07:35.418848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:34.999848Z digest=sha256:2f745f18efa384994bb4c08354ef43d682c2c27805d7ee3eaf5e621bd3a20e71

Observation ece80389-b587-4800-8f37-41c170e589d1 · outbound

This paper cites an unresolved cited work.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:07:35.412462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:35.002929Z digest=sha256:83354149d899dff330f1a91d713af2d4896ed462396e96c3dac2a1fe4a81b178

Observation b8ec1b39-c291-4b52-ba9d-2c16cba2ad62 · outbound

This paper cites ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.005076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.005076Z digest=sha256:a407f34e5fa1f828d26f7d5f80c20ba8403731f8ed5d1261c2d3c1501f058719

Observation 125480ec-2d46-4170-8c0a-0af4df8a7bf4 · outbound

This paper cites Ray: A Distributed Framework for Emerging AI Applications.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Ray: A Distributed Framework for Emerging AI Applications

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.008292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.008292Z digest=sha256:41b4ff2d22e19e30bb29ffec7d43f0034b7596740c2cc30943ad77c635d151d4

Observation 2dc228bb-0df8-4800-b3d4-13b550690708 · outbound

This paper cites an unresolved cited work.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:07:35.406430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:35.012485Z digest=sha256:acf2350c82f6addc2dcb76020a3b656b436b2063d3880ff319154e0a73cc869e

Observation bfb31e64-6bfc-42be-bdb2-7dd500da9784 · outbound

This paper cites an unresolved cited work.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:07:35.399985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:35.014719Z digest=sha256:2670d1c217a681d70289b7063cdec83b037de4e5d6caaa2f1c984e30df0ba734

Observation 587cf96f-62f1-403a-805c-4607dce9534f · outbound

This paper cites an unresolved cited work.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:07:35.392981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:35.017880Z digest=sha256:feb99b1dfd758eb3ee55c578cb39031fd6f12dd09c5c98562f6b64c7af531965

Observation 4838eca3-c3b4-4af8-9ef4-75008db0c8a0 · outbound

This paper cites GPT-4 Technical Report.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster GPT-4 Technical Report

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.021115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.021115Z digest=sha256:ab83882b5db059bc3cdc3aca70014b08ec456646325934850e6bad4c83a89629

Observation 9fb2e852-daad-490f-bb0a-cd05a80b2517 · outbound

This paper cites Training language models to follow instructions with human feedback.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Training language models to follow instructions with human feedback

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.023574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.023574Z digest=sha256:58a7faf9270b7124eeabe3a13768f80570b264d3c8bc97f0d34ce2685907fa6b

Observation beb524af-4d9c-4eff-839f-2d0f5deb3a0a · outbound

This paper cites Zero Bubble Pipeline Parallelism.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Zero Bubble Pipeline Parallelism

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.026073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.026073Z digest=sha256:b0e70166dc59392b8da9aa2206affe2e90e1ef5038b961dce3f4292ecec1e6ef

Observation e9f8e3a1-bdbe-4adc-94fd-9bb5e037feab · outbound

This paper cites DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.028705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.028705Z digest=sha256:2f66a610af1ddaf649704fff84c032f31a09ce596077dd86c54e51120b608e09

Observation f7bfac71-291a-42e4-8ef2-1a6395b6c13b · outbound

This paper cites an unresolved cited work.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:07:35.386734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:35.031930Z digest=sha256:2b1110b1b7de5ba1160c8ce8ee22a4c0abf49ee056aaa068d56de30fb54d2f87

Observation d0e9de4d-e451-409c-83a6-4cfc164297ac · outbound

This paper cites an unresolved cited work.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:07:35.380019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:35.034397Z digest=sha256:c5fd1a578bc6532585fb39b68d8007914b2f9281bcd3010d16925eb0ef54c700

Observation d46233d5-0f64-486a-a28d-424f1f570b32 · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.036822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.036822Z digest=sha256:c59c025b94d79c0021bb211ccdcaced864cb73cff08bfab48f253fdef99eb2d1

Observation 839a6ab6-b9d8-4d8e-8826-dd33bbeeee11 · outbound

This paper cites an unresolved cited work.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:07:35.373816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:35.040232Z digest=sha256:5498ffdb4d50b19e3acee61815a975b12d463bbc9779337ae93b054070cad09c

Observation ed56c3b8-b0e3-4188-9827-55d875e5e693 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.043416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.043416Z digest=sha256:e352e77384f557d6fdc327db6a54284608a78ad902056ca3751d7a83959a2378

Observation 30cb4c25-086f-498e-8596-70538033dbe2 · outbound

This paper cites GLU Variants Improve Transformer.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster GLU Variants Improve Transformer

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.046202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.046202Z digest=sha256:93b00b6d717ceba27bccbeb46694a29e1306d8f3bf78b41ab88d34baf4c94fa3

Observation 316d40cd-98bb-4086-a9b6-3e0f7ecd197e · outbound

This paper cites NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster NeMo-Aligner: Scalable Toolkit for Efficient Model Alignment

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.049613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.049613Z digest=sha256:12f2bd750f68ed23f14ba2b6acb033eddaa4d110ffae94871e4b6ac4895deb31

Observation faeb5e32-4a34-4bc2-b4f3-f386b777ee92 · outbound

This paper cites Exploring Data Scaling Trends and Effects in Reinforcement Learning from Human Feedback.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Exploring Data Scaling Trends and Effects in Reinforcement Learning from Human Feedback

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.052248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.052248Z digest=sha256:428a2fac94f579daaea52fd778710cbca204c4ed8aaf82093bcc06303d8c6388

Observation c2e4eec4-178c-404b-ada9-70485864a88b · outbound

This paper cites an unresolved cited work.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:07:35.366896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:35.054745Z digest=sha256:8a36007cc1c8ca95659b7a8176ebec62e1726905fe6a59dc5b0a13b767973b11

Observation c1625def-5d68-4474-ae3a-7cadbb169993 · outbound

This paper cites an unresolved cited work.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:07:35.360412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:35.057150Z digest=sha256:93c957f5fa8ff62ec106150484ccece52c0c63593252cc6054ff14f177f0f879

Observation 39c653dd-69d3-4796-8bee-cf5785049be4 · outbound

This paper cites Model Parallelism With Subnetwork Data Parallelism.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Model Parallelism With Subnetwork Data Parallelism

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-15T18:07:35.214451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:35.060903Z digest=sha256:43434cc7c3d3be44d4be6bbbeabb361d7465aa7b9fcf1b7e2e3e9b94bcb86f80

Observation 1000a28a-3c06-46a4-8452-a456b27a25d1 · outbound

This paper cites RoFormer: Enhanced Transformer with Rotary Position Embedding.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster RoFormer: Enhanced Transformer with Rotary Position Embedding

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.064355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.064355Z digest=sha256:2984db3e5ef807201382ee19d8d2ffb0468dee331d30dafeecee707f17457d40

Observation 5dec1695-e5b4-49b5-bc2a-87d7c0d64e38 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.067892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.067892Z digest=sha256:668101841a07fd0ecbea1173f407df8d3862cfcf323f0914f0651014439d7840

Observation 6994cb7c-e363-4dcc-b699-96031fada953 · outbound

This paper cites Reinforcement Learning Enhanced LLMs: A Survey.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Reinforcement Learning Enhanced LLMs: A Survey

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.071702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.071702Z digest=sha256:4c6683bd9f2fbcd544f347103b790c05b4ee9302978a1e29257d8619f4551db4

Observation b7d45633-a306-4f96-9379-b118a333a841 · outbound

This paper cites an unresolved cited work.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:07:35.352862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:35.074885Z digest=sha256:98f7a231a1061ec1c2de8dc54d37392b146628d35dd595687f9918c12f8a1645

Observation 8053a4b1-e1d9-44ac-abf5-475fd956475e · outbound

This paper cites HuggingFace's Transformers: State-of-the-art Natural Language Processing.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster HuggingFace's Transformers: State-of-the-art Natural Language Processing

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.078289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.078289Z digest=sha256:319676e2cc991aba98df123c00d676b83469bc324c613b726a8ab781cc4e77ce

Observation 492ba260-3140-4f48-82ac-df2ff2d4bcc0 · outbound

This paper cites Qwen3 Technical Report.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Qwen3 Technical Report

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.081382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.081382Z digest=sha256:7a0172470e07d1f81a6e2af93b491287d367d548b9b2d067e7183872241c3773

Observation 6ef14aef-05b1-4edc-9552-e013cbf37125 · outbound

This paper cites DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster DeepSpeed-Chat: Easy, Fast and Affordable RLHF Training of ChatGPT-like Models at All Scales

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.083833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.083833Z digest=sha256:70fa17372750c45f08cf02e13a9917170ccbf96717758326a976c317005749b9

Observation 1cc576a2-71c6-4e41-86a5-b8b5aabd5328 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.086295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.086295Z digest=sha256:6c2aa88ae5b73999c73b87c7ec093b668ea058d8f09cd95a4213cef58b1bbeed

Observation e5d02f0d-7ba9-4478-a393-3e2549207f83 · outbound

This paper cites Jenga: Effective Memory Management for Serving LLM with Heterogeneity.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Jenga: Effective Memory Management for Serving LLM with Heterogeneity

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.089407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.089407Z digest=sha256:64d873c848831c3d083f62312ad16d0d6f93cac97f9374d5912a448faca8805b

Observation 634afb3f-0b6e-4858-ad79-95e6cd71b7b3 · outbound

This paper cites an unresolved cited work.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:07:35.345951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:35.092869Z digest=sha256:ff3051833baf3057ed02f39874b44abc8b043bbef21c68dede23ccc089f38414

Observation cc08cee5-606e-4647-97b4-eb9220d84e53 · outbound

This paper cites an unresolved cited work.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:07:35.339291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:35.095520Z digest=sha256:e8c6f9db6a56e75b09c997e54502f4bec7cc112119e02c62c2698cf991b2b4f7

Observation d5c0019c-202e-4fd1-9df0-9237943381e4 · outbound

This paper cites C.; Xu, M.; Wright, L.; Shojanazeri, H.; Ott, M.; and Shleifer, S.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster C.; Xu, M.; Wright, L.; Shojanazeri, H.; Ott, M.; and Shleifer, S

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:07:35.332335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:35.098450Z digest=sha256:3e656207374cfedabba03ef88efc83daad2f183ec64a0d7ee18cabb646afe72d

Observation 129ae176-1c1b-434b-ab1b-c2ea728d7d3a · outbound

This paper cites PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster PyTorch FSDP: Experiences on Scaling Fully Sharded Data Parallel

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.101119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.101119Z digest=sha256:ce747643637a8016cc0a8a7ad9ee5dc2cdb6ab57545e25656443690bb5c1e4f6

Observation 2f9f3ec1-6e5b-4cf6-8a58-430ae5c9b106 · outbound

This paper cites SGLang: Efficient Execution of Structured Language Model Programs.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster SGLang: Efficient Execution of Structured Language Model Programs

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.103609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.103609Z digest=sha256:c45d4cb60eb766f4eefa67583d3503ffb7f8e46d5054adae41dac1b073a0a3d1

Observation 25c1c88f-5014-4811-8e10-4c1401053fdb · outbound

This paper cites StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster StreamRL: Scalable, Heterogeneous, and Elastic RL for LLMs with Disaggregated Stream Generation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.106165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.106165Z digest=sha256:ed17b6f0c161a828ff3dcc5ea8ee833d73a87df0cc6c73816f63a98d39300089

Observation c6cbf770-f489-4a1e-ae2a-6cdfe4ef4800 · outbound

This paper cites Optimizing RLHF Training for Large Language Models with Stage Fusion.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Optimizing RLHF Training for Large Language Models with Stage Fusion

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.109566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.109566Z digest=sha256:dea9ed780ce7f134b43e0fa84e6d1378f5ed49e578a62e07dc8c4bc8cdacf8e4

Observation bff32c95-feea-46f2-9c00-057fd1eb4a39 · outbound

This paper cites an unresolved cited work.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:07:35.324915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T18:07:35.112571Z digest=sha256:daabb3bab2744f0ab2b0abaf2b3a68c3de2023ffc5c4e80e0036cab0c8b9ca5f

Observation ba4981af-e42a-4176-be3d-a7b4fa52ea13 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster , " * write output.state after.block = add.period write newline

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.114695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.114695Z digest=sha256:eda261da8a6083864b4eafca0426c3475f004da7b790b453b38d8e8d4c84ca6b

Observation 97c41d6c-5a6b-483e-b26f-b708e3165583 · outbound

This paper cites write newline.

MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster write newline

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T18:07:35.118320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:07:35.118320Z digest=sha256:2e57d501572b8854442b31e5e69462fbbfa1f678c3f3b268779592f19e08e8a4

Pith citing papers

Observation e15eb777-934a-4431-ba3c-601a2b9b9378 · inbound

Schedule-and-Calibrate: Utility-Guided Multi-Task Reinforcement Learning for Code LLMs cites this paper.

Schedule-and-Calibrate: Utility-Guided Multi-Task Reinforcement Learning for Code LLMs MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:26:10.438554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-08T09:11:20.733960Z digest=sha256:c9e4a808c7b5bce797b7aaf8720a636012537009afc8d18e04b588bf5f1066dc

Observation a8b5541e-3e30-436d-804b-0ed94e9d8c6e · inbound

TideRL: Boosting Agentic RL Goodput with Readiness-Aware Scheduling cites this paper.

TideRL: Boosting Agentic RL Goodput with Readiness-Aware Scheduling MindSpeed RL: Distributed Dataflow for Scalable and Efficient RL Training on Ascend NPU Cluster

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T14:26:49.704505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:26:49.704505Z digest=sha256:002bbb9264b0a9c29bbbc58f3513210ae60d655749adef466a6d3dc8d0874c43