Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T13:52:30.795158Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2509.00217.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T13:52:30.795158Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
21 of 21 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 24514ff3-c30a-4254-a4ec-ed5a690fdb0a · outbound
Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference Helix Parallelism: Rethinking Sharding Strategies for Interactive Multi-Million-Token LLM Decoding
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a79779b-d657-4fdc-a734-2e37e0500b3e · outbound
Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference Beyond data and model parallelism for deep neural networks
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 40a14d2e-66cf-4780-9500-d1b0f6765978 · outbound
Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference DFModel: Design Space Optimization of Large-Scale Systems Exploiting Dataflow Mappings
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b76fa2b-c631-4d1f-91e3-164d2e7593a7 · outbound
Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference Sequence Parallelism: Long Sequence Training from System Perspective
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7f1be0c-89e7-48d9-b685-2dde9885f28f · outbound
Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference Uniap: Unifying inter-and intra-layer automatic parallelism by mixed integer quadratic programming
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7b101fe8-429b-4831-b6ba-58386779095b · outbound
Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference DeepSeek-V3 Technical Report
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f5b5377-1a15-4f07-84e9-84a7601fa0a8 · outbound
Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference GTC 2024 Presentation Slides
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2228811a-d1e5-4ec9-ba21-6477f006a408 · outbound
Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference XLA: A Machine Learning Compiler
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 595358c9-665f-4332-8376-9a535c1632b6 · outbound
Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference Pytorch: An imperative style, high-performance deep learning library
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b95ac045-07ad-4f81-933b-504736ffec3a · outbound
Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference Chimera: Communication fusion for hybrid parallelism in large language models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8807b12b-2a57-4ff6-86fc-7823c750932f · outbound
Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference Stable-baselines3: Reliable reinforcement learning implementations
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f477b404-937b-412b-9dc5-083a2c8a6dc7 · outbound
Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference COSMIC: Enabling Full-Stack Co-Design and Optimization of Distributed Machine Learning Systems
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c328cf8-f322-402f-ac26-24a7ec6b5976 · outbound
Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference TAPAS: Fast and Automatic Derivation of Tensor Parallel Strategies for Large Neural Networks
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad31324f-7da3-4822-8c0c-b32818347c4b · outbound
Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference Megatron-lm: Training multi-billion parameter language models using model parallelism
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f90f5743-a83d-4467-9673-c90090343477 · outbound
Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference Seesaw: High-throughput LLM Inference via Model Re-sharding
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71fae5b5-2e42-40e4-ba1c-1b99cd58f76c · outbound
Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference LLaMA: Open and Efficient Foundation Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f22e790-a78a-4de0-b34c-f9ac21ef6a9b · outbound
Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference Gspmd: General and scalable parallelization for ml computation graphs
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a7e2a1a5-b25a-4ed8-a26f-354291a9a81e · outbound
Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference E., and Stoica, I
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2db2af95-05f3-4256-babb-53c556d30641 · outbound
Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference E., Stoica, I., Jin, X., Xing, E., Chen, Q
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation adf31cff-bb2c-4e4e-8fdd-62fbfd84edf5 · outbound
Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference Transferable graph optimizers for ml compilers
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4cbc4e50-ff62-422c-bcf1-4b136548b95f · outbound
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.