Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2201.05596.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:40.349198Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
55
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 4a6e3fea-d3f0-420e-8fe3-141a6ee723d2 · inbound
THOR-MoE: Hierarchical Task-Guided and Context-Responsive Routing for Neural Machine Translation DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f142a64e-c741-441d-9f0f-2b83526640db · inbound
Mixture-of-Experts Can Surpass Dense LLMs Under Strictly Equal Resource DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation eedc4595-aabf-4780-b92a-0a605933d8d6 · inbound
ViFusion: In-Network Tensor Fusion for Scalable Video Feature Indexing DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e60790ab-ee19-4e09-a1ef-42a9d02db857 · inbound
Hecto: Modular Sparse Experts for Adaptive and Interpretable Reasoning DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1efa5c55-1c32-41a8-a2fd-f94b669458ff · inbound
Lilith: Developmental Modular LLMs with Chemical Signaling DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8bfb265-4748-4fe9-9caa-a729f3db85ad · inbound
MoE-Beyond: Learning-Based Expert Activation Prediction on Edge Devices DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation add94ac0-d2e8-450f-beb4-56798e2c52e7 · inbound
Taming the Chaos: Coordinated Autoscaling for Heterogeneous and Disaggregated LLM Inference DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2da6820-2894-4665-8d46-9c877e452924 · inbound
InfiniLoRA: Disaggregated Multi-LoRA Serving for Large Language Models DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d66fa626-19c4-4c42-ac0c-12dac90fa6be · inbound
Scaling Multi-Node Mixture-of-Experts Inference Using Expert Activation Patterns DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7ad34d7a-9293-4320-8ee3-62a0b1adc42d · inbound
Mamoda2.5: Enhancing Unified Multimodal Model with DiT-MoE DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2b3f315d-a421-439c-9a50-5f12bf42227b · inbound
Piper: Efficient Large-Scale MoE Training via Resource Modeling and Pipelined Hybrid Parallelism DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 107b3812-4841-44d2-82cf-c27f45b52dcc · inbound
UniPool: A Globally Shared Expert Pool for Mixture-of-Experts DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d0eb442a-f1cf-44d7-805a-660f5595b62f · inbound
MegaScale-Omni: A Hyper-Scale, Workload-Resilient System for MultiModal LLM Training in Production DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a1849866-3667-4f10-885c-d955ce1e7070 · inbound
Surviving Partial Rank Failures in Wide Expert-Parallel MoE Inference DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9e642b10-4d14-4661-9f30-87781363a849 · inbound
TIDE: Efficient and Lossless MoE Diffusion LLM Inference with I/O-aware Expert Offload DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3015ac02-6974-4910-8eef-975fc1f85fd0 · inbound
Hyperbolic and Evidence-Prioritized Experts for Large Vision-Language Models DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 96cb56dc-3ea7-4769-a6d4-f553726eeb32 · inbound
PCCL: Process Group-Aware Scalable and Generic Collective Algorithm Synthesizer DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bde65f2a-46c9-434d-b8be-446d7b679644 · inbound
EPnG: Adaptive Expert Prune-and-Grow for Parameter-Efficient MoE Fine-tuning DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 77eefc14-4106-4b35-b2b6-885d49af2fd6 · inbound
Communication-Aware Placement and Pruning for Efficient Mixture-of-Experts Inference DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scale
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.