Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2103.16716.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T14:29:49.048895Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-29T10:03:17.812816Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 422d557e-f2d9-4302-9520-2311b527c97e · inbound
ST-MoE: Designing Stable and Transferable Sparse Expert Models BASE Layers: Simplifying Training of Large, Sparse Models
Reference 170
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 016edb7f-bd71-4238-9d6b-453ca3201168 · inbound
OPT: Open Pre-trained Transformer Language Models BASE Layers: Simplifying Training of Large, Sparse Models
Reference 170
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation fb0dd2ee-7687-41a9-b492-3a32e9295655 · inbound
LLM.int8(): 8-bit Matrix Multiplication for Transformers at Scale BASE Layers: Simplifying Training of Large, Sparse Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 66d12c9b-7f6f-4926-8994-39f71ca8ff5a · inbound
A Minimal Bifurcation Model of Load Imbalance in a Softmax Mixture-of-Experts Router BASE Layers: Simplifying Training of Large, Sparse Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a98cf91f-94b7-4bdc-afd9-d0719badc588 · inbound
The Evolution of Mixture-of-Experts Architectures in Large Language Models: Routing, Topology, Load Balancing, and Expert Parallelism BASE Layers: Simplifying Training of Large, Sparse Models
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6a0a8a9-c102-4e9d-9621-8e5174d5fa11 · inbound
Riemann GeoResolver: A Non-Euclidean Attention Framework from Euclidean Resolver to Hyperbolic-Spherical Geometry BASE Layers: Simplifying Training of Large, Sparse Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.