Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2109.08668.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:04:38.115083Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-29T23:14:01.480491Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 1e1eadf9-6752-4c38-8a22-e808dfd44f84 · inbound
ST-MoE: Designing Stable and Transferable Sparse Expert Models Primer: Searching for Efficient Transformers for Language Modeling
Reference 200
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 152da4b4-8478-4165-8a0c-153ba16efdd4 · inbound
Flamingo: a Visual Language Model for Few-Shot Learning Primer: Searching for Efficient Transformers for Language Modeling
Reference 105
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7e631a03-93fe-489a-ae0b-77f276dc3fe6 · inbound
Fast Inference from Transformers via Speculative Decoding Primer: Searching for Efficient Transformers for Language Modeling
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ba483445-22e1-47fa-bc2e-426623c4cc27 · inbound
TriADA: Massively Parallel Trilinear Matrix-by-Tensor Multiply-Add Algorithm and Device Architecture for the Acceleration of 3D Discrete Transformations Primer: Searching for Efficient Transformers for Language Modeling
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b6b6d0c-110e-4537-ad71-4e3d70ec4747 · inbound
Resting Neurons, Active Insights: Robustifying Activation Sparsity in LLMs via Spontaneity Primer: Searching for Efficient Transformers for Language Modeling
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a97283ea-c4a4-401d-a35e-43b2ee9dd29d · inbound
Resting Neurons, Active Insights: Robustifying Activation Sparsity in LLMs via Spontaneity Primer: Searching for Efficient Transformers for Language Modeling
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2dc3b978-431c-4f0a-9962-9642290226d5 · inbound
Physics of Language Models: Part 4.1, Architecture Design and the Magic of Canon Layers Primer: Searching for Efficient Transformers for Language Modeling
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9ce9369-a7ab-4048-ab75-8ec05532f67d · inbound
NVIDIA Nemotron 3: Efficient and Open Intelligence Primer: Searching for Efficient Transformers for Language Modeling
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 31333250-5087-409d-9d92-d90c28a5c707 · inbound
From Competition to Collaboration: Designing Sustainable Mechanisms Between LLMs and Online Forums Primer: Searching for Efficient Transformers for Language Modeling
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a92a4ad1-0ee7-4fac-9646-899228e1fa5e · inbound
Three-Phase Transformer Primer: Searching for Efficient Transformers for Language Modeling
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 81a24bda-88dc-4f43-8406-f21e617dc975 · inbound
ELAS: Efficient Pre-Training of Low-Rank Large Language Models via 2:4 Activation Sparsity Primer: Searching for Efficient Transformers for Language Modeling
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9af5fca8-be7c-47d2-ab7e-c05513672db6 · inbound
On the global convergence of gradient descent for wide shallow models with bounded nonlinearities Primer: Searching for Efficient Transformers for Language Modeling
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9e074bf5-44df-4e0c-8516-be3272f0c9cb · inbound
Bug or Feature$^2$: Weight Drift, Activation Sparsity and Spikes Primer: Searching for Efficient Transformers for Language Modeling
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 41f62084-2ea2-415e-bf29-d466206687d0 · inbound
Bug or Feature$^2$: Weight Drift, Activation Sparsity and Spikes Primer: Searching for Efficient Transformers for Language Modeling
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 03ec595f-8600-461a-b6b7-3a795559c30a · inbound
Mapping the Schedule x Bit-Width Boundary in Sub-100M Quantisation-Aware Training Primer: Searching for Efficient Transformers for Language Modeling
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 430ad754-b9ad-4167-82e4-4d64d919e1c6 · inbound
Can Transformers Really Do It All? On the Compatibility of Inductive Biases Across Tasks Primer: Searching for Efficient Transformers for Language Modeling
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f1d067a-18da-46a7-b512-07118b323b7f · inbound
Domyn-Small: A European 10B Reasoning Language Model Primer: Searching for Efficient Transformers for Language Modeling
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.