Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2109.10686.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T16:20:38.370936Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T20:58:58.624838Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation f5c8e044-8d91-4c13-8ae6-394d6e6722d2 · inbound
ST-MoE: Designing Stable and Transferable Sparse Expert Models Scale Efficiently: Insights from Pre-training and Fine-tuning Transformers
Reference 203
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7cbb63ad-5a56-42f2-8554-9bf072061909 · inbound
BloombergGPT: A Large Language Model for Finance Scale Efficiently: Insights from Pre-training and Fine-tuning Transformers
Reference 114
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5bdfa959-e310-4485-baac-dd5456fe053f · inbound
Scaling Data-Constrained Language Models Scale Efficiently: Insights from Pre-training and Fine-tuning Transformers
Reference 113
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4ac6baf2-3cc7-48da-a810-1e916a881b4e · inbound
Chronos: Learning the Language of Time Series Scale Efficiently: Insights from Pre-training and Fine-tuning Transformers
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f22ca0db-c68d-4623-9a52-b12d456b3865 · inbound
PiKE: Adaptive Data Mixing for Large-Scale Multi-Task Learning Under Low Gradient Conflicts Scale Efficiently: Insights from Pre-training and Fine-tuning Transformers
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ceb5ec8e-0c97-4438-a924-013435778864 · inbound
MUDDFormer: Breaking Residual Bottlenecks in Transformers via Multiway Dynamic Dense Connections Scale Efficiently: Insights from Pre-training and Fine-tuning Transformers
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b3b7d66-e4a5-42fb-8d59-7cbebfbb1b7c · inbound
An Efficient Private GPT Never Autoregressively Decodes Scale Efficiently: Insights from Pre-training and Fine-tuning Transformers
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61d5f9ca-7202-40c5-8fdd-4c14bb08c6b2 · inbound
Progressive Scaling Visual Object Tracking Scale Efficiently: Insights from Pre-training and Fine-tuning Transformers
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd6fc5f9-c955-458c-9330-f74c042996d9 · inbound
Text-to-LoRA: Instant Transformer Adaption Scale Efficiently: Insights from Pre-training and Fine-tuning Transformers
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f3d18c2-6685-4b5f-a35a-2ac7f7ab6ca8 · inbound
Random Initialization Can't Catch Up: The Advantage of Language Model Transfer for Time Series Forecasting Scale Efficiently: Insights from Pre-training and Fine-tuning Transformers
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 348a92a7-7bae-4d58-bf8f-fbcd81cf3ceb · inbound
Can Interpretation Predict Behavior on Unseen Data? Scale Efficiently: Insights from Pre-training and Fine-tuning Transformers
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 080ef9b7-cddb-4152-bcc5-4922789c9383 · inbound
On the Fitness Landscape in the $NK$ Model Scale Efficiently: Insights from Pre-training and Fine-tuning Transformers
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 573986e9-fc42-4169-9db4-cb78c5561e47 · inbound
Large-Scale AI and Foundation Models for Neuroscience: A Comprehensive Review Scale Efficiently: Insights from Pre-training and Fine-tuning Transformers
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad6a4eb1-b724-46c8-9c8d-59507364db17 · inbound
Scaling Laws Meet Model Architecture: Toward Inference-Efficient LLMs Scale Efficiently: Insights from Pre-training and Fine-tuning Transformers
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 15cfc1ff-6bfc-4712-9e17-e679d507a60f · inbound
Large-scale Codec Avatars: The Unreasonable Effectiveness of Large-scale Avatar Pretraining Scale Efficiently: Insights from Pre-training and Fine-tuning Transformers
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 93b62383-9748-40ad-aebb-e83093829b24 · inbound
Variable-Width Transformers Scale Efficiently: Insights from Pre-training and Fine-tuning Transformers
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.