Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T14:53:28.952395Z
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 0 inbound Pith citation observations for arXiv:2608.13277.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T14:53:28.952395Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
30 of 30 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 98109fd7-6f8b-40cb-b188-b0a7ed773738 · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model Scaling Learning Algorithms Towards
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8dbfa3d5-3abe-47bc-91b7-33d60a19f9fc · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model and Osindero, Simon and Teh, Yee Whye , journal =
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9809a49-a920-4f2d-aaf9-b2d1ace22bea · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model 2016 , publisher=
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4b26545-0aba-4d8e-b207-040213d58538 · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model Scaling Laws for Neural Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60d5944c-e46d-4c57-8a01-fa4b810cc6d7 · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model 2022 , eprint=
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57dd31bf-f372-40a8-aa36-7e8140d730d5 · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model 2025 , eprint=
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation ea31c015-7823-468e-b558-74bde54c95c2 · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model 2023 , eprint=
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4de13d72-2e35-43eb-a6c1-01df85b31d50 · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model 2020 , eprint=
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd4cc876-0765-4bbc-bf72-8170333a7469 · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model 2022 , editor =
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bb9491f-5c44-4b53-9652-71c82b67cdca · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , month =
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 9fcb1430-c893-4a7b-a565-e5e18a7b021a · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model Routing Networks and the Challenges of Modular and Compositional Computation
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c80b79f3-cb58-4040-8048-0be01b59df24 · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model 2024 , eprint=
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8998f33e-0bee-4572-8c62-72568bff7f1d · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model Gemma: Open Models Based on Gemini Research and Technology
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3aef1a72-e43d-4464-9615-9cdf9654a132 · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model Distill , year =
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ad97f43-1212-435a-8428-8331de054987 · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model 2025 , eprint=
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 95693b35-454c-4dd4-b555-fa24d7b8c296 · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model 2022 , eprint=
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ec37f5a-8f8a-4761-abd8-267d527b58ad · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model bert2 BERT : Towards Reusable Pretrained Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 911d241e-5440-49f3-adad-26074b55ca2a · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model Revisiting Model Stitching to Compare Neural Representations
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8974cb8d-5f80-41bd-ac65-faedce62ac31 · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model 2023 , eprint=
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6ffa92f1-a89e-4ef2-a459-d30c012bc284 · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model Efficient Large-Scale Distributed Training of Conditional Maximum Entropy Models , url =
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 7b5f18d1-4900-4ca4-8629-ac78eac32f24 · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model Finetuned Language Models Are Zero-Shot Learners
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 118a7fe0-25b7-4844-9d6f-4861088b1b81 · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model 2022 , eprint=
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0b9d635-7744-49ad-be65-ee5b67080821 · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model 2022 , eprint=
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2686474-1188-493b-95da-65c1bf6241a1 · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model 2022 , eprint=
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a8ef394-e7e8-4f03-adaf-c0e050621e9d · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model Efficient Large-Scale Language Model Training on GPU Clusters Using Megatron-LM
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ae7bd33-5974-4081-a7a4-cb45b1d14471 · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model 2026 , eprint=
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 0ac3cd42-437f-4c24-a7e7-57e746b52d4b · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model The Twelfth International Conference on Learning Representations , year =
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 452539a7-ce9d-4dbc-b6df-63e6e4441c1e · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model Stacking Your Transformers: A Closer Look at Model Growth for Efficient
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation af331a98-4586-494d-8ef1-4f6823dc5c4d · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model Proceedings of the IEEE/CVF International Conference on Computer Vision , pages =
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f21f8c71-d134-4461-9602-267260c4bca0 · outbound
Mixture of Training: Recombining Small-Scale Scaffolded Pretraining Runs into a Larger Language Model Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
No inbound Pith citation observations are available.