Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 65 inbound Pith citation observations for arXiv:2203.03466.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T16:30:40.480336Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
22
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 49ebe1bb-df9a-4815-97d4-dcafa6184419 · inbound
The Falcon Series of Open Language Models Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c444c54f-998f-4ff6-8b4f-73b341ea2453 · inbound
MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation deffac94-4f6c-4eac-b5e7-54e88ba42668 · inbound
Quantum Machine Learning: A Hands-on Tutorial for Machine Learning Practitioners and Researchers Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 297
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5e1b34c-27a7-47fb-828b-3d361a3956a8 · inbound
Peri-LN: Revisiting Normalization Layer in the Transformer Architecture Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d732df23-9454-40e7-b4ec-9883a72a9547 · inbound
Learning Real-World Action-Video Dynamics with Heterogeneous Masked Autoregression Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04e73110-0a67-4513-93e1-ad65f68f6e77 · inbound
Training Deep Learning Models with Norm-Constrained LMOs Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 213
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f0f72734-4604-44e2-8874-80dae20de07e · inbound
Adaptive kernel predictors from feature-learning infinite limits of neural networks Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa20308c-6c33-4e18-b5bd-0f87e03f4ec8 · inbound
Dense Local Dependencies Induce Attention-Logit Explosion and Training Instability During Long-Sequence Transformer Training Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 697b68ac-94da-4602-88c9-c39d7c8cf124 · inbound
Eigenspectrum Analysis of Neural Networks without Aspect Ratio Bias Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52abb96c-b5eb-434a-a4a2-e9fb7259ebac · inbound
NeurIPS 2025 E2LM Competition : Early Training Evaluation of Language Models Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40a7be92-1fac-4728-bf5c-e254c7922771 · inbound
MiniCPM4: Ultra-Efficient LLMs on End Devices Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72c0ebd4-46d4-481f-a0f6-93c6d391586c · inbound
Scaling Transformers for Time Series Forecasting: Do Pretrained Large Models Outperform Small-Scale Alternatives? Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ce7b782-41b0-4714-80a4-1790ebce2626 · inbound
Decoupled Relative Learning Rate Schedules Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9fd95c1-acac-4044-b814-f73f22274a83 · inbound
SingLoRA: Low Rank Adaptation Using a Single Matrix Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c6d03a7-2de4-4203-a7a9-4740bcf55b6a · inbound
Simple Convergence Proof of Adam From a Sign-like Descent Perspective Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3427ea84-fa32-4cde-859f-82d61e64520b · inbound
Tree-Structured Parzen Estimator Can Solve Black-Box Combinatorial Optimization More Efficiently Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a97201bc-28ae-4c43-ab1a-f42024abd097 · inbound
BlockFFN: Towards End-Side Acceleration-Friendly Mixture-of-Experts with Chunk-Level Activation Sparsity Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc378ec0-4c39-468d-b6b5-76884dd7a08f · inbound
Sub-Scaling Laws: On the Role of Data Density and Training Strategies in LLMs Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34726096-14df-4e98-91c6-80915f8efca6 · inbound
Language Models Improve When Pretraining Data Matches Target Tasks Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 114
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 483f52a0-8dc1-42fe-bf33-4d97f6bd4c49 · inbound
Geometry of Neural Reinforcement Learning in Continuous State and Action Spaces Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 135
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0743b662-ab65-4238-a029-20da4a2bda73 · inbound
Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 117
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81b581f8-cae7-4270-9ef3-5eca51b731dd · inbound
FM4NPP: A Scaling Foundation Model for Nuclear and Particle Physics Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0af2bd64-eb77-4913-9748-358a415f6665 · inbound
Customizing the Inductive Biases of Softmax Attention using Structured Matrices Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab5518ee-fc94-4220-a46e-72af56d69016 · inbound
Why Low-Precision Transformer Training Fails: An Analysis on Flash Attention Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fd660cdf-3628-46a6-b947-574615fd1ace · inbound
Scaling depth capacity via zero/one-layer model expansion Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb368725-c50d-45c7-8f62-05e13dca0e1a · inbound
Deep Delta Learning Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f573360b-9b5b-4e98-adca-cbf73a1091b1 · inbound
Deep Delta Learning Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b000a58-37e8-4306-b5c5-7c2ae9ac8dca · inbound
SPARKLING: Balancing Signal Preservation and Symmetry Breaking for Width-Progressive Learning Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3214c141-a4e9-4554-9b9e-31637c460081 · inbound
Theory of Optimal Learning Rate Schedules and Scaling Laws for a Random Feature Model Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 67f86891-955d-4825-a120-c12d2d5ae5e3 · inbound
Spectral Condition for $\mu$P under Width-Depth Scaling Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d1182c22-33cd-4657-afeb-7823b8e9844c · inbound
Rethinking Language Model Scaling under Transferable Hypersphere Optimization Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e90e9a0c-317b-4ffb-b84e-5ab5945b74d3 · inbound
MixAtlas: Uncertainty-aware Data Mixture Optimization for Multimodal LLM Midtraining Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation eecc8385-3b1a-4813-964e-1f18c989a4de · inbound
There Will Be a Scientific Theory of Deep Learning Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 54776053-ae13-4975-8faf-474c8e3bcb47 · inbound
Feature Starvation as Geometric Instability in Sparse Autoencoders Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 74854a07-b373-4121-8df6-0403657081e1 · inbound
OrScale: Orthogonalised Optimization with Layer-Wise Trust-Ratio Scaling Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b34976b3-b3b6-439e-bef8-05741a924671 · inbound
Spectral Dynamics in Deep Networks: Feature Learning, Outlier Escape, and Learning Rate Transfer Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 72b3403e-7121-4245-9d98-d3e8d63d4b50 · inbound
Spectral Dynamics in Deep Networks: Feature Learning, Outlier Escape, and Learning Rate Transfer Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9f20b3bb-8813-4816-bbe7-768aa250bd67 · inbound
Sparse Layers are Critical to Scaling Looped Language Models Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8e1563b3-a639-4f80-b3cb-96f1a24bc552 · inbound
Sparse Layers are Critical to Scaling Looped Language Models Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 223ee6e0-6e30-416c-bd99-f75729fde684 · inbound
Intrinsic Muon: Spectral Optimization on Riemannian Matrix Manifolds Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bbd3b439-e241-4331-b0e6-ac119083b36a · inbound
Mix, Don't Tune: Bilingual Pre-Training Outperforms Hyperparameter Search in Data-Constrained Settings Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 137d2869-5d3a-495f-be56-6ce14f851813 · inbound
How to Scale Mixture-of-Experts: From muP to the Maximally Scale-Stable Parameterization Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 116
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ea5e42c0-568f-4e84-8ab1-52c434573633 · inbound
GQA-{\mu}P: The maximal parameterization update for grouped query attention Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 179ecece-900d-401e-86c9-a166afde2f41 · inbound
Simply Stabilizing the Loop via Fully Looped Transformer Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7a499c47-f64b-49f4-a754-0378549c499c · inbound
Block-Based Double Decoders Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e40a3231-baa4-4812-8478-49d6dfdc0cdb · inbound
Block-Based Double Decoders Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8c8492ca-96bf-428a-b8bd-ec0762b8846e · inbound
Unified Neural Scaling Laws Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d0fdc726-c70c-4df4-9a57-e1a44dd9b2bc · inbound
MuCon: Clipped Muon Updates for LLM Training Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f3b1de82-4626-42fd-9ac3-884dae64da54 · inbound
Spectral Reach: Understanding Neural Scaling as Progress into the Spectral Tail Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de0c3d54-50eb-47db-8ec4-798875eb26da · inbound
On the Scaling of PEFT: Towards Million Personal Models of Trillion Parameters Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7ae09695-3ace-4ddd-970f-0f69a7e1452b · inbound
Unlocking Feature Learning in Gated Delta Networks at Scale Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation aee8d983-2224-4d30-94c8-a2bf7f7e5443 · inbound
Spectral Scaling Laws of Muon Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation de8c3ed4-101f-4eba-bafc-97276805c1e0 · inbound
Double Preconditioning (DoPr): Optimization for Test-Time Performance, not Validation Loss Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 102
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 344187bf-160c-4d78-a7e9-d048c05d745f · inbound
On the Residual Scaling of Looped Transformers: Stability and Transferability Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 14a3a96b-c447-4fd7-b4ff-5620e21c9628 · inbound
LLM Evolution as an Industry-Scale Ecosystem: A Lifecycle Perspective on Continual Learning Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 125
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ad024d55-dee2-4067-9e3c-0b8a1ac4c6bd · inbound
Improving Neural Network Training by Decoupling the Magnitude and Direction of Weight Vectors Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 145
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d0ce3013-56c9-42f3-8993-87acf7fe0c60 · inbound
Improving Neural Network Training by Decoupling the Magnitude and Direction of Weight Vectors Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 323bfc35-c383-4d2e-84ee-2b2faada5d28 · inbound
Geometric Dyson Brownian Motions and the Free Log-Normal Limit for a Non-Square Gaussian Matrix Product Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cd3b708c-ec2c-4b61-bcfe-e8f332637497 · inbound
Geometric Dyson Brownian Motions and the Free Log-Normal Limit for a Non-Square Gaussian Matrix Product Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e15db8af-203a-4c7f-9dfb-862f7c416441 · inbound
The Role of Rigor in Artificial Intelligence Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6143a722-f2b9-467f-bd1d-61bb7bc87469 · inbound
Index SLM Technical Report Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38cbf6cf-6290-4515-aa64-80b98712a8be · inbound
Scale Weight Decay and Train Better Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4feb1af7-058a-4472-a586-d05d034f1dd2 · inbound
Bridging Compute- and Data-Optimal Pretraining Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9cc9133-94d9-4698-a52f-d1cec4bca3cb · inbound
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d27a86f-b256-4327-8e34-4a6e96894d2e · inbound
Sign compression for Muon: SignMuon, MuonSign, and the Limits of Error Feedback Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.