REVIEW 7 cited by
Predicting Neural Network Accuracy from Weights
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
We show experimentally that the accuracy of a trained neural network can be predicted surprisingly well by looking only at its weights, without evaluating it on input data. We motivate this task and introduce a formal setting for it. Even when using simple statistics of the weights, the predictors are able to rank neural networks by their performance with very high accuracy (R2 score more than 0.98). Furthermore, the predictors are able to rank networks trained on different, unobserved datasets and with different architectures. We release a collection of 120k convolutional neural networks trained on four different datasets to encourage further research in this area, with the goal of understanding network training and performance better.
Forward citations
Cited by 7 Pith papers
-
On the Expressive Power of Permutation-Equivariant Weight-Space Networks
Permutation-equivariant weight-space networks are all equally expressive, and universality holds when hidden-layer biases are pairwise distinct.
-
Can this Model Also Recognize Dogs? Zero-Shot Model Search from Weights
ProbeLog represents each classifier output by its responses to fixed probe images and uses CLIP to answer text queries, achieving 43.8% top-1 accuracy when searching 1,500 ImageNet-trained models for a concept.
-
WeightCLIP: Aligning Datasets and Models for Weight Space Learning
Contrastive dataset–weight alignment reshapes weight-space latents so dataset prompts retrieve, generate, and refine neural nets better than prior weight-space methods.
-
Suitability Filter: A Statistical Framework for Classifier Evaluation in Real-World Deployment Settings
A statistical non-inferiority test on estimated per-sample correctness probabilities flags when a classifier's accuracy on unlabeled user data drops by more than a chosen margin relative to its test set.
-
An open dataset of neural networks for hypernetwork research
A public dataset of 10,000 LeNet-5 networks split into 10 Imagenette classes is released, with a 72% Naive Bayes baseline for classifying networks by their weights.
-
NeuroVoxel-LM: Language-Aligned 3D Perception via Dynamic Voxelization and Meta-Embedding
NeuroVoxel-LM combines dynamic multi-resolution voxelization with attention-based pooling of NeRF weights, reporting faster 3D feature extraction and modestly better NeRF captioning than fixed-resolution and max-pooli...
-
Predicting Deep Neural Network Training Outcomes from Early Training Telemetry
After only a few epochs, a run's own loss, accuracy, gradient, and weight-norm telemetry predicts its final accuracy (R² = 0.92 to 0.99) and relative performance (AUC = 0.983 to 0.998) across six image-classification domains.
Discussion (0). Continue with ORCID to comment.