REVIEW 5 cited by
Surrogate NAS Benchmarks: Going Beyond the Limited Search Spaces of Tabular NAS Benchmarks
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
abstract
The most significant barrier to the advancement of Neural Architecture Search (NAS) is its demand for large computational resources, which hinders scientifically sound empirical evaluations of NAS methods. Tabular NAS benchmarks have alleviated this problem substantially, making it possible to properly evaluate NAS methods in seconds on commodity machines. However, an unintended consequence of tabular NAS benchmarks has been a focus on extremely small architectural search spaces since their construction relies on exhaustive evaluations of the space. This leads to unrealistic results that do not transfer to larger spaces. To overcome this fundamental limitation, we propose a methodology to create cheap NAS surrogate benchmarks for arbitrary search spaces. We exemplify this approach by creating surrogate NAS benchmarks on the existing tabular NAS-Bench-101 and on two widely used NAS search spaces with up to $10^{21}$ architectures ($10^{13}$ times larger than any previous tabular NAS benchmark). We show that surrogate NAS benchmarks can model the true performance of architectures better than tabular benchmarks (at a small fraction of the cost), that they lead to faithful estimates of how well different NAS methods work on the original non-surrogate benchmark, and that they can generate new scientific insight. We open-source all our code and believe that surrogate NAS benchmarks are an indispensable tool to extend scientifically sound work on NAS to large and exciting search spaces.
Forward citations
Cited by 5 Pith papers
-
Rethinking Performance Analysis for Configurable Software Systems: A Case Study from a Fitness Landscape Perspective
Modeling configurable software performance as a spatial fitness landscape reveals highly rugged terrain, many scattered local optima, few consistently important options, and prevalent high-order interactions.
-
Architecture-Aware Learning Curve Extrapolation via Graph Ordinary Differential Equation
LC-GODE uses a graph-encoded architecture embedding inside a latent neural ODE to extrapolate learning curves, reporting lower error and better model ranking than architecture-agnostic baselines on four benchmarks.
-
Deep Electromagnetic Structure Design Under Limited Evaluation Budgets
A progressive quadtree search with consistency-based sample selection designs electromagnetic structures with 1000 simulations, beating baselines that use up to 7000.
-
Refining Salience-Aware Sparse Fine-Tuning Strategies for Language Models
A static sparsity mask chosen by plain gradients matches or beats LoRA and second-order salience metrics across NLP fine-tuning benchmarks.
-
On Accelerating Edge AI: Optimizing Resource-Constrained Environments
The paper argues that model compression, neural architecture search, and compiler optimizations work together to accelerate edge AI, but it provides no new experimental evidence.
Discussion (0). Continue with ORCID to comment.