Pith. sign in

REVIEW 5 cited by

Surrogate NAS Benchmarks: Going Beyond the Limited Search Spaces of Tabular NAS Benchmarks

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2008.09777 v4 pith:HOXTTLSI submitted 2020-08-22 cs.LG

classification cs.LG
keywords benchmarkssearchspacestabularsurrogatemethodsarchitecturesbenchmark
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

The most significant barrier to the advancement of Neural Architecture Search (NAS) is its demand for large computational resources, which hinders scientifically sound empirical evaluations of NAS methods. Tabular NAS benchmarks have alleviated this problem substantially, making it possible to properly evaluate NAS methods in seconds on commodity machines. However, an unintended consequence of tabular NAS benchmarks has been a focus on extremely small architectural search spaces since their construction relies on exhaustive evaluations of the space. This leads to unrealistic results that do not transfer to larger spaces. To overcome this fundamental limitation, we propose a methodology to create cheap NAS surrogate benchmarks for arbitrary search spaces. We exemplify this approach by creating surrogate NAS benchmarks on the existing tabular NAS-Bench-101 and on two widely used NAS search spaces with up to $10^{21}$ architectures ($10^{13}$ times larger than any previous tabular NAS benchmark). We show that surrogate NAS benchmarks can model the true performance of architectures better than tabular benchmarks (at a small fraction of the cost), that they lead to faithful estimates of how well different NAS methods work on the original non-surrogate benchmark, and that they can generate new scientific insight. We open-source all our code and believe that surrogate NAS benchmarks are an indispensable tool to extend scientifically sound work on NAS to large and exciting search spaces.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 5 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Rethinking Performance Analysis for Configurable Software Systems: A Case Study from a Fitness Landscape Perspective

    cs.PF 2024-12 conditional novelty 6.0 of 10

    Modeling configurable software performance as a spatial fitness landscape reveals highly rugged terrain, many scattered local optima, few consistently important options, and prevalent high-order interactions.

  2. Architecture-Aware Learning Curve Extrapolation via Graph Ordinary Differential Equation

    cs.LG 2024-12 conditional novelty 6.0 of 10

    LC-GODE uses a graph-encoded architecture embedding inside a latent neural ODE to extrapolate learning curves, reporting lower error and better model ranking than architecture-agnostic baselines on four benchmarks.

  3. Deep Electromagnetic Structure Design Under Limited Evaluation Budgets

    cs.LG 2025-06 conditional novelty 5.0 of 10

    A progressive quadtree search with consistency-based sample selection designs electromagnetic structures with 1000 simulations, beating baselines that use up to 7000.

  4. Refining Salience-Aware Sparse Fine-Tuning Strategies for Language Models

    cs.CL 2024-12 conditional novelty 5.0 of 10

    A static sparsity mask chosen by plain gradients matches or beats LoRA and second-order salience metrics across NLP fine-tuning benchmarks.

  5. On Accelerating Edge AI: Optimizing Resource-Constrained Environments

    cs.LG 2025-01 conditional novelty 2.0 of 10

    The paper argues that model compression, neural architecture search, and compiler optimizations work together to accelerate edge AI, but it provides no new experimental evidence.

Pith tools