Pith. sign in

REVIEW 4 cited by

Random Search and Reproducibility for Neural Architecture Search

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1902.07638 v3 pith:K2XZIC5D submitted 2019-02-20 cs.LG stat.ML

classification cs.LGstat.ML
keywords searchrandomresultscompetitiveearly-stoppingweight-sharingarchitecturebaseline
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Neural architecture search (NAS) is a promising research direction that has the potential to replace expert-designed networks with learned, task-specific architectures. In this work, in order to help ground the empirical results in this field, we propose new NAS baselines that build off the following observations: (i) NAS is a specialized hyperparameter optimization problem; and (ii) random search is a competitive baseline for hyperparameter optimization. Leveraging these observations, we evaluate both random search with early-stopping and a novel random search with weight-sharing algorithm on two standard NAS benchmarks---PTB and CIFAR-10. Our results show that random search with early-stopping is a competitive NAS baseline, e.g., it performs at least as well as ENAS, a leading NAS method, on both benchmarks. Additionally, random search with weight-sharing outperforms random search with early-stopping, achieving a state-of-the-art NAS result on PTB and a highly competitive result on CIFAR-10. Finally, we explore the existing reproducibility issues of published NAS results. We note the lack of source material needed to exactly reproduce these results, and further discuss the robustness of published results given the various sources of variability in NAS experimental setups. Relatedly, we provide all information (code, random seeds, documentation) needed to exactly reproduce our results, and report our random search with weight-sharing results for each benchmark on multiple runs.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. HM-NAS: Efficient Neural Architecture Search via Hierarchical Masking

    cs.LG 2019-08 conditional novelty 6.0 of 10

    HM-NAS reaches 2.41% test error on CIFAR-10 with 1.8M parameters and 1.8 GPU days, and 73.4% top-1 on ImageNet, by learning hierarchical masks over a weight-sharing supernet.

  2. Feature Partitioning for Efficient Multi-Task Architectures

    cs.LG 2019-08 conditional novelty 6.0 of 10

    A channel-level feature-partitioning search space with a distillation proxy lets multi-task architecture search quickly find efficient sharing patterns.

  3. MANAS: Multi-Agent Neural Architecture Search

    cs.CV 2019-09 reject novelty 5.0 of 10

    MANAS frames DARTS-style neural architecture search as parallel adversarial bandits per edge, but its theoretical guarantee requires other agents' actions to be fixed and it omits the key weight-sharing random-search ...

  4. Scalable Reinforcement-Learning-Based Neural Architecture Search for Cancer Deep Learning Research

    cs.LG 2019-09 conditional novelty 5.0 of 10

    Reinforcement-learning-based neural architecture search finds smaller and faster neural networks with accuracy comparable to, or better than, manually designed networks on three cancer drug-response benchmarks.

Pith tools