REVIEW 4 cited by
Random Search and Reproducibility for Neural Architecture Search
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Neural architecture search (NAS) is a promising research direction that has the potential to replace expert-designed networks with learned, task-specific architectures. In this work, in order to help ground the empirical results in this field, we propose new NAS baselines that build off the following observations: (i) NAS is a specialized hyperparameter optimization problem; and (ii) random search is a competitive baseline for hyperparameter optimization. Leveraging these observations, we evaluate both random search with early-stopping and a novel random search with weight-sharing algorithm on two standard NAS benchmarks---PTB and CIFAR-10. Our results show that random search with early-stopping is a competitive NAS baseline, e.g., it performs at least as well as ENAS, a leading NAS method, on both benchmarks. Additionally, random search with weight-sharing outperforms random search with early-stopping, achieving a state-of-the-art NAS result on PTB and a highly competitive result on CIFAR-10. Finally, we explore the existing reproducibility issues of published NAS results. We note the lack of source material needed to exactly reproduce these results, and further discuss the robustness of published results given the various sources of variability in NAS experimental setups. Relatedly, we provide all information (code, random seeds, documentation) needed to exactly reproduce our results, and report our random search with weight-sharing results for each benchmark on multiple runs.
Forward citations
Cited by 4 Pith papers
-
HM-NAS: Efficient Neural Architecture Search via Hierarchical Masking
HM-NAS reaches 2.41% test error on CIFAR-10 with 1.8M parameters and 1.8 GPU days, and 73.4% top-1 on ImageNet, by learning hierarchical masks over a weight-sharing supernet.
-
Feature Partitioning for Efficient Multi-Task Architectures
A channel-level feature-partitioning search space with a distillation proxy lets multi-task architecture search quickly find efficient sharing patterns.
-
MANAS: Multi-Agent Neural Architecture Search
MANAS frames DARTS-style neural architecture search as parallel adversarial bandits per edge, but its theoretical guarantee requires other agents' actions to be fixed and it omits the key weight-sharing random-search ...
-
Scalable Reinforcement-Learning-Based Neural Architecture Search for Cancer Deep Learning Research
Reinforcement-learning-based neural architecture search finds smaller and faster neural networks with accuracy comparable to, or better than, manually designed networks on three cancer drug-response benchmarks.
Discussion (0). Continue with ORCID to comment.