Pith. sign in

REVIEW 1 cited by

AutoHAS: Efficient Hyperparameter and Architecture Search

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2006.03656 v3 pith:VNQHHTZP submitted 2020-06-05 cs.CV

classification cs.CV
keywords autohassearcharchitecturearchitecturescontrollerefficientweightaccuracy
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Efficient hyperparameter or architecture search methods have shown remarkable results, but each of them is only applicable to searching for either hyperparameters (HPs) or architectures. In this work, we propose a unified pipeline, AutoHAS, to efficiently search for both architectures and hyperparameters. AutoHAS learns to alternately update the shared network weights and a reinforcement learning (RL) controller, which learns the probability distribution for the architecture candidates and HP candidates. A temporary weight is introduced to store the updated weight from the selected HPs (by the controller), and a validation accuracy based on this temporary weight serves as a reward to update the controller. In experiments, we show AutoHAS is efficient and generalizable to different search spaces, baselines and datasets. In particular, AutoHAS can improve the accuracy over popular network architectures, such as ResNet and EfficientNet, on CIFAR-10/100, ImageNet, and four more other datasets.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. PreNeT: Leveraging Computational Features to Predict Deep Neural Network Training Time

    cs.LG 2024-12 conditional novelty 4.0 of 10

    PreNeT predicts per-epoch training time for deep learning layers, including attention and embedding, using computational operations, memory, and GPU peak FLOPs features, and reports up to 72% improvement over a 2018 baseline.

Pith tools