BANANAS: Bayesian Optimization with Neural Architectures for Neural Architecture Search

White, Colin; Neiswanger, Willie; Savani, Yash

Computer Science > Machine Learning

arXiv:1910.11858 (cs)

[Submitted on 25 Oct 2019 (v1), last revised 2 Nov 2020 (this version, v3)]

Title:BANANAS: Bayesian Optimization with Neural Architectures for Neural Architecture Search

Authors:Colin White, Willie Neiswanger, Yash Savani

View PDF

Abstract:Over the past half-decade, many methods have been considered for neural architecture search (NAS). Bayesian optimization (BO), which has long had success in hyperparameter optimization, has recently emerged as a very promising strategy for NAS when it is coupled with a neural predictor. Recent work has proposed different instantiations of this framework, for example, using Bayesian neural networks or graph convolutional networks as the predictive model within BO. However, the analyses in these papers often focus on the full-fledged NAS algorithm, so it is difficult to tell which individual components of the framework lead to the best performance.
In this work, we give a thorough analysis of the "BO + neural predictor" framework by identifying five main components: the architecture encoding, neural predictor, uncertainty calibration method, acquisition function, and acquisition optimization strategy. We test several different methods for each component and also develop a novel path-based encoding scheme for neural architectures, which we show theoretically and empirically scales better than other encodings. Using all of our analyses, we develop a final algorithm called BANANAS, which achieves state-of-the-art performance on NAS search spaces. We adhere to the NAS research checklist (Lindauer and Hutter 2019) to facilitate best practices, and our code is available at this https URL.

Subjects:	Machine Learning (cs.LG); Neural and Evolutionary Computing (cs.NE); Machine Learning (stat.ML)
Cite as:	arXiv:1910.11858 [cs.LG]
	(or arXiv:1910.11858v3 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1910.11858
Journal reference:	AAAI Conference on Artificial Intelligence 2021

Submission history

From: Colin White [view email]
[v1] Fri, 25 Oct 2019 17:35:49 UTC (612 KB)
[v2] Wed, 19 Feb 2020 08:39:44 UTC (503 KB)
[v3] Mon, 2 Nov 2020 15:28:47 UTC (413 KB)

Computer Science > Machine Learning

Title:BANANAS: Bayesian Optimization with Neural Architectures for Neural Architecture Search

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:BANANAS: Bayesian Optimization with Neural Architectures for Neural Architecture Search

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators