Fetching the paper…
Reading the bibliography…
We present a simple and powerful algorithm for parallel black box optimization called Successive Halving and Classification (SHAC).
Portfolio selection
H. Markowitz · 1952
Earlier work this paper cites.
Designing neural networks using genetic algorithms
G. F. Miller, P. M. Todd, and S. U. Hegde · 1989
Earlier work this paper cites.
Adaptation in natural and artificial systems: an introductory analysis with applications to biology, control, and artificial intelligence
J. H. Holland · 1992
Earlier work this paper cites.
Introduction to Reinforcement Learning
R. S. Sutton and A. G. Barto · 1998
Earlier work this paper cites.
Gradient-based optimization of hyperparameters
Y. Bengio · 2000
Earlier work this paper cites.
Greedy function approximation: a gradient boosting machine
J. H. Friedman · 2001
Earlier work this paper cites.
Classification and regression by randomforest
A. Liaw, M. Wiener, et al · 2002
Earlier work this paper cites.
Evolving neural networks through augmenting topologies
K. O. Stanley and R. Miikkulainen · 2002
Earlier work this paper cites.
Algorithms for hyper-parameter optimization
J. S. Bergstra, R. Bardenet, Y. Bengio, and B. Kégl · 2011
Earlier work this paper cites.
Sequential model-based optimization for general algorithm configuration
F. Hutter, H. H. Hoos, and K. Leyton-Brown · 2011
Earlier work this paper cites.
Random search for hyper-parameter optimization
J. Bergstra and Y. Bengio · 2012
Earlier work this paper cites.
Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups
G. Hinton, L. Deng, D. Yu, G. E. Dahl, A.-r. Mohamed, N. Jaitly, A. Senior, V. Vanhoucke, P. Nguyen, T. N. Sainath, et al · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
Practical bayesian optimization of machine learning algorithms
J. Snoek, H. Larochelle, and R. P. Adams · 2012
Earlier work this paper cites.
Introductory lectures on convex optimization: A basic course
Y. Nesterov · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Cited alongside, same era.
Input warping for bayesian optimization of non-stationary functions
J. Snoek, K. Swersky, R. Zemel, and R. Adams · 2014
Cited alongside, same era.
Dropout: A simple way to prevent neural networks from overfitting
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov · 2014
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
D. Bahdanau, K. Cho, and Y. Bengio · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
S. Ioffe and C. Szegedy · 2015
Cited alongside, same era.
Gradient-based hyperparameter optimization through reversible learning
Population based training of neural networks
M. Jaderberg, V. Dalibard, S. Osindero, W. M. Czarnecki, J. Donahue, A. Razavi, O. Vinyals, T. Green, I. Dunning, K. Simonyan, et al · 2017
Later among the works it cites.
R. Miikkulainen, J. Liang, E. Meyerson, A. Rawal, D. Fink, O. Francon, B. Raju, A. Navruzyan, N. Duffy, and B. Hodjat · 2017
Later among the works it cites.
Deeparchitect: Automatically designing and training deep architectures
R. Negrinho and G. Gordon · 2017
Later among the works it cites.
Large-scale evolution of image classifiers
E. Real, S. Moore, A. Selle, S. Saxena, Y. L. Suematsu, Q. Le, and A. Kurakin · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. Maclaurin, D. Duvenaud, and R. Adams · 2015
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2015
Cited alongside, same era.
Scalable bayesian optimization using deep neural networks
J. Snoek, O. Rippel, K. Swersky, R. Kiros, N. Satish, N. Sundaram, M. Patwary, M. Prabhat, and R. Adams · 2015
Cited alongside, same era.
Going deeper with convolutions
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, A. Rabinovich, et al · 2015
Cited alongside, same era.
Xgboost: A scalable tree boosting system
T. Chen and C. Guestrin · 2016
Cited alongside, same era.
Hypernetworks
D. Ha, A. Dai, and Q. V. Le · 2016
Cited alongside, same era.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Cited alongside, same era.
J. Schulman, F. Wolski, P. Dhariwal, A. Radford, and O. Klimov · 2017
Later among the works it cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin · 2017
Later among the works it cites.
Genetic cnn
L. Xie and A. Yuille · 2017
Later among the works it cites.
Neural architecture search with reinforcement learning
B. Zoph and Q. V. Le · 2017
Later among the works it cites.
Learning transferable architectures for scalable image recognition
B. Zoph, V. Vasudevan, J. Shlens, and Q. V. Le · 2017
Later among the works it cites.
Smash: one-shot model architecture search through hypernetworks
A. Brock, T. Lim, J. M. Ritchie, and N. Weston · 2018
Closest in time.
Derivative free optimization via repeated classification
T. B. Hashimoto, S. Yadlowsky, and J. C. Duchi · 2018
Closest in time.
Neural architecture search with bayesian optimisation and optimal transport
K. Kandasamy, W. Neiswanger, J. Schneider, B. Poczos, and E. Xing · 2018
Closest in time.
Hierarchical representations for efficient architecture search
H. Liu, K. Simonyan, O. Vinyals, C. Fernando, and K. Kavukcuoglu · 2018
Closest in time.
Practical network blocks design with q-learning
Z. Zhong, J. Yan, and C.-L. Liu · 2018
Closest in time.