Fetching the paper…
Reading the bibliography…
Modern deep learning methods are very sensitive to many hyperparameters, and, due to the long training times of state-of-the-art models, vanilla Bayesian hyperparameter optimization is typically computationally infeasible.
Letter recognition using holland-style adaptive classifiers
Frey, P. W. and Slate, D. J · 1991
Earlier work this paper cites.
Scaling up the accuracy of naive-bayes classifiers: A decision-tree hybrid
Kohavi, R · 1996
Earlier work this paper cites.
Gradient-based learning applied to document recognition
LeCun, Y., Bottou, L., Bengio, Y., and Haffner, P · 2001
Earlier work this paper cites.
Evolutionary data mining with automatic rule generalization
Cattral, R., Oppacher, F., and Deugo, D · 2002
Earlier work this paper cites.
Brochu, E., Cora, V., and de Freitas, N · 2010
Earlier work this paper cites.
Statsmodels: Econometric and statistical modeling with python
Seabold, S. and Perktold, J · 2010
Earlier work this paper cites.
Statsmodels: Econometric and statistical modeling with python
Seabold, S. and Perktold, J · 2010
Earlier work this paper cites.
Algorithms for hyper-parameter optimization
Bergstra, J., Bardenet, R., Bengio, Y., and Kégl, B · 2011
Earlier work this paper cites.
Sequential model-based optimization for general algorithm configuration
Hutter, F., Hoos, H., and Leyton-Brown, K · 2011
Earlier work this paper cites.
Practical Bayesian optimization of machine learning algorithms
Snoek, J., Larochelle, H., and Adams, R. P · 2012
Earlier work this paper cites.
Towards an empirical foundation for assessing Bayesian optimization of hyperparameters
Eggensperger, K., Feurer, M., Hutter, F., Bergstra, J., Snoek, J., Hoos, H., and Leyton-Brown, K · 2013
Earlier work this paper cites.
UCI machine learning repository, 2013
Lichman, M · 2013
Earlier work this paper cites.
Multi-task Bayesian optimization
Swersky, K., Snoek, J., and Adams, R · 2013
Earlier work this paper cites.
Searching for exotic particles in high-energy physics with deep learning
Baldi, P., Sadowski, P., and Whiteson, D · 2014
Earlier work this paper cites.
Making a science of model search: Hyperparameter optimization in hundreds of dimensions for vision architectures
Bergstra, J., Yamins, D., and Cox, D · 2014
Earlier work this paper cites.
Stochastic gradient Hamiltonian Monte Carlo
Chen, T., Fox, E., and Guestrin, C · 2014
Earlier work this paper cites.
Freeze-thaw bayesian optimization
Swersky, K., Snoek, J., and Adams, R · 2014
Earlier work this paper cites.
OpenML: Networked science in machine learning
Vanschoren, J., van Rijn, J., Bischl, B., and Torgo, L · 2014
Cited alongside, same era.
Proceedings of the 32nd International Conference on Machine Learning (ICML’15) , volume 37, 2015. Omnipress
Bach, F. and Blei, D. (eds.) · 2015
Cited alongside, same era.
Efficient benchmarking of hyperparameter optimizers via surrogates
Eggensperger, K., Hutter, F., Hoos, H., and Leyton-Brown, K · 2015
Cited alongside, same era.
Probabilistic backpropagation for scalable learning of Bayesian neural networks
Hernández-Lobato, J. and Adams, R · 2015
Cited alongside, same era.
Scalable Bayesian optimization using deep neural networks
Snoek, J., Rippel, O., Swersky, K., Kiros, R., Satish, N., Sundaram, N., Patwary, M., Prabhat, and Adams, R · 2015
Cited alongside, same era.
Openai gym, 2016
Brockman, G., Cheung, V., Pettersson, L., Schneider, J., Schulman, J., Tang, J., and Zaremba, W · 2016
Learning curve prediction with Bayesian neural networks
Klein, A., Falkner, S., Springenberg, J. T., and Hutter, F · 2017
Later among the works it cites.
Hyperband: Bandit-based configuration evaluation for hyperparameter optimization
Li, L., Jamieson, K., DeSalvo, G., Rostamizadeh, A., and Talwalkar, A · 2017
Later among the works it cites.
Progressive neural architecture search
Liu, C., Zoph, B., Shlens, J., Hua, W., Li, L.-J., Fei-Fei, L., Yuille, A., Huang, J., and Murphy, K · 2017
Later among the works it cites.
On the state of the art of evaluation in neural language models
Melis, G., Dyer, C., and Blunsom, P · 2017
Later among the works it cites.
Multiple adaptive bayesian linear regression for scalable bayesian optimization with warm start
Perrone, V., Jenatton, R., Seeger, M., and Archambeau, C · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Non-stochastic best arm identification and hyperparameter optimization
Jamieson, K. and Talwalkar, A · 2016
Cited alongside, same era.
Towards automatically-tuned neural networks
Mendoza, H., Klein, A., Feurer, M., Springenberg, J., and Hutter, F · 2016
Cited alongside, same era.
Taking the human out of the loop: A review of Bayesian optimization
Shahriari, B., Swersky, K., Wang, Z., Adams, R., and de Freitas, N · 2016
Cited alongside, same era.
Bayesian optimization with robust bayesian neural networks
Springenberg, J., Klein, A., Falkner, S., and Hutter, F · 2016
Cited alongside, same era.
Published online:
2017
Cited alongside, same era.
Hyperparameter optimization of deep neural networks: Combining hyperband with Bayesian model selection
Bertrand, H., Ardon, R., Perrot, M., and Bloch, I · 2017
Cited alongside, same era.
Later among the works it cites.
Multi-information source optimization
Poloczek, M., Wang, J., and Frazier, P · 2017
Later among the works it cites.
Tensorforce: A tensorflow library for applied reinforcement learning
Schaarschmidt, M., Kuhnle, A., and Fricke, K · 2017
Later among the works it cites.
Proximal policy optimization algorithms
Schulman, J., Wolski, F., Dhariwal, P., Radford, A., and Klimov, O · 2017
Later among the works it cites.
Neural architecture search with reinforcement learning
Zoph, B. and Le, Q. V · 2017
Later among the works it cites.
Learning transferable architectures for scalable image recognition
Zoph, B., Vasudevan, V., Shlens, J., and Le, Q. V · 2017
Later among the works it cites.
Hyperparameter optimization of deep neural networks: Combining hyperband with Bayesian model selection
Bertrand, H., Ardon, R., Perrot, M., and Bloch, I · 2017
Later among the works it cites.
Fast Bayesian optimization of machine learning hyperparameters on large datasets
Klein, A., Falkner, S., Bartels, S., Hennig, P., and Hutter, F · 2017
Later among the works it cites.
Hyperband: Bandit-based configuration evaluation for hyperparameter optimization
Li, L., Jamieson, K., DeSalvo, G., Rostamizadeh, A., and Talwalkar, A · 2017
Later among the works it cites.
Massively parallel hyperparameter tuning, 2018
Li, L., Jamieson, K., Rostamizadeh, A., Gonina, K., Hardt, M., Recht, B., and Talwalkar, A · 2018
Closest in time.
Regularized Evolution for Image Classifier Architecture Search
Real, E., Aggarwal, A., Huang, Y., and Le, Q. V · 2018
Closest in time.
Combination of hyperband and bayesian optimization for hyperparameter optimization in deep learning
Wang, J., Xu, J., and Wang, X · 2018
Closest in time.
Combination of hyperband and bayesian optimization for hyperparameter optimization in deep learning
Wang, J., Xu, J., and Wang, X · 2018
Closest in time.