Fetching the paper…
Reading the bibliography…
Neural networks dominate the modern machine learning landscape, but their training and success still suffer from sensitivity to empirical choices of hyperparameters such as model architecture, loss function, and optimisation algorithm.
The generalization of student’s’ problem when several different population variances are involved
Bernard L Welch · 1947
Earlier work this paper cites.
Adapting crossover in evolutionary algorithms
William M. Spears · 1995
Earlier work this paper cites.
Empirical investigation of the benefits of partial Lamarckianism
Christopher R Houck, Jeffery A Joines, Michael G Kay, and James R Wilson · 1997
Earlier work this paper cites.
Exploring the effects of Lamarckian and Baldwinian learning in evolving recurrent neural networks
Kim WC Ku and Man-Wai Mak · 1997
Earlier work this paper cites.
Probabilistic incremental program evolution: Stochastic search through program space
Rafal Salustowicz and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
An overview of parameter control methods by self-adaptation in evolutionary algorithms
Thomas Bäck · 1998
Earlier work this paper cites.
G-prop-iii: Global optimization of multilayer perceptrons using an evolutionary algorithm
PA Castillo, V Rivas, JJ Merelo, Jesús González, Alberto Prieto, and Gustavo Romero · 1999
Earlier work this paper cites.
Optimization for problem classes-neural networks that learn to learn
Michael Husken, Jens E Gayko, and Bernhard Sendhoff · 2000
Earlier work this paper cites.
Self adaptive evolutionary algorithms, 2004
Bartlomiej Gloger · 2004
Earlier work this paper cites.
Evolutionary optimization of neural systems: The use of strategy adaptation
Christian Igel, Stefan Wiegand, and Frauke Friedrichs · 2005
Earlier work this paper cites.
Lamarckian evolution and the Baldwin effect in evolutionary neural networks
PA Castillo, MG Arenas, JG Castellano, JJ Merelo, A Prieto, V Rivas, and G Romero · 2006
Earlier work this paper cites.
Natural selection fails to optimize mutation rates for long-term adaptation on rugged fitness landscapes
Jeff Clune, Dusan Misevic, Charles Ofria, Richard E Lenski, Santiago F Elena, and Rafael Sanjuán · 2008
Earlier work this paper cites.
Gaussian process optimization in the bandit setting: No regret and experimental design
Niranjan Srinivas, Andreas Krause, Sham M Kakade, and Matthias Seeger · 2009
Earlier work this paper cites.
Oracle inequalities for computationally budgeted model selection
Alekh Agarwal, John C Duchi, Peter L Bartlett, and Clement Levrard · 2011
Earlier work this paper cites.
Algorithms for hyper-parameter optimization
James S Bergstra, Rémi Bardenet, Yoshua Bengio, and Balázs Kégl · 2011
Earlier work this paper cites.
Efficient multi-start strategies for local search algorithms
András György and Levente Kocsis · 2011
Earlier work this paper cites.
Sequential model-based optimization for general algorithm configuration
Frank Hutter, Holger H Hoos, and Kevin Leyton-Brown · 2011
Earlier work this paper cites.
Evolutionary computation meets machine learning: A survey
Jun Zhang, Zhi-hui Zhan, Ying Lin, Ni Chen, Yue-jiao Gong, Jing-hui Zhong, Henry SH Chung, Yun Li, and Yu-hui Shi · 2011
Earlier work this paper cites.
Random search for hyper-parameter optimization
James Bergstra and Yoshua Bengio · 2012
Earlier work this paper cites.
Practical Bayesian optimization of machine learning algorithms
Jasper Snoek, Hugo Larochelle, and Ryan P Adams · 2012
Cited alongside, same era.
Lecture 6.5-rmsprop: Divide the gradient by a running average of its recent magnitude
Tijmen Tieleman and Geoffrey Hinton · 2012
Cited alongside, same era.
The arcade learning environment: An evaluation platform for general agents
Marc G Bellemare, Yavar Naddaf, Joel Veness, and Michael Bowling · 2013
Cited alongside, same era.
Multi-task Bayesian optimization
Kevin Swersky, Jasper Snoek, and Ryan P Adams · 2013
Cited alongside, same era.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Cited alongside, same era.
Asynchronous methods for deep reinforcement learning
Volodymyr Mnih, Adria Puigdomenech Badia, Mehdi Mirza, Alex Graves, Timothy Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu · 2016
Later among the works it cites.
Unsupervised representation learning with deep convolutional generative adversarial networks
Alec Radford, Luke Metz, and Soumith Chintala · 2016
Later among the works it cites.
Selecting near-optimal learners via incremental data allocation
Ashish Sabharwal, Horst Samulowitz, and Gerald Tesauro · 2016
Later among the works it cites.
Improved techniques for training GANs
Tim Salimans, Ian Goodfellow, Wojciech Zaremba, Vicki Cheung, Alec Radford, and Xi Chen · 2016
Later among the works it cites.
Bayesian optimization with robust Bayesian neural networks
Jost Tobias Springenberg, Aaron Klein, Stefan Falkner, and Frank Hutter · 2016
Later among the works it cites.
The parallel knowledge gradient method for batch Bayesian optimization
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kevin Swersky, Jasper Snoek, and Ryan Prescott Adams · 2014
Cited alongside, same era.
Speeding up automatic hyperparameter optimization of deep neural networks by extrapolation of learning curves
Tobias Domhan, Jost Tobias Springenberg, and Frank Hutter · 2015
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba · 2015
Cited alongside, same era.
Pierre-Yves Massé and Yann Ollivier · 2015
Cited alongside, same era.
Parallel predictive entropy search for batch global optimization of expensive objective functions
Amar Shah and Zoubin Ghahramani · 2015
Cited alongside, same era.
Scalable Bayesian optimization using deep neural networks
Jasper Snoek, Oren Rippel, Kevin Swersky, Ryan Kiros, Nadathur Satish, Narayanan Sundaram, Mostofa Patwary, Mr Prabhat, and Ryan Adams · 2015
Cited alongside, same era.
Optimizing deep learning hyper-parameters through an evolutionary algorithm
Steven R Young, Derek C Rose, Thomas P Karnowski, Seung-Hwan Lim, and Robert M Patton · 2015
Cited alongside, same era.
Jian Wu and Peter Frazier · 2016
Later among the works it cites.
A survey on evolutionary computation approaches to feature selection
Bing Xue, Mengjie Zhang, Will N Browne, and Xin Yao · 2016
Later among the works it cites.
Google vizier: A service for black-box optimization
Daniel Golovin, Benjamin Solnik, Subhodeep Moitra, Greg Kochanski, John Karro, and D Sculley · 2017
Closest in time.
Improved training of Wasserstein GANs
Ishaan Gulrajani, Faruk Ahmed, Martin Arjovsky, Vincent Dumoulin, and Aaron Courville · 2017
Closest in time.
Hierarchical representations for efficient architecture search
Hanxiao Liu, Karen Simonyan, Oriol Vinyals, Chrisantha Fernando, and Koray Kavukcuoglu · 2017
Closest in time.
Hyperdrive: Exploring hyperparameters with POP scheduling
Jeff Rasley, Yuxiong He, Feng Yan, Olatunji Ruwase, and Rodrigo Fonseca · 2017
Closest in time.
Large-scale evolution of image classifiers
Esteban Real, Sherry Moore, Andrew Selle, Saurabh Saxena, Yutaka Leon Suematsu, Quoc Le, and Alex Kurakin · 2017
Closest in time.
Variational approaches for auto-encoding generative adversarial networks
Mihaela Rosca, Balaji Lakshminarayanan, David Warde-Farley, and Shakir Mohamed · 2017
Closest in time.
Cyclical learning rates for training neural networks
Leslie N Smith · 2017
Closest in time.
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Closest in time.
Feudal networks for hierarchical reinforcement learning
Alexander Sasha Vezhnevets, Simon Osindero, Tom Schaul, Nicolas Heess, Max Jaderberg, David Silver, and Koray Kavukcuoglu · 2017
Closest in time.
Starcraft II: A new challenge for reinforcement learning
Oriol Vinyals, Timo Ewalds, Sergey Bartunov, Petko Georgiev, Alexander Sasha Vezhnevets, Michelle Yeo, Alireza Makhzani, Heinrich Küttler, John Agapiou, Julian Schrittwieser, et al · 2017
Closest in time.
LR-GAN: Layered recursive generative adversarial networks for image generation
Jianwei Yang, Anitha Kanna, Dhruv Batra, and Devi Parikh · 2017
Closest in time.