Fetching the paper…
Reading the bibliography…
Automatically determining the optimal size of a neural network for a given task without prior information currently requires an expensive global search and training many networks from scratch.
Three constructive algorithms for network learning
Stephen Gallant · 1986
Earlier work this paper cites.
Dynamic node creation in backpropagation networks
Timur Ash · 1989
Earlier work this paper cites.
The cascade-correlation learning architecture
Scott Fahlman and Christian Lebiere · 1990
Earlier work this paper cites.
A practical bayesian framework for backpropagation networks
David McKay · 1992
Earlier work this paper cites.
Regression shrinkage and selection via the lasso
Robert Tibshirani · 1996
Earlier work this paper cites.
Computing with infinite networks
Christopher K. I. Williams · 1997
Earlier work this paper cites.
Adaptive regularization in neural network modeling
Jan Larsen, Claus Svarer, Lars Nonboe Andersen, and Lars Kai Hansen · 1998
Earlier work this paper cites.
Bayesian methods for neural networks
Juan F. De Freitas · 2003
Earlier work this paper cites.
A fast iterative shrinkage-thresholding algorithm for linear inverse problems
Amir Back and Marc Teboulle · 2006
Earlier work this paper cites.
Model selection and estimation in regression with grouped variables
Ming Yuan and Yin Lin · 2006
Earlier work this paper cites.
Sequential model-based optimization for general algorithm configuration (extended version)
Frank Hutter, Holger H. Hoos, and Kevin Leyton-Brown · 2009
Earlier work this paper cites.
Adaptive subgradient methods for online learning and stochastic optimization
John Duchi, Elad Hazan, and Yoram Singer · 2011
Earlier work this paper cites.
Computing with infinite networks
Andrew Saxe, Pang Wei Koh, Zhenghao Chen, Maneesh Bhand, Bipin Suresh, and Andrew Ng · 2011
Earlier work this paper cites.
Random search for hyper-parameter optimization
James Bergstra and Yoshua Bengio · 2012
Earlier work this paper cites.
Practical bayesian optimization of machine learning algorithms
Jasper Snoek, Hugo Larochelle, and Ryan P. Adams · 2012
Cited alongside, same era.
Lecture 6.5 - rmsprop, coursera: Neural networks for machine learning
Tijmen Tieleman and Geoffrey Hinton · 2012
Cited alongside, same era.
Adadelta: An adaptive learning rate method
Matthew D. Zeiler · 2012
Cited alongside, same era.
Improving deep neural networks for lvcsr using rectified linear units and dropout
George E. Dahl, Tara N. Sainath, and Geoffrey E. Hinton · 2013
Cited alongside, same era.
On the importance of initialization and momentum in deep learning
Ilya Sutskever, James Martens, George Dahl, and Geoffrey Hinton · 2013
Cited alongside, same era.
Do deep nets really need to be deep?
Fitnets: Hints for thin deep nets
Adriana Romero, Nicolas Ballas, Samira Ebrahimi Kahou, Antoine Chassang, Carlo Gatta, and Yoshua Bengio · 2015
Later among the works it cites.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman · 2015
Later among the works it cites.
Scalable bayesian optimization using deep neural networks
Jasper Snoek, Oren Rippel, Kevin Swersky, Ryan Kiros, Nadathur Satish, Narayanan Sundaram, Md. Mostofa Ali Patwary, Prabhat, and Ryan P. Adams · 2015
Later among the works it cites.
Learning the number of neurons in deep networks
Jose M. Alvarez and Mathieu Salzmann · 2016
Later among the works it cites.
Net2net: accelerating learning via knowledge transfer
Tianqi Chen, Ian Goodfellow, and Jonathon Shlens · 2016
Later among the works it cites.
Perforatedcnns: Acceleration through elimination of redundant convolutions
Michael Figurnov, Aijan Ibraimova, Dmitry Vetrov, and Pushmeet Kohli · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Lei Jimmy Ba and Rich Caruana · 2014
Cited alongside, same era.
Memory bounded deep convolutional networks
Maxwell D. Collins and Pushmeet Kohli · 2014
Cited alongside, same era.
Learning by stretching deep networks
Gaurav Pandey and Ambedkar Dukkipati · 2014
Cited alongside, same era.
Learning the structure of deep convolutional networks
Jiashi Feng and Trevor Darrell · 2015
Cited alongside, same era.
Steps toward deep kernel methods from infinite neural networks
Tamir Hazan and Tommi Jaakkola · 2015
Cited alongside, same era.
Distilling the knowledge in a neural network
Geoffrey Hinton, Oriol Vinyals, and Jeff Dean · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy · 2015
Cited alongside, same era.
Later among the works it cites.
Dynamic network durgery for efficient dnns
Yiwen Guo, Anbang Yao, and Yurong Chen · 2016
Later among the works it cites.
Scalable gradient-based tuning of continuous regularization hyperparameters
Jelena Luketina, Mathias Berglund, Klaus Greff, and Raiko Tapani · 2016
Later among the works it cites.
Bayesian optimization with robust bayesian neural networks
Jost Tobias Springenberg, Aaron Klein, Stefan Falkner, and Frank Hutter · 2016
Later among the works it cites.
Network morphism
Tao Wei, Changhu Wang, Yong Rui, and Chang Wen Chen · 2016
Later among the works it cites.
Learning structured sparsity in deep neural networks
Wei Wen, Chunpeng Wu, Wandan Wang, Yiran Chen, and Hai Li · 2016
Later among the works it cites.
Dsd: Dense-sparse-dense training for deep neural networks
Aaron Klein, Stefan Falkner, Jost Tobias Springenberg, and Frank Hutter · 2017
Closest in time.
Pruning convolutional neural networks for efficient inference
Pavlo Molchanov, Stephen Tyree, Tero Karras, Timo Aila, and Jan Kautz · 2017
Closest in time.
Neural architecture search with reinforcement learning
Barret Zoph and Quoc V. Le · 2017
Closest in time.