Fetching the paper…
Reading the bibliography…
In this work, we investigate the use of sparsity-inducing regularizers during training of Convolution Neural Networks (CNNs).
A simple weight decay can improve generalization
Anders Krogh and John A. Hertz · 1992
Earlier work this paper cites.
Pruning algorithms-a survey
Russell Reed · 1993
Earlier work this paper cites.
Bagging predictors
Leo Breiman · 1996
Earlier work this paper cites.
Regression shrinkage and selection via the lasso
Robert Tibshirani · 1996
Earlier work this paper cites.
Sparse coding with an overcomplete basis set: A strategy employed by V1?
Bruno A. Olshausen and David J. Field · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann LeCun, Léon Bottou, Yoshua Bengio, and Patrick Haffner · 1998
Earlier work this paper cites.
Model compression
Cristian Buciluǎ, Rich Caruana, and Alexandru Niculescu-Mizil · 2006
Earlier work this paper cites.
Compressed sensing
David L. Donoho · 2006
Earlier work this paper cites.
Lessons from the netflix prize challenge
Robert M. Bell and Yehuda Koren · 2007
Earlier work this paper cites.
80 million tiny images: A large data set for nonparametric object and scene recognition
Antonio Torralba, Rob Fergus, and William T. Freeman · 2008
Earlier work this paper cites.
Extracting and composing robust features with denoising autoencoders
Pascal Vincent, Hugo Larochelle, Yoshua Bengio, and Pierre-Antoine Manzagol · 2008
Cited alongside, same era.
ImageNet: A large-scale hierarchical image database
Jia Deng, Wei Dong, R. Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Cited alongside, same era.
Learning multiple layers of features from tiny images
Alex Krizhevsky · 2009
Cited alongside, same era.
Stochastic gradient descent training for L1-regularized log-linear models with cumulative penalty
Yoshimasa Tsuruoka, Jun’ichi Tsujii, and Sophia Ananiadou · 2009
Cited alongside, same era.
Sparse reconstruction by separable approximation
Stephen J. Wright, Robert D. Nowak, and Mário A. T. Figueiredo · 2009
Cited alongside, same era.
Large-scale machine learning with stochastic gradient descent
Léon Bottou · 2010
ImageNet classification with deep convolutional neural networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton · 2012
Later among the works it cites.
Practical bayesian optimization of machine learning algorithms
Jasper Snoek, Hugo Larochelle, and Ryan P. Adams · 2012
Later among the works it cites.
Min Lin, Qiang Chen, and Shuicheng Yan · 2013
Later among the works it cites.
Parallel stochastic gradient algorithms for large-scale matrix completion
Benjamin Recht and Christopher Ré · 2013
Later among the works it cites.
On the importance of initialization and momentum in deep learning
Ilya Sutskever, James Martens, George Dahl, and Geoffrey E. Hinton · 2013
Later among the works it cites.
Caffe: Convolutional architecture for fast feature embedding
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Convergence rates of inexact proximal-gradient methods for convex optimization
Mark Schmidt, Nicolas L. Roux, and Francis R. Bach · 2011
Cited alongside, same era.
Pegasos: Primal estimated sub-gradient solver for SVM
Shai Shalev-Shwartz, Yoram Singer, Nathan Srebro, and Andrew Cotter · 2011
Cited alongside, same era.
Large scale distributed deep networks
Jeffrey Dean, Greg Corrado, Rajat Monga, Kai Chen, Matthieu Devin, Mark Mao, Marc’Aurelio Ranzato, Andrew Senior, Paul Tucker, Ke Yang, Quoc V. Le, and Andrew Y. Ng · 2012
Cited alongside, same era.
Improving neural networks by preventing co-adaptation of feature detectors
Geoffrey E. Hinton, Nitish Srivastava, Alex Krizhevsky, Ilya Sutskever, and Ruslan R. Salakhutdinov · 2012
Cited alongside, same era.
Yangqing Jia, Evan Shelhamer, Jeff Donahue, Sergey Karayev, Jonathan Long, Ross Girshick, Sergio Guadarrama, and Trevor Darrell · 2014
Closest in time.
Speeding up convolutional neural networks with low rank expansions
Andrew Zisserman Max Jaderberg, Andrea Vedaldi · 2014
Closest in time.
Overfeat: Integrated recognition, localization and detection using convolutional networks
Pierre Sermanet, David Eigen, Xiang Zhang, Michaël Mathieu, Rob Fergus, and Yann LeCun · 2014
Closest in time.
Going deeper with convolutions
Christian Szegedy, Wei Liu, Yangqing Jia, Pierre Sermanet, Scott Reed, Dragomir Anguelov, Dumitru Erhan, Vincent Vanhoucke, and Andrew Rabinovich · 2014
Closest in time.
Efficient and accurate approximations of nonlinear convolutional networks
Xiangyu Zhang, Jianhua Zou, Xiang Ming, Kaiming He, and Jian Sun · 2014
Closest in time.