Fetching the paper…
Reading the bibliography…
Currently, deep neural networks are the state of the art on problems such as speech recognition and computer vision.
Approximation by superpositions of a sigmoidal function
George Cybenko · 1989
Earlier work this paper cites.
Speaker-independent phone recognition using hidden markov models
K-F Lee and H-W Hon · 1989
Earlier work this paper cites.
Model compression
Cristian Buciluǎ, Rich Caruana, and Alexandru Niculescu-Mizil · 2006
Earlier work this paper cites.
Reducing the dimensionality of data with neural networks
G.E. Hinton and R.R. Salakhutdinov · 2006
Earlier work this paper cites.
80 million tiny images: A large data set for nonparametric object and scene recognition
Antonio Torralba, Robert Fergus, and William T Freeman · 2008
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Alex Krizhevsky and Geoffrey Hinton · 2009
Earlier work this paper cites.
Why does unsupervised pre-training help deep learning?
Dumitru Erhan, Yoshua Bengio, Aaron Courville, Pierre-Antoine Manzagol, Pascal Vincent, and Samy Bengio · 2010
Earlier work this paper cites.
Rectified linear units improve restricted boltzmann machines
V. Nair and G.E. Hinton · 2010
Cited alongside, same era.
Stacked denoising autoencoders: Learning useful representations in a deep network with a local denoising criterion
P. Vincent, H. Larochelle, I. Lajoie, Y. Bengio, and P.A. Manzagol · 2010
Cited alongside, same era.
Conversational speech transcription using context-dependent deep neural networks
Frank Seide, Gang Li, and Dong Yu · 2011
Cited alongside, same era.
Applying convolutional neural networks concepts to hybrid nn-hmm model for speech recognition
Ossama Abdel-Hamid, Abdel-rahman Mohamed, Hui Jiang, and Gerald Penn · 2012
Cited alongside, same era.
Improving neural networks by preventing co-adaptation of feature detectors
G.E. Hinton, N. Srivastava, A. Krizhevsky, I. Sutskever, and R.R. Salakhutdinov · 2012
Cited alongside, same era.
Big neural networks waste capacity
Yann N Dauphin and Yoshua Bengio · 2013
Closest in time.
Recent advances in deep learning for speech research at microsoft
Li Deng, Jinyu Li, Jui-Ting Huang, Kaisheng Yao, Dong Yu, Frank Seide, Michael Seltzer, Geoff Zweig, Xiaodong He, Jason Williams, et al · 2013
Closest in time.
Understanding deep architectures using a recursive convolutional network
David Eigen, Jason Rolfe, Rob Fergus, and Yann LeCun · 2013
Closest in time.
Maxout networks
Ian Goodfellow, David Warde-Farley, Mehdi Mirza, Aaron Courville, and Yoshua Bengio · 2013
Closest in time.
Low-rank matrix factorization for deep neural network training with high-dimensional output targets
Tara N Sainath, Brian Kingsbury, Vikas Sindhwani, Ebru Arisoy, and Bhuvana Ramabhadran · 2013
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Acoustic modeling using deep belief networks
Abdel-rahman Mohamed, George E Dahl, and Geoffrey Hinton · 2012
Cited alongside, same era.
Exploring convolutional neural network structures and optimization techniques for speech recognition
Ossama Abdel-Hamid, Li Deng, and Dong Yu · 2013
Cited alongside, same era.
Restructuring of deep neural network acoustic models with singular value decomposition
Jian Xue, Jinyu Li, and Yifan Gong · 2013
Closest in time.
Stochastic pooling for regularization of deep convolutional neural networks
Matthew D Zeiler and Rob Fergus · 2013
Closest in time.