Fetching the paper…
Reading the bibliography…
Deep Neural Networks (DNN) have achieved state-of-the-art results in a wide range of tasks, with the best results obtained with large training sets and large models.
A method for unconstrained convex minimization problem with the rate of convergence o ( 1 / k 2 ) o(1/k^{2})
Yu Nesterov · 1983
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann LeCun, Leon Bottou, Yoshua Bengio, and Patrick Haffner · 1998
Earlier work this paper cites.
Expectation propagation for approximate bayesian inference
Thomas P Minka · 2001
Earlier work this paper cites.
A neural probabilistic language model
Yoshua Bengio, Réjean Ducharme, Pascal Vincent, and Christian Jauvin · 2003
Earlier work this paper cites.
Large Scale Machine Learning
R. Collobert · 2004
Earlier work this paper cites.
Hardware complexity of modular multiplication and exponentiation
J.P. David, K. Kalach, and N. Tittley · 2007
Earlier work this paper cites.
Large-scale deep unsupervised learning using graphics processors
Rajat Raina, Anand Madhavan, and Andrew Y. Ng · 2009
Earlier work this paper cites.
A highly scalable restricted Boltzmann machine FPGA implementation
Sang Kyun Kim, Lawrence C McAfee, Peter Leonard McMahon, and Kunle Olukotun · 2009
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Xavier Glorot and Yoshua Bengio · 2010
Earlier work this paper cites.
Rectified linear units improve restricted Boltzmann machines
V. Nair and G.E. Hinton · 2010
Earlier work this paper cites.
Theano: a CPU and GPU math expression compiler
James Bergstra, Olivier Breuleux, Frédéric Bastien, Pascal Lamblin, Razvan Pascanu, Guillaume Desjardins, Joseph Turian, David Warde-Farley, and Yoshua Bengio · 2010
Earlier work this paper cites.
Practical variational inference for neural networks
Alex Graves · 2011
Earlier work this paper cites.
Deep sparse rectifier neural networks
X. Glorot, A. Bordes, and Y. Bengio · 2011
Earlier work this paper cites.
Deep neural networks for acoustic modeling in speech recognition
Geoffrey Hinton, Li Deng, George E. Dahl, Abdel-rahman Mohamed, Navdeep Jaitly, Andrew Senior, Vincent Vanhoucke, Patrick Nguyen, Tara Sainath, and Brian Kingsbury · 2012
Earlier work this paper cites.
ImageNet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. Hinton · 2012
Earlier work this paper cites.
Large scale distributed deep networks
J. Dean, G.S Corrado, R. Monga, K. Chen, M. Devin, Q.V. Le, M.Z. Mao, M.A. Ranzato, A. Senior, P. Tucker, K. Yang, and A. Y. Ng · 2012
Cited alongside, same era.
Theano: new features and speed improvements
Frédéric Bastien, Pascal Lamblin, Razvan Pascanu, James Bergstra, Ian J. Goodfellow, Arnaud Bergeron, Nicolas Bouchard, and Yoshua Bengio · 2012
Cited alongside, same era.
Deep convolutional neural networks for LVCSR
Tara Sainath, Abdel rahman Mohamed, Brian Kingsbury, and Bhuvana Ramabhadran · 2013
Cited alongside, same era.
Improving neural networks with dropout
Nitish Srivastava · 2013
Cited alongside, same era.
Regularization of neural networks using dropconnect
Li Wan, Matthew Zeiler, Sixin Zhang, Yann LeCun, and Rob Fergus · 2013
Cited alongside, same era.
Maxout networks
Ian J. Goodfellow, David Warde-Farley, Mehdi Mirza, Aaron Courville, and Yoshua Bengio · 2013
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba · 2014
Later among the works it cites.
Chen-Yu Lee, Saining Xie, Patrick Gallagher, Zhengyou Zhang, and Zhuowen Tu · 2014
Later among the works it cites.
Spatially-sparse convolutional neural networks
Benjamin Graham · 2014
Later among the works it cites.
Expectation backpropagation: Parameter-free training of multilayer neural networks with continuous or discrete weights
Daniel Soudry, Itay Hubara, and Ron Meir · 2014
Later among the works it cites.
Fixed-point feedforward deep neural network design using weights+ 1, 0, and- 1
Kyuyeon Hwang and Wonyong Sung · 2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Deep learning using linear support vector machines
Yichuan Tang · 2013
Cited alongside, same era.
Min Lin, Qiang Chen, and Shuicheng Yan · 2013
Cited alongside, same era.
Pylearn2: a machine learning research library
Ian J. Goodfellow, David Warde-Farley, Pascal Lamblin, Vincent Dumoulin, Mehdi Mirza, Razvan Pascanu, James Bergstra, Frédéric Bastien, and Yoshua Bengio · 2013
Cited alongside, same era.
Going deeper with convolutions
Christian Szegedy, Wei Liu, Yangqing Jia, Pierre Sermanet, Scott Reed, Dragomir Anguelov, Dumitru Erhan, Vincent Vanhoucke, and Andrew Rabinovich · 2014
Cited alongside, same era.
Fast and robust neural network joint models for statistical machine translation
Jacob Devlin, Rabih Zbib, Zhongqiang Huang, Thomas Lamar, Richard Schwartz, and John Makhoul · 2014
Cited alongside, same era.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V. Le · 2014
Cited alongside, same era.
X1000 real-time phoneme recognition vlsi using feed-forward deep neural networks
Jonghong Kim, Kyuyeon Hwang, and Wonyong Sung · 2014
Later among the works it cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2015
Closest in time.
Rounding methods for neural networks with low resolution synaptic weights
Lorenz K Muller and Giacomo Indiveri · 2015
Closest in time.
Deep learning with limited numerical precision
Suyog Gupta, Ankur Agrawal, Kailash Gopalakrishnan, and Pritish Narayanan · 2015
Closest in time.
Low precision arithmetic for deep learning
Matthieu Courbariaux, Yoshua Bengio, and Jean-Pierre David · 2015
Closest in time.
Hippocampal spine head sizes are highly precise
Thomas M Bartol, Cailey Bromer, Justin P Kinney, Michael A Chirillo, Jennifer N Bourne, Kristen M Harris, and Terrence J Sejnowski · 2015
Closest in time.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy · 2015
Closest in time.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman · 2015
Closest in time.
Training binary multilayer neural networks for image classification using expectation backpropgation
Zhiyong Cheng, Daniel Soudry, Zexi Mao, and Zhenzhong Lan · 2015
Closest in time.
Lasagne: First release., August 2015
Sander Dieleman, Jan Schlüter, Colin Raffel, Eben Olson, Søren Kaae Sønderby, Daniel Nouri, Daniel Maturana, Martin Thoma, Eric Battenberg, Jack Kelly, Jeffrey De Fauw, Michael Heilman, diogo149, Brian McFee, Hendrik Weideman, takacsg84, peterderivaz, Jon, instagibbs, Dr. Kashif Rasul, CongLiu, Britefury, and Jonas Degrave · 2015
Closest in time.