Fetching the paper…
Reading the bibliography…
The driving force behind deep networks is their ability to compactly represent rich classes of functions.
Introduction to matrix analysis , volume 960
Richard Bellman · 1970
Earlier work this paper cites.
Speaker-independent phone recognition using hidden markov models
K-F Lee and H-W Hon · 1989
Earlier work this paper cites.
Darpa timit acoustic-phonetic continous speech corpus cd-rom. nist speech disc 1-1.1
John S Garofolo, Lori F Lamel, William M Fisher, Jonathon G Fiscus, and David S Pallett · 1993
Earlier work this paper cites.
Convolutional networks for images, speech, and time series
Yann LeCun and Yoshua Bengio · 1995
Earlier work this paper cites.
Heterogeneous acoustic measurements and multiple classifiers for speech recognition
Andrew K Halberstadt · 1998
Earlier work this paper cites.
Lebesgue integration on Euclidean space
Frank Jones · 2001
Earlier work this paper cites.
The zero set of a polynomial
Richard Caron and Tim Traynor · 2005
Earlier work this paper cites.
Rectified linear units improve restricted boltzmann machines
Vinod Nair and Geoffrey E Hinton · 2010
Earlier work this paper cites.
Shallow vs. deep sum-product networks
Olivier Delalleau and Yoshua Bengio · 2011
Earlier work this paper cites.
Tensor Spaces and Numerical Tensor Calculus , volume 42 of
Wolfgang Hackbusch · 2012
Earlier work this paper cites.
ImageNet Classification with Deep Convolutional Neural Networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton · 2012
Earlier work this paper cites.
On the number of inference regions of deep feed forward networks with piece-wise linear activations
Razvan Pascanu, Guido Montufar, and Yoshua Bengio · 2013
Cited alongside, same era.
Caffe: Convolutional architecture for fast feature embedding
Yangqing Jia, Evan Shelhamer, Jeff Donahue, Sergey Karayev, Jonathan Long, Ross Girshick, Sergio Guadarrama, and Trevor Darrell · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba · 2014
Cited alongside, same era.
On the number of linear regions of deep neural networks
Guido F Montufar, Razvan Pascanu, Kyunghyun Cho, and Yoshua Bengio · 2014
Cited alongside, same era.
The power of depth for feedforward neural networks
Ronen Eldan and Ohad Shamir · 2015
Cited alongside, same era.
Liang-Chieh Chen, George Papandreou, Iasonas Kokkinos, Kevin Murphy, and Alan L Yuille · 2016
Later among the works it cites.
Convolutional rectifier networks as generalized tensor decompositions
Nadav Cohen and Amnon Shashua · 2016
Later among the works it cites.
Neural machine translation in linear time
Nal Kalchbrenner, Lasse Espeholt, Karen Simonyan, Aaron van den Oord, Alex Graves, and Koray Kavukcuoglu · 2016
Later among the works it cites.
Learning real and boolean functions: When is deep better than shallow
Hrushikesh Mhaskar, Qianli Liao, and Tomaso Poggio · 2016
Later among the works it cites.
Exponential expressivity in deep neural networks through transient chaos
Ben Poole, Subhaneil Lahiri, Maithreyi Raghu, Jascha Sohl-Dickstein, and Surya Ganguli · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2015
Cited alongside, same era.
Beating the Perils of Non-Convexity: Guaranteed Training of Neural Networks using Tensor Methods
Majid Janzamin, Hanie Sedghi, and Anima Anandkumar · 2015
Cited alongside, same era.
Deep learning
Yann LeCun, Yoshua Bengio, and Geoffrey Hinton · 2015
Cited alongside, same era.
I-theory on depth vs width: hierarchical function composition
Tomaso Poggio, Fabio Anselmi, and Lorenzo Rosasco · 2015
Cited alongside, same era.
Going Deeper with Convolutions
Christian Szegedy, Wei Liu, Yangqing Jia, Pierre Sermanet, Scott Reed, Dragomir Anguelov, Dumitru Erhan, Vincent Vanhoucke, and Andrew Rabinovich · 2015
Cited alongside, same era.
Representation benefits of deep feedforward networks
Matus Telgarsky · 2015
Cited alongside, same era.
Deep simnets
Nadav Cohen, Or Sharir, and Amnon Shashua
Cited in the paper.
Later among the works it cites.
On the expressive power of deep neural networks
Maithra Raghu, Ben Poole, Jon Kleinberg, Surya Ganguli, and Jascha Sohl-Dickstein · 2016
Later among the works it cites.
Training input-output recurrent neural networks through spectral methods
Hanie Sedghi and Anima Anandkumar · 2016
Later among the works it cites.
Or Sharir, Ronen Tamari, Nadav Cohen, and Amnon Shashua · 2016
Later among the works it cites.
Wavenet: A generative model for raw audio
Aäron van den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew Senior, and Koray Kavukcuoglu · 2016
Later among the works it cites.
Inductive bias of deep convolutional networks through pooling geometry
Nadav Cohen and Amnon Shashua · 2017
Closest in time.