Fetching the paper…
Reading the bibliography…
Deep learning has been widely applied and brought breakthroughs in speech recognition, computer vision, and many other domains.
Cybenko G (1989) Approximations by superpositions of sigmoidal functions. Mathematics of Control, Signals, and Systems
1989
Earlier work this paper cites.
Hornik K, Stinchcombe M, White H (1989) Multilayer feedforward networks are universal approximators. Neural networks
1989
Earlier work this paper cites.
Daubechies I (1992) Ten Lectures on Wavelets, SIAM
1992
Earlier work this paper cites.
Barron A (1993) Universal approximation bounds for superpositions of a sigmoidal function. IEEE Transactions on Information Theory
1993
Earlier work this paper cites.
Leshno M, Lin Y, Pinkus A, Schocken S (1993) Multilayer feedforward networks with a non-polynomial activation function can approximate any function. Neural Networks
1993
Earlier work this paper cites.
Mhaskar H (1993) Approximation properties of a multilayered feedforward artificial neural network. Advances in Computational Mathematics
1993
Earlier work this paper cites.
Chui C, Li X, Mhaskar H (1996) Limitations of the approximation capabilities of neural networks with one hidden layer. Advances in Computational Mathematics
1996
Earlier work this paper cites.
LeCun Y, Bottou L, Bengio Y, Haffner P (1998) Gradient-based learning applied to document recognition. Proceedings of the IEEE
1998
Earlier work this paper cites.
Pinkus A (1999) Approximation theory of the MLP model in neural networks. Acta Numerica
1999
Earlier work this paper cites.
Smale S, Zhou D (2004) Shannon sampling and function reconstruction from point values, Bulletins of the American Mathematical Society
2004
Earlier work this paper cites.
Hinton G, Osindero S, Teh Y (2006) A fast learning algorithm for deep belief nets. Neural Computation
2006
Cited alongside, same era.
Bruna J, Mallat S (2013) Invariant scattering convolution networks. IEEE Transactions on Pattern Analysis and Machine Intelligence
2013
Cited alongside, same era.
LeCun Y, Bengio Y, Hinton G (2015) Deep learning. Nature
2015
Cited alongside, same era.
Goodfellow I, Bengio Y, Courville A (2016) Deep Learning. MIT Press
2016
Cited alongside, same era.
Mallat S (2016) Understanding deep convolutional networks. Philosophical Transactions of the Royal Society A
2016
Cited alongside, same era.
Mhaskar H, Poggio T (2016) Deep vs. shallow networks: An approximation theory perspective. Analysis and Applications
Guo Z, Lin S, Zhou D (2017) Learning theory of distributed spectral algorithms. Inverse Problems
2017
Later among the works it cites.
Yarotsky D (2017) Error bounds for approximations with deep ReLU networks. Neural Networks
2017
Later among the works it cites.
Zhou D (2018) Deep distributed convolutional neural networks: universality. Analysis and Applications
2018
Closest in time.
Strang G (2018) Linear Algebra and Learning from Data. Book manuscipt
2018
Closest in time.
Shaham U, Cloninger A, Coifman R (2018) Provable approximation properties for deep neural networks. Applied and Computational Harmonic Analysis
2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2016
Cited alongside, same era.
Fan J, Hu T, Wu Q, Zhou D (2016) Consistency analysis of an empirical minimum error entropy algorithm. Applied and Computational Harmonic Analysis
2016
Cited alongside, same era.
Telgarsky M (2016) Benefits of depth in neural networks. 29th Annual Conference on Learning Theory
2016
Cited alongside, same era.
Eldan R, Shamir O (2016) The power of depth for feedforward neural networks. COLT: 907-940
2016
Cited alongside, same era.
Lin S, Guo X, Zhou D (2017) Distributed learning with regularized least squares. Journal of Machine Learning Research
2017
Cited alongside, same era.
2018
Closest in time.
2018
Closest in time.
2018
Closest in time.
Krizhevsky A, Sutskever I, Hinton G (2012) Imagenet classification with deep convolutional neural networks NIPS
2097
Closest in time.