Fetching the paper…
Reading the bibliography…
The Structure of Non-Enumerable Sets of Points
A. J. Ward · 1933
Earlier work this paper cites.
A logical calculus of ideas immanent in nervous activity
W. McCulloch and W. Pitts · 1943
Earlier work this paper cites.
Foundations of modern analysis
J. Dieudonné · 1960
Earlier work this paper cites.
Translation invariant subspaces of finite dimension
P. M. Anselone and J. Korevaar · 1964
Earlier work this paper cites.
Learning in networks is hard
J. Judd · 1987
Earlier work this paper cites.
Real and complex analysis
W. Rudin · 1987
Earlier work this paper cites.
Neural networks and principal component analysis: Learning from examples without local minima
P. Baldi and K. Hornik · 1988
Earlier work this paper cites.
Training a 3-node neural network is NP-complete
A. Blum and R. Rivest · 1989
Earlier work this paper cites.
Approximation by superpositions of a sigmoidal function
G. Cybenko · 1989
Earlier work this paper cites.
Multilayer feedforward networks are universal approximators
K. Hornik, M. Stinchcombe, and H. White · 1989
Earlier work this paper cites.
Networks and the best approximation property
F. Girosi and T. Poggio · 1990
Earlier work this paper cites.
Functional analysis
W. Rudin · 1991
Earlier work this paper cites.
Universal approximation bounds for superpositions of a sigmoidal function
A. Barron · 1993
Earlier work this paper cites.
Multilayer feedforward networks with a nonpolynomial activation function can approximate any function
M. Leshno, V. Y. Lin, A. Pinkus, and S. Schocken · 1993
Earlier work this paper cites.
Approximation properties of a multilayered feedforward artificial neural network
H. N. Mhaskar · 1993
Earlier work this paper cites.
Neural networks for optimal approximation of smooth and analytic functions
H. Mhaskar · 1996
Earlier work this paper cites.
Neural Networks: A Comprehensive Foundation
S. Haykin · 1998
Earlier work this paper cites.
Artificial neural networks for solving ordinary and partial differential equations
I. E. Lagaris, A. Likas, and D. I. Fotiadis · 1998
Earlier work this paper cites.
Neural network learning: theoretical foundations
M. Anthony and P. L. Bartlett · 1999
Earlier work this paper cites.
Real Analysis
G. Folland · 1999
Earlier work this paper cites.
Approximation by neural networks is not continuous
P. C. Kainen, V. Kurková, and A. Vogt · 1999
Earlier work this paper cites.
Best approximation by Heaviside perceptron networks
P. Kainen, V. Kurková, and A. Vogt · 2000
Earlier work this paper cites.
Hardness results for neural network approximation problems
P. L. Bartlett and S. Ben-David · 2002
Earlier work this paper cites.
Rademacher and Gaussian complexities: Risk bounds and structural results
P. L. Bartlett and S. Mendelson · 2002
Earlier work this paper cites.
On the mathematical foundations of learning
F. Cucker and S. Smale · 2002
Earlier work this paper cites.
Classical Fourier analysis
L. Grafakos · 2008
Cited alongside, same era.
Analysis III
H. Amann and J. Escher · 2009
Cited alongside, same era.
Quadratic polynomials learn better image features
J. Bergstra, G. Desjardins, P. Lamblin, and Y. Bengio · 2009
Cited alongside, same era.
Best approximation by ridge functions in L p L_{p} -spaces
V. E. Maiorov · 2010
Cited alongside, same era.
Rectified linear units improve restricted Boltzmann machines
V. Nair and G. Hinton · 2010
Cited alongside, same era.
Deep sparse rectifier neural networks
X. Glorot, A. Bordes, and Y. Bengio · 2011
Cited alongside, same era.
Introduction to topological manifolds
J. M. Lee · 2011
Deep learning-based numerical methods for high-dimensional parabolic partial differential equations and backward stochastic differential equations
W. E, J. Han, and A. Jentzen · 2017
Later among the works it cites.
Topology and geometry of half-rectified network optimization
C. D. Freeman and J. Bruna · 2017
Later among the works it cites.
Densely connected convolutional networks
G. Huang, Z. Liu, L. van der Maaten, and K. Q. Weinberger · 2017
Later among the works it cites.
An arctan-activated WASD neural network approach to the prediction of Dow Jones Industrial Average
B. Liao, C. Ma, L. Xiao, R. Lu, and L. Ding · 2017
Later among the works it cites.
The loss surface of deep and wide neural networks
Q. Nguyen and M. Hein · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Context-dependent pre-trained deep neural networks for large-vocabulary speech recognition
G. E. Dahl, D. Yu, L. Deng, and A. Acero · 2012
Cited alongside, same era.
Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups
G. Hinton, L. Deng, D. Yu, G. E. Dahl, A.-R. Mohamed, N. Jaitly, A. Senior, V. Vanhoucke, P. Nguyen, T. N. Sainath, et al · 2012
Cited alongside, same era.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. Hinton · 2012
Cited alongside, same era.
Foundations of machine learning
M. Mohri, A. Rostamizadeh, and A. Talwalkar · 2012
Cited alongside, same era.
Measure theory
D. L. Cohn · 2013
Cited alongside, same era.
Q. Nguyen and M. Hein · 2017
Later among the works it cites.
Depth-width tradeoffs in approximating natural functions with neural networks
I. Safran and O. Shamir · 2017
Later among the works it cites.
Mastering chess and shogi by self-play with a general reinforcement learning algorithm
D. Silver, T. Hubert, J. Schrittwieser, I. Antonoglou, M. Lai, A. Guez, M. Lanctot, L. Sifre, D. Kumaran, T. Graepel, et al · 2017
Later among the works it cites.
Episodic Exploration for Deep Deterministic Policies for StarCraft Micromanagement
N. Usunier, G. Synnaeve, Z. Lin, and S. Chintala · 2017
Later among the works it cites.
Artificial Intelligence and Games
G. N. Yannakakis and J. Togelius · 2017
Later among the works it cites.
Error bounds for approximations with deep ReLU networks
D. Yarotsky · 2017
Later among the works it cites.
Convexified convolutional neural networks
Y. Zhang, P. Liang, and M. J. Wainwright · 2017
Later among the works it cites.
On the global convergence of gradient descent for over-parameterized models using optimal transport
L. Chizat and F. Bach · 2018
Closest in time.
Neural tangent kernel: Convergence and generalization in neural networks
A. Jacot, F. Gabriel, and C. Hongler · 2018
Closest in time.
A mean field view of the landscape of two-layer neural networks
S. Mei, A. Montanari, and P.-M. Nguyen · 2018
Closest in time.
Optimal approximation of piecewise smooth functions using deep ReLU neural networks
P. Petersen and F. Voigtlaender · 2018
Closest in time.
G. M. Rotskoff and E. Vanden-Eijnden · 2018
Closest in time.
Neural networks with finite intrinsic dimension have no spurious valleys
L. Venturi, A. S. Bandeira, and J. Bruna · 2018
Closest in time.
The deep Ritz method: a deep learning-based numerical algorithm for solving variational problems
E. Weinan and B. Yu · 2018
Closest in time.
A convergence theory for deep learning via over-parameterization
Z. Allen-Zhu, Y. Li, and Z. Song · 2019
Closest in time.
Nearly-tight VC-dimension and pseudodimension bounds for piecewise linear neural networks
P. L. Bartlett, N. Harvey, C. Liaw, and A. Mehrabian · 2019
Closest in time.
Optimal approximation with sparsely connected deep neural networks
H. Bölcskei, P. Grohs, G. Kutyniok, and P. C. Petersen · 2019
Closest in time.
The phase diagram of approximation rates for deep neural networks
D. Yarotsky and A. Zhevnerchuk · 2019
Closest in time.
Uncountable closed set A A , existence of point at which A A accumulates ”from two sides” of a hyperplane
PhoemueX ( https://math.stackexchange.com/users/151552/phoemuex ) · 2020
Closest in time.