Fetching the paper…
Reading the bibliography…
We discuss the expressive power of neural networks which use the non-smooth ReLU activation function $\varrho(x) = \max\{0,x\}$ by analyzing the approximation theoretic properties of such networks.
Entropies of several sets of real valued functions
G.F. Clements · 1963
Earlier work this paper cites.
Deep learning
Y. LeCun, Y. Bengio, and G. Hinton · 2015
Earlier work this paper cites.
Error bounds for approximations with deep ReLU networks
D. Yarotsky · 2017
Cited alongside, same era.
Optimal approximation of piecewise smooth functions using deep ReLU neural networks
P. Petersen and F. Voigtlaender · 2018
Cited alongside, same era.
Nearly-tight VC-dimension and pseudodimension bounds for piecewise linear neural networks
P. L. Bartlett, N. Harvey, C. Liaw, A. Mehrabian
Cited in the paper.
Equivalence of approximation by convolutional neural networks and fully-connected networks
P. Petersen and F. Voigtlaender
Cited in the paper.
Depth-width tradeoffs in approximating natural functions with neural networks
I. Safran and O. Shamir
Cited in the paper.
Optimal approximation of continuous functions by very deep ReLU networks
D. Yarotsky
Cited in the paper.
Optimal Approximation with Sparsely Connected Deep Neural Networks
H. Boelcskei, P. Grohs, G. Kutyniok, and P. Petersen · 2019
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…