Fetching the paper…
Reading the bibliography…
There is a longstanding debate whether the Kolmogorov-Arnold representation theorem can explain the use of more than one hidden layer in neural networks.
Deep Neural Network Approximation Theory
Elbrächter, D., Perekrestenko, D., Grohs, P., and Bölcskei, H · 1901
Earlier work this paper cites.
The phase diagram of approximation rates for deep neural networks
Yarotsky, D., and Zhevnerchuk, A · 1906
Earlier work this paper cites.
Estimation of a function of low local dimensionality by deep neural networks
Kohler, M., Krzyzak, A., and Langer, S · 1908
Earlier work this paper cites.
Deep ReLU network approximation of functions on a manifold
Schmidt-Hieber, J · 1908
Earlier work this paper cites.
On the representation of continuous functions of many variables by superposition of continuous functions of one variable and addition
Kolmogorov, A. N · 1957
Earlier work this paper cites.
On the structure of continuous functions of several variables
Sprecher, D. A · 1965
Earlier work this paper cites.
Kolmogorov’s mapping neural network existence theorem
Hecht-Nielsen, R · 1987
Earlier work this paper cites.
Representation properties of networks: Kolmogorov’s theorem is irrelevant
Girosi, F., and Poggio, T · 1989
Earlier work this paper cites.
Kolmogorov’s theorem is relevant
Kurkova, V · 1991
Earlier work this paper cites.
Kolmogorov’s theorem and multilayer neural networks
Kurkova, V · 1992
Earlier work this paper cites.
Analog computation via neural networks
Siegelmann, H. T., and Sontag, E. D · 1994
Earlier work this paper cites.
A numerical implementation of Kolmogorov’s superpositions
Sprecher, D. A · 1996
Earlier work this paper cites.
A numerical implementation of Kolmogorov’s superpositions ii
Sprecher, D. A · 1997
Earlier work this paper cites.
On the approximation of functional classes equipped with a uniform measure using ridge functions
Maiorov, V., Meir, R., and Ratsaby, J · 1999
Earlier work this paper cites.
Lower bounds for approximation by MLP neural networks
Maiorov, V., and Pinkus, A · 1999
Cited alongside, same era.
On best approximation by ridge functions
Maiorov, V. E · 1999
Cited alongside, same era.
Approximation theory of the MLP model in neural networks
Pinkus, A · 1999
Cited alongside, same era.
Deep Network Approximation for Smooth Functions
Lu, J., Shen, Z., Yang, H., and Zhang, S · 2001
Cited alongside, same era.
On the best approximation by ridge functions in the uniform norm
Gordon, Y., Maiorov, V., Meyer, M., and Reisner, S · 2002
Cited alongside, same era.
Deep Network with Approximation Error Being Reciprocal of Width to Power of Square Root of Depth
Nonparametric regression based on hierarchical interaction models
Kohler, M., and Krzyżak, A · 2017
Later among the works it cites.
Why and when can deep-but not shallow-networks avoid the curse of dimensionality: A review
Poggio, T., Mhaskar, H., Rosasco, L., Miranda, B., and Liao, Q · 2017
Later among the works it cites.
Approximation capability of two hidden layer feedforward neural networks with fixed weights
Guliyev, N. J., and Ismailov, V. E · 2018
Later among the works it cites.
Optimal approximation of continuous functions by very deep ReLU networks
Yarotsky, D · 2018
Later among the works it cites.
Nearly-tight VC-dimension and pseudodimension bounds for piecewise linear neural networks
Bartlett, P., Harvey, N., Liaw, C., and Mehrabian, A · 2019
Later among the works it cites.
On deep learning as a remedy for the curse of dimensionality in nonparametric regression
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Shen, Z., Yang, H., and Zhang, S · 2006
Cited alongside, same era.
Rate-optimal estimation for a general class of nonparametric regression models with unknown link functions
Horowitz, J. L., and Mammen, E · 2007
Cited alongside, same era.
An Application of Kolmogorov’s Superposition Theorem to Function Reconstruction in Higher Dimensions
Braun, J · 2009
Cited alongside, same era.
On a constructive proof of Kolmogorov’s superposition theorem
Braun, J., and Griebel, M · 2009
Cited alongside, same era.
The Elements of Statistical Learning
Hastie, T., Tibshirani, R., and Friedman, J · 2009
Cited alongside, same era.
Space-filling curves
Bader, M · 2013
Cited alongside, same era.
Visualizing and understanding convolutional networks
Zeiler, M. D., and Fergus, R · 2014
Cited alongside, same era.
Bauer, B., and Kohler, M · 2019
Later among the works it cites.
A comparison of deep networks with ReLU activation function and linear spline-type methods
Eckle, K., and Schmidt-Hieber, J · 2019
Later among the works it cites.
Rethinking imagenet pre-training
He, K., Girshick, R., and Dollar, P · 2019
Later among the works it cites.
A theoretical analysis of deep q-learning
Fan, J., Wang, Z., Xie, Y., and Yang, Z · 2020
Closest in time.
Error bounds for deep ReLU networks using the Kolmogorov–Arnold superposition theorem
Montanelli, H., and Yang, H · 2020
Closest in time.
Adaptive approximation and generalization of deep neural network with intrinsic dimensionality
Nakada, R., and Imaizumi, M · 2020
Closest in time.
Nonparametric regression using deep neural networks with ReLU activation function
Schmidt-Hieber, J · 2020
Closest in time.
Deep network approximation characterized by number of neurons
Shen, Z · 2020
Closest in time.