Fetching the paper…
Reading the bibliography…
A three-hidden-layer neural network with super approximation power is introduced.
Error bounds for approximations with deep ReLU neural networks in W s , p W^{s,p} norms
Gühring, I., Kutyniok, G., Petersen, P., 2019 · 1902
Earlier work this paper cites.
A consensus-based global optimization method for high dimensional machine learning problems
Carrillo, J.A.T., Jin, S., Li, L., Zhu, Y., 2019 · 1909
Earlier work this paper cites.
How much over-parameterization is sufficient to learn deep ReLU networks?
Chen, Z., Cao, Y., Zou, D., Gu, Q., 2019b · 1911
Earlier work this paper cites.
On the representation of continuous functions of several variables by superposition of continuous functions of a smaller number of variables
Kolmogorov, A.N., 1956 · 1956
Earlier work this paper cites.
On functions of three variables
Arnold, V.I., 1957 · 1957
Earlier work this paper cites.
On the representation of continuous functions of several variables by superposition of continuous functions of one variable and addition
Kolmogorov, A.N., 1957 · 1957
Earlier work this paper cites.
A simplex method for function minimization
Nelder, J., Mead, R., 1965 · 1965
Earlier work this paper cites.
Optimization by simulated annealing
Kirkpatrick, S., Gelatt, C.D., Vecchi, M.P., 1983 · 1983
Earlier work this paper cites.
Kolmogorov’s theorem is relevant
Kůrková, V., 1991 · 1991
Earlier work this paper cites.
Genetic algorithms
Holland, J.H., 1992 · 1992
Earlier work this paper cites.
Kolmogorov’s theorem and multilayer neural networks
Kůrková, V., 1992 · 1992
Earlier work this paper cites.
Universal approximation bounds for superpositions of a sigmoidal function
Barron, A.R., 1993 · 1993
Earlier work this paper cites.
Particle swarm optimization, in: Proceedings of ICNN’95 - International Conference on Neural Networks, pp. 1942–1948 vol.4
Kennedy, J., Eberhart, R., 1995 · 1995
Earlier work this paper cites.
Almost linear VC-dimension bounds for piecewise polynomial networks
Bartlett, P., Maiorov, V., Meir, R., 1998 · 1998
Earlier work this paper cites.
Lower bounds for approximation by MLP neural networks
Maiorov, V., Pinkus, A., 1999 · 1999
Earlier work this paper cites.
Deep network approximation for smooth functions
Lu, J., Shen, Z., Yang, H., Zhang, S., 2020 · 2001
Earlier work this paper cites.
Kolmogorov’s spline network
Igelnik, B., Parikh, N., 2003 · 2003
Earlier work this paper cites.
Lu, Y., Ma, C., Lu, Y., Lu, J., Ying, L., 2020 · 2003
Earlier work this paper cites.
Approximation in shift-invariant spaces with deep ReLU neural networks
Yang, Y., Wang, Y., 2020 · 2005
Cited alongside, same era.
Quantized neural networks: Characterization and holistic optimization
Boo, Y., Shin, S., Sung, W., 2020 · 2006
Cited alongside, same era.
Two-Layer Neural Networks for Partial Differential Equations: Optimization and Generalization Theory
Luo, T., Yang, H., 2020 · 2006
Cited alongside, same era.
On a constructive proof of Kolmogorov’s superposition theorem
Braun, J., Griebel, M., 2009 · 2009
Cited alongside, same era.
Deep neural network approximation via function compositions
Zhang, S., 2020 · 2010
How sgd selects the global minima in over-parameterized learning: A dynamical stability perspective, in: Bengio, S., Wallach, H., Larochelle, H., Grauman, K., Cesa-Bianchi, N., Garnett, R. (Eds.), Advances in Neural Information Processing Systems 31. Curran Associates, Inc., pp. 8279–8288
Wu, L., Ma, C., E, W., 2018 · 2018
Later among the works it cites.
Optimal approximation of continuous functions by very deep ReLU networks, in: Bubeck, S., Perchet, V., Rigollet, P. (Eds.), Proceedings of the 31st Conference On Learning Theory, PMLR. pp. 639–649
Yarotsky, D., 2018 · 2018
Later among the works it cites.
A note on the expressive power of deep rectified linear unit networks in high-dimensional spaces
Chen, L., Wu, C., 2019 · 2019
Later among the works it cites.
Gradient descent provably optimizes over-parameterized neural networks, in: International Conference on Learning Representations
Du, S.S., Zhai, X., Poczos, B., Singh, A., 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Estimating or propagating gradients through stochastic neurons for conditional computation
Bengio, Y., Léonard, N., Courville, A., 2013 · 2013
Cited alongside, same era.
Nearly-tight VC-dimension bounds for piecewise linear neural networks, in: Kale, S., Shamir, O. (Eds.), Proceedings of the 2017 Conference on Learning Theory, PMLR, Amsterdam, Netherlands. pp. 1064–1068
Harvey, N., Liaw, C., Mehrabian, A., 2017 · 2017
Cited alongside, same era.
Quantized neural networks: Training neural networks with low precision weights and activations
Hubara, I., Courbariaux, M., Soudry, D., El-Yaniv, R., Bengio, Y., 2017 · 2017
Cited alongside, same era.
A consensus-based model for global optimization and its mean-field limit
Pinnau, R., Totzeck, C., Tse, O., Martin, S., 2017 · 2017
Cited alongside, same era.
Why and when can deep—but not shallow—networks avoid the curse of dimensionality: A review
Poggio, T., Mhaskar, H.N., Rosasco, L., Miranda, B., Liao, Q., 2017 · 2017
Cited alongside, same era.
Error bounds for approximations with deep ReLU networks
Yarotsky, D., 2017 · 2017
Cited alongside, same era.
Approximation and estimation for high-dimensional deep learning networks
Barron, A.R., Klusowski, J.M., 2018 · 2018
Cited alongside, same era.
A priori estimates of the population risk for two-layer neural networks
E, W., Ma, C., Wu, L., 2019 · 2019
Later among the works it cites.
Optimization strategies in quantized neural networks: A review, in: 2019 International Conference on Data Mining Workshops (ICDMW), pp. 385–390
Lin, Y., Lei, M., Niu, L., 2019 · 2019
Later among the works it cites.
New error bounds for deep ReLU networks using sparse grids
Montanelli, H., Du, Q., 2019 · 2019
Later among the works it cites.
Error bounds for deep ReLU networks using the Kolmogorov-Arnold superposition theorem
Montanelli, H., Yang, H., 2020 · 2019
Later among the works it cites.
Exponential ReLU DNN expression of holomorphic maps in high dimension
Opschoor, J.A., Schwab, C., Zech, J., 2019 · 2019
Later among the works it cites.
Nonlinear approximation via compositions
Shen, Z., Yang, H., Zhang, S., 2019 · 2019
Later among the works it cites.
Understanding straight-through estimator in training activation quantized neural nets URL: https://openreview.net/forum?id=Skh4jRcKQ
Yin, P., Lyu, J., Zhang, S., Osher, S.J., Qi, Y., Xin, J., 2019 · 2019
Later among the works it cites.
Overcoming the curse of dimensionality in the approximative pricing of financial derivatives with default risks
Hutzenthaler, M., Jentzen, A., Wurstemberger, v.W., 2020 · 2020
Closest in time.
Deep ReLU networks overcome the curse of dimensionality for bandlimited functions
Montanelli, H., Yang, H., Du, Q., 2020 · 2020
Closest in time.
Nonparametric regression using deep neural networks with ReLU activation function
Schmidt-Hieber, J., 2020 · 2020
Closest in time.
Deep network approximation characterized by number of neurons
Shen, Z., Yang, H., Zhang, S., 2020 · 2020
Closest in time.
The phase diagram of approximation rates for deep neural networks 33, 13005–13015
Yarotsky, D., Zhevnerchuk, A., 2020 · 2020
Closest in time.
The Kolmogorov–Arnold representation theorem revisited
Schmidt-Hieber, J., 2021 · 2021
Closest in time.
Deep network with approximation error being reciprocal of width to power of square root of depth
Shen, Z., Yang, H., Zhang, S., 2021 · 2021
Closest in time.