Fetching the paper…
Reading the bibliography…
A new network with super approximation power is introduced.
Error bounds for approximations with deep ReLU neural networks in W s , p W^{s,p} norms
Gühring, I., Kutyniok, G., and Petersen, P. (2019) · 1902
Earlier work this paper cites.
Understanding straight-through estimator in training activation quantized neural nets
Yin, P., Lyu, J., Zhang, S., Osher, S., Qi, Y., and Xin, J. (2019) · 1903
Earlier work this paper cites.
Generalization bounds of stochastic gradient descent for wide and deep neural networks
Cao, Y. and Gu, Q. (2019) · 1905
Earlier work this paper cites.
Approximation spaces of deep neural networks
Gribonval, R., Kutyniok, G., Nielsen, M., and Voigtlaender, F. (2019) · 1905
Earlier work this paper cites.
Deep network approximation characterized by number of neurons
Shen, Z., Yang, H., and Zhang, S. (2019b) · 1906
Earlier work this paper cites.
The phase diagram of approximation rates for deep neural networks
Yarotsky, D. and Zhevnerchuk, A. (2019) · 1906
Earlier work this paper cites.
Adaptive approximation and estimation of deep neural network with intrinsic dimensionality
Nakada, R. and Imaizumi, M. (2019) · 1907
Earlier work this paper cites.
A consensus-based global optimization method for high dimensional machine learning problems
Carrillo, J. A. T., Jin, S., Li, L., and Zhu, Y. (2019) · 1909
Earlier work this paper cites.
Ji, Z. and Telgarsky, M. (2020) · 1909
Earlier work this paper cites.
How much over-parameterization is sufficient to learn deep ReLU networks?
Chen, Z., Cao, Y., Zou, D., and Gu, Q. (2019b) · 1911
Earlier work this paper cites.
Deep learning via dynamical systems: An approximation perspective
Li, Q., Lin, T., and Shen, Z. (2019) · 1912
Earlier work this paper cites.
Particle swarm optimization
Kennedy, J. and Eberhart, R. (1995) · 1948
Earlier work this paper cites.
On the representation of continuous functions of several variables by superposition of continuous functions of a smaller number of variables
Kolmogorov, A. N. (1956) · 1956
Earlier work this paper cites.
On functions of three variables
Arnold, V. I. (1957) · 1957
Earlier work this paper cites.
On the representation of continuous functions of several variables by superposition of continuous functions of one variable and addition
Kolmogorov, A. N. (1957) · 1957
Earlier work this paper cites.
A simplex method for function minimization
Nelder, J. and Mead, R. (1965) · 1965
Earlier work this paper cites.
Optimization by simulated annealing
Kirkpatrick, S., Gelatt, C. D., and Vecchi, M. P. (1983) · 1983
Earlier work this paper cites.
Approximation by superpositions of a sigmoidal function
Cybenko, G. (1989) · 1989
Earlier work this paper cites.
Optimal nonlinear approximation
Devore, R. A. (1989) · 1989
Earlier work this paper cites.
Multilayer feedforward networks are universal approximators
Hornik, K., Stinchcombe, M., and White, H. (1989) · 1989
Earlier work this paper cites.
Genetic algorithms
Holland, J. H. (1992) · 1992
Cited alongside, same era.
Kolmogorov’s theorem and multilayer neural networks
Kůrková, V. (1992) · 1992
Cited alongside, same era.
Universal approximation bounds for superpositions of a sigmoidal function
Barron, A. R. (1993) · 1993
Cited alongside, same era.
Almost linear VC-dimension bounds for piecewise polynomial networks
Bartlett, P., Maiorov, V., and Meir, R. (1998) · 1998
Cited alongside, same era.
Lower bounds for approximation by MLP neural networks
Maiorov, V. and Pinkus, A. (1999) · 1999
Cited alongside, same era.
Deep network approximation for smooth functions
Lu, J., Shen, Z., Yang, H., and Zhang, S. (2020) · 2001
Cited alongside, same era.
Exponential convergence of the deep neural network approximation for analytic functions
E, W. and Wang, Q. (2018) · 2018
Later among the works it cites.
Approximation capability of two hidden layer feedforward neural networks with fixed weights
Guliyev, N. J. and Ismailov, V. E. (2018) · 2018
Later among the works it cites.
Neural tangent kernel: Convergence and generalization in neural networks
Jacot, A., Gabriel, F., and Hongler, C. (2018) · 2018
Later among the works it cites.
Optimal approximation of piecewise smooth functions using deep ReLU neural networks
Petersen, P. and Voigtlaender, F. (2018) · 2018
Later among the works it cites.
Two-step quantization for low-bit neural networks
Wang, P., Hu, Q., Zhang, Y., Zhang, C., Liu, Y., and Cheng, J. (2018) · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kolmogorov’s spline network
Igelnik, B. and Parikh, N. (2003) · 2003
Cited alongside, same era.
Approximation in shift-invariant spaces with deep ReLU neural networks
Yang, Y. and Wang, Y. (2020) · 2005
Cited alongside, same era.
Quantized neural networks: Characterization and holistic optimization
Boo, Y., Shin, S., and Sung, W. (2020) · 2006
Cited alongside, same era.
Two-layer neural networks for partial differential equations: Optimization and generalization theory
Luo, T. and Yang, H. (2020) · 2006
Cited alongside, same era.
On a constructive proof of kolmogorov’s superposition theorem
Braun, J. and Griebel, M. (2009) · 2009
Cited alongside, same era.
Estimating or propagating gradients through stochastic neurons for conditional computation
Bengio, Y., Léonard, N., and Courville, A. (2013) · 2013
Cited alongside, same era.
Optimal approximation of continuous functions by very deep ReLU networks
Yarotsky, D. (2018) · 2018
Later among the works it cites.
Learning and generalization in overparameterized neural networks, going beyond two layers
Allen-Zhu, Z., Li, Y., and Liang, Y. (2019) · 2019
Later among the works it cites.
Fine-grained analysis of optimization and generalization for overparameterized two-layer neural networks
Arora, S., Du, S. S., Hu, W., Li, Z., and Wang, R. (2019) · 2019
Later among the works it cites.
Approximation analysis of convolutional neural networks
Bao, C., Li, Q., Shen, Z., Tai, C., Wu, L., and Xiang, X. (2019) · 2019
Later among the works it cites.
Optimal approximation with sparsely connected deep neural networks
Bölcskei, H., Grohs, P., Kutyniok, G., and Petersen, P. (2019) · 2019
Later among the works it cites.
A note on the expressive power of deep rectified linear unit networks in high-dimensional spaces
Chen, L. and Wu, C. (2019) · 2019
Later among the works it cites.
A priori estimates of the population risk for two-layer neural networks
E, W., Ma, C., and Wu, L. (2019) · 2019
Later among the works it cites.
Optimization strategies in quantized neural networks: A review
Lin, Y., Lei, M., and Niu, L. (2019) · 2019
Later among the works it cites.
New error bounds for deep ReLU networks using sparse grids
Montanelli, H. and Du, Q. (2019) · 2019
Later among the works it cites.
Exponential ReLU DNN expression of holomorphic maps in high dimension
Opschoor, J. A. A., Schwab, C., and Zech, J. (2019) · 2019
Later among the works it cites.
Adaptivity of deep ReLU network for learning in Besov and mixed smooth Besov spaces: optimal rate and curse of dimensionality
Suzuki, T. (2019) · 2019
Later among the works it cites.
Representation formulas and pointwise properties for barron functions
E, W. and Wojtowytsch, S. (2020) · 2020
Closest in time.
Error bounds for deep ReLU networks using the Kolmogorov-Arnold superposition theorem
Montanelli, H. and Yang, H. (2020) · 2020
Closest in time.
Deep ReLU networks overcome the curse of dimensionality for bandlimited functions
Montanelli, H., Yang, H., and Du, Q. (2020) · 2020
Closest in time.
Universality of deep convolutional neural networks
Zhou, D.-X. (2020) · 2020
Closest in time.