Fetching the paper…
Reading the bibliography…
Deep learning has been applied to various tasks in the field of machine learning and has shown superiority to other common procedures such as kernel methods.
New thoughts on Besov spaces
Peetre, J., 1976 · 1976
Earlier work this paper cites.
Approximation by superpositions of a sigmoidal function
Cybenko, G., 1989 · 1989
Earlier work this paper cites.
Minimax risk over hyperrectangles, and implications
Donoho, D.L., Liu, R.C., MacGibbon, B., 1990 · 1990
Earlier work this paper cites.
Ten lectures on wavelets
Daubechies, I., 1992 · 1992
Earlier work this paper cites.
Unconditional bases are optimal bases for data compression and for statistical estimation
Donoho, D.L., 1993 · 1993
Earlier work this paper cites.
Minimax theory of image reconstruction
Korostelev, A.P., Tsybakov, A.B., 1993 · 1993
Earlier work this paper cites.
Ideal spatial adaptation by wavelet shrinkage
Donoho, D.L., Johnstone, J.M., 1994 · 1994
Earlier work this paper cites.
Unconditional bases and bit-level compression
Donoho, D.L., 1996 · 1996
Earlier work this paper cites.
Weak convergence and empirical processes
van der Vaart, A.W., Wellner, J.A., 1996 · 1996
Earlier work this paper cites.
Minimax estimation via wavelet shrinkage
Donoho, D.L., Johnstone, I.M., 1998 · 1998
Earlier work this paper cites.
Information-theoretic determination of minimax rates of convergence
Yang, Y., Barron, A., 1999 · 1999
Earlier work this paper cites.
The elements of statistical learning
Friedman, J., Hastie, T., Tibshirani, R., 2001 · 2001
Earlier work this paper cites.
Wavelet threshold estimation of a regression function with random design
Zhang, S., Wong, M.Y., Zheng, Z., 2002 · 2002
Cited alongside, same era.
Minimax estimation of linear functionals over nonconvex parameter spaces
Cai, T.T., Low, M.G., 2004 · 2004
Cited alongside, same era.
Real Analysis: Measure Theory, Integration, and Hilbert Spaces (Princeton Lectures in Analysis, Book 3)
Stein, E.M., Shakarchi, R., 2005 · 2005
Cited alongside, same era.
Pattern Recognition and Machine Learning (Information Science and Statistics)
Bishop, C.M., 2006 · 2006
Cited alongside, same era.
Concentration inequalities and asymptotic results for ratio type empirical processes
Giné, E., Koltchinskii, V., 2006 · 2006
Cited alongside, same era.
A distribution-free theory of nonparametric regression
Györfi, L., Kohler, M., Krzyzak, A., Walk, H., 2006 · 2006
Adaptive minimax regression estimation over sparse ℓ q \ell_{q} -hulls
Wang, Z., Paterlini, S., Gao, F., Yang, Y., 2014 · 2014
Later among the works it cites.
Deep learning in neural networks: An overview
Schmidhuber, J., 2015 · 2015
Later among the works it cites.
Deep learning
Goodfellow, I., Bengio, Y., Courville, A., Bengio, Y., 2016 · 2016
Later among the works it cites.
Optimal approximation with sparsely connected deep neural networks
Bölcskei, H., Grohs, P., Kutyniok, G., Petersen, P., 2017 · 2017
Later among the works it cites.
DGD approximation theory workshop
Keiper, S., Kutyniok, G., Petersen, P., 2017 · 2017
Later among the works it cites.
Nonparametric regression using deep neural networks with ReLU activation function
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Local rademacher complexities and oracle inequalities in risk minimization
Koltchinskii, V., 2006 · 2006
Cited alongside, same era.
An introduction to compressive sampling [a sensing/sampling paradigm that goes against the common knowledge in data acquisition]
Candès, E.J., Wakin, M.B., 2008 · 2008
Cited alongside, same era.
Concentration on measure
Lafferty, J., Liu, H., Wasserman, L., 2008 · 2008
Cited alongside, same era.
Introduction to Nonparametric Estimation
Tsybakov, A.B., 2008 · 2008
Cited alongside, same era.
Deep sparse rectifier neural networks, in: Proceedings of the 14th International Conference on Artificial Intelligence and Statistics, pp. 315–323
Glorot, X., Bordes, A., Bengio, Y., 2011 · 2011
Cited alongside, same era.
Minimax rates of estimation for high-dimensional linear regression over ℓ q \ell_{q} -balls
Raskutti, G., Wainwright, M.J., Yu, B., 2011 · 2011
Cited alongside, same era.
Schmidt-Hieber, J., 2017 · 2017
Later among the works it cites.
Neural network with unbounded activation functions is universal approximator
Sonoda, S., Murata, N., 2017 · 2017
Later among the works it cites.
Error bounds for approximations with deep ReLU networks
Yarotsky, D., 2017 · 2017
Later among the works it cites.
Optimal approximation of piecewise smooth functions using deep ReLU neural networks
Petersen, P., Voigtlaender, F., 2018 · 2018
Later among the works it cites.
Deep neural networks learn non-smooth functions effectively, in: Proceedings of the 22nd International Conference on Artificial Intelligence and Statistics
Imaizumi, M., Fukumizu, K., 2019 · 2019
Closest in time.
Adaptivity of deep ReLU network for learning in Besov and mixed smooth Besov spaces: optimal rate and curse of dimensionality, in: ICLR 2019
Suzuki, T., 2019 · 2019
Closest in time.