Fetching the paper…
Reading the bibliography…
Understanding the impact of data structure on the computational tractability of learning is a key challenge for the theory of neural networks.
Geometrical and Statistical Properties of Systems of Linear Inequalities with Applications in Pattern Recognition
T.M. Cover · 1965
Earlier work this paper cites.
Typical distributions of linear functionals in finite-dimensional spaces of high dimension
V. N. Sudakov · 1978
Earlier work this paper cites.
Neocognitron: A new algorithm for pattern recognition tolerant of deformations and shifts in position
K. Fukushima and S. Miyake · 1982
Earlier work this paper cites.
Asymptotics of graphical projection pursuit
P. Diaconis and D. Freedman · 1984
Earlier work this paper cites.
Neural networks and principal component analysis: Learning from examples without local minima
P. Baldi and K. Hornik · 1989
Earlier work this paper cites.
Three unfinished works on the optimal storage capacity of networks
E. Gardner and B. Derrida · 1989
Earlier work this paper cites.
Handwritten digit recognition with a back-propagation network
Y. LeCun, B.E. Boser, J.S. Denker, D. Henderson, R.E. Howard, W.E. Hubbard, and L.D. Jackel · 1990
Earlier work this paper cites.
Eigenvalues of covariance matrices: Application to neural-network learning
Y. Le Cun, I. Kanter, and S.A. Solla · 1991
Earlier work this paper cites.
Generalization in a linear perceptron in the presence of noise
A. Krogh and J.A. Hertz · 1992
Earlier work this paper cites.
Statistical mechanics of learning from examples
H. S. Seung, H. Sompolinsky, and N. Tishby · 1992
Earlier work this paper cites.
On almost linearity of low dimensional projections from high dimensional data
P. Hall and K.-C. Li · 1993
Earlier work this paper cites.
The statistical mechanics of learning a rule
T.L.H. Watkin, A. Rau, and M. Biehl · 1993
Earlier work this paper cites.
Learning by on-line gradient descent
M. Biehl and H. Schwarze · 1995
Earlier work this paper cites.
Bayesian Learning for Neural Networks
R.M. Neal · 1995
Earlier work this paper cites.
Statistical Mechanics of Learning
A. Engel and C. Van den Broeck · 2001
Earlier work this paper cites.
On concentration of distributions of random weighted sums
S. G. Bobkov · 2003
Earlier work this paper cites.
Deterministic equivalents for certain functionals of large random matrices
W. Hachem, P. Loubaton, and J. Najim · 2007
Earlier work this paper cites.
Random features for large-scale kernel machines
A. Rahimi and B. Recht · 2008
Earlier work this paper cites.
Learning multiple layers of features from tiny images
A. Krizhevsky, G. Hinton, et al · 2009
Earlier work this paper cites.
Weighted sums of random kitchen sinks: Replacing minimization with randomization in learning
A. Rahimi and B. Recht · 2009
Earlier work this paper cites.
On-line learning in neural networks , volume 17
D. Saad · 2009
Earlier work this paper cites.
Approximation of projections of random vectors
E. Meckes · 2010
Earlier work this paper cites.
Density estimation by dual ascent of the log-likelihood
E. G Tabak, E. Vanden-Eijnden, et al · 2010
Earlier work this paper cites.
The singular values and vectors of low rank perturbations of large rectangular random matrices
F. Benaych-Georges and R.R. Nadakuditi · 2012
Earlier work this paper cites.
Matrix analysis
R.A. Horn and C.R. Johnson · 2012
Earlier work this paper cites.
Foundations of Machine Learning
M. Mohri, A. Rostamizadeh, and A. Talwalkar · 2012
Earlier work this paper cites.
Invariant scattering convolution networks
J. Bruna and S. Mallat · 2013
Earlier work this paper cites.
The spectrum of random inner-product kernel matrices
X. Cheng and A. Singer · 2013
Earlier work this paper cites.
A family of nonparametric density estimation algorithms
E. G Tabak and C.V. Turner · 2013
Earlier work this paper cites.
The nature of statistical learning theory
V. Vapnik · 2013
Earlier work this paper cites.
Generative adversarial nets
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio · 2014
Cited alongside, same era.
Auto-encoding variational bayes
D.P. Kingma and M. Welling · 2014
Cited alongside, same era.
Conditional generative adversarial nets
M. Mirza and S. Osindero · 2014
Cited alongside, same era.
Analysis of Boolean Functions
R. O’Donnell · 2014
Cited alongside, same era.
Exact solutions to the nonlinear dynamics of learning in deep linear neural networks
A.M. Saxe, J.L. McClelland, and S. Ganguli · 2014
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
S. Ioffe and C. Szegedy · 2015
Cited alongside, same era.
Theoretical insights into the optimization landscape of over-parameterized shallow neural networks
M. Soltanolkotabi, A. Javanmard, and J.D. Lee · 2018
Later among the works it cites.
Intrinsic dimension of data representations in deep neural networks
A. Ansuini, A. Laio, J.H. Macke, and D. Zoccolan · 2019
Later among the works it cites.
The spiked matrix model with generative priors
B. Aubin, B. Loureiro, A. Maillard, F. Krzakala, and L. Zdeborová · 2019
Later among the works it cites.
Generalization from correlated sets of patterns in the perceptron
F. Borra, M.C. Lagomarsino, P. Rotondo, and M. Gherardi · 2019
Later among the works it cites.
Large scale GAN training for high fidelity natural image synthesis
A. Brock, J. Donahue, and K. Simonyan · 2019
Later among the works it cites.
On lazy training in differentiable programming
L. Chizat, E. Oyallon, and F. Bach · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Variational inference with normalizing flows
D. Rezende and S. Mohamed · 2015
Cited alongside, same era.
Deep learning and hierarchical generative models
E. Mossel · 2016
Cited alongside, same era.
A probabilistic framework for deep learning
A.B. Patel, M.T. Nguyen, and R. Baraniuk · 2016
Cited alongside, same era.
Unsupervised representation learning with deep convolutional generative adversarial networks
A. Radford, L. Metz, and S. Chintala · 2016
Cited alongside, same era.
Statistical physics of inference: thresholds and algorithms
L. Zdeborová and F. Krzakala · 2016
Cited alongside, same era.
Density estimation using real NVP
L. Dinh, J. Sohl-Dickstein, and S. Bengio · 2017
Cited alongside, same era.
Later among the works it cites.
The spectral norm of random inner-product kernel matrices
Z. Fan and A. Montanari · 2019
Later among the works it cites.
Limitations of lazy training of two-layers neural network
B. Ghorbani, S. Mei, T. Misiakiewicz, and A. Montanari · 2019
Later among the works it cites.
Dynamics of stochastic gradient descent for two-layer neural networks in the teacher-student setup
S. Goldt, M.S. Advani, A.M. Saxe, F. Krzakala, and L. Zdeborová · 2019
Later among the works it cites.
The comparative power of reLU networks and polynomial kernels in the presence of sparse latent structure
F. Koehler and A. Risteski · 2019
Later among the works it cites.
Generalized sliced wasserstein distances
S. Kolouri, K. Nadjahi, U. Simsekli, R. Badeau, and G. Rohde · 2019
Later among the works it cites.
The generalization error of random features regression: Precise asymptotics and double descent curve
S. Mei and A. Montanari · 2019
Later among the works it cites.
A. Montanari, F. Ruan, Y. Sohn, and J. Yan · 2019
Later among the works it cites.
Normalizing Flows for Probabilistic Modeling and Inference
G. Papamakarios, E. Nalisnick, D.J. Rezende, S. Mohamed, and B. Lakshminarayanan · 2019
Later among the works it cites.
Pytorch: An imperative style, high-performance deep learning library
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga, A. Desmaison, A. Kopf, E. Yang, Z. DeVito, M. Raison, A. Tejani, S. Chilamkurthy, B. Steiner, L. Fang, J. Bai, and S. Chintala · 2019
Later among the works it cites.
Kernel random matrices of large concentrated data: the example of gan-generated images
M.E.A. Seddik, M. Tamaazousti, and R. Couillet · 2019
Later among the works it cites.
Mean field analysis of neural networks: A central limit theorem
J. Sirignano and K. Spiliopoulos · 2019
Later among the works it cites.
Data-dependence of plateau phenomenon in learning with neural network — statistical mechanical analysis
Y. Yoshida and M. Okada · 2019
Later among the works it cites.
High-dimensional dynamics of generalization error in neural networks
M.S. Advani, A.M. Saxe, and H. Sompolinsky · 2020
Closest in time.
Statistical Mechanics of Deep Learning
Y. Bahri, J. Kadmon, J. Pennington, S.S. Schoenholz, J. Sohl-Dickstein, and S. Ganguli · 2020
Closest in time.
Separability and geometry of object manifolds in deep neural networks
U. Cohen, SY Chung, D.D. Lee, and H. Sompolinsky · 2020
Closest in time.
A precise performance analysis of learning with random features
O. Dhifallah and Y. M. Lu · 2020
Closest in time.
Mean-field inference methods for neural networks
M. Gabrié · 2020
Closest in time.
Generalisation error in learning with random features and the hidden manifold model
F. Gerace, B. Loureiro, F. Krzakala, M. Mézard, and L. Zdeborová · 2020
Closest in time.
Modeling the influence of data structure on learning in neural networks: The hidden manifold model
S. Goldt, M. Mézard, F. Krzakala, and L. Zdeborová · 2020
Closest in time.
Universality laws for high-dimensional learning with random features
H. Hu and Y.M. Lu · 2020
Closest in time.
Normalizing flows: An introduction and review of current methods
I. Kobyzev, S. Prince, and M. Brubaker · 2020
Closest in time.
Counting the learnable functions of geometrically structured data
P. Rotondo, M. C. Lagomarsino, and M. Gherardi · 2020
Closest in time.
Random matrix theory proves that deep learning representations of gan-data behave as gaussian mixtures
M.E.A. Seddik, C. Louart, M. Tamaazousti, and R. Couillet · 2020
Closest in time.
Understanding deep learning is also a job for physicists
L. Zdeborová · 2020
Closest in time.