Fetching the paper…
Reading the bibliography…
While classical in many theoretical settings - and in particular in statistical physics-inspired works - the assumption of Gaussian i.i.d.
A problem in geometric probability
James G Wendel · 1962
Earlier work this paper cites.
Geometrical and statistical properties of systems of linear inequalities with applications in pattern recognition
Thomas M Cover · 1965
Earlier work this paper cites.
Optimal storage properties of neural network models
Elizabeth Gardner and Bernard Derrida · 1988
Earlier work this paper cites.
Three unfinished works on the optimal storage capacity of networks
Elizabeth Gardner and Bernard Derrida · 1989
Earlier work this paper cites.
Storage capacity of memory networks with binary couplings
Werner Krauth and Marc Mézard · 1989
Earlier work this paper cites.
Finite-size effects and bounds for perceptron models
Bernard Derrida, RB Griffiths, and A Prugel-Bennett · 1991
Earlier work this paper cites.
Statistical mechanics of learning from examples
Hyunjune Sebastian Seung, Haim Sompolinsky, and Naftali Tishby · 1992
Earlier work this paper cites.
Information capacity of a perceptron
N Brunel, J-P Nadal, and G Toulouse · 1992
Earlier work this paper cites.
Statistical mechanics of the maximum-likelihood density estimation
N Barkai and Haim Sompolinsky · 1994
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. Lecun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
The nature of statistical learning theory
Vladimir Vapnik · 1999
Earlier work this paper cites.
Random features for large-scale kernel machines
Ali Rahimi, Benjamin Recht, et al · 2007
Earlier work this paper cites.
Support vector machines
Ingo Steinwart and Andreas Christmann · 2008
Earlier work this paper cites.
Observed universality of phase transitions in high-dimensional geometry, with implications for modern data analysis and signal processing
David Donoho and Jared Tanner · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Alex Krizhevsky · 2009
Earlier work this paper cites.
Universality Laws for High-Dimensional Learning with Random Features
Hong Hu and Yue M. Lu · 2009
Earlier work this paper cites.
Observed universality of phase transitions in high-dimensional geometry, with implications for modern data analysis and signal processing
David Donoho and Jared Tanner · 2009
Earlier work this paper cites.
The spectrum of kernel random matrices
Noureddine El Karoui · 2010
Earlier work this paper cites.
Transfer learning
Lisa Torrey and Jude Shavlik · 2010
Earlier work this paper cites.
Applications of the lindeberg principle in communications and statistical learning
Satish Babu Korada and Andrea Montanari · 2011
Earlier work this paper cites.
Scikit-learn: Machine learning in Python
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V. Dubourg, J. Vanderplas, A. Passos, D. Cournapeau, M. Brucher, M. Perrot, and E. Duchesnay · 2011
Earlier work this paper cites.
The mnist database of handwritten digit images for machine learning research
Li Deng · 2012
Earlier work this paper cites.
An introduction to statistical learning
Gareth James, Daniela Witten, Trevor Hastie, and Robert Tibshirani · 2013
Earlier work this paper cites.
Invariant scattering convolution networks
Joan Bruna and Stéphane Mallat · 2013
Earlier work this paper cites.
Understanding machine learning: From theory to algorithms
Shai Shalev-Shwartz and Shai Ben-David · 2014
Earlier work this paper cites.
On large-batch training for deep learning: Generalization gap and sharp minima
Nitish Shirish Keskar, Dheevatsa Mudigere, Jorge Nocedal, Mikhail Smelyanskiy, and Ping Tak Peter Tang · 2016
Earlier work this paper cites.
Phase transitions and optimal algorithms in high-dimensional gaussian mixture clustering
Thibault Lesieur, Caterina De Bacco, Jess Banks, Florent Krzakala, Cris Moore, and Lenka Zdeborová · 2016
Earlier work this paper cites.
CVXPY: A Python-embedded modeling language for convex optimization
Steven Diamond and Stephen Boyd · 2016
Earlier work this paper cites.
Universality of the sat-unsat (jamming) threshold in non-convex continuous constraint satisfaction problems
Silvio Franz, Giorgio Parisi, Maxime Sevelev, Pierfrancesco Urbani, and Francesco Zamponi · 2017
Earlier work this paper cites.
Fashion-mnist: a novel image dataset for benchmarking machine learning algorithms, 2017
Han Xiao, Kashif Rasul, and Roland Vollgraf · 2017
Earlier work this paper cites.
A closer look at memorization in deep networks
Devansh Arpit, Stanisław Jastrzębski, Nicolas Ballas, David Krueger, Emmanuel Bengio, Maxinder S Kanwal, Tegan Maharaj, Asja Fischer, Aaron Courville, Yoshua Bengio, et al · 2017
Cited alongside, same era.
Exploring generalization in deep learning
Behnam Neyshabur, Srinadh Bhojanapalli, David McAllester, and Nathan Srebro · 2017
Cited alongside, same era.
Nonlinear random matrix theory for deep learning
Jeffrey Pennington and Pratik Worah · 2017
Cited alongside, same era.
Universality of the elastic net error
Andrea Montanari and Phan-Minh Nguyen · 2017
Cited alongside, same era.
A universal analysis of large-scale regularized least squares solutions
Ashkan Panahi and Babak Hassibi · 2017
Cited alongside, same era.
Algorithmic pure states for the negative spherical perceptron
Ahmed El Alaoui and Mark Sellke · 2020
Later among the works it cites.
What do neural networks learn when trained with random labels?
Hartmut Maennel, Ibrahim Alabdulmohsin, Ilya Tolstikhin, Robert JN Baldock, Olivier Bousquet, Sylvain Gelly, and Daniel Keysers · 2020
Later among the works it cites.
Kymatio: Scattering transforms in python
Mathieu Andreux, Tomás Angles, Georgios Exarchakis, Roberto Leonarduzzi, Gaspar Rochette, Louis Thiry, John Zarka, Stéphane Mallat, Joakim Andén, Eugene Belilovsky, et al · 2020
Later among the works it cites.
Generalisation error in learning with random features and the hidden manifold model
Federica Gerace, Bruno Loureiro, Florent Krzakala, Marc Mézard, and Lenka Zdeborová · 2020
Later among the works it cites.
On the optimal weighted ℓ 2 \ell_{2} regularization in overparameterized linear regression
Denny Wu and Ji Xu · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Lenaic Chizat, Edouard Oyallon, and Francis Bach · 2018
Cited alongside, same era.
Neural tangent kernel: Convergence and generalization in neural networks
Arthur Jacot, Franck Gabriel, and Clément Hongler · 2018
Cited alongside, same era.
Precise error analysis of regularized m m -estimators in high dimensions
Christos Thrampoulidis, Ehsan Abbasi, and Babak Hassibi · 2018
Cited alongside, same era.
High-dimensional asymptotics of prediction: Ridge regression and classification
Edgar Dobriban, Stefan Wager, et al · 2018
Cited alongside, same era.
Co-teaching: Robust training of deep neural networks with extremely noisy labels
Bo Han, Quanming Yao, Xingrui Yu, Gang Niu, Miao Xu, Weihua Hu, Ivor Tsang, and Masashi Sugiyama · 2018
Cited alongside, same era.
Generalized cross entropy loss for training deep neural networks with noisy labels
Zhilu Zhang and Mert R Sabuncu · 2018
Cited alongside, same era.
A random matrix approach to neural networks
Cosme Louart, Zhenyu Liao, and Romain Couillet · 2018
Cited alongside, same era.
Later among the works it cites.
Kernel alignment risk estimator: Risk prediction from training data
Arthur Jacot, Berfin Şimşek, Francesco Spadaro, Clément Hongler, and Franck Gabriel · 2020
Later among the works it cites.
The lasso with general gaussian designs with applications to hypothesis testing
Michael Celentano, Andrea Montanari, and Yuting Wei · 2020
Later among the works it cites.
Generalization error in high-dimensional perceptrons: Approaching bayes error with convex optimization
Benjamin Aubin, Florent Krzakala, Yue M Lu, and Lenka Zdeborová · 2020
Later among the works it cites.
The role of regularization in classification of high-dimensional noisy gaussian mixture
Francesca Mignacco, Florent Krzakala, Yue Lu, Pierfrancesco Urbani, and Lenka Zdeborova · 2020
Later among the works it cites.
Optimality of least-squares for classification in gaussian-mixture models
Hossein Taheri, Ramtin Pedarsani, and Christos Thrampoulidis · 2020
Later among the works it cites.
Provable tradeoffs in adversarially robust classification
Edgar Dobriban, Hamed Hassani, David Hong, and Alexander Robey · 2020
Later among the works it cites.
Prevalence of neural collapse during the terminal phase of deep learning training
Vardan Papyan, XY Han, and David L Donoho · 2020
Later among the works it cites.
Learning curves of generic features maps for realistic datasets with a teacher-student model
Bruno Loureiro, Cedric Gerbelot, Hugo Cui, Sebastian Goldt, Florent Krzakala, Marc Mezard, and Lenka Zdeborová · 2021
Later among the works it cites.
Tractability from overparametrization: The example of the negative perceptron
Andrea Montanari, Yiqiao Zhong, and Kangjie Zhou · 2021
Later among the works it cites.
Understanding deep learning (still) requires rethinking generalization
Chiyuan Zhang, Samy Bengio, Moritz Hardt, Benjamin Recht, and Oriol Vinyals · 2021
Later among the works it cites.
Learning gaussian mixtures with generalized linear models: Precise asymptotics in high-dimensions
Bruno Loureiro, Gabriele Sicuro, Cedric Gerbelot, Alessandro Pacco, Florent Krzakala, and Lenka Zdeborová · 2021
Later among the works it cites.
Phase transitions in transfer learning for high-dimensional perceptrons
Oussama Dhifallah and Yue M. Lu · 2021
Later among the works it cites.
Statistical mechanics of deep linear neural networks: The backpropagating kernel renormalization
Qianyi Li and Haim Sompolinsky · 2021
Later among the works it cites.
Generalization error rates in kernel regression: The crossover from the noiseless to noisy regime
Hugo Cui, Bruno Loureiro, Florent Krzakala, and Lenka Zdeborová · 2021
Later among the works it cites.
Phase transitions for one-vs-one and one-vs-all linear separability in multiclass gaussian mixtures
Ganesh Ramachandra Kini and Christos Thrampoulidis · 2021
Later among the works it cites.
The unexpected deterministic and universal behavior of large softmax classifiers
Mohamed El Amine Seddik, Cosme Louart, Romain Couillet, and Mohamed Tamaazousti · 2021
Later among the works it cites.
Benign overfitting in binary classification of gaussian mixtures
Ke Wang and Christos Thrampoulidis · 2021
Later among the works it cites.
The gaussian equivalence of generative models for learning with shallow neural networks
Sebastian Goldt, Bruno Loureiro, Galen Reeves, Florent Krzakala, Marc Mezard, and Lenka Zdeborova · 2022
Closest in time.
Universality of empirical risk minimization
Andrea Montanari and Basil Saeed · 2022
Closest in time.
Data-driven emergence of convolutional structure in neural networks, 2022
Alessandro Ingrosso and Sebastian Goldt · 2022
Closest in time.
Probing transfer learning with a model of synthetic correlated datasets
Federica Gerace, Luca Saglietti, Stefano Sarao Mannelli, Andrew Saxe, and Lenka Zdeborová · 2022
Closest in time.
Fluctuations, bias, variance & ensemble of learners: Exact asymptotics for convex losses in high-dimension, 2022
Bruno Loureiro, Cédric Gerbelot, Maria Refinetti, Gabriele Sicuro, and Florent Krzakala · 2022
Closest in time.
Umberto M Tomasini, Antonio Sclocchi, and Matthieu Wyart · 2022
Closest in time.
High-dimensional asymptotics of feature learning: How one gradient step improves the representation, 2022
Jimmy Ba, Murat A. Erdogdu, Taiji Suzuki, Zhichao Wang, Denny Wu, and Greg Yang · 2022
Closest in time.
Universality laws for gaussian mixtures in generalized linear models
Yatin Dandi, Florent Krzakala, Bruno Loureiro, Ludovic Stephan, and Lenka Zdeborova · 2023
Closest in time.