Fetching the paper…
Reading the bibliography…
We study universal traits which emerge both in real-world complex datasets, as well as in artificially generated ones.
Traditional and heavy-tailed self-regularization in neural network models
Charles H Martin and Michael W Mahoney · 1901
Earlier work this paper cites.
Beitrag zur theorie des ferromagnetismus
Ernst Ising · 1925
Earlier work this paper cites.
Generalised product moment distribution in samples from an indefinitely large population
John Wishart Wishart · 1928
Earlier work this paper cites.
A mathematical theory of communication
Claude Elwood Shannon · 1948
Earlier work this paper cites.
Information theory and statistics
Solomon Kullback and Richard A Leibler · 1951
Earlier work this paper cites.
On the classical motion of charged particles
M. A. Smorodinsky · 1953
Earlier work this paper cites.
On measures of information and entropy
Alfred Rényi · 1956
Earlier work this paper cites.
Statistical mechanics of charged particles
T. D. Lee and C. N. Yang · 1966
Earlier work this paper cites.
Random matrix models in nuclear physics
T. A. Brody · 1981
Earlier work this paper cites.
Random matrix theory and quantum chaos
A. Pandey · 1983
Earlier work this paper cites.
Characterization of chaotic quantum spectra and universality of level fluctuation laws
O. Bohigas, M. J. Giannoni, and C. Schmit · 1984
Earlier work this paper cites.
Random matrices , volume 111
M. L. Mehta · 1991
Earlier work this paper cites.
Statistical mechanics of learning from examples
Hyunjune Sebastian Seung, Haim Sompolinsky, and Naftali Tishby · 1992
Earlier work this paper cites.
The statistical mechanics of learning a rule
Timothy LH Watkin, Albrecht Rau, and Michael Biehl · 1993
Earlier work this paper cites.
On the empirical distribution of eigenvalues of a class of large dimensional random matrices
Jack W Silverstein and Z. D. Bai · 1995
Earlier work this paper cites.
Statistical modeling of natural images with wavelets
David L Donoho · 1995
Earlier work this paper cites.
Origins of scaling in natural images
Daniel L. Ruderman · 1997
Earlier work this paper cites.
Supersymmetry and disorder in quantum mechanics
K. B. Efetov · 1997
Earlier work this paper cites.
Random-matrix theories in quantum physics: common concepts
Thomas Guhr, Axel Müller–Groeling, and Hans A. Weidenmüller · 1998
Earlier work this paper cites.
Random matrix theory and financial markets
V. Plerou, P. Gopikrishnan, B. Rosenow, L. A. N. Amaral, H. E. Stanley, and Stanley M. S · 1999
Earlier work this paper cites.
Statistical mechanics of learning
Andreas Engel and Christian Van den Broeck · 2001
Earlier work this paper cites.
A simple formula for the average gate fidelity of a quantum dynamical operation
Michael A Nielsen · 2002
Earlier work this paper cites.
Phase transition of the largest eigenvalue for non-null complex sample covariance matrices, 2004
Jinho Baik, Gerard Ben Arous, and Sandrine Peche · 2004
Earlier work this paper cites.
Random Matrices
M. L. Mehta · 2004
Earlier work this paper cites.
Toeplitz and circulant matrices: A review
Robert M. Gray · 2006
Earlier work this paper cites.
Optimal rates for the regularized least-squares algorithm
A. Caponnetto and E. De Vito · 2007
Earlier work this paper cites.
Localization of interacting fermions at high temperature
Vadim Oganesyan and David A. Huse · 2007
Cited alongside, same era.
Tiny imagenet: A benchmark for evaluation of image classification algorithms
Antonio Torralba, Andreas A Efros, and Christopher Anderson · 2008
Cited alongside, same era.
Spectral analysis of large dimensional random matrices , volume 20
Zhidong Bai and Jack W Silverstein · 2010
Cited alongside, same era.
The mnist database of handwritten digits
Yann LeCun, Léon Bottou, Yoshua Bengio, and Pierre Haffner · 2010
Cited alongside, same era.
Topics in random matrix theory , volume 132
Terence Tao · 2012
Cited alongside, same era.
Distribution of the ratio of consecutive level spacings in random matrix ensembles
Y. Y. Atas, E. Bogomolny, O. Giraud, and G. Roux · 2013
Cited alongside, same era.
Generalization error in high-dimensional perceptrons: Approaching bayes error with convex optimization
Benjamin Aubin, Florent Krzakala, Yue M Lu, and Lenka Zdeborov’a · 2020
Later among the works it cites.
The performance analysis of generalized margin maximizers on separable data
Fariborz Salehi, Ehsan Abbasi, and Babak Hassibi · 2020
Later among the works it cites.
Random matrix theory proves that deep learning representations of gan-data behave as gaussian mixtures, 2020
Mohamed El Amine Seddik, Cosme Louart, Mohamed Tamaazousti, and Romain Couillet · 2020
Later among the works it cites.
A random matrix analysis of random fourier features: beyond the gaussian kernel, a precise phase transition, and the corresponding double descent
Zhenyu Liao, Romain Couillet, and Michael W Mahoney · 2021
Later among the works it cites.
Hessian eigenspectra of more realistic nonlinear models, 2021
Zhenyu Liao and Michael W. Mahoney · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
On robust regression with high-dimensional predictors
Noureddine El Karoui, Derek Bean, Peter J Bickel, Chinghway Lim, and Bin Yu · 2013
Cited alongside, same era.
Exact solutions to the nonlinear dynamics of learning in deep linear neural networks
Andrew M Saxe, James L McClelland, and Surya Ganguli · 2014
Cited alongside, same era.
Universality for the largest eigenvalue of sample covariance matrices with general population
Zhigang Bao, Guangming Pan, and Wang Zhou · 2015
Cited alongside, same era.
Celeba: A large-scale celebrity face attribute dataset
Ziwei Liu, Zihang Luo, Xiaogang Wang, and Xiaoou Tang · 2015
Cited alongside, same era.
Statistical physics of inference: Thresholds and algorithms
Lenka Zdeborov’a and Florent Krzakala · 2016
Cited alongside, same era.
High dimensional robust m-estimation: Asymptotic variance via approximate message passing
David Donoho and Andrea Montanari · 2016
Cited alongside, same era.
Generalisation error in learning with random features and the hidden manifold model
Federica Gerace, Bruno Loureiro, Florent Krzakala, Marc Mezard, and Lenka Zdeborova · 2021
Later among the works it cites.
Learning curves of generic features maps for realistic datasets with a teacher-student model
Bruno Loureiro, Cedric Gerbelot, Hugo Cui, Sebastian Goldt, Florent Krzakala, Marc Mezard, and Lenka Zdeborová · 2021
Later among the works it cites.
A solvable model of neural scaling laws, 2022
Alexander Maloney, Daniel A. Roberts, and James Sully · 2022
Later among the works it cites.
Scaling laws and interpretability of learning from repeated data, 2022
Danny Hernandez, Tom Brown, Tom Conerly, Nova DasSarma, Dawn Drain, Sheer El-Showk, Nelson Elhage, Zac Hatfield-Dodds, Tom Henighan, Tristan Hume, Scott Johnston, Ben Mann, Chris Olah, Catherine Olsson, Dario Amodei, Nicholas Joseph, Jared Kaplan, and Sam McCandlish · 2022
Later among the works it cites.
Scaling laws under the microscope: Predicting transformer performance from small scale experiments
Maor Ivgi, Yair Carmon, and Jonathan Berant · 2022
Later among the works it cites.
Revisiting neural scaling laws in language and vision
Ibrahim M. Alabdulmohsin, Behnam Neyshabur, and Xiaohua Zhai · 2022
Later among the works it cites.
Scaling laws from the data manifold dimension
Utkarsh Sharma and Jared Kaplan · 2022
Later among the works it cites.
Beyond neural scaling laws: beating power law scaling via data pruning
Ben Sorscher, Robert Geirhos, Shashank Shekhar, Surya Ganguli, and Ari Morcos · 2022
Later among the works it cites.
Random matrix analysis of deep neural network weight matrices
Matthias Thamm, Max Staats, and Bernd Rosenow · 2022
Later among the works it cites.
Random Matrix Methods for Machine Learning
Romain Couillet and Zhenyu Liao · 2022
Later among the works it cites.
Universality laws for high-dimensional learning with random features, 2022
Hong Hu and Yue M. Lu · 2022
Later among the works it cites.
The generalization error of random features regression: Precise asymptotics and the double descent curve
Song Mei and Andrea Montanari · 2022
Later among the works it cites.
The gaussian equivalence of generative models for learning with shallow neural networks
Sebastian Goldt, Bruno Loureiro, Galen Reeves, Florent Krzakala, Marc Mezard, and Lenka Zdeborova · 2022
Later among the works it cites.
More than a toy: Random matrix models predict how real-world neural representations generalize
Alexander Wei, Wei Hu, and Jacob Steinhardt · 2022
Later among the works it cites.
The universal statistical structure and scaling laws of chaos and turbulence, 2023
Noam Levi and Yaron Oz · 2023
Closest in time.
A simplistic model of neural scaling laws: Multiperiodic santa fe processes
Lukasz Debowski · 2023
Closest in time.
Scaling laws for multilingual neural machine translation
Patrick Fernandes, Behrooz Ghorbani, Xavier Garcia, Markus Freitag, and Orhan Firat · 2023
Closest in time.
Quantum chaos and circuit parameter optimization
Joonho Kim, Yaron Oz, and Dario Rosa · 2023
Closest in time.
Gaussian universality of perceptrons with random labels
Federica Gerace, Florent Krzakala, Bruno Loureiro, Ludovic Stephan, and Lenka Zdeborová · 2023
Closest in time.
Are gaussian data all you need? extents and limits of universality in high-dimensional generalized linear estimation
Luca Pesce, Florent Krzakala, Bruno Loureiro, and Ludovic Stephan · 2023
Closest in time.