More data can hurt for linear regression: Sample-wise double descent
Original
Preetum Nakkiran · 2019
Later among the works it cites.
High-dimensional dynamics of generalization error in neural networks
Madhu S Advani, Andrew M Saxe, and Haim Sompolinsky · 2020
Closest in time.
Generalization of two-layer neural networks: An asymptotic viewpoint
Jimmy Ba, Murat A. Erdogdu, Taiji Suzuki, Denny Wu, and Tianzong Zhang · 2020
Closest in time.
Benign overfitting in linear regression
Peter L Bartlett, Philip M Long, Gábor Lugosi, and Alexander Tsigler · 2020
Closest in time.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Closest in time.
Multiple descent: Design your own generalization curve
Original
Lin Chen, Y. Min, M. Belkin, and Amin Karbasi · 2020
Closest in time.
Double trouble in double descent: Bias and variance (s) in the lazy regime
Stéphane d’Ascoli, Maria Refinetti, Giulio Biroli, and Florent Krzakala · 2020
Closest in time.
Wonder: Weighted one-shot distributed ridge regression in high dimensions
Edgar Dobriban and Yue Sheng · 2020
Closest in time.
Spectra of the conjugate kernel and neural tangent kernel for linear-width neural networks
Zhou Fan and Zhichao Wang · 2020
Closest in time.
Scaling description of generalization with number of parameters in deep learning
Mario Geiger, Arthur Jacot, Stefano Spigler, Franck Gabriel, Levent Sagun, Stéphane d’Ascoli, Giulio Biroli, Clément Hongler, and Matthieu Wyart · 2020
Closest in time.
Generalisation error in learning with random features and the hidden manifold model
Federica Gerace, Bruno Loureiro, Florent Krzakala, Marc Mézard, and Lenka Zdeborová · 2020
Closest in time.
Provable benefit of orthogonal initialization in optimizing deep linear networks
Wei Hu, Lechao Xiao, and Jeffrey Pennington · 2020
Closest in time.
Implicit regularization of random feature models
Arthur Jacot, Berfin Simsek, Francesco Spadaro, Clément Hongler, and Franck Gabriel · 2020
Closest in time.
The optimal ridge penalty for real-world high-dimensional data can be zero or negative due to the implicit ridge regularization
Dmitry Kobak, Jonathan Lomond, and Benoit Sanchez · 2020
Closest in time.
Provable more data hurt in high dimensional least squares estimator
Original
Z. Li, Chuanlong Xie, and Qinwen Wang · 2020
Closest in time.
On the multiple descent of minimum-norm interpolants and restricted lower isometry of kernels
Tengyuan Liang, Alexander Rakhlin, and Xiyu Zhai · 2020
Closest in time.
A random matrix analysis of random fourier features: beyond the gaussian kernel, a precise phase transition, and the corresponding double descent
Zhenyu Liao, Romain Couillet, and Michael W Mahoney · 2020
Closest in time.
Ridge regression: Structure, cross-validation, and sketching
Sifan Liu and Edgar Dobriban · 2020
Closest in time.
A brief prehistory of double descent
Marco Loog, Tom Viering, Alexander Mey, Jesse H Krijthe, and David MJ Tax · 2020
Closest in time.
Harmless interpolation of noisy data in regression
Vidya Muthukumar, Kailas Vodrahalli, Vignesh Subramanian, and Anant Sahai · 2020
Closest in time.
Deep double descent: Where bigger models and more data hurt
Preetum Nakkiran, Gal Kaplun, Yamini Bansal, Tristan Yang, Boaz Barak, and Ilya Sutskever · 2020
Closest in time.
Deep isometric learning for visual recognition
Haozhi Qi, Chong You, Xiaolong Wang, Yi Ma, and Jitendra Malik · 2020
Closest in time.
Memorizing without overfitting: Bias, variance, and interpolation in over-parameterized models
Original
Jason W Rocks and Pankaj Mehta · 2020
Closest in time.
On the optimal weighted ℓ 2 \ell_{2} regularization in overparameterized linear regression
Denny Wu and Ji Xu · 2020
Closest in time.
Weighted optimization: better generalization by smoother interpolation
Original
Yuege Xie, Rachel Ward, Holger Rauhut, and Hung-Hsu Chou · 2020
Closest in time.
Rethinking bias-variance trade-off for generalization of neural networks
Zitong Yang, Yaodong Yu, Chong You, Jacob Steinhardt, and Yi Ma · 2020
Closest in time.
Linearized two-layers neural networks in high dimension
Behrooz Ghorbani, Song Mei, Theodor Misiakiewicz, and Andrea Montanari · 2021
Closest in time.
Optimal regularization can mitigate double descent
Preetum Nakkiran, Prayaag Venkat, Sham M. Kakade, and Tengyu Ma · 2021
Closest in time.