Deep double descent: Where bigger models and more data hurt
Original
Preetum Nakkiran, Gal Kaplun, Yamini Bansal, Tristan Yang, Boaz Barak, and Ilya Sutskever · 1912
Earlier work this paper cites.
Structural risk minimization over data-dependent hierarchies
John Shawe-Taylor, Peter L. Bartlett, Robert C. Williamson, and Martin Anthony · 1998
Earlier work this paper cites.
Convex optimization
Stephen Boyd and Lieven Vandenberghe · 2004
Earlier work this paper cites.
Margin maximizing loss functions
Saharon Rosset, Ji Zhu, and Trevor J Hastie · 2004
Earlier work this paper cites.
Increasing-margin adversarial (IMA) training to improve adversarial robustness of neural networks
Original
Linhai Ma and Liang Liang · 2005
Earlier work this paper cites.
Robust optimization , volume 28
Aharon Ben-Tal, Laurent El Ghaoui, and Arkadi Nemirovski · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Alex Krizhevsky, Geoffrey Hinton, et al · 2009
Earlier work this paper cites.
Structured sparsity through convex optimization
Francis Bach, Rodolphe Jenatton, Julien Mairal, Guillaume Obozinski, et al · 2012
Earlier work this paper cites.
Intriguing properties of neural networks
Original
Christian Szegedy, Wojciech Zaremba, Ilya Sutskever, Joan Bruna, Dumitru Erhan, Ian Goodfellow, and Rob Fergus · 2013
Earlier work this paper cites.
Margins, shrinkage, and boosting
Matus Telgarsky · 2013
Earlier work this paper cites.
Proximal algorithms
Neal Parikh and Stephen Boyd · 2013
Earlier work this paper cites.
Explaining and harnessing adversarial examples
Original
Ian J Goodfellow, Jonathon Shlens, and Christian Szegedy · 2014
Earlier work this paper cites.
A unified gradient regularization family for adversarial examples
Chunchuan Lyu, Kaizhu Huang, and Hai-Ning Liang · 2015
Earlier work this paper cites.
Understanding deep learning requires rethinking generalization
Original
Chiyuan Zhang, Samy Bengio, Moritz Hardt, Benjamin Recht, and Oriol Vinyals · 2016
Earlier work this paper cites.
CVXPY: A Python-embedded modeling language for convex optimization
Steven Diamond and Stephen Boyd · 2016
Earlier work this paper cites.
Towards deep learning models resistant to adversarial attacks
Original
Aleksander Madry, Aleksandar Makelov, Ludwig Schmidt, Dimitris Tsipras, and Adrian Vladu · 2017
Earlier work this paper cites.
Adversarial Patch
Original
Tom B. Brown, Dandelion Mané, Aurko Roy, Martín Abadi, and Justin Gilmer · 2017
Earlier work this paper cites.
Formal guarantees on the robustness of a classifier against adversarial manipulation
Matthias Hein and Maksym Andriushchenko · 2017
Earlier work this paper cites.
Robust large margin deep neural networks
Jure Sokolić, Raja Giryes, Guillermo Sapiro, and Miguel RD Rodrigues · 2017
Earlier work this paper cites.
Understanding deep learning requires rethinking generalization
Chiyuan Zhang, Samy Bengio, Moritz Hardt, Benjamin Recht, and Oriol Vinyals · 2017
Earlier work this paper cites.
Towards evaluating the robustness of neural networks
Nicholas Carlini and David Wagner · 2017
Earlier work this paper cites.
Robustness may be at odds with accuracy
Original
Dimitris Tsipras, Shibani Santurkar, Logan Engstrom, Alexander Turner, and Aleksander Madry · 2018
Earlier work this paper cites.
Adversarially robust generalization requires more data
Ludwig Schmidt, Shibani Santurkar, Dimitris Tsipras, Kunal Talwar, and Aleksander Madry · 2018
Earlier work this paper cites.
Max-margin adversarial (MMA) training: Direct input space margin maximization through adversarial training
Original
Gavin Weiguang Ding, Yash Sharma, Kry Yik Chau Lui, and Ruitong Huang · 2018
Earlier work this paper cites.
Large margin deep networks for classification, 2018
Gamaleldin F. Elsayed, Dilip Krishnan, Hossein Mobahi, Kevin Regan, and Samy Bengio · 2018
Earlier work this paper cites.
Gradient descent aligns the layers of deep linear networks
Original
Ziwei Ji and Matus Telgarsky · 2018
Earlier work this paper cites.