Fetching the paper…
Reading the bibliography…
We initiate the study of the inherent tradeoffs between the size of a neural network and its robustness, as measured by its Lipschitz constant.
On the capabilities of multilayer perceptrons
Eric B Baum · 1988
Earlier work this paper cites.
Approximation by superpositions of a sigmoidal function
George Cybenko · 1989
Earlier work this paper cites.
Sum of even powers of real linear forms , volume 463
Bruce Arie Reznick · 1992
Earlier work this paper cites.
Multilayer feedforward networks with a nonpolynomial activation function can approximate any function
Moshe Leshno, Vladimir Ya Lin, Allan Pinkus, and Shimon Schocken · 1993
Earlier work this paper cites.
Polynomial interpolation in several variables
James Alexander and André Hirschowitz · 1995
Earlier work this paper cites.
Interior point polynomial time methods in convex programming
Arkadi Nemirovski · 2004
Earlier work this paper cites.
Symmetric tensors and symmetric tensor rank
Pierre Comon, Gene Golub, Lek-Heng Lim, and Bernard Mourrain · 2008
Earlier work this paper cites.
Introduction to the non-asymptotic analysis of random matrices
Roman Vershynin · 2012
Earlier work this paper cites.
Intriguing properties of neural networks
Christian Szegedy, Wojciech Zaremba, Ilya Sutskever, Joan Bruna, Dumitru Erhan, Ian Goodfellow, and Rob Fergus · 2013
Cited alongside, same era.
Explaining and harnessing adversarial examples
Ian Goodfellow, Jonathon Shlens, and Christian Szegedy · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy · 2015
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2015
Cited alongside, same era.
Random version of dvoretzky’s theorem in lpn
Grigoris Paouris, Petros Valettas, and Joel Zinn · 2017
Cited alongside, same era.
Nuclear norm of higher-order tensors
Shmuel Friedland and Lek-Heng Lim · 2018
Cited alongside, same era.
Adversarially robust generalization requires more data
Ludwig Schmidt, Shibani Santurkar, Dimitris Tsipras, Kunal Talwar, and Aleksander Madry · 2018
Later among the works it cites.
Adversarial examples from computational constraints
Sébastien Bubeck, Yin Tat Lee, Eric Price, and Ilya Razenshteyn · 2019
Later among the works it cites.
Computational limitations in robust classification and win-win results
Akshay Degwekar, Preetum Nakkiran, and Vinod Vaikuntanathan · 2019
Later among the works it cites.
Adversarial training can hurt generalization
Aditi Raghunathan, Sang Michael Xie, Fanny Yang, John C Duchi, and Percy Liang · 2019
Later among the works it cites.
Small relu networks are powerful memorizers: a tight analysis of memorization capacity
Chulhee Yun, Suvrit Sra, and Ali Jadbabaie · 2019
Later among the works it cites.
Feature purification: How adversarial training performs robust deep learning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
On orthogonal tensors and best rank-one approximation ratio
Zhening Li, Yuji Nakatsukasa, Tasuku Soma, and André Uschmajew · 2018
Cited alongside, same era.
Towards deep learning models resistant to adversarial attacks
Aleksander Madry, Aleksandar Makelov, Ludwig Schmidt, Dimitris Tsipras, and Adrian Vladu · 2018
Cited alongside, same era.
Zeyuan Allen-Zhu and Yuanzhi Li · 2020
Closest in time.
A corrective view of neural networks: Representation, memorization and learning
Guy Bresler and Dheeraj Nagaraj · 2020
Closest in time.
Network size and weights size for memorization with two-layers neural networks
Sébastien Bubeck, Ronen Eldan, Yin Tat Lee, and Dan Mikulincer · 2020
Closest in time.