Fine-grained analysis of optimization and generalization for overparameterized two-layer neural networks
Original
Sanjeev Arora, Simon S. Du, Wei Hu, Zhiyuan Li, and Ruosong Wang · 1901
Earlier work this paper cites.
Can SGD Learn Recurrent Neural Networks with Provable Generalization?
Original
Zeyuan Allen-Zhu and Yuanzhi Li · 1902
Earlier work this paper cites.
On exact computation with an infinitely wide neural net
Original
Sanjeev Arora, Simon S Du, Wei Hu, Zhiyuan Li, Ruslan Salakhutdinov, and Ruosong Wang · 1904
Earlier work this paper cites.
What Can ResNet Learn Efficiently, Going Beyond Kernels?
Original
Zeyuan Allen-Zhu and Yuanzhi Li · 1905
Earlier work this paper cites.
High frequency component helps explain the generalization of convolutional neural networks
Original
Haohan Wang, Xindi Wu, Pengcheng Yin, and Eric P Xing · 1905
Earlier work this paper cites.
On a lemma of littlewood and offord
Paul Erdös · 1945
Earlier work this paper cites.
Polynomial threshold functions, ac
Jehoshua Bruck and Roman Smolensky · 1992
Earlier work this paper cites.
The polynomial method in circuit complexity
Richard Beigel · 1993
Earlier work this paper cites.
Spectral properties of threshold functions
Craig Gotsman and Nathan Linial · 1994
Earlier work this paper cites.
Sparse coding with an overcomplete basis set: A strategy employed by v1?
Bruno A Olshausen and David J Field · 1997
Earlier work this paper cites.
Sparse coding and decorrelation in primary visual cortex during natural vision
William E Vinje and Jack L Gallant · 2000
Earlier work this paper cites.
Non-negative sparse coding
Patrik O Hoyer · 2002
Earlier work this paper cites.
Sparse coding of sensory inputs
Bruno A Olshausen and David J Field · 2004
Earlier work this paper cites.
Stochastic gradient descent optimizes over-parameterized deep relu networks
Original
Difan Zou, Yuan Cao, Dongruo Zhou, and Quanquan Gu · 2005
Earlier work this paper cites.
Efficient sparse coding algorithms
Honglak Lee, Alexis Battle, Rajat Raina, and Andrew Y Ng · 2007
Earlier work this paper cites.
Visualizing higher-layer features of a deep network
Dumitru Erhan, Y. Bengio, Aaron Courville, and Pascal Vincent · 2009
Earlier work this paper cites.
Online dictionary learning for sparse coding
Julien Mairal, Francis Bach, Jean Ponce, and Guillermo Sapiro · 2009
Earlier work this paper cites.
Linear spatial pyramid matching using sparse coding for image classification
Jianchao Yang, Kai Yu, Yihong Gong, and Thomas Huang · 2009
Earlier work this paper cites.
Learning fast approximations of sparse coding
Karol Gregor and Yann LeCun · 2010
Earlier work this paper cites.
Online learning for matrix factorization and sparse coding
Julien Mairal, Francis Bach, Jean Ponce, and Guillermo Sapiro · 2010
Earlier work this paper cites.
New degree bounds for polynomial threshold functions
Ryan O’Donnell and Rocco A Servedio · 2010
Earlier work this paper cites.
The sign-rank of ac
Alexander A Razborov and Alexander A Sherstov · 2010
Earlier work this paper cites.
Robust sparse coding for face recognition
Meng Yang, Lei Zhang, Jian Yang, and David Zhang · 2011
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton · 2012
Earlier work this paper cites.
Exact recovery of sparsely-used dictionaries
Daniel A Spielman, Huan Wang, and John Wright · 2012
Earlier work this paper cites.
Exact recovery of sparsely used overcomplete dictionaries
Alekh Agarwal, Animashree Anandkumar, and Praneeth Netrapalli · 2013
Earlier work this paper cites.
Evasion attacks against machine learning at test time
Battista Biggio, Igino Corona, Davide Maiorca, Blaine Nelson, Nedim Šrndić, Pavel Laskov, Giorgio Giacinto, and Fabio Roli · 2013
Earlier work this paper cites.
Speech recognition with deep recurrent neural networks
Alex Graves, Abdel-rahman Mohamed, and Geoffrey Hinton · 2013
Earlier work this paper cites.
Intriguing properties of neural networks
Original
Christian Szegedy, Wojciech Zaremba, Ilya Sutskever, Joan Bruna, Dumitru Erhan, Ian Goodfellow, and Rob Fergus · 2013
Earlier work this paper cites.
New algorithms for learning incoherent and overcomplete dictionaries
Sanjeev Arora, Rong Ge, and Ankur Moitra · 2014
Earlier work this paper cites.
On the local correctness of
Quan Geng and John Wright · 2014
Earlier work this paper cites.
Explaining and harnessing adversarial examples
Original
Ian J Goodfellow, Jonathon Shlens, and Christian Szegedy · 2014
Earlier work this paper cites.
On the identifiability of overcomplete dictionaries via the minimisation principle underlying k-svd
Karin Schnass · 2014
Earlier work this paper cites.
Simple, efficient, and neural algorithms for sparse coding
Sanjeev Arora, Rong Ge, Tengyu Ma, and Ankur Moitra · 2015
Earlier work this paper cites.
Dictionary learning and tensor decomposition via the sum-of-squares method
Boaz Barak, Jonathan A Kelner, and David Steurer · 2015
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Original
Sergey Ioffe and Christian Szegedy · 2015
Earlier work this paper cites.
Understanding deep image representations by inverting them
Aravindh Mahendran and Andrea Vedaldi · 2015
Earlier work this paper cites.