Gradient-based learning applied to document recognition
Y. Lecun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Model selection and estimation in regression with grouped variables
M. Yuan and Y. Lin · 2006
Earlier work this paper cites.
A fast iterative shrinkage-thresholding algorithm for linear inverse problems
A. Beck and M. Teboulle · 2009
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L. Li, K. Li, and F.-F. Li · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
A. Krizhevsky and G. Hinton · 2009
Earlier work this paper cites.
Distributed optimization and statistical learning via the alternating direction method of multipliers
S. Boyd, N. Parikh, E. Chu, B. Peleato, J. Eckstein, et al · 2011
Earlier work this paper cites.
Optimization with sparsity-inducing penalties
F. Bach, R. Jenatton, J. Mairal, and G. Obozinski · 2012
Earlier work this paper cites.
Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups
G. Hinton, L. Deng, D. Yu, G. Dahl, A.-r. Mohamed, N. Jaitly, A. Senior, V. Vanhoucke, P. Nguyen, T. Sainath, et al · 2012
Earlier work this paper cites.
Stochastic alternating direction method of multipliers
H. Ouyang, N. He, L. Tran, and A. Gray · 2013
Earlier work this paper cites.
Dual averaging and proximal gradient descent for online alternating direction multiplier method
T. Suzuki · 2013
Earlier work this paper cites.
Intriguing properties of neural networks
Original
C. Szegedy, W. Zaremba, I. Sutskever, J. Bruna, D. Erhan, I. Goodfellow, and R. Fergus · 2013
Earlier work this paper cites.
Explaining and harnessing adversarial examples
Original
I. Goodfellow, J. Shlens, and C. Szegedy · 2014
Earlier work this paper cites.
Proximal algorithms
N. Parikh, S. Boyd, et al · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Original
D. Kingma and J. Ba · 2015
Earlier work this paper cites.
Sparsity-aware sensor collaboration for linear coherent estimation
S. Liu, S. Kar, M. Fardad, and P. K. Varshney · 2015
Earlier work this paper cites.
Deep neural networks are easily fooled: High confidence predictions for unrecognizable images
A. Nguyen, J. Yosinski, and J. Clune · 2015
Earlier work this paper cites.