Fetching the paper…
Reading the bibliography…
The residual neural network (ResNet) is a popular deep network architecture which has the ability to obtain high-accuracy results on several image processing problems.
Splitting algorithms for the sum of two nonlinear operators
Pierre-Louis Lions and Bertrand Mercier · 1979
Earlier work this paper cites.
Backpropagation applied to handwritten zip code recognition
Yann LeCun, Bernhard Boser, John S. Denker, Donnie Henderson, Richard E. Howard, Wayne Hubbard, and Lawrence D. Jackel · 1989
Earlier work this paper cites.
Learning long-term dependencies with gradient descent is difficult
Yoshua Bengio, Patrice Simard, and Paolo Frasconi · 1994
Earlier work this paper cites.
Nonsmooth sequential analysis in asplund spaces
Boris S. Mordukhovich and Yongheng Shao · 1996
Earlier work this paper cites.
Prox-regular functions in variational analysis
R. A. Poliquin and R. T. Rockafellar · 1996
Earlier work this paper cites.
Some Gronwall type inequalities and applications
Sever Silvestru Dragomir · 2003
Earlier work this paper cites.
Relaxation of an optimal control problem involving a perturbed sweeping process
Jean Fenel Edmond and Lionel Thibault · 2005
Earlier work this paper cites.
Learning deep architectures for AI
Yoshua Bengio · 2009
Earlier work this paper cites.
Efficient learning using forward-backward splitting
Yoram Singer and John C. Duchi · 2009
Earlier work this paper cites.
ImageNet classification with deep convolutional neural networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E. Hinton · 2012
Earlier work this paper cites.
Evasion attacks against machine learning at test time
Battista Biggio, Igino Corona, Davide Maiorca, Blaine Nelson, Nedim Šrndić, Pavel Laskov, Giorgio Giacinto, and Fabio Roli · 2013
Earlier work this paper cites.
Intriguing properties of neural networks
Christian Szegedy, Wojciech Zaremba, Ilya Sutskever, Joan Bruna, Dumitru Erhan, Ian Goodfellow, and Rob Fergus · 2013
Earlier work this paper cites.
A field guide to forward-backward splitting with a FASTA implementation
Tom Goldstein, Christoph Studer, and Richard Baraniuk · 2014
Earlier work this paper cites.
Generative adversarial nets
Ian J. Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman · 2014
Earlier work this paper cites.
Random walk initialization for training very deep feedforward networks
David Sussillo and L. F. Abbott · 2014
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2015
Cited alongside, same era.
Delving deep into rectifiers: Surpassing human-level performance on ImageNet classification
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy · 2015
Cited alongside, same era.
Deep learning
Yann LeCun, Yoshua Bengio, and Geoffrey Hinton · 2015
Cited alongside, same era.
ImageNet large scale visual recognition challenge
Olga Russakovsky, Jia Deng, Hao Su, Jonathan Krause, Sanjeev Satheesh, Sean Ma, Zhiheng Huang, Andrej Karpathy, Aditya Khosla, Michael Bernstein, Alexander C. Berg, and Li Fei-Fei · 2015
Cited alongside, same era.
Going deeper with convolutions
Christian Szegedy, Wei Liu, Yangqing Jia, Pierre Sermanet, Scott Reed, Dragomir Anguelov, Dumitru Erhan, Vincent Vanhoucke, and Andrew Rabinovich · 2015
The reversible residual network: Backpropagation without storing activations
Aidan N. Gomez, Mengye Ren, Raquel Urtasun, and Roger B. Grosse · 2017
Later among the works it cites.
Stable architectures for deep neural networks
Eldad Haber and Lars Ruthotto · 2017
Later among the works it cites.
Global stability of almost periodic solutions of monotone sweeping processes and their response to non-monotone perturbations
Mikhail Kamenskii, Oleg Makarenkov, Lakmi Niwanthi, and Paul Raynaud de Fitte · 2017
Later among the works it cites.
Visualizing the loss landscape of neural nets
Hao Li, Zheng Xu, Gavin Taylor, and Tom Goldstein · 2017
Later among the works it cites.
Deep residual learning and PDEs on manifold
Zhen Li and Zuoqiang Shi · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
An l 1 l^{1} penalty method for general obstacle problems
Giang Tran, Hayden Schaeffer, William M. Feldman, and Stanley J. Osher · 2015
Cited alongside, same era.
Entropy-SGD: Biasing gradient descent into wide valleys
Pratik Chaudhari, Anna Choromanska, Stefano Soatto, Yann LeCun, Carlo Baldassi, Christian Borgs, Jennifer Chayes, Levent Sagun, and Riccardo Zecchina · 2016
Cited alongside, same era.
Identity mappings in deep residual networks
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Cited alongside, same era.
Densely connected convolutional networks
Gao Huang, Zhuang Liu, Laurens van der Maaten, and Kilian Q. Weinberger · 2016
Cited alongside, same era.
On large-batch training for deep learning: Generalization gap and sharp minima
Nitish Shirish Keskar, Dheevatsa Mudigere, Jorge Nocedal, Mikhail Smelyanskiy, and Ping Tak Peter Tang · 2016
Cited alongside, same era.
FractalNet: Ultra-deep neural networks without residuals
Gustav Larsson, Michael Maire, and Gregory Shakhnarovich · 2016
Cited alongside, same era.
Mathematics of deep learning
Rene Vidal, Joan Bruna, Raja Giryes, and Stefano Soatto · 2017
Later among the works it cites.
Optimization methods for large-scale machine learning
Léon Bottou, Frank E. Curtis, and Jorge Nocedal · 2018
Closest in time.
Deep relaxation: partial differential equations for optimizing deep neural networks
Pratik Chaudhari, Adam Oberman, Stanley Osher, Stefano Soatto, and Guillaume Carlier · 2018
Closest in time.
Gradient descent provably optimizes over-parameterized neural networks
Simon S. Du, Xiyu Zhai, Barnabas Poczos, and Aarti Singh · 2018
Closest in time.
A mean-field optimal control formulation of deep learning
Weinan E, Jiequn Han, and Qianxiao Li · 2018
Closest in time.
Lipschitz regularized deep neural networks converge and generalize
Adam M. Oberman and Jeff Calder · 2018
Closest in time.
Deep neural networks motivated by partial differential equations
Lars Ruthotto and Eldad Haber · 2018
Closest in time.
A penalty method for some nonlinear variational obstacle problems
Hayden Schaeffer · 2018
Closest in time.
Deep Limits of Residual Neural Networks
Matthew Thorpe and Yves van Gennip · 2018
Closest in time.
Deep neural nets with interpolating function as output activation
Bao Wang, Xiyang Luo, Zhen Li, Wei Zhu, Zuoqiang Shi, and Stanley J. Osher · 2018
Closest in time.