Fetching the paper…
Reading the bibliography…
Deep residual networks were shown to be able to scale up to thousands of layers and still have improving performance.
Learning complex, extended sequences using the principle of history compression
J. Schmidhuber · 1992
Earlier work this paper cites.
Scaling learning algorithms towards AI
Yoshua Bengio and Yann LeCun · 2007
Earlier work this paper cites.
An empirical evaluation of deep architectures on problems with many factors of variation
Hugo Larochelle, Dumitru Erhan, Aaron Courville, James Bergstra, and Yoshua Bengio · 2007
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Yoshua Bengio and Xavier Glorot · 2010
Earlier work this paper cites.
Torch7: A matlab-like environment for machine learning
R. Collobert, K. Kavukcuoglu, and C. Farabet · 2011
Earlier work this paper cites.
Deep learning made easier by linear transformations in perceptrons
Tapani Raiko, Harri Valpola, and Yann Lecun · 2012
Earlier work this paper cites.
Maxout networks
Ian J. Goodfellow, David Warde-Farley, Mehdi Mirza, Aaron Courville, and Yoshua Bengio · 2013
Earlier work this paper cites.
Min Lin, Qiang Chen, and Shuicheng Yan · 2013
Earlier work this paper cites.
On the importance of initialization and momentum in deep learning
Ilya Sutskever, James Martens, George E. Dahl, and Geoffrey E. Hinton · 2013
Earlier work this paper cites.
On the complexity of shallow and deep neural network classifiers
Monica Bianchini and Franco Scarselli · 2014
Earlier work this paper cites.
Benjamin Graham · 2014
Cited alongside, same era.
Deeply-Supervised Nets
C.-Y. Lee, S. Xie, P. Gallagher, Z. Zhang, and Z. Tu · 2014
Cited alongside, same era.
On the number of linear regions of deep neural networks
Guido F. Montúfar, Razvan Pascanu, KyungHyun Cho, and Yoshua Bengio · 2014
Cited alongside, same era.
FitNets: Hints for thin deep nets
Adriana Romero, Nicolas Ballas, Samira Ebrahimi Kahou, Antoine Chassang, Carlo Gatta, and Yoshua Bengio · 2014
Cited alongside, same era.
Dropout: A simple way to prevent neural networks from overfitting
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov · 2014
Cited alongside, same era.
Fast and accurate deep network learning by exponential linear units (elus)
Going deeper with convolutions
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich · 2015
Later among the works it cites.
Net2net: Accelerating learning via knowledge transfer
T. Chen, I. Goodfellow, and J. Shlens · 2016
Closest in time.
Locnet: Improving localization accuracy for object detection
Spyros Gidaris and Nikos Komodakis · 2016
Closest in time.
Training and investigating residual nets, 2016
Sam Gross and Michael Wilber · 2016
Closest in time.
Identity mappings in deep residual networks
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Djork-Arné Clevert, Thomas Unterthiner, and Sepp Hochreiter · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy · 2015
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2015
Cited alongside, same era.
Rupesh Kumar Srivastava, Klaus Greff, and Jürgen Schmidhuber · 2015
Cited alongside, same era.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun
Cited in the paper.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun
Cited in the paper.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. Hinton
Cited in the paper.
Gao Huang, Yu Sun, Zhuang Liu, Daniel Sedra, and Kilian Q. Weinberger · 2016
Closest in time.
Optnet - reducing memory usage in torch neural networks, 2016
Francisco Massa · 2016
Closest in time.
Inception-v4, inception-resnet and the impact of residual connections on learning
Christian Szegedy, Sergey Ioffe, and Vincent Vanhoucke · 2016
Closest in time.
A multipath network for object detection
S. Zagoruyko, A. Lerer, T.-Y. Lin, P. O. Pinheiro, S. Gross, S. Chintala, and P. Dollár · 2016
Closest in time.