Fetching the paper…
Reading the bibliography…
There has been a strong demand for algorithms that can execute machine learning as faster as possible and the speed of deep learning has accelerated by 30 times only in the past two years.
ImageNet Classification with Deep Convolutional Neural Networks
Alex Krizhevsky, Sutskever Ilya, and Hinton Geoffrey E · 2012
Earlier work this paper cites.
Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift
S. Ioffe and C. Szegedy · 2015
Earlier work this paper cites.
Deep Residual Learning for Image Recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
Accurate, Large Minibatch SGD: Training ImageNet in 1 Hour
P. Goyal, P. Dollar, R. Girshick, P. Noordhuis, L. Wesolowski, A. Kyrola, A. Tulloch, Y. Jia, and K. He · 2017
Earlier work this paper cites.
Don’t Decay the Learning Rate, Increase the Batch Size
S. L. Smith, P.-J. Kindermans, C. Ying, and Q. V. Le · 2017
Cited alongside, same era.
Extremely Large Minibatch SGD: Training ResNet-50 on ImageNet in 15 Minutes
T. Akiba, S. Suzuki, and K. Fukuda · 2017
Cited alongside, same era.
Large Batch Training Of Convolutional Networks
Y. You, I. Gitman, and B. Ginsburg · 2017
Cited alongside, same era.
X. Jia, S. Song, W. He, Y. Wang, H. Rong, F. Zhou, L. Xie, Z. Guo, Y. Yang, L. Yu, T. Chen, G. Hu, S. Shi, and X. Chu · 2018
Cited alongside, same era.
Image Classification at Supercomputer Scale
C. Ying, S. Kumar, D. Chen, T. Wang, and Y. Cheng · 2018
Later among the works it cites.
Massively Distributed SGD: ImageNet/ResNet-50 Training in a Flash
H. Mikami, H. Suganuma, P. U-chupala, Y. Tanaka, and Y. Kageyama · 2019
Closest in time.
Rethinking the Inception Architecture for Computer Vision
Christian Szegedy, Vincent Vanhoucke, Sergey Ioffe, Jonathon Shlens, and Zbigniew Wojna · 2048
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…