Fetching the paper…
Reading the bibliography…
Data loading can dominate deep neural network training time on large-scale systems.
M. Raab and A. Steger, “Balls into bins—a simple and tight analysis,” in International Workshop on Randomization and Approximation Techniques in Computer Science . Springer, 1998, pp. 159–170
1998
Earlier work this paper cites.
X. Chen, L. Liu, Z. Liu, and T. Jiang, “On the minimum common integer partition problem,” ACM Transactions on Algorithms (TALG) , vol. 5, no. 1, p. 12, 2008
2008
Earlier work this paper cites.
O. Dekel, R. Gilad-Bachrach, O. Shamir, and L. Xiao, “Optimal distributed online prediction using mini-batches,” Journal of Machine Learning Research , vol. 13, no. Jan, pp. 165–202, 2012
2012
Earlier work this paper cites.
K. Soomro, A. Roshan Zamir, and M. Shah, “UCF101: A dataset of 101 human actions classes from videos in the wild,” in CRCV-TR-12-01 , 2012
2012
Earlier work this paper cites.
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
M. Abadi et al. , “TensorFlow: Large-scale machine learning on heterogeneous systems,” 2015, software available from tensorflow.org. [Online]. Available: https://www.tensorflow.org/
2015
Earlier work this paper cites.
O. Russakovsky et al. , “ImageNet Large Scale Visual Recognition Challenge,” International Journal of Computer Vision (IJCV) , vol. 115, no. 3, pp. 211–252, 2015
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Cited alongside, same era.
N. P. Jouppi et al. , “In-datacenter performance analysis of a tensor processing unit,” in Computer Architecture (ISCA), 2017 ACM/IEEE 44th Annual International Symposium on . IEEE, 2017, pp. 1–12
2017
Cited alongside, same era.
A. Paszke, S. Gross, S. Chintala, G. Chanan, E. Yang, Z. DeVito, Z. Lin, A. Desmaison, L. Antiga, and A. Lerer, “Automatic differentiation in pytorch,” in NIPS-W , 2017
2017
Cited alongside, same era.
2017
Cited alongside, same era.
H. Mikami, H. Suganuma, P. U-chupala, Y. Tanaka, and Y. Kageyama, “Imagenet/resnet-50 training in 224 seconds,” 2018
2018
Later among the works it cites.
T. Kurth, S. Treichler, J. Romero, M. Mudigonda, N. Luehr, E. Phillips, A. Mahesh, M. Matheson, J. Deslippe et al. , “Exascale deep learning for climate analytics,” in Proceedings of the International Conference for High Performance Computing, Networking, Storage, and Analysis . IEEE Press, 2018, p. 51
2018
Later among the works it cites.
A. Sergeev and M. D. Balso, “Horovod: fast and easy distributed deep learning in tensorflow,” 2018
2018
Later among the works it cites.
Y. Zhu, F. Chowdhury, H. Fu, A. Moody, K. Mohror, K. Sato, and W. Yu, “Entropy-aware i/o pipelining for large-scale deep learning on hpc systems,” in 2018 IEEE 26th International Symposium on Modeling, Analysis, and Simulation of Computer and Telecommunication Systems (MASCOTS) . IEEE, 2018, pp. 145–156
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. Kurth et al. , “Deep learning at 15pf,” Proceedings of the International Conference for High Performance Computing, Networking, Storage and Analysis on - SC ’17 , 2017. [Online]. Available: http://dx.doi.org/10.1145/3126908.3126916
2017
Cited alongside, same era.
Y. You, Z. Zhang, C.-J. Hsieh, J. Demmel, and K. Keutzer, “Imagenet training in minutes,” Proceedings of the 47th International Conference on Parallel Processing - ICPP 2018 , 2018. [Online]. Available: http://dx.doi.org/10.1145/3225058.3225069
2018
Cited alongside, same era.
X. Jia et al. , “Highly scalable deep learning training system with mixed-precision: Training imagenet in four minutes,” 2018
2018
Cited alongside, same era.
C. Ying, S. Kumar, D. Chen, T. Wang, and Y. Cheng, “Image classification at supercomputer scale,” 2018
2018
Cited alongside, same era.
Intel, “MKL-DNN,” https://github.com/intel/mkl-dnn , 2019
2019
Closest in time.
Python Wiki contributors, “Global interpreter lock,” 2019, [Online; accessed 01-March-2019]. [Online]. Available: https://wiki.python.org/moin/GlobalInterpreterLock
2019
Closest in time.
Facebook, “pytorch/examples,” https://github.com/pytorch/examples , 2019
2019
Closest in time.
F. Di Natale et al. , “A massively parallel infrastructure for adaptive multiscale simulations: Modeling RAS initiation pathway for cancer,” in Supercomputing: The International Conference for High Performance Computing, Networking, Storage, and Analysis (To Appear) . ACM, 2019
2019
Closest in time.