Order statistics
David, H. A. and Nagaraja, H. N · 2003
Earlier work this paper cites.
MPI for python
Dalcín, L., Paz, R., and Storti, M · 2005
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, A · 2009
Earlier work this paper cites.
Large scale distributed deep networks
Dean, J., Corrado, G., Monga, R., Chen, K., Devin, M., Mao, M., Senior, A., Tucker, P., Yang, K., Le, Q. V., et al · 2012
Earlier work this paper cites.
Optimal distributed online prediction using mini-batches
Dekel, O., Gilad-Bachrach, R., Shamir, O., and Xiao, L · 2012
Earlier work this paper cites.
Stochastic first-and zeroth-order methods for nonconvex stochastic programming
Ghadimi, S. and Lan, G · 2013
Earlier work this paper cites.
Exploiting bounded staleness to speed up big data analytics
Cui, H., Cipar, J., Ho, Q., Kim, J. K., Lee, S., Kumar, A., Wei, J., Dai, W., Ganger, G. R., Gibbons, P. B., et al · 2014
Earlier work this paper cites.
Scaling distributed machine learning with the parameter server
Li, M., Andersen, D. G., Park, J. W., Smola, A. J., Ahmed, A., Josifovski, V., Long, J., Shekita, E. J., and Su, B.-Y · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Original
Simonyan, K. and Zisserman, A · 2014
Earlier work this paper cites.
SparkNet: Training deep networks in spark
Original
Moritz, P., Nishihara, R., Stoica, I., and Jordan, M. I · 2015
Earlier work this paper cites.
Experiments on parallel training of deep neural network using model averaging
Original
Su, H. and Chen, H · 2015
Earlier work this paper cites.
Deep learning with elastic averaging SGD
Zhang, S., Choromanska, A. E., and LeCun, Y · 2015
Earlier work this paper cites.