Fetching the paper…
Reading the bibliography…
We propose a novel robust aggregation rule for distributed synchronous Stochastic Gradient Descent~(SGD) under a general Byzantine failure model.
Time bounds for selection
M. Blum, R. W. Floyd, V. Pratt, R. L. Rivest, and R. E. Tarjan · 1973
Earlier work this paper cites.
The byzantine generals problem
L. Lamport, R. E. Shostak, and M. C. Pease · 1982
Earlier work this paper cites.
Training invariant support vector machines using selective sampling
G. Loosli, S. Canu, and L. Bottou · 2007
Earlier work this paper cites.
Learning multiple layers of features from tiny images
A. Krizhevsky and G. Hinton · 2009
Earlier work this paper cites.
Large scale distributed deep networks
J. Dean, G. S. Corrado, R. Monga, K. Chen, M. Devin, Q. V. Le, M. Z. Mao, M. Ranzato, A. W. Senior, P. A. Tucker, K. Yang, and A. Y. Ng · 2012
Earlier work this paper cites.
More effective distributed ml via a stale synchronous parallel parameter server
Q. Ho, J. Cipar, H. Cui, S. Lee, J. K. Kim, P. B. Gibbons, G. A. Gibson, G. R. Ganger, and E. P. Xing · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Cited alongside, same era.
Convex optimization: Algorithms and complexity
S. Bubeck et al · 2015
Cited alongside, same era.
Machine learning with adversaries: Byzantine tolerant gradient descent
P. Blanchard, R. Guerraoui, J. Stainer, et al · 2017
Cited alongside, same era.
Distributed statistical machine learning in adversarial settings: Byzantine gradient descent
Y. Chen, L. Su, and J. Xu · 2017
Cited alongside, same era.
A review on security issues and attacks in distributed systems
D. Harinath, P. Satyanarayana, and M. R. Murthy · 2017
Cited alongside, same era.
Scaling distributed machine learning with the parameter server
Risk analysis and countermeasure for bit-flipping attack in lorawan
J. Lee, D. Hwang, J. Park, and K.-H. Kim · 2017
Later among the works it cites.
Communication-efficient learning of deep networks from decentralized data
H. B. McMahan, E. Moore, D. Ramage, S. Hampson, and B. A. y Arcas · 2017
Later among the works it cites.
Variants of rmsprop and adagrad with logarithmic regret bounds
M. C. Mukkamala and M. Hein · 2017
Later among the works it cites.
Byzantine stochastic gradient descent
D. Alistarh, Z. Allen-Zhu, and J. Li · 2018
Closest in time.
Byzantine-robust distributed learning: Towards optimal statistical rates
D. Yin, Y. Chen, K. Ramchandran, and P. Bartlett · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
M. Li, D. G. Andersen, J. W. Park, A. J. Smola, A. Ahmed, V. Josifovski, J. Long, E. J. Shekita, and B.-Y. Su
Cited in the paper.
Communication efficient distributed machine learning with the parameter server
M. Li, D. G. Andersen, A. J. Smola, and K. Yu
Cited in the paper.