Fetching the paper…
Reading the bibliography…
We propose three new robust aggregation rules for distributed synchronous Stochastic Gradient Descent~(SGD) under a general Byzantine failure model.
Time bounds for selection
Blum, Manuel, Floyd, Robert W, Pratt, Vaughan, Rivest, Ronald L, and Tarjan, Robert E · 1973
Earlier work this paper cites.
The byzantine generals problem
Lamport, Leslie, Shostak, Robert E., and Pease, Marshall C · 1982
Earlier work this paper cites.
Distributed algorithms
Lynch, Nancy A · 1996
Earlier work this paper cites.
Training invariant support vector machines using selective sampling
Loosli, Gaëlle, Canu, Stéphane, and Bottou, Léon · 2007
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, Alex and Hinton, Geoffrey · 2009
Earlier work this paper cites.
Large scale distributed deep networks
Dean, Jeffrey, Corrado, Gregory S., Monga, Rajat, Chen, Kai, Devin, Matthieu, Le, Quoc V., Mao, Mark Z., Ranzato, Marc’Aurelio, Senior, Andrew W., Tucker, Paul A., Yang, Ke, and Ng, Andrew Y · 2012
Earlier work this paper cites.
More effective distributed ml via a stale synchronous parallel parameter server
Ho, Qirong, Cipar, James, Cui, Henggang, Lee, Seunghak, Kim, Jin Kyu, Gibbons, Phillip B., Gibson, Garth A., Ganger, Gregory R., and Xing, Eric P · 2013
Cited alongside, same era.
Adam: A method for stochastic optimization
Kingma, Diederik P. and Ba, Jimmy · 2014
Cited alongside, same era.
Mxnet: A flexible and efficient machine learning library for heterogeneous distributed systems
Chen, Tianqi, Li, Mu, Li, Yutian, Lin, Min, Wang, Naiyan, Wang, Minjie, Xiao, Tianjun, Xu, Bing, Zhang, Chiyuan, and Zhang, Zheng · 2015
Cited alongside, same era.
Geometric median and robust estimation in banach spaces
Minsker, Stanislav et al · 2015
Cited alongside, same era.
Tensorflow: A system for large-scale machine learning
Abadi, Martín, Barham, Paul, Chen, Jianmin, Chen, Zhifeng, Davis, Andy, Dean, Jeffrey, Devin, Matthieu, Ghemawat, Sanjay, Irving, Geoffrey, Isard, Michael, Kudlur, Manjunath, Levenberg, Josh, Monga, Rajat, Moore, Sherry, Murray, Derek Gordon, Steiner, Benoit, Tucker, Paul A., Vasudevan, Vijay, Warden, Pete, Wicke, Martin, Yu, Yuan, and Zhang, Xiaoqiang · 2016
Cntk: Microsoft’s open-source deep-learning toolkit
Seide, Frank and Agarwal, Amit · 2016
Later among the works it cites.
Machine learning with adversaries: Byzantine tolerant gradient descent
Blanchard, Peva, Guerraoui, Rachid, Stainer, Julien, et al · 2017
Later among the works it cites.
Distributed statistical machine learning in adversarial settings: Byzantine gradient descent
Chen, Yudong, Su, Lili, and Xu, Jiaming · 2017
Later among the works it cites.
A review on security issues and attacks in distributed systems
Harinath, Depavath, Satyanarayana, P, and Murthy, MV Ramana · 2017
Later among the works it cites.
Communication-efficient learning of deep networks from decentralized data
McMahan, H. Brendan, Moore, Eider, Ramage, Daniel, Hampson, Seth, and y Arcas, Blaise Aguera · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Geometric median in nearly linear time
Cohen, Michael B, Lee, Yin Tat, Miller, Gary, Pachocki, Jakub, and Sidford, Aaron · 2016
Cited alongside, same era.
Scaling distributed machine learning with the parameter server
Li, Mu, Andersen, David G., Park, Jun Woo, Smola, Alexander J., Ahmed, Amr, Josifovski, Vanja, Long, James, Shekita, Eugene J., and Su, Bor-Yiing
Cited in the paper.
Communication efficient distributed machine learning with the parameter server
Li, Mu, Andersen, David G., Smola, Alexander J., and Yu, Kai
Cited in the paper.
Mukkamala, Mahesh Chandra and Hein, Matthias · 2017
Later among the works it cites.