2019

Asynchronous Federated Optimization

Xie, Cong, Koyejo, Sanmi, Gupta, Indranil

Understand

Federated learning enables training on a massive number of edge devices.

  • To improve flexibility and scalability, we propose a new asynchronous federated optimization algorithm.
  • We prove that the proposed approach has near-linear convergence to a global optimum, for both strongly convex and a restricted family of non-convex problems.
  • Empirical results show that the proposed algorithm converges quickly and tolerates staleness in various applications.

Built on

  • Health insurance portability and accountability act of 1996

    Steve Anderson: HealthInsurance.org · 1996

    Earlier work this paper cites.

  • Learning multiple layers of features from tiny images

    Alex Krizhevsky and Geoffrey Hinton · 2009

    Earlier work this paper cites.

  • Slow learners are fast

    Martin Zinkevich, John Langford, and Alex J Smola · 2009

    Earlier work this paper cites.

  • More effective distributed ml via a stale synchronous parallel parameter server

    Qirong Ho, James Cipar, Henggang Cui, Seunghak Lee, Jin Kyu Kim, Phillip B Gibbons, Garth A Gibson, Greg Ganger, and Eric P Xing · 2013

    Earlier work this paper cites.

  • Federated learning: Strategies for improving communication efficiency

    Original

    Jakub Konevcnỳ, H Brendan McMahan, Felix X Yu, Peter Richtárik, Ananda Theertha Suresh, and Dave Bacon · 2016

    Earlier work this paper cites.

Similar

  • Communication-efficient learning of deep networks from decentralized data

    Original

    H Brendan McMahan, Eider Moore, Daniel Ramage, Seth Hampson, et al · 2016

    Cited alongside, same era.

  • Pointer sentinel mixture models

    Original

    Stephen Merity, Caiming Xiong, James Bradbury, and Richard Socher · 2016

    Cited alongside, same era.

  • Practical secure aggregation for privacy-preserving machine learning

    Keith Bonawitz, Vladimir Ivanov, Ben Kreuter, Antonio Marcedone, H Brendan McMahan, Sarvar Patel, Daniel Ramage, Aaron Segal, and Karn Seth · 2017

    Cited alongside, same era.

  • Asynchronous decentralized parallel stochastic gradient descent

    Original

    Xiangru Lian, Wei Zhang, Ce Zhang, and Ji Liu · 2017

    Cited alongside, same era.

  • Scaling distributed machine learning with the parameter server

    Mu Li, David G Andersen, Jun Woo Park, Alexander J Smola, Amr Ahmed, Vanja Josifovski, James Long, Eugene J Shekita, and Bor-Yiing Su

    Cited in the paper.

  • Communication efficient distributed machine learning with the parameter server

    Mu Li, David G Andersen, Alexander J Smola, and Kai Yu

    Cited in the paper.

Then

  • Asynchronous stochastic gradient descent with delay compensation

    Shuxin Zheng, Qi Meng, Taifeng Wang, Wei Chen, Nenghai Yu, Zhi-Ming Ma, and Tie-Yan Liu · 2017

    Later among the works it cites.

  • European Union’s General Data Protection Regulation (GDPR)

    EU · 2018

    Later among the works it cites.

  • Towards federated learning at scale: System design

    Original

    Keith Bonawitz, Hubert Eichner, Wolfgang Grieskamp, Dzmitry Huba, Alex Ingerman, Vladimir Ivanov, Chloe Kiddon, Jakub Konecny, Stefano Mazzocchi, H Brendan McMahan, et al · 2019

    Closest in time.

  • Family Educational Rights and Privacy Act (FERPA)

    US Department of Education · 2019

    Closest in time.

Beyond the bibliography

alphaXiv searches the wider corpus for related work and actual follow-ups.

Open on alphaXiv

alphaXiv is searching for related work…