Parallelized stochastic gradient descent
Martin Zinkevich, Markus Weimer, Lihong Li, and Alex J Smola · 2010
Earlier work this paper cites.
Stochastic first-and zeroth-order methods for nonconvex stochastic programming
Saeed Ghadimi and Guanghui Lan · 2013
Earlier work this paper cites.
Accelerating stochastic gradient descent using predictive variance reduction
Rie Johnson and Tong Zhang · 2013
Earlier work this paper cites.
On the importance of initialization and momentum in deep learning
Ilya Sutskever, James Martens, George Dahl, and Geoffrey Hinton · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Original
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Communication-efficient distributed optimization using an approximate newton-type method
Ohad Shamir, Nati Srebro, and Tong Zhang · 2014
Earlier work this paper cites.
Communication complexity of distributed convex learning and optimization
Yossi Arjevani and Ohad Shamir · 2015
Earlier work this paper cites.
Linear convergence of gradient and proximal-gradient methods under the polyak-łojasiewicz condition
Hamed Karimi, Julie Nutini, and Mark Schmidt · 2016
Earlier work this paper cites.
Federated optimization: Distributed machine learning for on-device intelligence
Original
Jakub Konečnỳ, H. Brendan McMahan, Daniel Ramage, and Peter Richtárik · 2016
Earlier work this paper cites.
Federated learning: Strategies for improving communication efficiency
Original
Jakub Konečnỳ, H. Brendan McMahan, Felix X. Yu, Peter Richtárik, Ananda Theertha Suresh, and Dave Bacon · 2016
Earlier work this paper cites.
Aide: Fast and communication efficient distributed optimization
Original
Sashank J. Reddi, Jakub Konečnỳ, Peter Richtárik, Barnabás Póczós, and Alex Smola · 2016
Earlier work this paper cites.
Qsgd: Communication-efficient sgd via gradient quantization and encoding
Dan Alistarh, Demjan Grubic, Jerry Li, Ryota Tomioka, and Milan Vojnovic · 2017
Earlier work this paper cites.
Practical secure aggregation for privacy-preserving machine learning
Keith Bonawitz, Vladimir Ivanov, Ben Kreuter, Antonio Marcedone, H. Brendan McMahan, Sarvar Patel, Daniel Ramage, Aaron Segal, and Karn Seth · 2017
Earlier work this paper cites.
Emnist: Extending mnist to handwritten letters
Gregory Cohen, Saeed Afshar, Jonathan Tapson, and Andre Van Schaik · 2017
Earlier work this paper cites.
Differentially private federated learning: A client level perspective
Original
Robin C Geyer, Tassilo Klein, and Moin Nabi · 2017
Earlier work this paper cites.
Less than a single pass: Stochastically controlled stochastic gradient
Lihua Lei and Michael Jordan · 2017
Earlier work this paper cites.
Communication-efficient learning of deep networks from decentralized data
Brendan McMahan, Eider Moore, Daniel Ramage, Seth Hampson, and Blaise Agüera y Arcas · 2017
Earlier work this paper cites.
Distributed mean estimation with limited communication
Ananda Theertha Suresh, Felix X. Yu, Sanjiv Kumar, and H. Brendan McMahan · 2017
Earlier work this paper cites.
Large batch training of convolutional networks
Original
Yang You, Igor Gitman, and Boris Ginsburg · 2017
Earlier work this paper cites.
cpSGD: Communication-efficient and differentially-private distributed SGD
Naman Agarwal, Ananda Theertha Suresh, Felix X. Yu, Sanjiv Kumar, and Brendan McMahan · 2018
Earlier work this paper cites.
Expanding the reach of federated learning by reducing client resource requirements
Original
Sebastian Caldas, Jakub Konečny, H Brendan McMahan, and Ameet Talwalkar · 2018
Earlier work this paper cites.
Leaf: A benchmark for federated settings
Original
Sebastian Caldas, Peter Wu, Tian Li, Jakub Konečnỳ, H Brendan McMahan, Virginia Smith, and Ameet Talwalkar · 2018
Earlier work this paper cites.
Compiling machine learning programs via high-level tracing
Roy Frostig, Matthew James Johnson, and Chris Leary · 2018
Earlier work this paper cites.
On the convergence of federated optimization in heterogeneous networks
Original
Tian Li, Anit Kumar Sahu, Maziar Sanjabi, Manzil Zaheer, Ameet Talwalkar, and Virginia Smith · 2018
Earlier work this paper cites.
Lectures on convex optimization
Yurii Nesterov · 2018
Earlier work this paper cites.