Fetching the paper…
Reading the bibliography…
While training a machine learning model using multiple workers, each of which collects data from their own data sources, it would be most useful when the data collected from different workers can be {\em unique} and {\em different}.
Quantized consensus
A. Kashyap, T. Başar, and R. Srikant · 2007
Earlier work this paper cites.
Distributed subgradient methods for multi-agent optimization
A. Nedic and A. Ozdaglar · 2009
Earlier work this paper cites.
On distributed averaging algorithms and quantization effects
A. Nedic, A. Olshevsky, A. Ozdaglar, and J. N. Tsitsiklis · 2009
Earlier work this paper cites.
Robust stochastic approximation approach to stochastic programming
A. Nemirovski, A. Juditsky, G. Lan, and A. Shapiro · 2009
Earlier work this paper cites.
Non-asymptotic analysis of stochastic approximation algorithms for machine learning
E. Moulines and F. R. Bach · 2011
Earlier work this paper cites.
Dual averaging for distributed optimization: Convergence analysis and network scaling
J. C. Duchi, A. Agarwal, and M. J. Wainwright · 2012
Earlier work this paper cites.
Quantized consensus by means of gossip algorithm
J. Lavaei and R. M. Murray · 2012
Earlier work this paper cites.
Stochastic first- and zeroth-order methods for nonconvex stochastic programming
S. Ghadimi and G. Lan · 2013
Earlier work this paper cites.
Accelerating stochastic gradient descent using predictive variance reduction
R. Johnson and T. Zhang · 2013
Earlier work this paper cites.
Saga: A fast incremental gradient method with support for non-strongly convex composite objectives
A. Defazio, F. Bach, and S. Lacoste-Julien · 2014
Earlier work this paper cites.
On the linear convergence of the admm in decentralized consensus optimization
W. Shi, Q. Ling, K. Yuan, G. Wu, and W. Yin · 2014
Earlier work this paper cites.
Mxnet: A flexible and efficient machine learning library for heterogeneous distributed systems
T. Chen, M. Li, Y. Li, M. Lin, N. Wang, M. Wang, T. Xiao, B. Xu, C. Zhang, and Z. Zhang · 2015
Cited alongside, same era.
Incremental majorization-minimization optimization with application to large-scale machine learning
J. Mairal · 2015
Cited alongside, same era.
Decentralized double stochastic averaging gradient
A. Mokhtari and A. Ribeiro · 2015
Cited alongside, same era.
Tensorflow: A system for large-scale machine learning
M. Abadi, P. Barham, J. Chen, Z. Chen, A. Davis, J. Dean, M. Devin, S. Ghemawat, G. Irving, M. Isard, et al · 2016
Cited alongside, same era.
Gossip dual averaging for decentralized optimization of pairwise functions
I. Colin, A. Bellet, J. Salmon, and S. Clémençon · 2016
Cited alongside, same era.
Z. Li and M. Yan · 2017
Later among the works it cites.
Z. Li, W. Shi, and M. Yan · 2017
Later among the works it cites.
Dynamic safe interruptibility for decentralized multi-agent reinforcement learning
E. Mhamdi, E. Mahdi, H. Hendrikx, R. Guerraoui, and A. D. O. Maurer · 2017
Later among the works it cites.
Network topology and communication-computation tradeoffs in decentralized optimization
A. Nedić, A. Olshevsky, and M. G. Rabbat · 2017
Later among the works it cites.
Deep decentralized multi-task multi-agent rl under partial observability
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Konečnỳ, J. Liu, P. Richtárik, and M. Takáč · 2016
Cited alongside, same era.
Dsa: Decentralized double stochastic averaging gradient algorithm
A. Mokhtari and A. Ribeiro · 2016
Cited alongside, same era.
Cntk: Microsoft’s open-source deep-learning toolkit
F. Seide and A. Agarwal · 2016
Cited alongside, same era.
On the convergence of decentralized gradient descent
K. Yuan, Q. Ling, and W. Yin · 2016
Cited alongside, same era.
Fully decentralized policies for multi-agent systems: An information theoretic approach
R. Dobbe, D. Fridovich-Keil, and C. Tomlin · 2017
Cited alongside, same era.
Communication-efficient algorithms for decentralized and stochastic optimization
G. Lan, S. Lee, and Y. Zhou · 2017
Cited alongside, same era.
Can decentralized algorithms outperform centralized algorithms? a case study for decentralized parallel stochastic gradient descent
X. Lian, C. Zhang, H. Zhang, C.-J. Hsieh, W. Zhang, and J. Liu
Cited in the paper.
S. Omidshafiei, J. Pazis, C. Amato, J. P. How, and J. Vian · 2017
Later among the works it cites.
Minimizing finite sums with the stochastic average gradient
M. Schmidt, N. Le Roux, and F. Bach · 2017
Later among the works it cites.
Distributed online optimization in dynamic environments using mirror descent
S. Shahrampour and A. Jadbabaie · 2017
Later among the works it cites.
Distributed mean estimation with limited communication
A. T. Suresh, F. X. Yu, S. Kumar, and H. B. McMahan · 2017
Later among the works it cites.
Exact diffusion for distributed optimization and learning—part i: Algorithm development
K. Yuan, B. Ying, X. Zhao, and A. H. Sayed · 2017
Later among the works it cites.
Projection-free distributed online learning in networks
W. Zhang, P. Zhao, W. Zhu, S. C. Hoi, and T. Zhang · 2017
Later among the works it cites.