Fetching the paper…
Reading the bibliography…
Distributed optimization methods for large-scale machine learning suffer from a communication bottleneck.
Efficient Large-Scale Distributed Training of Conditional Maximum Entropy Models
Mann, G., McDonald, R., Mohri, M., Silberman, N., and Walker, D. D · 2009
Earlier work this paper cites.
Consensus-Based Distributed Support Vector Machines
Forero, P. A., Cano, A., and Giannakis, G. B · 2010
Earlier work this paper cites.
Parallelized Stochastic Gradient Descent
Zinkevich, M. A., Weimer, M., Smola, A. J., and Li, L · 2010
Earlier work this paper cites.
Distributed optimization and statistical learning via the alternating direction method of multipliers
Boyd, S., Parikh, N., Chu, E., Peleato, B., and Eckstein, J · 2011
Earlier work this paper cites.
Hogwild!: A Lock-Free Approach to Parallelizing Stochastic Gradient Descent
Niu, F., Recht, B., Ré, C., and Wright, S. J · 2011
Earlier work this paper cites.
Solving Large Scale Linear SVM with Distributed Block Minimization
Pechyony, D., Shen, L., and Jones, R · 2011
Earlier work this paper cites.
Distributed Learning, Communication Complexity and Privacy
Balcan, M.-F., Blum, A., Fine, S., and Mansour, Y · 2012
Earlier work this paper cites.
Large Linear Classification When Data Cannot Fit in Memory
Yu, H.-F., Hsieh, C.-J., Chang, K.-W., and Lin, C.-J · 2012
Earlier work this paper cites.
Resilient Distributed Datasets: A Fault-Tolerant Abstraction for In-Memory Cluster Computing
Zaharia, M., Chowdhury, M., Das, T., Dave, A., McCauley, M., Franklin, M. J., Shenker, S., and Stoica, I · 2012
Earlier work this paper cites.
Estimation, Optimization, and Parallelism when Data is Sparse
Duchi, J. C., Jordan, M. I., and McMahan, H. B · 2013
Earlier work this paper cites.
Accelerated, parallel and proximal coordinate descent
Fercoq, O. and Richtárik, P · 2013
Earlier work this paper cites.
On the complexity analysis of randomized block-coordinate descent methods
Lu, Z. and Xiao, L · 2013
Cited alongside, same era.
Distributed coordinate descent method for learning with big data
Richtárik, P. and Takáč, M · 2013
Cited alongside, same era.
Trading Computation for Communication: Distributed Stochastic Dual Coordinate Ascent
Yang, T · 2013
Cited alongside, same era.
On Theoretical Analysis of Distributed Stochastic Dual Coordinate Ascent
Yang, T., Zhu, S., Jin, R., and Lin, Y · 2013
Cited alongside, same era.
Communication-Efficient Algorithms for Statistical Optimization
Zhang, Y., Duchi, J. C., and Wainwright, M. J · 2013
Cited alongside, same era.
LOCO: Distributing Ridge Regression with Random Projections
McWilliams, B., Heinze, C., Meinshausen, N., Krummenacher, G., and Vanchinathan, H. P · 2014
Later among the works it cites.
Coordinate descent with arbitrary sampling I: Algorithms and complexity
Qu, Z. and Richtárik, P · 2014
Later among the works it cites.
Randomized dual coordinate ascent with arbitrary sampling
Qu, Z., Richtárik, P., and Zhang, T · 2014
Later among the works it cites.
Iteration complexity of randomized block-coordinate descent methods for minimizing a composite function
Richtárik, P. and Takáč, M · 2014
Later among the works it cites.
Distributed Stochastic Optimization and Learning
Shamir, O. and Srebro, N · 2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Fast distributed coordinate descent for non-strongly convex losses
Fercoq, O., Qu, Z., Richtárik, P., and Takáč, M · 2014
Cited alongside, same era.
Communication-efficient distributed dual coordinate ascent
Jaggi, M., Smith, V., Takáč, M., Terhorst, J., Krishnan, S., Hofmann, T., and Jordan, M. I · 2014
Cited alongside, same era.
Asynchronous stochastic coordinate descent: Parallelism and convergence properties
Liu, J. and Wright, S. J · 2014
Cited alongside, same era.
An Asynchronous Parallel Stochastic Coordinate Descent Algorithm
Liu, J., Wright, S. J., Ré, C., Bittorf, V., and Sridhar, S · 2014
Cited alongside, same era.
Distributed block coordinate descent for minimizing partially separable functions
Mareček, J., Richtárik, P., and Takáč, M · 2014
Cited alongside, same era.
Accelerated mini-batch stochastic dual coordinate ascent
Shalev-Shwartz, S. and Zhang, T
Cited in the paper.
Accelerated proximal stochastic dual coordinate ascent for regularized loss minimization
Shalev-Shwartz, S. and Zhang, T
Cited in the paper.
Communication efficient distributed optimization using an approximate newton-type method
Shamir, O., Srebro, N., and Zhang, T · 2014
Later among the works it cites.
Distributed Box-Constrained Quadratic Optimization for Dual Linear SVM
Lee, C.-P. and Roth, D · 2015
Closest in time.
Parallel coordinate descent methods for big data optimization
Richtárik, P. and Takáč, M · 2015
Closest in time.
On the complexity of parallel coordinate descent
Tappenden, R., Takáč, M., and Richtárik, P · 2015
Closest in time.
DiSCO: Distributed Optimization for Self-Concordant Empirical Loss
Zhang, Y. and Lin, X · 2015
Closest in time.