Fetching the paper…
Reading the bibliography…
Stochastic algorithms are efficient approaches to solving machine learning and optimization problems.
Expectation propagation for approximate Bayesian inference
T. P. Minka · 2001
Earlier work this paper cites.
Latent Dirichlet allocation
D. M. Blei, A. Y. Ng, and M. I. Jordan · 2003
Earlier work this paper cites.
Finding scientific topics
T. L. Griffiths and M. Steyvers · 2004
Earlier work this paper cites.
Solving large scale linear prediction problems using stochastic gradient descent algorithms
T. Zhang · 2004
Earlier work this paper cites.
Numerical Optimization
J. Nocedal and S. J. Wright · 2006
Earlier work this paper cites.
The netflix prize
J. Bennett and S. Lanning · 2007
Earlier work this paper cites.
Training invariant support vector machines using selective sampling
G. Loosli, S. Canu, and L. Bottou · 2007
Earlier work this paper cites.
Distributed inference for latent Dirichlet allocation
D. Newman, P. Smyth, M. Welling, and A. U. Asuncion · 2007
Earlier work this paper cites.
Mapreduce: simplified data processing on large clusters
J. Dean and S. Ghemawat · 2008
Earlier work this paper cites.
Extracting and composing robust features with denoising autoencoders
P. Vincent, H. Larochelle, Y. Bengio, and P.-A. Manzagol · 2008
Earlier work this paper cites.
Matrix factorization techniques for recommender systems
Y. Koren, R. Bell, and C. Volinsky · 2009
Earlier work this paper cites.
Dual averaging method for regularized stochastic learning and online optimization
L. Xiao · 2009
Earlier work this paper cites.
Large-scale machine learning with stochastic gradient descent
L. Bottou · 2010
Earlier work this paper cites.
Online importance weight aware updates
N. Karampatziakis and J. Langford · 2010
Earlier work this paper cites.
Distributed nonnegative matrix factorization for web-scale dyadic data analysis on mapreduce
C. Liu, H.-c. Yang, J. Fan, L.-W. He, and Y.-M. Wang · 2010
Earlier work this paper cites.
Piccolo: Building fast, distributed programs with partitioned tables
R. Power and J. Li · 2010
Cited alongside, same era.
Parallelized stochastic gradient descent
M. Zinkevich, M. Weimer, L. Li, and A. J. Smola · 2010
Cited alongside, same era.
Distributed delayed stochastic optimization
A. Agarwal and J. C. Duchi · 2011
Cited alongside, same era.
Adaptive subgradient methods for online learning and stochastic optimization
J. Duchi, E. Hazan, and Y. Singer · 2011
Cited alongside, same era.
Large-scale matrix factorization with distributed stochastic gradient descent
R. Gemulla, E. Nijkamp, P. J. Haas, and Y. Sismanis · 2011
Cited alongside, same era.
Making gradient descent optimal for strongly convex stochastic optimization
A. Rakhlin, O. Shamir, and K. Sridharan · 2011
More effective distributed ML via a stale synchronous parallel parameter server
Q. Ho, J. Cipar, H. Cui, S. Lee, J. K. Kim, P. B. Gibbons, G. A. Gibson, G. Ganger, and E. P. Xing · 2013
Later among the works it cites.
Stochastic variational inference
M. D. Hoffman, D. M. Blei, C. Wang, and J. Paisley · 2013
Later among the works it cites.
Accelerating stochastic gradient descent using predictive variance reduction
R. Johnson and T. Zhang · 2013
Later among the works it cites.
UCI machine learning repository, 2013
M. Lichman · 2013
Later among the works it cites.
An asynchronous parallel stochastic coordinate descent algorithm
J. Liu, S. J. Wright, C. Ré, V. Bittorf, and S. Sridhar · 2013
Later among the works it cites.
Naiad: a timely dataflow system
D. G. Murray, F. McSherry, R. Isaacs, M. Isard, P. Barham, and M. Abadi · 2013
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Hogwild: A lock-free approach to parallelizing stochastic gradient descent
B. Recht, C. Re, S. Wright, and F. Niu · 2011
Cited alongside, same era.
Scalable inference in latent variable models
A. Ahmed, M. Aly, J. Gonzalez, S. Narayanamurthy, and E. Smola · 2012
Cited alongside, same era.
Large scale distributed deep networks
J. Dean, G. Corrado, R. Monga, K. Chen, M. Devin, M. Mao, A. Senior, P. Tucker, K. Yang, Q. V. Le, et al · 2012
Cited alongside, same era.
Dual averaging for distributed optimization: convergence analysis and network scaling
J. C. Duchi, A. Agarwal, and M. J. Wainwright · 2012
Cited alongside, same era.
Mahout project
T. A. S. Foundation · 2012
Cited alongside, same era.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Cited alongside, same era.
Later among the works it cites.
Minimizing finite sums with the stochastic average gradient
M. Schmidt, N. L. Roux, and F. Bach · 2013
Later among the works it cites.
Mli: An api for distributed machine learning
E. R. Sparks, A. Talwalkar, V. Smith, J. Kottalam, X. Pan, J. Gonzalez, M. J. Franklin, M. Jordan, T. Kraska, et al · 2013
Later among the works it cites.
Petuum: A new platform for distributed machine learning on big data
E. P. Xing, Q. Ho, W. Dai, J. K. Kim, J. Wei, S. Lee, X. Zheng, P. Xie, A. Kumar, and Y. Yu · 2013
Later among the works it cites.
A fast parallel SGD for matrix factorization in shared memory systems
Y. Zhuang, W.-S. Chin, Y.-C. Juan, and C.-J. Lin · 2013
Later among the works it cites.
Communication-efficient distributed dual coordinate ascent
M. Jaggi, V. Smith, M. Takác, J. Terhorst, S. Krishnan, T. Hofmann, and M. I. Jordan · 2014
Later among the works it cites.
Scaling distributed machine learning with the parameter server
M. Li, D. G. Andersen, J. W. Park, A. J. Smola, A. Ahmed, V. Josifovski, J. Long, E. J. Shekita, and B.-Y. Su · 2014
Later among the works it cites.
Same but different: Fast and high-quality gibbs parameter estimation
H. Zhao, B. Jiang, and J. Canny · 2014
Later among the works it cites.
Mllib: Machine learning in apache spark
X. Meng, J. Bradley, B. Yavuz, E. Sparks, S. Venkataraman, D. Liu, J. Freeman, D. Tsai, M. Amde, S. Owen, et al · 2015
Closest in time.