Fetching the paper…
Reading the bibliography…
The minimization of convex objectives coming from linear supervised learning problems, such as penalized generalized linear models, can be formulated as finite sums of convex functions.
A stochastic approximation method
H. Robbins and S. Monro · 1951
Earlier work this paper cites.
A cluster process representation of a self-exciting process
A. G. Hawkes and D. Oakes · 1974
Earlier work this paper cites.
Updating quasi-newton matrices with limited storage
J. Nocedal · 1980
Earlier work this paper cites.
Interior-point polynomial algorithms in convex programming
Y. Nesterov and A. Nemirovskii · 1994
Earlier work this paper cites.
Nonlinear programming
D. P. Bertsekas · 1999
Earlier work this paper cites.
Seismicity analysis through point-process modeling: A review
Y. Ogata · 1999
Earlier work this paper cites.
Nonlinear Equations
J. Nocedal and S. J. Wright · 2006
Earlier work this paper cites.
An introduction to the theory of point processes: volume II: general theory and structure
D. J. Daley and D. Vere-Jones · 2007
Earlier work this paper cites.
A fast iterative shrinkage-thresholding algorithm for linear inverse problems
A. Beck and M. Teboulle · 2009
Earlier work this paper cites.
Image deblurring with poisson data: from cells to galaxies
M. Bertero, P. Boccacci, G. Desiderà, and G. Vicidomini · 2009
Earlier work this paper cites.
Large-scale behavioral targeting
Y. Chen, D. Pavlov, and J. F. Canny · 2009
Earlier work this paper cites.
Modeling wine preferences by data mining from physicochemical properties
P. Cortez, A. Cerdeira, F. Almeida, T. Matos, and J. Reis · 2009
Earlier work this paper cites.
Self-concordant analysis for logistic regression
F. Bach et al · 2010
Earlier work this paper cites.
Fitting additive poisson models
H. C. Boshuizen and E. J. Feskens · 2010
Earlier work this paper cites.
This is spiral-tap: Sparse poisson intensity reconstruction algorithms—theory and practice
Z. T. Harmany, R. F. Marcia, and R. M. Willett · 2012
Cited alongside, same era.
Accelerating stochastic gradient descent using predictive variance reduction
R. Johnson and T. Zhang · 2013
Cited alongside, same era.
UCI machine learning repository, 2013
M. Lichman · 2013
Cited alongside, same era.
Modeling and estimation of multi-source clustering in crime and security data
G. Mohler et al · 2013
Cited alongside, same era.
Introductory lectures on convex optimization: A basic course
Y. Nesterov · 2013
Cited alongside, same era.
Minimizing finite sums with the stochastic average gradient
M. Schmidt, N. L. Roux, and F. Bach · 2013
Stochastic optimization with importance sampling for regularized loss minimization
P. Zhao and T. Zhang · 2015
Later among the works it cites.
Estimation of slowly decreasing hawkes kernels: application to high-frequency order book dynamics
E. Bacry, T. Jaisson, and J.-F. Muzy · 2016
Later among the works it cites.
A descent lemma beyond lipschitz gradient continuity: first-order methods revisited and applications
H. H. Bauschke, J. Bolte, and M. Teboulle · 2016
Later among the works it cites.
Learning and forecasting opinion dynamics in social networks
A. De, I. Valera, N. Ganguly, S. Bhattacharya, and M. G. Rodriguez · 2016
Later among the works it cites.
Variance-reduced and projection-free stochastic optimization
E. Hazan and H. Luo · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Stochastic dual coordinate ascent methods for regularized loss minimization
S. Shalev-Shwartz and T. Zhang · 2013
Cited alongside, same era.
Saga: A fast incremental gradient method with support for non-strongly convex composite objectives
A. Defazio, F. Bach, and S. Lacoste-Julien · 2014
Cited alongside, same era.
Accelerated proximal stochastic dual coordinate ascent for regularized loss minimization
S. Shalev-Shwartz and T. Zhang · 2014
Cited alongside, same era.
A proximal stochastic gradient method with progressive variance reduction
L. Xiao and T. Zhang · 2014
Cited alongside, same era.
Hawkes processes in finance
E. Bacry, I. Mastromatteo, and J.-F. Muzy · 2015
Cited alongside, same era.
Coevolve: A joint point process model for information diffusion and network co-evolution
M. Farajtabar, Y. Wang, M. G. Rodriguez, S. Li, H. Zha, and L. Song · 2015
Cited alongside, same era.
Hawkes processes for continuous time sequence classification: an application to rumour stance classification in twitter
M. Lukasik, P. Srijith, D. Vu, K. Bontcheva, A. Zubiaga, and T. Cohn · 2016
Later among the works it cites.
Predicting social media performance metrics and evaluation of the impact on brand building: A data mining approach
S. Moro, P. Rita, and B. Vala · 2016
Later among the works it cites.
Sdna: Stochastic dual newton ascent for empirical risk minimization
Z. Qu, P. Richtárik, M. Takác, and O. Fercoq · 2016
Later among the works it cites.
Barzilai-borwein step size for stochastic gradient descent
C. Tan, S. Ma, Y.-H. Dai, and Y. Qian · 2016
Later among the works it cites.
Stripping customers’ feedback on hotels through data mining: the case of las vegas strip
S. Moro, P. Rita, and J. Coelho · 2017
Later among the works it cites.
The role of volume in order book dynamics: a multivariate hawkes process analysis
M. Rambaldi, E. Bacry, and F. Lillo · 2017
Later among the works it cites.
Generalized self-concordant functions: a recipe for newton-type methods
T. Sun and Q. Tran-Dinh · 2017
Later among the works it cites.
Relatively smooth convex optimization by first-order methods, and applications
H. Lu, R. M. Freund, and Y. Nesterov · 2018
Closest in time.
Adaptive stochastic dual coordinate ascent for conditional random fields
R. L. Priol, A. Touati, and S. Lacoste-Julien · 2018
Closest in time.