Fetching the paper…
Reading the bibliography…
Machine learning models are not static and may need to be retrained on slightly changed datasets, for instance, with the addition or deletion of a set of data points.
A stochastic approximation method
Robbins, H. and Monro, S · 1951
Earlier work this paper cites.
Notes on bias in estimation
Quenouille, M. H · 1956
Earlier work this paper cites.
On stochastic approximation
Gladyshev, E · 1965
Earlier work this paper cites.
A theory of adaptive pattern classifiers
Amari, S · 1967
Earlier work this paper cites.
Iterative solution of nonlinear equations in several variables , volume 30
Ortega, J. M. and Rheinboldt, W. C · 1970
Earlier work this paper cites.
Detection of influential observation in linear regression
Cook, R. D · 1977
Earlier work this paper cites.
The solution of nonlinear finite element equations
Matthies, H. and Strang, G · 1979
Earlier work this paper cites.
Updating quasi-newton matrices with limited storage
Nocedal, J · 1980
Earlier work this paper cites.
Testing a class of methods for solving minimization problems with simple bounds on the variables
Conn, A. R., Gould, N. I., and Toint, P. L · 1988
Earlier work this paper cites.
Convergence of quasi-newton matrices generated by the symmetric rank one update
Conn, A. R., Gould, N. I., and Toint, P. L · 1991
Earlier work this paper cites.
Estimation of convergence rate for robust identification algorithms
Kul’chitskiy, O. Y. and Mozgovoy, A · 1992
Earlier work this paper cites.
Representations of quasi-newton matrices and their use in limited memory methods
Byrd, R. H., Nocedal, J., and Schnabel, R. B · 1994
Earlier work this paper cites.
A limited memory algorithm for bound constrained optimization
Byrd, R. H., Lu, P., Nocedal, J., and Zhu, C · 1995
Earlier work this paper cites.
Neuro-dynamic programming , volume 5
Bertsekas, D. P. and Tsitsiklis, J. N · 1996
Earlier work this paper cites.
Algorithm 778: L-bfgs-b: Fortran subroutines for large-scale bound-constrained optimization
Zhu, C., Byrd, R. H., Lu, P., and Nocedal, J · 1997
Earlier work this paper cites.
Online learning and stochastic approximations
Bottou, L · 1998
Earlier work this paper cites.
Gradient-based learning applied to document recognition
LeCun, Y., Bottou, L., Bengio, Y., and Haffner, P · 1998
Earlier work this paper cites.
Lazy learning meets the recursive least squares algorithm
Birattari, M., Bontempi, G., and Bersini, H · 1999
Earlier work this paper cites.
Comparative accuracies of artificial neural networks and discriminant analysis in predicting forest cover types from cartographic variables
Blackard, J. A. and Dean, D. J · 1999
Earlier work this paper cites.
Subsampling
Politis, D. N., Romano, J. P., and Wolf, M · 1999
Earlier work this paper cites.
Incremental learning with support vector machines
Syed, N. A., Huan, S., Kah, L., and Sung, K · 1999
Earlier work this paper cites.
Incremental and decremental support vector machine learning
Cauwenberghs, G. and Poggio, T · 2001
Cited alongside, same era.
Stochastic learning
Bottou, L · 2003
Cited alongside, same era.
Convex optimization
Boyd, S. and Vandenberghe, L · 2004
Cited alongside, same era.
Rcv1: A new benchmark collection for text categorization research
Lewis, D. D., Yang, Y., Rose, T. G., and Li, F · 2004
Cited alongside, same era.
Solving large scale linear prediction problems using stochastic gradient descent algorithms
Zhang, T · 2004
Cited alongside, same era.
Numerical optimization
Nocedal, J. and Wright, S · 2006
Cited alongside, same era.
The tradeoffs of large scale learning
Optimization methods for large-scale machine learning
Bottou, L., Curtis, F. E., and Nocedal, J · 2016
Later among the works it cites.
Council regulation (eu) no 2016/679
European Union, C. o · 2016
Later among the works it cites.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S., and Sun, J · 2016
Later among the works it cites.
Linear convergence of gradient and proximal-gradient methods under the polyak-łojasiewicz condition
Karimi, H., Nutini, J., and Schmidt, M · 2016
Later among the works it cites.
The expected norm of a sum of independent random matrices: An elementary approach
Tropp, J. A · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Bousquet, O. and Bottou, L · 2008
Cited alongside, same era.
Evaluating derivatives: principles and techniques of algorithmic differentiation , volume 105
Griewank, A. and Walther, A · 2008
Cited alongside, same era.
A tutorial on conformal prediction
Shafer, G. and Vovk, V · 2008
Cited alongside, same era.
Privacy-preserving logistic regression
Chaudhuri, K. and Monteleoni, C · 2009
Cited alongside, same era.
Learning multiple layers of features from tiny images
Krizhevsky, A. and Hinton, G · 2009
Cited alongside, same era.
Concentration of the adjacency matrix and of the laplacian in random graphs with independent edges
Oliveira, R. I · 2009
Cited alongside, same era.
Doshi-Velez, F. and Kim, B · 2017
Later among the works it cites.
Understanding black-box predictions via influence functions
Koh, P. W. and Liang, P · 2017
Later among the works it cites.
Palm: Machine learning explanations for iterative debugging
Krishnan, S. and Wu, E · 2017
Later among the works it cites.
Certified defenses for data poisoning attacks
Steinhardt, J., Koh, P. W. W., and Liang, P. S · 2017
Later among the works it cites.
Robust linear regression: A review and comparison
Yu, C. and Yao, W · 2017
Later among the works it cites.
Optimization methods for large-scale machine learning
Bottou, L., Curtis, F. E., and Nocedal, J · 2018
Later among the works it cites.
A modern maximum-likelihood theory for high-dimensional logistic regression
Sur, P. and Candès, E. J · 2018
Later among the works it cites.
Bourtoule, L., Chandrasekaran, V., Choquette-Choo, C., Jia, H., Travers, A., Zhang, B., Lie, D., and Papernot, N · 2019
Later among the works it cites.
Data shapley: Equitable valuation of data for machine learning
Ghorbani, A. and Zou, J · 2019
Later among the works it cites.
Making ai forget you: Data deletion in machine learning
Ginart, A., Guan, M., Valiant, G., and Zou, J. Y · 2019
Later among the works it cites.
A unified theory of sgd: Variance reduction, sampling, quantization and coordinate descent
Gorbunov, E., Hanzely, F., and Richtárik, P · 2019
Later among the works it cites.
Sgd: General analysis and improved rates
Gower, R. M., Loizou, N., Qian, X., Sailanbayev, A., Shulgin, E., and Richtárik, P · 2019
Later among the works it cites.
Certified data removal from machine learning models
Guo, C., Goldstein, T., Hannun, A., and van der Maaten, L · 2019
Later among the works it cites.
“amnesia”–towards machine learning models that can forget user data very fast
Schelter, S · 2019
Later among the works it cites.
Priu: A provenance-based approach for incrementally updating regression models
Wu, Y., Tannen, V., and Davidson, S. B · 2020
Closest in time.