Fetching the paper…
Reading the bibliography…
Many areas of machine learning and science involve large linear algebra problems, such as eigendecompositions, solving linear systems, computing matrix exponentials, and trace estimation.
Inverting modified matrices
Max A Woodbury · 1950
Earlier work this paper cites.
The principle of minimized iterations in the solution of the matrix eigenvalue problem
Walter Edwin Arnoldi · 1951
Earlier work this paper cites.
Gmres: A generalized minimal residual algorithm for solving nonsymmetric linear systems
Youcef Saad and Martin H Schultz · 1986
Earlier work this paper cites.
Improving the convergence of back-propagation learning with second-order methods
S Becker and Yann Lecun · 1989
Earlier work this paper cites.
A stochastic estimator of the trace of the influence matrix for laplacian smoothing splines
Michael F Hutchinson · 1989
Earlier work this paper cites.
Numerical Linear Algebra
Lloyd N. Trefethen and David Bau · 1997
Earlier work this paper cites.
Toward The Optimal Preconditioned Eigensolver: Locally Optimal Block Preconditioned Conjugate Gradient Method
Andrew Knyazev · 2000
Earlier work this paper cites.
On Spectral Clustering: Analysis and an algorithm
Andrew Y. Ng, Michael I. Jordan, and Yair Weiss · 2001
Earlier work this paper cites.
Iterative methods for sparse linear systems
Yousef Saad · 2003
Earlier work this paper cites.
Adaptive Smoothed Aggregation ( α \alpha SA) Multigrid
M. Brezina, R. Falgout, S. MacLachlan, T. Manteuffel, S. McCormick, and J. Ruge · 2005
Earlier work this paper cites.
Sparse gaussian processes using pseudo-inputs
Edward Snelson and Zoubin Ghahramani · 2005
Earlier work this paper cites.
Direct methods for sparse linear systems
Timothy A Davis · 2006
Earlier work this paper cites.
Multi-task Gaussian Process Prediction
Edwin V. Bonilla, Kian Ming A. Chai, and Christopher K. I. Williams · 2007
Earlier work this paper cites.
Random Features for Large-Scale Kernel Machines
Ali Rahimi and Ben Recht · 2007
Earlier work this paper cites.
Topmoumoute online natural gradient algorithm
Nicolas Roux, Pierre-Antoine Manzagol, and Yoshua Bengio · 2007
Earlier work this paper cites.
Fast gaussian process methods for point process intensity estimation
John P Cunningham, Krishna V Shenoy, and Maneesh Sahani · 2008
Earlier work this paper cites.
Deep learning via hessian-free optimization
James Martens · 2010
Earlier work this paper cites.
Sinkhorn Distances: Lightspeed Computation of Optimal Transport
Marco Cuturi · 2013
Earlier work this paper cites.
Accelerating Stochastic Gradient Descent using Predictive Variance Reduction
Rie Johnshon and Tong Zhang · 2013
Cited alongside, same era.
Spot – A Linear-Operator Toolbox
Ewout van den Berg and Michael P. Friedlander · 2013
Cited alongside, same era.
Julia: A Fresh Approach to Numerical Computing
Jeff Bezanson, Alan Edelman, Stefan Karpinski, and Viral B. Shah · 2014
Cited alongside, same era.
Fast kernel learning for multidimensional pattern extrapolation
Andrew G Wilson, Elad Gilboa, Arye Nehorai, and John P Cunningham · 2014
Cited alongside, same era.
TensorFlow: Large-scale machine learning on heterogeneous systems, 2015
Martín Abadi, Ashish Agarwal, Paul Barham, Eugene Brevdo, Zhifeng Chen, Craig Citro, Greg S. Corrado, Andy Davis, Jeffrey Dean, Matthieu Devin, Sanjay Ghemawat, Ian Goodfellow, Andrew Harp, Geoffrey Irving, Michael Isard, Yangqing Jia, Rafal Jozefowicz, Lukasz Kaiser, Manjunath Kudlur, Josh Levenberg, Dandelion Mané, Rajat Monga, Sherry Moore, Derek Murray, Chris Olah, Mike Schuster, Jonathon Shlens, Benoit Steiner, Ilya Sutskever, Kunal Talwar, Paul Tucker, Vincent Vanhoucke, Vijay Vasudevan, Fernanda Viégas, Oriol Vinyals, Pete Warden, Martin Wattenberg, Martin Wicke, Yuan Yu, and Xiaoqiang Zheng · 2015
PyLops—A linear-operator Python library for scalable algebra and optimization
Matteo Ravasi and Ivan Vasconcelos · 2019
Later among the works it cites.
Exact gaussian processes on a million data points
Ke Wang, Geoff Pleiss, Jacob Gardner, Stephen Tyree, Kilian Q Weinberger, and Andrew Gordon Wilson · 2019
Later among the works it cites.
Scalable Second Order Optimization for Deep Learning
Rohan Anil, Vineet Gupta, Tomer Koren, Kevin Regan, and Yoram Singer · 2020
Later among the works it cites.
Randomized Numerical Linear Algebra: Foundations & Algorithms
Per-Gunnar Martinsson and Joel Tropp · 2020
Later among the works it cites.
SciPy 1.0: Fundamental Algorithms for Scientific Computing in Python
Pauli Virtanen, Ralf Gommers, Travis E. Oliphant, Matt Haberland, Tyler Reddy, David Cournapeau, Evgeni Burovski, Pearu Peterson, Warren Weckesser, Jonathan Bright, Stéfan J. van der Walt, Matthew Brett, Joshua Wilson, K. Jarrod Millman, Nikolay Mayorov, Andrew R. J. Nelson, Eric Jones, Robert Kern, Eric Larson, C J Carey, İlhan Polat, Yu Feng, Eric W. Moore, Jake VanderPlas, Denis Laxalde, Josef Perktold, Robert Cimrman, Ian Henriksen, E. A. Quintero, Charles R. Harris, Anne M. Archibald, Antônio H. Ribeiro, Fabian Pedregosa, Paul van Mulbregt, and SciPy 1.0 Contributors · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Gradient-based hyperparameter optimization through reversible learning
Dougal Maclaurin, David Duvenaud, and Ryan Adams · 2015
Cited alongside, same era.
Optimizing neural networks with kronecker-factored approximate curvature
James Martens and Roger Grosse · 2015
Cited alongside, same era.
A Stochastic PCA and SVD Algorithm with an Exponential Convergence Rate
Ohad Shamir · 2015
Cited alongside, same era.
Chainer: a next-generation open source framework for deep learning
Seiya Tokui, Kenta Oono, Shohei Hido, and Justin Clayton · 2015
Cited alongside, same era.
Kernel interpolation for scalable structured gaussian processes (kiss-gp)
Andrew Wilson and Hannes Nickisch · 2015
Cited alongside, same era.
Modeling, inference and optimization with composable differentiable procedures
Dougal Maclaurin · 2016
Cited alongside, same era.
JAX: composable transformations of Python+NumPy programs
James Bradbury, Roy Frostig, Peter Hawkins, Matthew James Johnson, Chris Leary, Dougal Maclaurin, George Necula, Adam Paszke, Jake VanderPlas, Skye Wanderman-Milne, and Qiao Zhang · 2018
Cited alongside, same era.
Later among the works it cites.
Kernel operations on the GPU, with autodiff, without memory overflows
Benjamin Charlier, Jean Feydy, Joan Alexis Glaunes, François-David Collin, and Ghislain Durif · 2021
Later among the works it cites.
Evolutional Deep Neural Network
Yifan Du and Tamer A Zaki · 2021
Later among the works it cites.
A Practical Method for Constructing Equivariant Multilayer Perceptrons for Arbitrary Matrix Groups
Marc Finzi, Max Welling, and Andrew Gordon Wilson · 2021
Later among the works it cites.
SKIing on Simplices: Kernel Interpolation on the Permutohedral Lattice for Scalable Gaussian Processes
Sanyam Kapoor, Marc Finzi, Ke Alexander Wang, and Andrew Gordon Gordon Wilson · 2021
Later among the works it cites.
A general framework for vecchia approximations of gaussian processes
Matthias Katzfuss and Joseph Guinness · 2021
Later among the works it cites.
Neural Operator: Learning Maps Between Function Spaces
Nikola Kovachki, Zongyi Li, Burigede Liu, Kamyar Azizzadenesheli, Kaushik Bhattacharya, Andrew Stuart, and Anima Anandkumar · 2021
Later among the works it cites.
A general linear-time inference method for gaussian processes on one dimension
Jackson Loper, David Blei, John P Cunningham, and Liam Paninski · 2021
Later among the works it cites.
Low-Precision Arithmetic for Fast Gaussian Processes
Wesley J. Maddox, Andres Potapczynski, and Andrew Gordon Wilson · 2022
Later among the works it cites.
S4ND: Modeling Images and Videos as Multidimensional Signals Using State Spaces
Eric Nguyen, Karan Goel, Albert Gu, Gordon W. Downs, Preey Shah, Tri Dao, Stephen A. Baccus, and Christopher Ré · 2022
Later among the works it cites.
PyAMG: Algebraic Multigrid Solvers in Python
Nathan Bell, Luke N. Olson, Jacob Schroder, and Ben Southworth · 2023
Closest in time.
A Stable and Scalable Method for Solving Initial Value PDEs with Neural Networks
Marc Finzi, Andres Potapczynski, Matthew Choptuik, and Andrew Gordon Wilson · 2023
Closest in time.
Hungry Hungry Hippos: Towards Language Modeling with State Space Models
Daniel Y. Fu, Tri Dao, Khaled K. Saab, Armin W. Thomas, Atri Rudra, and Christopher Ré · 2023
Closest in time.