Fetching the paper…
Reading the bibliography…
The study of modern machine learning models often necessitates storing vast quantities of gradients or Hessian vector products (HVPs).
Extensions of lipschitz mappings into a hilbert space, 1984
William B Johnson and Joram Lindenstrauss · 1984
Earlier work this paper cites.
Improved approximation algorithms for large matrices via random projections
Tamas Sarlos · 2006
Earlier work this paper cites.
The fast johnson-lindenstrauss transform and approximate nearest neighbors
Nir Ailon and Bernard Chazelle · 2009
Earlier work this paper cites.
Fastfood - approximating kernel expansions in loglinear time
Quoc Le, Tamas Sarlos, and Alex Smola · 2013
Earlier work this paper cites.
Sketching as a tool for numerical linear algebra, November 2014
David P Woodruff · 2014
Earlier work this paper cites.
Fast nonlinear embeddings via structured matrices, 2016
Krzysztof Choromanski and Francois Fagan · 2016
Earlier work this paper cites.
Lecture notes on randomized linear algebra
Michael W. Mahoney · 2016
Earlier work this paper cites.
Understanding black-box predictions via influence functions
Pang Wei Koh and Percy Liang · 2017
Earlier work this paper cites.
Gradient descent happens in a tiny subspace, 2018
Guy Gur-Ari, Daniel A. Roberts, and Ethan Dyer · 2018
Earlier work this paper cites.
Measuring the intrinsic dimension of objective landscapes
Chunyuan Li, Heerad Farkhoor, Rosanne Liu, and Jason Yosinski · 2018
Earlier work this paper cites.
Empirical analysis of the hessian of over-parametrized neural networks, 2018
Levent Sagun, Utku Evci, V. Ugur Guney, Yann Dauphin, and Leon Bottou · 2018
Cited alongside, same era.
High-Dimensional Probability
Roman Vershynin · 2018
Cited alongside, same era.
An investigation into neural net optimization via hessian eigenvalue density
Shankar Krishnan, Behrooz Ghorbani, and Xiao Ying · 2019
Cited alongside, same era.
Random features for kernel approximation: A survey on algorithms, theory, and beyond
Fanghui Liu, Xiaolin Huang, Yudong Chen, and Johan A K Suykens · 2020
Cited alongside, same era.
Estimating training data influence by tracing gradient descent
Garima Pruthi, Frederick Liu, Satyen Kale, and Mukund Sundararajan · 2020
Cited alongside, same era.
Intrinsic dimensionality explains the effectiveness of language model Fine-Tuning
Armen Aghajanyan, Sonal Gupta, and Luke Zettlemoyer · 2021
Scaling up influence functions
Andrea Schioppa, Polina Zablotskaia, David Vilar Torres, and Artem Sokolov · 2022
Later among the works it cites.
First is better than last for language data influence
Chih-Kuan Yeh, Ankur Taly, Mukund Sundararajan, Frederick Liu, and Pradeep Ravikumar · 2022
Later among the works it cites.
High-dimensional sgd aligns with emerging outlier eigenspaces, 2023
Gerard Ben Arous, Reza Gheissari, Jiaoyang Huang, and Aukosh Jagannath · 2023
Later among the works it cites.
Influence diagnostics under self-concordance
Jillian Fisher, Lang Liu, Krishna Pillutla, Yejin Choi, and Zaid Harchaoui · 2023
Later among the works it cites.
Studying large language model generalization with influence functions, 2023
Roger Grosse, Juhan Bae, Cem Anil, Nelson Elhage, Alex Tamkin, Amirhossein Tajdini, Benoit Steiner, Dustin Li, Esin Durmus, Ethan Perez, Evan Hubinger, Kamilė Lukošiūtė, Karina Nguyen, Nicholas Joseph, Sam McCandlish, Jared Kaplan, and Samuel R. Bowman · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
FastIF: Scalable influence functions for efficient model interpretation and debugging
Han Guo, Nazneen Rajani, Peter Hase, Mohit Bansal, and Caiming Xiong · 2021
Cited alongside, same era.
Few-shot parameter-efficient fine-tuning is better and cheaper than in-context learning
Haokun Liu, Derek Tam, Muqeeth Mohammed, Jay Mohta, Tenghao Huang, Mohit Bansal, and Colin Raffel · 2022
Cited alongside, same era.
PAC-bayes compression bounds so tight that they can explain generalization
Sanae Lotfi, Marc Anton Finzi, Sanyam Kapoor, Andres Potapczynski, Micah Goldblum, and Andrew Gordon Wilson · 2022
Cited alongside, same era.
Sung Min Park, Kristian Georgiev, Andrew Ilyas, Guillaume Leclerc, and Aleksander Madry · 2023
Later among the works it cites.
Theoretical and practical perspectives on what influence functions do, 2023 (also accepted in Neurips23)
Andrea Schioppa, Katja Filippova, Ivan Titov, and Polina Zablotskaia · 2023
Later among the works it cites.
Optimal eigenvalue approximation via sketching
William Swartworth and David P Woodruff · 2023
Later among the works it cites.
Faithful and efficient explanations for neural networks via neural tangent kernel surrogate models, 2024 (also accepted in ICLR24)
Andrew Engel, Zhichao Wang, Natalie S. Frank, Ioana Dumitriu, Sutanay Choudhury, Anand Sarwate, and Tony Chiang · 2024
Closest in time.