Fetching the paper…
Reading the bibliography…
We propose a general class of sample based explanations of machine learning models, which we term generalized representers.
Xvi. functions of positive and negative type, and their connection the theory of integral equations
James Mercer · 1909
Earlier work this paper cites.
A value for n-person games
Lloyd S Shapley · 1953
Earlier work this paper cites.
Residuals and influence in regression
R Dennis Cook and Sanford Weisberg · 1982
Earlier work this paper cites.
A value for n-person games , page 31–40
Lloyd S. Shapley · 1988
Earlier work this paper cites.
A generalized representer theorem
Bernhard Schölkopf, Ralf Herbrich, and Alex J Smola · 2001
Earlier work this paper cites.
Influence functions in deep learning are fragile
Samyadeep Basu, Philip Pope, and Soheil Feizi · 2006
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Alex Krizhevsky, Geoffrey Hinton, et al · 2009
Earlier work this paper cites.
MNIST handwritten digit database
Yann LeCun and Corinna Cortes · 2010
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Why should i trust you?: Explaining the predictions of any classifier
Marco Tulio Ribeiro, Sameer Singh, and Carlos Guestrin · 2016
Earlier work this paper cites.
Second-order stochastic optimization in linear time
Naman Agarwal, Brian Bullins, and Elad Hazan · 2016
Earlier work this paper cites.
Grad-cam: Visual explanations from deep networks via gradient-based localization
Ramprasaath R Selvaraju, Michael Cogswell, Abhishek Das, Ramakrishna Vedantamand, Devi Parikh, and Dhruv Parikh · 2017
Earlier work this paper cites.
Smoothgrad: removing noise by adding noise
Daniel Smilkov, Nikhil Thorat, Been Kim, Fernanda Viégas, and Martin Wattenberg · 2017
Earlier work this paper cites.
A unified approach to interpreting model predictions
Scott M Lundberg and Su-In Lee · 2017
Earlier work this paper cites.
Understanding black-box predictions via influence functions
Pang Wei Koh and Percy Liang · 2017
Earlier work this paper cites.
Towards a rigorous science of interpretable machine learning
Finale Doshi-Velez and Been Kim · 2017
Earlier work this paper cites.
Axiomatic attribution for deep networks
Mukund Sundararajan, Ankur Taly, and Qiqi Yan · 2017
Earlier work this paper cites.
Deep neural networks as gaussian processes
Jaehoon Lee, Yasaman Bahri, Roman Novak, Samuel S Schoenholz, Jeffrey Pennington, and Jascha Sohl-Dickstein · 2017
Earlier work this paper cites.
Representer point selection for explaining deep neural networks
Chih-Kuan Yeh, Joon Kim, Ian En-Hsu Yen, and Pradeep K Ravikumar · 2018
Earlier work this paper cites.
Neural tangent kernel: Convergence and generalization in neural networks
Arthur Jacot, Franck Gabriel, and Clément Hongler · 2018
Earlier work this paper cites.
Consistent individualized feature attribution for tree ensembles
Scott M Lundberg, Gabriel G Erion, and Su-In Lee · 2018
Earlier work this paper cites.
Connecting optimization and regularization paths
Arun Suggala, Adarsh Prasad, and Pradeep K Ravikumar · 2018
Cited alongside, same era.
The unreasonable effectiveness of deep features as a perceptual metric
Richard Zhang, Phillip Isola, Alexei A Efros, Eli Shechtman, and Oliver Wang · 2018
Cited alongside, same era.
Data shapley: Equitable valuation of data for machine learning
Amirata Ghorbani and James Zou · 2019
Cited alongside, same era.
Definitions, methods, and applications in interpretable machine learning
W James Murdoch, Chandan Singh, Karl Kumbier, Reza Abbasi-Asl, and Bin Yu · 2019
Cited alongside, same era.
On the (in) fidelity and sensitivity of explanations
Chih-Kuan Yeh, Cheng-Yu Hsieh, Arun Suggala, David I Inouye, and Pradeep K Ravikumar · 2019
Cited alongside, same era.
Towards efficient data valuation based on the shapley value
Ruoxi Jia, David Dao, Boxin Wang, Frances Ann Hubis, Nick Hynes, Nezihe Merve Gürel, Bo Li, Ce Zhang, Dawn Song, and Costas J Spanos · 2019
Representer point selection via local jacobian expansion for post-hoc classifier explanation of deep neural networks and ensemble models
Yi Sui, Ga Wu, and Scott Sanner · 2021
Later among the works it cites.
Properties of the after kernel
Philip M Long · 2021
Later among the works it cites.
The disagreement problem in explainable machine learning: A practitioner’s perspective
Satyapriya Krishna, Tessa Han, Alex Gu, Javin Pombra, Shahin Jabbari, Steven Wu, and Himabindu Lakkaraju · 2022
Later among the works it cites.
Impossibility theorems for feature attribution
Blair Bilodeau, Natasha Jaques, Pang Wei Koh, and Been Kim · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Data cleansing for models trained with sgd
Satoshi Hara, Atsushi Nitanda, and Takanori Maehara · 2019
Cited alongside, same era.
Estimating training data influence by tracing gradient descent
Garima Pruthi, Frederick Liu, Satyen Kale, and Mukund Sundararajan · 2020
Cited alongside, same era.
The shapley taylor interaction index
Mukund Sundararajan, Kedar Dhamdhere, and Ashish Agarwal · 2020
Cited alongside, same era.
On completeness-aware concept-based explanations in deep neural networks
Chih-Kuan Yeh, Been Kim, Sercan Arik, Chun-Liang Li, Tomas Pfister, and Pradeep Ravikumar · 2020
Cited alongside, same era.
Characterizing structural regularities of labeled data in overparameterized models
Ziheng Jiang, Chiyuan Zhang, Kunal Talwar, and Michael C Mozer · 2020
Cited alongside, same era.
Data valuation using reinforcement learning
Jinsung Yoon, Sercan Arik, and Tomas Pfister · 2020
Cited alongside, same era.
Valerie Chen, Nari Johnson, Nicholay Topin, Gregory Plumb, and Ameet Talwalkar · 2022
Later among the works it cites.
On the importance of application-grounded experimental design for evaluating explainable ml methods
Kasun Amarasinghe, Kit T Rodolfa, Sérgio Jesus, Valerie Chen, Vladimir Balayan, Pedro Saleiro, Pedro Bizarro, Ameet Talwalkar, and Rayid Ghani · 2022
Later among the works it cites.
Training data influence analysis and estimation: A survey
Zayd Hammoudeh and Daniel Lowd · 2022
Later among the works it cites.
Data banzhaf: A data valuation framework with maximal robustness to learning stochasticity
Tianhao Wang and Ruoxi Jia · 2022
Later among the works it cites.
Datamodels: Predicting predictions from training data
Andrew Ilyas, Sung Min Park, Logan Engstrom, Guillaume Leclerc, and Aleksander Madry · 2022
Later among the works it cites.
Adapting and evaluating influence-estimation methods for gradient-boosted decision trees
Jonathan Brophy, Zayd Hammoudeh, and Daniel Lowd · 2022
Later among the works it cites.
Cross-loss influence functions to explain deep network representations
Andrew Silva, Rohit Chopra, and Matthew Gombolay · 2022
Later among the works it cites.
Scaling up influence functions
Andrea Schioppa, Polina Zablotskaia, David Vilar, and Artem Sokolov · 2022
Later among the works it cites.
Rethinking influence functions of neural networks in the over-parameterized regime
Rui Zhang and Shihua Zhang · 2022
Later among the works it cites.
Tracinad: Measuring influence for anomaly detection
Hugo Thimonier, Fabrice Popineau, Arpad Rimmel, Bich-Liên Doan, and Fabrice Daniel · 2022
Later among the works it cites.
First is better than last for training data influence
Chih-Kuan Yeh, Ankur Taly, Mukund Sundararajan, Frederick Liu, and Pradeep Ravikumar · 2022
Later among the works it cites.
More than a toy: Random matrix models predict how real-world neural representations generalize
Alexander Wei, Wei Hu, and Jacob Steinhardt · 2022
Later among the works it cites.
A kernel-based view of language model fine-tuning
Sadhika Malladi, Alexander Wettig, Dingli Yu, Danqi Chen, and Sanjeev Arora · 2022
Later among the works it cites.
Faith-shap: The faithful shapley interaction index
Che-Ping Tsai, Chih-Kuan Yeh, and Pradeep Ravikumar · 2023
Closest in time.
Trak: Attributing model behavior at scale
Sung Min Park, Kristian Georgiev, Andrew Ilyas, Guillaume Leclerc, and Aleksander Madry · 2023
Closest in time.
Gradient-based automated iterative recovery for parameter-efficient tuning
Maximilian Mozes, Tolga Bolukbasi, Ann Yuan, Frederick Liu, Nithum Thain, and Lucas Dixon · 2023
Closest in time.