Fetching the paper…
Reading the bibliography…
Explainable AI (XAI) aims to make AI systems more transparent, yet many practices emphasise mathematical rigour over practical user needs.
The principle of minimized iterations in the solution of the matrix eigenvalue problem
Walter Edwin Arnoldi. 1951 · 1951
Earlier work this paper cites.
The Influence Curve and its Role in Robust Estimation
Frank R. Hampel. 1974 · 1974
Earlier work this paper cites.
Statlog (German Credit Data)
Hans Hofmann. 1994 · 1994
Earlier work this paper cites.
Scenario-based design: envisioning work and technology in system development
John M. Carroll (Ed.). 1995 · 1995
Earlier work this paper cites.
Bonferroni correction
Eric W Weisstein. 2004 · 2004
Earlier work this paper cites.
Computing Krippendorff’s alpha-reliability
Klaus Krippendorff. 2011 · 2011
Earlier work this paper cites.
Scikit-learn: Machine Learning in Python
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V. Dubourg, J. Vanderplas, A. Passos, D. Cournapeau, M. Brucher, M. Perrot, and E. Duchesnay. 2011 · 2011
Earlier work this paper cites.
The Caltech-UCSD Birds-200-2011 Dataset
C. Wah, S. Branson, P. Welinder, P. Perona, and S. Belongie. 2011 · 2011
Earlier work this paper cites.
Thematic analysis
Virginia Braun and Victoria Clarke. 2012 · 2012
Earlier work this paper cites.
Deep inside convolutional networks: Visualising image classification models and saliency maps
Karen Simonyan, Andrea Vedaldi, and Andrew Zisserman. 2013 · 2013
Earlier work this paper cites.
Using entropy-related measures in categorical data visualization. In 2014 IEEE Pacific Visualization Symposium . IEEE, 81–88
Jamal Alsakran, Xiaoke Huang, Ye Zhao, Jing Yang, and Karl Fast. 2014 · 2014
Earlier work this paper cites.
Deep residual learning for image recognition. In Proceedings of the IEEE conference on computer vision and pattern recognition . 770–778
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Earlier work this paper cites.
" Why should i trust you?" Explaining the predictions of any classifier. In Proceedings of the 22nd ACM SIGKDD international conference on knowledge discovery and data mining . 1135–1144
Marco Tulio Ribeiro, Sameer Singh, and Carlos Guestrin. 2016 · 2016
Earlier work this paper cites.
Second-order stochastic optimization for machine learning in linear time
Naman Agarwal, Brian Bullins, and Elad Hazan. 2017 · 2017
Earlier work this paper cites.
Towards a rigorous science of interpretable machine learning
Finale Doshi-Velez and Been Kim. 2017 · 2017
Earlier work this paper cites.
Understanding Black-box Predictions via Influence Functions. In Proceedings of the 34th International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 70) , Doina Precup and Yee Whye Teh (Eds.). PMLR, 1885–1894
Pang Wei Koh and Percy Liang. 2017 · 2017
Earlier work this paper cites.
A unified approach to interpreting model predictions
Scott M Lundberg and Su-In Lee. 2017 · 2017
Earlier work this paper cites.
Smoothgrad: removing noise by adding noise
Daniel Smilkov, Nikhil Thorat, Been Kim, Fernanda Viégas, and Martin Wattenberg. 2017 · 2017
Earlier work this paper cites.
Axiomatic attribution for deep networks. In International conference on machine learning . PMLR, 3319–3328
Mukund Sundararajan, Ankur Taly, and Qiqi Yan. 2017 · 2017
Earlier work this paper cites.
Sanity checks for saliency maps
Julius Adebayo, Justin Gilmer, Michael Muelly, Ian Goodfellow, Moritz Hardt, and Been Kim. 2018 · 2018
Earlier work this paper cites.
A Survey of Methods for Explaining Black Box Models
Riccardo Guidotti, Anna Monreale, Salvatore Ruggieri, Franco Turini, Fosca Giannotti, and Dino Pedreschi. 2018 · 2018
Earlier work this paper cites.
Understanding the origins of bias in word embeddings. In International conference on machine learning . PMLR, 803–811
Marc-Etienne Brunet, Colleen Alkalay-Houlihan, Ashton Anderson, and Richard Zemel. 2019 · 2019
Earlier work this paper cites.
Input Similarity from the Neural Network Perspective. In Advances in Neural Information Processing Systems , H. Wallach, H. Larochelle, A. Beygelzimer, F. d'Alché-Buc, E. Fox, and R. Garnett (Eds.), Vol. 32. Curran Associates, Inc
Guillaume Charpiat, Nicolas Girard, Loris Felardos, and Yuliya Tarabalka. 2019 · 2019
Earlier work this paper cites.
Data shapley: Equitable valuation of data for machine learning. In International conference on machine learning . PMLR, 2242–2251
Amirata Ghorbani and James Zou. 2019 · 2019
Earlier work this paper cites.
Data cleansing for models trained with SGD
Satoshi Hara, Atsushi Nitanda, and Takanori Maehara. 2019 · 2019
Earlier work this paper cites.
On the accuracy of influence functions for measuring group effects
Pang Wei W Koh, Kai-Siang Ang, Hubert Teo, and Percy S Liang. 2019 · 2019
Earlier work this paper cites.
Explanation in artificial intelligence: Insights from the social sciences
Tim Miller. 2019 · 2019
Earlier work this paper cites.
PyTorch: An Imperative Style, High-Performance Deep Learning Library
Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, Alban Desmaison, Andreas Kopf, Edward Yang, Zachary DeVito, Martin Raison, Alykhan Tejani, Sasank Chilamkurthy, Benoit Steiner, Lu Fang, Junjie Bai, and Soumith Chintala. 2019 · 2019
Earlier work this paper cites.
Considerations on explainable AI and users’ mental models. In CHI 2019 Workshop: Where is the human? Bridging the gap between AI and HCI . Association for Computing Machinery, Inc
Heleen Rutjes, Martijn Willemsen, and Wijnand IJsselsteijn. 2019 · 2019
Earlier work this paper cites.
Designing theory-driven user-centric explainable AI. In Proceedings of the 2019 CHI conference on human factors in computing systems . 1–15
Danding Wang, Qian Yang, Ashraf Abdul, and Brian Y Lim. 2019 · 2019
Earlier work this paper cites.
Explainability scenarios: towards scenario-based XAI design. In Proceedings of the 24th International Conference on Intelligent User Interfaces . 252–257
Christine T Wolf. 2019 · 2019
Earlier work this paper cites.
On second-order group influence functions for black-box predictions. In International Conference on Machine Learning . PMLR, 715–724
Samyadeep Basu, Xuchen You, and Soheil Feizi. 2020 · 2020
Earlier work this paper cites.
What Do People Really Want When They Say They Want "Explainable AI?" We Asked 60 Stakeholders.. In Extended Abstracts of the 2020 CHI Conference on Human Factors in Computing Systems (Honolulu, HI, USA) (CHI EA ’20) . Association for Computing Machinery, New York, NY, USA, 1–7
Andrea Brennen. 2020 · 2020
Cited alongside, same era.
Human-Centered Explainable AI: Towards a Reflective Sociotechnical Approach. In HCI International 2020 - Late Breaking Papers: Multimodality and Intelligence , Constantine Stephanidis, Masaaki Kurosu, Helmut Degen, and Lauren Reinerman-Jones (Eds.). Springer International Publishing, Cham, 449–466
Upol Ehsan and Mark O. Riedl. 2020 · 2020
Cited alongside, same era.
What neural networks memorize and why: Discovering the long tail via influence estimation
Vitaly Feldman and Chiyuan Zhang. 2020 · 2020
Cited alongside, same era.
Understanding and Visualizing Data Iteration in Machine Learning. In Proceedings of the 2020 CHI Conference on Human Factors in Computing Systems (Honolulu, HI, USA) (CHI ’20) . Association for Computing Machinery, New York, NY, USA, 1–13
"Help Me Help the AI": Understanding How Explainability Can Support Human-AI Interaction. In Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems (Hamburg, Germany) (CHI ’23) . Association for Computing Machinery, New York, NY, USA, Article 250, 17 pages
Sunnie S. Y. Kim, Elizabeth Anne Watkins, Olga Russakovsky, Ruth Fong, and Andrés Monroy-Hernández. 2023 · 2023
Later among the works it cites.
Datainf: Efficiently estimating data influence in lora-tuned llms and diffusion models
Yongchan Kwon, Eric Wu, Kevin Wu, and James Zou. 2023 · 2023
Later among the works it cites.
Users are the north star for ai transparency
Alex Mei, Michael Saxon, Shiyu Chang, Zachary C Lipton, and William Yang Wang. 2023 · 2023
Later among the works it cites.
From Anecdotal Evidence to Quantitative Evaluation Methods: A Systematic Review on Evaluating Explainable AI
Meike Nauta, Jan Trienes, Shreyasi Pathak, Elisa Nguyen, Michelle Peters, Yasmin Schmitt, Jörg Schlötterer, Maurice van Keulen, and Christin Seifert. 2023 · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Fred Hohman, Kanit Wongsuphasawat, Mary Beth Kery, and Kayur Patel. 2020 · 2020
Cited alongside, same era.
Human Factors in Model Interpretability: Industry Practices, Challenges, and Needs
Sungsoo Ray Hong, Jessica Hullman, and Enrico Bertini. 2020 · 2020
Cited alongside, same era.
Interpreting Interpretability: Understanding Data Scientists’ Use of Interpretability Tools for Machine Learning. In Proceedings of the 2020 CHI Conference on Human Factors in Computing Systems (Honolulu, HI, USA) (CHI ’20) . Association for Computing Machinery, New York, NY, USA, 1–14
Harmanpreet Kaur, Harsha Nori, Samuel Jenkins, Rich Caruana, Hanna Wallach, and Jennifer Wortman Vaughan. 2020 · 2020
Cited alongside, same era.
Questioning the AI: Informing Design Practices for Explainable AI User Experiences. In Proceedings of the 2020 CHI Conference on Human Factors in Computing Systems (Honolulu, HI, USA) (CHI ’20) . Association for Computing Machinery, New York, NY, USA, 1–15
Q. Vera Liao, Daniel Gruen, and Sarah Miller. 2020 · 2020
Cited alongside, same era.
Estimating Training Data Influence by Tracing Gradient Descent. In Advances in Neural Information Processing Systems , H. Larochelle, M. Ranzato, R. Hadsell, M.F. Balcan, and H. Lin (Eds.), Vol. 33. Curran Associates, Inc., 19920–19930
Garima Pruthi, Frederick Liu, Satyen Kale, and Mukund Sundararajan. 2020 · 2020
Cited alongside, same era.
Influence Functions in Deep Learning Are Fragile. In International Conference on Learning Representations
Samyadeep Basu, Phil Pope, and Soheil Feizi. 2021 · 2021
Cited alongside, same era.
Expanding explainability: Towards social transparency in ai systems. In Proceedings of the 2021 CHI conference on human factors in computing systems . 1–19
Upol Ehsan, Q Vera Liao, Michael Muller, Mark O Riedl, and Justin D Weisz. 2021 · 2021
Cited alongside, same era.
FastIF: Scalable Influence Functions for Efficient Model Interpretation and Debugging. In Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, Online and Punta Cana, Dominican Republic, 10333–10350
Han Guo, Nazneen Rajani, Peter Hase, Mohit Bansal, and Caiming Xiong. 2021 · 2021
Cited alongside, same era.
“Everyone wants to do the model work, not the data work”: Data Cascades in High-Stakes AI. In Proceedings of the 2021 CHI Conference on Human Factors in Computing Systems (<conf-loc>, <city>Yokohama</city>, <country>Japan</country>, </conf-loc>) (CHI ’21) . Association for Computing Machinery, New York, NY, USA, Article 39, 15 pages
Nithya Sambasivan, Shivani Kapania, Hannah Highfill, Diana Akrong, Praveen Paritosh, and Lora M Aroyo. 2021 · 2021
Cited alongside, same era.
Later among the works it cites.
A comprehensive overview of large language models
Humza Naveed, Asad Ullah Khan, Shi Qiu, Muhammad Saqib, Saeed Anwar, Muhammad Usman, Naveed Akhtar, Nick Barnes, and Ajmal Mian. 2023 · 2023
Later among the works it cites.
A Bayesian Approach To Analysing Training Data Attribution in Deep Learning. In Proceedings of the 2023 Conference on Neural Information Processing Systems
Elisa Nguyen, Minjoon Seo, and Seong Joon Oh. 2023 · 2023
Later among the works it cites.
TRAK: Attributing Model Behavior at Scale. In International Conference on Machine Learning (ICML)
Sung Min Park, Kristian Georgiev, Andrew Ilyas, Guillaume Leclerc, and Aleksander Madry. 2023 · 2023
Later among the works it cites.
AI Act, Annex III
European Parliament. 2023 · 2023
Later among the works it cites.
AIMEE: An Exploratory Study of How Rules Support AI Developers to Explain and Edit Models
David Piorkowski, Inge Vejsbjerg, Owen Cornec, Elizabeth M. Daly, and Öznur Alkan. 2023 · 2023
Later among the works it cites.
Robust Speech Recognition via Large-Scale Weak Supervision. In Proceedings of the 40th International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 202) , Andreas Krause, Emma Brunskill, Kyunghyun Cho, Barbara Engelhardt, Sivan Sabato, and Jonathan Scarlett (Eds.). PMLR, 28492–28518
Alec Radford, Jong Wook Kim, Tao Xu, Greg Brockman, Christine Mcleavey, and Ilya Sutskever. 2023 · 2023
Later among the works it cites.
Towards Human-Centered Explainable AI: A Survey of User Studies for Model Explanations
Yao Rong, Tobias Leemann, Thai-Trang Nguyen, Lisa Fiedler, Peizhu Qian, Vaibhav Unhelkar, Tina Seidel, Gjergji Kasneci, and Enkelejda Kasneci. 2023 · 2023
Later among the works it cites.
ConvXAI: Delivering Heterogeneous AI Explanations via Conversations to Support Human-AI Scientific Writing. In Companion Publication of the 2023 Conference on Computer Supported Cooperative Work and Social Computing (Minneapolis, MN, USA) (CSCW ’23 Companion) . Association for Computing Machinery, New York, NY, USA, 384–387
Hua Shen, Chieh-Yang Huang, Tongshuang Wu, and Ting-Hao Kenneth Huang. 2023 · 2023
Later among the works it cites.
Intriguing properties of data attribution on diffusion models
Xiaosen Zheng, Tianyu Pang, Chao Du, Jing Jiang, and Min Lin. 2023 · 2023
Later among the works it cites.
Training Data Attribution via Approximate Unrolled Differentation
Juhan Bae, Wu Lin, Jonathan Lorraine, and Roger Grosse. 2024 · 2024
Closest in time.
Quanda: An Interpretability Toolkit for Training Data Attribution Evaluation and Beyond
Dilyara Bareeva, Galip Ümit Yolcu, Anna Hedström, Niklas Schmolenski, Thomas Wiegand, Wojciech Samek, and Sebastian Lapuschkin. 2024 · 2024
Closest in time.
Impossibility theorems for feature attribution
Blair Bilodeau, Natasha Jaques, Pang Wei Koh, and Been Kim. 2024 · 2024
Closest in time.
What is Your Data Worth to GPT? LLM-Scale Data Valuation with Influence Functions
Sang Keun Choe, Hwijeen Ahn, Juhan Bae, Kewen Zhao, Minsoo Kang, Youngseog Chung, Adithya Pratapa, Willie Neiswanger, Emma Strubell, Teruko Mitamura, et al · 2024
Closest in time.
dattri: A Library for Efficient Data Attribution
Junwei Deng, Ting-Wei Li, Shiyuan Zhang, Shixuan Liu, Yijun Pan, Hao Huang, Xinhe Wang, Pingbang Hu, Xingjian Zhang, and Jiaqi Ma. 2024 · 2024
Closest in time.
The Who in XAI: How AI Background Shapes Perceptions of AI Explanations. In Proceedings of the CHI Conference on Human Factors in Computing Systems (Honolulu, HI, USA) (CHI ’24) . Association for Computing Machinery, New York, NY, USA, Article 316, 32 pages
Upol Ehsan, Samir Passi, Q. Vera Liao, Larry Chan, I-Hsiang Lee, Michael Muller, and Mark O Riedl. 2024 · 2024
Closest in time.
Training data influence analysis and estimation: A survey
Zayd Hammoudeh and Daniel Lowd. 2024 · 2024
Closest in time.
Most Influential Subset Selection: Challenges, Promises, and Beyond. In Advances in Neural Information Processing Systems , A. Globerson, L. Mackey, D. Belgrave, A. Fan, U. Paquet, J. Tomczak, and C. Zhang (Eds.), Vol. 37. Curran Associates, Inc., 119778–119810
Yuzheng Hu, Pingbang Hu, Han Zhao, and Jiaqi Ma. 2024 · 2024
Closest in time.
Data Debiasing with Datamodels (D3M): Improving Subgroup Robustness via Data Selection
Saachi Jain, Kimia Hamidieh, Kristian Georgiev, Andrew Ilyas, Marzyeh Ghassemi, and Aleksander Madry. 2024 · 2024
Closest in time.
DCAI: Data-centric Artificial Intelligence. In Companion Proceedings of the ACM Web Conference 2024 (Singapore, Singapore) (WWW ’24) . Association for Computing Machinery, New York, NY, USA, 1482–1485
Wei Jin, Haohan Wang, Daochen Zha, Qiaoyu Tan, Yao Ma, Sharon Li, and Su-In Lee. 2024 · 2024
Closest in time.
Dissecting users’ needs for search result explanations. In Proceedings of the CHI Conference on Human Factors in Computing Systems (Honolulu, HI, USA) (CHI ’24) . Association for Computing Machinery, New York, NY, USA, Article 841, 17 pages
Prerna Juneja, Wenjuan Zhang, Alison Marie Smith-Renner, Hemank Lamba, Joel Tetreault, and Alex Jaimes. 2024 · 2024
Closest in time.
Delta-Influence: Unlearning Poisons via Influence Functions
Wenjie Li, Jiawei Li, Christian Schroeder de Witt, Ameya Prabhu, and Amartya Sanyal. 2024 · 2024
Closest in time.
Efficient Shapley Values for Attributing Global Properties of Diffusion Models to Data Group
Chris Lin, Mingyu Lu, Chanwoo Kim, and Su-In Lee. 2024 · 2024
Closest in time.
K-Alpha Calculator–Krippendorff’s Alpha Calculator: A user-friendly tool for computing Krippendorff’s Alpha inter-rater reliability coefficient
Giacomo Marzi, Marco Balzano, and Davide Marchiori. 2024 · 2024
Closest in time.
Error discovery by clustering influence embeddings. In Proceedings of the 37th International Conference on Neural Information Processing Systems (New Orleans, LA, USA) (NIPS ’23) . Curran Associates Inc., Red Hook, NY, USA, Article 1809, 13 pages
Fulton Wang, Julius Adebayo, Sarah Tan, Diego Garcia-Olano, and Narine Kokhlikyan. 2024 · 2024
Closest in time.
Optimizing ml training with metagradient descent
Logan Engstrom, Andrew Ilyas, Benjamin Chen, Axel Feldmann, William Moses, and Aleksander Madry. 2025 · 2025
Closest in time.
MAGIC: Near-Optimal Data Attribution for Deep Learning
Andrew Ilyas and Logan Engstrom. 2025 · 2025
Closest in time.
Better Training Data Attribution via Better Inverse Hessian-Vector Products
Andrew Wang, Elisa Nguyen, Runshi Yang, Juhan Bae, Sheila A McIlraith, and Roger Grosse. 2025a · 2025
Closest in time.