Fetching the paper…
Reading the bibliography…
In recent years, the field of explainable AI (XAI) has produced a vast collection of algorithms, providing a useful toolbox for researchers and practitioners to build XAI applications.
Logic and conversation
Herbert P Grice. 1975 · 1975
Earlier work this paper cites.
The presentation of self in everyday life . Vol. 21
Erving Goffman et al · 1978
Earlier work this paper cites.
On user studies and information needs
Tom D Wilson. 1981 · 1981
Earlier work this paper cites.
The elaboration likelihood model of persuasion
Richard E Petty and John T Cacioppo. 1986 · 1986
Earlier work this paper cites.
Conversational processes and causal explanation
Denis J Hilton. 1990 · 1990
Earlier work this paper cites.
Impression management: A literature review and two-component model
Mark R Leary and Robin M Kowalski. 1990 · 1990
Earlier work this paper cites.
Grounding in communication
Herbert H Clark and Susan E Brennan. 1991 · 1991
Earlier work this paper cites.
Explaining in conversation: Towards an argument model
Charles Antaki and Ivan Leudar. 1992 · 1992
Earlier work this paper cites.
Explanation and interaction: the computer generation of explanatory dialogues
Alison Cawsey. 1992 · 1992
Earlier work this paper cites.
Extracting tree-structured representations of trained networks
Mark Craven and Jude Shavlik. 1995 · 1995
Earlier work this paper cites.
Sense-making theory and practice: An overview of user interests in knowledge seeking and use
Brenda Dervin. 1998 · 1998
Earlier work this paper cites.
Relational agents: a model and implementation of building user trust. In Proceedings of the SIGCHI conference on Human factors in computing systems . 396–403
Timothy Bickmore and Justine Cassell. 2001 · 2001
Earlier work this paper cites.
A new dialectical theory of explanation
Douglas Walton. 2004 · 2004
Earlier work this paper cites.
How the mind explains behavior: Folk explanations, meaning, and social interaction
Bertram F Malle. 2006 · 2006
Earlier work this paper cites.
Credibility: A multidisciplinary framework
Soo Young Rieh and David R Danielson. 2007 · 2007
Earlier work this paper cites.
The elements of statistical learnin
Trevor Hastie, Robert Tibshirani, and Jerome Friedman. 2009 · 2009
Earlier work this paper cites.
Assessing demand for intelligibility in context-aware applications. In Proceedings of the 11th international conference on Ubiquitous computing . 195–204
Brian Y Lim and Anind K Dey. 2009 · 2009
Earlier work this paper cites.
Thinking, fast and slow
Daniel Kahneman. 2011 · 2011
Earlier work this paper cites.
The design of everyday things: Revised and expanded edition
Don Norman. 2013 · 2013
Earlier work this paper cites.
On pixel-wise explanations for non-linear classifier decisions by layer-wise relevance propagation
Sebastian Bach, Alexander Binder, Grégoire Montavon, Frederick Klauschen, Klaus-Robert Müller, and Wojciech Samek. 2015 · 2015
Earlier work this paper cites.
Intelligible models for healthcare: Predicting pneumonia risk and hospital 30-day readmission. In Proceedings of KDD
Rich Caruana, Yin Lou, Johannes Gehrke, Paul Koch, Marc Sturm, and Noemie Elhadad. 2015 · 2015
Earlier work this paper cites.
Peeking inside the black box: Visualizing statistical learning with plots of individual conditional expectation
Alex Goldstein, Adam Kapelner, Justin Bleich, and Emil Pitkin. 2015 · 2015
Earlier work this paper cites.
Variable importance analysis: a comprehensive review
Pengfei Wei, Zhenzhou Lu, and Jingwen Song. 2015 · 2015
Earlier work this paper cites.
Examples are not enough, learn to criticize! criticism for interpretability. In Proceedings of NIPS
Been Kim, Rajiv Khanna, and Oluwasanmi O Koyejo. 2016 · 2016
Earlier work this paper cites.
Interacting with predictions: Visual inspection of black-box machine learning models. In Proceedings of the 2016 CHI Conference on Human Factors in Computing Systems . ACM, 5686–5697
Josua Krause, Adam Perer, and Kenney Ng. 2016 · 2016
Earlier work this paper cites.
Interpretable decision sets: A joint framework for description and prediction. In Proceedings of the 22nd ACM SIGKDD international conference on knowledge discovery and data mining . 1675–1684
Himabindu Lakkaraju, Stephen H Bach, and Jure Leskovec. 2016 · 2016
Earlier work this paper cites.
Understanding neural networks through representation erasure
Jiwei Li, Will Monroe, and Dan Jurafsky. 2016 · 2016
Earlier work this paper cites.
Why should i trust you?: Explaining the predictions of any classifier. In Proceedings of KDD
Marco Tulio Ribeiro, Sameer Singh, and Carlos Guestrin. 2016 · 2016
Earlier work this paper cites.
H2O.ai Machine Learning Interpretability
2017 · 2017
Earlier work this paper cites.
Towards a rigorous science of interpretable machine learning
Finale Doshi-Velez and Been Kim. 2017 · 2017
Earlier work this paper cites.
Interpretable & explorable approximations of black box models
Himabindu Lakkaraju, Ece Kamar, Rich Caruana, and Jure Leskovec. 2017 · 2017
Earlier work this paper cites.
A unified approach to interpreting model predictions. In Proceedings of the 31st international conference on neural information processing systems . 4768–4777
Scott M Lundberg and Su-In Lee. 2017 · 2017
Earlier work this paper cites.
Tim Miller, Piers Howe, and Liz Sonenberg. 2017 · 2017
Earlier work this paper cites.
Grad-cam: Visual explanations from deep networks via gradient-based localization. In Proceedings of the IEEE international conference on computer vision . 618–626
Ramprasaath R Selvaraju, Michael Cogswell, Abhishek Das, Ramakrishna Vedantam, Devi Parikh, and Dhruv Batra. 2017 · 2017
Earlier work this paper cites.
Counterfactual explanations without opening the black box: Automated decisions and the GDPR
Sandra Wachter, Brent Mittelstadt, and Chris Russell. 2017 · 2017
Earlier work this paper cites.
Model Interpretation with Skater
2018 · 2018
Earlier work this paper cites.
Peeking inside the black-box: A survey on Explainable Artificial Intelligence (XAI)
Amina Adadi and Mohammed Berrada. 2018 · 2018
Earlier work this paper cites.
Towards robust interpretability with self-explaining neural networks
David Alvarez-Melis and Tommi S Jaakkola. 2018 · 2018
Cited alongside, same era.
Explanations based on the missing: Towards contrastive explanations with pertinent negatives
Amit Dhurandhar, Pin-Yu Chen, Ronny Luss, Chun-Chen Tu, Paishun Ting, Karthikeyan Shanmugam, and Payel Das. 2018a · 2018
Cited alongside, same era.
Improving Simple Models with Confidence Profiles
Amit Dhurandhar, Karthikeyan Shanmugam, Ronny Luss, and Peder A Olsen. 2018b · 2018
Cited alongside, same era.
Bringing transparency design into practice. In 23rd international conference on intelligent user interfaces . 211–223
Malin Eiband, Hanna Schneider, Mark Bilandzic, Julian Fazekas-Con, Mareike Haug, and Heinrich Hussmann. 2018 · 2018
Cited alongside, same era.
Interpretability beyond feature attribution: Quantitative testing with concept activation vectors (tcav). In International conference on machine learning . PMLR, 2668–2677
Explainable Artificial Intelligence (XAI): Concepts, taxonomies, opportunities and challenges toward responsible AI
Alejandro Barredo Arrieta, Natalia Díaz-Rodríguez, Javier Del Ser, Adrien Bennetot, Siham Tabik, Alberto Barbado, Salvador García, Sergio Gil-López, Daniel Molina, Richard Benjamins, et al · 2020
Later among the works it cites.
AI Explainability 360: An Extensible Toolkit for Understanding Data and Machine Learning Models
Vijay Arya, Rachel K. E. Bellamy, Pin-Yu Chen, Amit Dhurandhar, Michael Hind, Samuel C. Hoffman, Stephanie Houde, Q. Vera Liao, Ronny Luss, Aleksandra Mojsilovic, Sami Mourad, Pablo Pedemonte, Ramya Raghavendra, John Richards, Prasanna Sattigeri, Karthikeyan Shanmugam, Moninder Singh, Kush R. Varshney, Dennis Wei, and Yunfeng Zhang. 2020 · 2020
Later among the works it cites.
Proxy tasks and subjective measures can be misleading in evaluating explainable AI systems. In Proceedings of the 25th International Conference on Intelligent User Interfaces . 454–464
Zana Buçinca, Phoebe Lin, Krzysztof Z Gajos, and Elena L Glassman. 2020 · 2020
Later among the works it cites.
Human-centered explainable ai: Towards a reflective sociotechnical approach. In International Conference on Human-Computer Interaction . Springer, 449–466
Upol Ehsan and Mark O Riedl. 2020 · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Been Kim, Martin Wattenberg, Justin Gilmer, Carrie Cai, James Wexler, Fernanda Viegas, et al · 2018
Cited alongside, same era.
Distribution-free predictive inference for regression
Jing Lei, Max G’Sell, Alessandro Rinaldo, Ryan J Tibshirani, and Larry Wasserman. 2018 · 2018
Cited alongside, same era.
The mythos of model interpretability
Zachary C Lipton. 2018 · 2018
Cited alongside, same era.
Explanation in artificial intelligence: Insights from the social sciences
Tim Miller. 2018 · 2018
Cited alongside, same era.
Deep k-nearest neighbors: Towards confident, interpretable and robust deep learning
Nicolas Papernot and Patrick McDaniel. 2018 · 2018
Cited alongside, same era.
Stakeholders in explainable AI
Alun Preece, Dan Harborne, Dave Braines, Richard Tomsett, and Supriyo Chakraborty. 2018 · 2018
Cited alongside, same era.
Anchors: High-precision model-agnostic explanations. In Thirty-Second AAAI Conference on Artificial Intelligence
Marco Tulio Ribeiro, Sameer Singh, and Carlos Guestrin. 2018 · 2018
Cited alongside, same era.
Learning global additive explanations for neural nets using model distillation
Sarah Tan, Rich Caruana, Giles Hooker, Paul Koch, and Albert Gordo. 2018 · 2018
Cited alongside, same era.
Later among the works it cites.
Evaluating Explainable AI: Which Algorithmic Explanations Help Users Predict Model Behavior?. In Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics . 5540–5552
Peter Hase and Mohit Bansal. 2020 · 2020
Later among the works it cites.
Human factors in model interpretability: Industry practices, challenges, and needs
Sungsoo Ray Hong, Jessica Hullman, and Enrico Bertini. 2020 · 2020
Later among the works it cites.
Interpreting Interpretability: Understanding Data Scientists’ Use of Interpretability Tools for Machine Learning. In Proceedings of the 2020 CHI Conference on Human Factors in Computing Systems . 1–14
Harmanpreet Kaur, Harsha Nori, Samuel Jenkins, Rich Caruana, Hanna Wallach, and Jennifer Wortman Vaughan. 2020 · 2020
Later among the works it cites.
Questioning the AI: informing design practices for explainable AI user experiences. In Proceedings of the 2020 CHI Conference on Human Factors in Computing Systems . 1–15
Q Vera Liao, Daniel Gruen, and Sarah Miller. 2020 · 2020
Later among the works it cites.
Why does my model fail? contrastive local explanations for retail forecasting. In Proceedings of the 2020 Conference on Fairness, Accountability, and Transparency . 90–98
Ana Lucic, Hinda Haned, and Maarten de Rijke. 2020 · 2020
Later among the works it cites.
From local explanations to global understanding with explainable AI for trees
Scott M Lundberg, Gabriel Erion, Hugh Chen, Alex DeGrave, Jordan M Prutkin, Bala Nair, Ronit Katz, Jonathan Himmelfarb, Nisha Bansal, and Su-In Lee. 2020 · 2020
Later among the works it cites.
A human-centered agenda for intelligible machine learning
Jennifer Wortman Vaughan and Hanna Wallach. 2020 · 2020
Later among the works it cites.
CheXplain: Enabling Physicians to Explore and Understand Data-Driven, AI-Enabled Medical Imaging Analysis. In Proceedings of the 2020 CHI Conference on Human Factors in Computing Systems . 1–13
Yao Xie, Melody Chen, David Kao, Ge Gao, and Xiang ‘Anthony’ Chen. 2020 · 2020
Later among the works it cites.
Effect of confidence and explanation on accuracy and trust calibration in AI-assisted decision making. In Proceedings of the 2020 Conference on Fairness, Accountability, and Transparency . 295–305
Yunfeng Zhang, Q Vera Liao, and Rachel KE Bellamy. 2020 · 2020
Later among the works it cites.
From human explanation to model interpretability: A framework based on weight of evidence. In AAAI Conference on Human Computation and Crowdsourcing (HCOMP)
David Alvarez-Melis, Harmanpreet Kaur, Hal Daumé III, Hanna Wallach, and Jennifer Wortman Vaughan. 2021 · 2021
Closest in time.
Does the whole exceed its parts? the effect of ai explanations on complementary team performance. In Proceedings of the 2021 CHI Conference on Human Factors in Computing Systems . 1–16
Gagan Bansal, Tongshuang Wu, Joyce Zhou, Raymond Fok, Besmira Nushi, Ece Kamar, Marco Tulio Ribeiro, and Daniel Weld. 2021 · 2021
Closest in time.
To trust or to think: cognitive forcing functions can reduce overreliance on AI in AI-assisted decision-making
Zana Buçinca, Maja Barbara Malaya, and Krzysztof Z Gajos. 2021 · 2021
Closest in time.
I Think I Get Your Point, AI! The Illusion of Explanatory Depth in Explainable AI. In 26th International Conference on Intelligent User Interfaces . 307–317
Michael Chromik, Malin Eiband, Felicitas Buchner, Adrian Krüger, and Andreas Butz. 2021 · 2021
Closest in time.
Expanding explainability: Towards social transparency in ai systems. In Proceedings of the 2021 CHI Conference on Human Factors in Computing Systems . 1–19
Upol Ehsan, Q Vera Liao, Michael Muller, Mark O Riedl, and Justin D Weisz. 2021a · 2021
Closest in time.
The Who in Explainable AI: How AI Background Shapes Perceptions of AI Explanations
Upol Ehsan, Samir Passi, Q Vera Liao, Larry Chan, I Lee, Michael Muller, Mark O Riedl, et al · 2021
Closest in time.
Operationalizing Human-Centered Perspectives in Explainable AI. In Extended Abstracts of the 2021 CHI Conference on Human Factors in Computing Systems . 1–6
Upol Ehsan, Philipp Wintersberger, Q Vera Liao, Martina Mara, Marc Streit, Sandra Wachter, Andreas Riener, and Mark O Riedl. 2021c · 2021
Closest in time.
Datasheets for Datasets
Timnit Gebru, Jamie Morgenstern, Briana Vecchione, Jennifer Wortman Vaughan, Hanna Wallach, Hal Daumé, III, and Kate Crawford. 2021 · 2021
Closest in time.
Explainable active learning (xal) toward ai explanations as interfaces for machine teachers
Bhavya Ghai, Q Vera Liao, Yunfeng Zhang, Rachel Bellamy, and Klaus Mueller. 2021 · 2021
Closest in time.
The false hope of current approaches to explainable artificial intelligence in health care
Marzyeh Ghassemi, Luke Oakden-Rayner, and Andrew L Beam. 2021 · 2021
Closest in time.
Soumya Ghosh, Q Vera Liao, Karthikeyan Natesan Ramamurthy, Jiri Navratil, Prasanna Sattigeri, Kush R Varshney, and Yunfeng Zhang. 2021 · 2021
Closest in time.
Question-Driven Design Process for Explainable AI User Experiences
Q Vera Liao, Milena Pribić, Jaesik Han, Sarah Miller, and Daby Sow. 2021 · 2021
Closest in time.
Interpretable counterfactual explanations guided by prototypes. In Joint European Conference on Machine Learning and Knowledge Discovery in Databases . Springer, 650–665
Arnaud Van Looveren and Janis Klaise. 2021 · 2021
Closest in time.
Model LineUpper: Supporting Interactive Model Comparison at Multiple Levels for AutoML. In 26th International Conference on Intelligent User Interfaces . 170–174
Shweta Narkar, Yunfeng Zhang, Q Vera Liao, Dakuo Wang, and Justin D Weisz. 2021 · 2021
Closest in time.
Anchoring Bias Affects Mental Model Formation and User Reliance in Explainable AI Systems. In 26th International Conference on Intelligent User Interfaces . 340–350
Mahsan Nourani, Chiradeep Roy, Jeremy E Block, Donald R Honeycutt, Tahrima Rahman, Eric Ragan, and Vibhav Gogate. 2021 · 2021
Closest in time.
Manipulating and measuring model interpretability. In Proceedings of the 2021 CHI Conference on Human Factors in Computing Systems . 1–52
Forough Poursabzi-Sangdeh, Daniel G Goldstein, Jake M Hofman, Jennifer Wortman Wortman Vaughan, and Hanna Wallach. 2021 · 2021
Closest in time.
CoFrNets: Interpretable neural architecture inspired by continued fractions. In Advances in Neural Information Processing Systems
Isha Puri, Amit Dhurandhar, Tejaswini Pedapati, Karthikeyan Shanmugam, Dennis Wei, and Kush R. Varshney. 2021 · 2021
Closest in time.
Wait, But Why?: Assessing Behavior Explanation Strategies for Real-Time Strategy Games. In 26th International Conference on Intelligent User Interfaces . 32–42
Justus Robertson, Athanasios Vasileios Kokkinakis, Jonathan Hook, Ben Kirman, Florian Block, Marian F Ursu, Sagarika Patra, Simon Demediuk, Anders Drachen, and Oluseyi Olarewaju. 2021 · 2021
Closest in time.
Beyond Expertise and Roles: A Framework to Characterize the Stakeholders of Interpretable Machine Learning and their Needs. In Proceedings of the 2021 CHI Conference on Human Factors in Computing Systems . 1–16
Harini Suresh, Steven R Gomez, Kevin K Nam, and Arvind Satyanarayan. 2021 · 2021
Closest in time.
Visual, textual or hybrid: the effect of user expertise on different explanations. In 26th International Conference on Intelligent User Interfaces . 109–119
Maxwell Szymanski, Martijn Millecamp, and Katrien Verbert. 2021 · 2021
Closest in time.
Are Explanations Helpful? A Comparative Study of the Effects of Explanations in AI-Assisted Decision-Making. In 26th International Conference on Intelligent User Interfaces . 318–328
Xinru Wang and Ming Yin. 2021 · 2021
Closest in time.
Human-AI Collaboration for UX Evaluation: Effects of Explanation and Synchronization
Mingming Fan, Xianyou Yang, TszTung Yu, Q Vera Liao, and Jian Zhao. 2022 · 2022
Closest in time.
Deciding Fast and Slow: The Role of Cognitive Biases in AI-Assisted Decision-Making. In Proceedings of the ACM Conference on Computer Supported Cooperative Work and Social Computing
Charvi Rastogi, Yunfeng Zhang, Dennis Wei, Kush R. Varshney, Amit Dhurandhar, and Richard Tomsett. 2022 · 2022
Closest in time.