Fetching the paper…
Reading the bibliography…
There is a disconnect between explanatory artificial intelligence (XAI) methods and the types of explanations that are useful for and demanded by society (policy makers, government officials, etc.) Questions that experts in artificial intelligence (AI) ask opaque systems provide inside explanations, focused on debugging, reliability, and validation.
Survey and critique of techniques for extracting rules from trained artificial neural networks
Robert Andrews, Joachim Diederich, and Alan B Tickle · 1995
Earlier work this paper cites.
Greedy function approximation: a gradient boosting machine
Jerome H Friedman · 2001
Earlier work this paper cites.
Leakage in data mining: Formulation, detection, and avoidance
Shachar Kaufman, Saharon Rosset, Claudia Perlich, and Ori Stitelman · 2012
Earlier work this paper cites.
Understanding variable importances in forests of randomized trees
Gilles Louppe, Louis Wehenkel, Antonio Sutera, and Pierre Geurts · 2013
Earlier work this paper cites.
The bayesian case model: A generative approach for case-based reasoning and prototype classification
Been Kim, Cynthia Rudin, and Julie A Shah · 2014
Earlier work this paper cites.
Cnn features off-the-shelf: an astounding baseline for recognition
Ali Sharif Razavian, Hossein Azizpour, Josephine Sullivan, and Stefan Carlsson · 2014
Earlier work this paper cites.
Algorithms for interpretable machine learning
Cynthia Rudin · 2014
Earlier work this paper cites.
Explaining prediction models and individual predictions with feature contributions
Erik Štrumbelj and Igor Kononenko · 2014
Earlier work this paper cites.
Machines without principals: liability rules and artificial intelligence
David C Vladeck · 2014
Earlier work this paper cites.
How transferable are features in deep neural networks?
Jason Yosinski, Jeff Clune, Yoshua Bengio, and Hod Lipson · 2014
Earlier work this paper cites.
Droid-sec: deep learning in android malware detection
Zhenlong Yuan, Yongqiang Lu, Zhaoguo Wang, and Yibo Xue · 2014
Earlier work this paper cites.
Control use of data to protect privacy
Susan Landau · 2015
Earlier work this paper cites.
Interpretable classifiers using rules and bayesian analysis: Building a better stroke prediction model
Benjamin Letham, Cynthia Rudin, Tyler H McCormick, David Madigan, et al · 2015
Earlier work this paper cites.
End to end learning for self-driving cars
Mariusz Bojarski, Davide Del Testa, Daniel Dworakowski, Bernhard Firner, Beat Flepp, Prasoon Goyal, Lawrence D Jackel, Mathew Monfort, Urs Muller, Jiakai Zhang, et al · 2016
Cited alongside, same era.
European union regulations on algorithmic decision-making and a" right to explanation"
Bryce Goodman and Seth Flaxman · 2016
Cited alongside, same era.
Privacy is an essentially contested concept: a multi-dimensional analytic for mapping privacy
Deirdre K Mulligan, Colin Koopman, and Nick Doty · 2016
Cited alongside, same era.
Grad-cam: Visual explanations from deep networks via gradient-based localization
Ramprasaath R Selvaraju, Michael Cogswell, Abhishek Das, Ramakrishna Vedantam, Devi Parikh, and Dhruv Batra · 2016
Cited alongside, same era.
Learning deep features for discriminative localization
Detecting bias in black-box models using transparent model distillation
Sarah Tan, Rich Caruana, Giles Hooker, and Yin Lou · 2017
Later among the works it cites.
Causal interpretations of black-box models
Qingyuan Zhao and Trevor Hastie · 2017
Later among the works it cites.
This looks like that: deep learning for interpretable image recognition
Chaofan Chen, Oscar Li, Alina Barnett, Jonathan Su, and Cynthia Rudin · 2018
Later among the works it cites.
Autonomous vehicles: No driver… no regulation?
Joan Claybrook and Shaun Kildare · 2018
Later among the works it cites.
Amazon scraps secret AI recruiting tool that showed bias against women
Jeffrey Dastin · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Bolei Zhou, Aditya Khosla, Agata Lapedriza, Aude Oliva, and Antonio Torralba · 2016
Cited alongside, same era.
Deepred–rule extraction from deep neural networks
Jan Ruben Zilke, Eneldo Loza Mencía, and Frederik Janssen · 2016
Cited alongside, same era.
Network dissection: Quantifying interpretability of deep visual representations
David Bau, Bolei Zhou, Aditya Khosla, Aude Oliva, and Antonio Torralba · 2017
Cited alongside, same era.
Towards a rigorous science of interpretable machine learning
Finale Doshi-Velez and Been Kim · 2017
Cited alongside, same era.
Accountability of AI under the law: The role of explanation
Finale Doshi-Velez, Mason Kortz, Ryan Budish, Chris Bavitz, Sam Gershman, David O’Brien, Stuart Schieber, James Waldo, David Weinberger, and Alexandra Wood · 2017
Cited alongside, same era.
Explainable artificial intelligence (xai)
David Gunning · 2017
Cited alongside, same era.
Tcav: Relative concept importance testing with linear concept activation vectors
Been Kim, Justin Gilmer, Fernanda Viegas, Ulfar Erlingsson, and Martin Wattenberg · 2017
Cited alongside, same era.
Understanding black-box predictions via influence functions
Pang Wei Koh and Percy Liang · 2017
Cited alongside, same era.
Leilani H Gilpin, David Bau, Ben Z Yuan, Ayesha Bajwa, Michael Specter, and Lalana Kagal · 2018
Later among the works it cites.
Report: Software bug led to death in Uber’s self-driving crash, May 2018
Timothy B. Lee · 2018
Later among the works it cites.
The Uber Crash Won’t Be the Last Shocking Self-Driving Death
Aarian Marshall · 2018
Later among the works it cites.
Saving governance-by-design
Deirdre K Mulligan and Kenneth A Bamberger · 2018
Later among the works it cites.
Multimodal explanations: Justifying decisions and pointing to the evidence
Dong Huk Park, Lisa Anne Hendricks, Zeynep Akata, Anna Rohrbach, Bernt Schiele, Trevor Darrell, and Marcus Rohrbach · 2018
Later among the works it cites.
Defining explainable ai for requirements analysis
Raymond Sheh and Isaac Monteath · 2018
Later among the works it cites.
Interpreting neural network judgments via minimal, stable, and symbolic corrections
Xin Zhang, Armando Solar-Lezama, and Rishabh Singh · 2018
Later among the works it cites.