Fetching the paper…
Reading the bibliography…
As machine learning black boxes are increasingly being deployed in domains such as healthcare and criminal justice, there is growing emphasis on building tools and techniques for explaining these black boxes in an interpretable manner.
Repository of Machine Learning
C Blake, E Koegh, and CJ Mertz. 1999 · 1999
Earlier work this paper cites.
A data-driven software tool for enabling cooperative information sharing among police departments
Michael Redmond and Alok Baveja. 2002 · 2002
Earlier work this paper cites.
UCI machine learning repository, 2007
Arthur Asuncion and David Newman. 2007 · 2007
Earlier work this paper cites.
Communities and crime unnormalized data set
M Redmond. 2011 · 2011
Earlier work this paper cites.
Machine bias
Julia Angwin, Jeff Larson, Surya Mattu, and Lauren Kirchner. 2016 · 2016
Earlier work this paper cites.
Accountable algorithms
Joshua A Kroll, Solon Barocas, Edward W Felten, Joel R Reidenberg, David G Robinson, and Harlan Yu. 2016 · 2016
Earlier work this paper cites.
How we analyzed the COMPAS recidivism algorithm
Jeff Larson, Surya Mattu, Lauren Kirchner, and Julia Angwin. 2016 · 2016
Earlier work this paper cites.
Regulation (EU) 2016/679 of the European Parliament and of the Council of 27 April 2016 on the protection of natural persons with regard to the processing of personal data and on the free movement of such data, and repealing Directive 95/46
General Data Protection Regulation. 2016 · 2016
Earlier work this paper cites.
"Why Should I Trust You?": Explaining the Predictions of Any Classifier. In Knowledge Discovery and Data Mining (KDD)
Marco Tulio Ribeiro, Sameer Singh, and Carlos Guestrin. 2016 · 2016
Earlier work this paper cites.
Mapping chemical performance on molecular structures using locally interpretable explanations
Leanne S Whitmore, Anthe George, and Corey M Hudson. 2016 · 2016
Cited alongside, same era.
Towards a rigorous science of interpretable machine learning
Finale Doshi-Velez and Been Kim. 2017 · 2017
Cited alongside, same era.
A Unified Approach to Interpreting Model Predictions
Scott M Lundberg and Su-In Lee. 2017 · 2017
Cited alongside, same era.
Interpretability Beyond Feature Attribution: Quantitative Testing with Concept Activation Vectors (TCAV). In Proceedings of the 35th International Conference on Machine Learning (Proceedings of Machine Learning Research) , Jennifer Dy and Andreas Krause (Eds.), Vol. 80. PMLR, Stockholmsmässan, Stockholm Sweden, 2668–2677
Been Kim, Martin Wattenberg, Justin Gilmer, Carrie Cai, James Wexler, Fernanda Viegas, and Rory S ayres. 2018 · 2018
Cited alongside, same era.
Fairwashing: the risk of rationalization. In International Conference on Machine Learning . 161–170
Ulrich Aivodji, Hiromi Arai, Olivier Fortineau, Sébastien Gambs, Satoshi Hara, and Alain Tapp. 2019 · 2019
Closest in time.
Explanations can be manipulated and geometry is to blame
Ann-Kathrin Dombrowski, Maximilian Alber, Christopher J Anders, Marcel Ackermann, Klaus-Robert Müller, and Pan Kessel. 2019 · 2019
Closest in time.
On the interpretability of machine learning-based model for predicting hypertension
Radwa Elshawi, Mouaz H Al-Mallah, and Sherif Sakr. 2019 · 2019
Closest in time.
Interpretation of neural networks is fragile. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 33. 3681–3688
Amirata Ghorbani, Abubakar Abid, and James Zou. 2019 · 2019
Closest in time.
Fooling Neural Network Interpretations via Adversarial Model Manipulation
Juyeon Heo, Sunghwan Joo, and Taesup Moon. 2019 · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
The Mythos of Model Interpretability
Zachary C. Lipton. 2018 · 2018
Cited alongside, same era.
Anchors: High-precision model-agnostic explanations. In Thirty-Second AAAI Conference on Artificial Intelligence
Marco Tulio Ribeiro, Sameer Singh, and Carlos Guestrin. 2018 · 2018
Cited alongside, same era.
The intuitive appeal of explainable machines
Andrew D Selbst and Solon Barocas. 2018 · 2018
Cited alongside, same era.
Distill-and-compare: auditing black-box models using transparent model distillation. In Proceedings of the 2018 AAAI/ACM Conference on AI, Ethics, and Society . ACM, 303–310
Sarah Tan, Rich Caruana, Giles Hooker, and Yin Lou. 2018 · 2018
Cited alongside, same era.
Closest in time.
Global Explanations of Neural Networks: Mapping the Landscape of Predictions. In Proceedings of the 2019 AAAI/ACM Conference on AI, Ethics, and Society (AIES ’19) . 279–287
Mark Ibrahim, Melissa Louie, Ceena Modarres, and John Paisley. 2019 · 2019
Closest in time.
Explaining explanations in AI. In Proceedings of the conference on fairness, accountability, and transparency . ACM, 279–288
Brent Mittelstadt, Chris Russell, and Sandra Wachter. 2019 · 2019
Closest in time.
Stop explaining black box machine learning models for high stakes decisions and use interpretable models instead
Cynthia Rudin. 2019 · 2019
Closest in time.