Fetching the paper…
Reading the bibliography…
Since Artificial Intelligence (AI) software uses techniques like deep lookahead search and stochastic optimization of huge neural networks to fit mammoth datasets, it often results in complex behavior that is difficult for people to understand.
Some Philosophical Problems from the Standpoint of Artificial Intelligence. In Machine Intelligence
J. Mccarthy and P. Hayes. 1969 · 1969
Earlier work this paper cites.
Logic and Conversation
P. Grice. 1975 · 1975
Earlier work this paper cites.
XPLAIN: a system for creating and explaining expert consulting programs
W. Swartout. 1983 · 1983
Earlier work this paper cites.
Intelligent tutoring systems
J. R Anderson, F. Boyle, and B. Reiser. 1985 · 1985
Earlier work this paper cites.
Causal explanation
D. Lewis. 1986 · 1986
Earlier work this paper cites.
Conversational processes and causal explanation
D. Hilton. 1990 · 1990
Earlier work this paper cites.
Explanation, imagination, and confidence in judgment
Derek J Koehler. 1991 · 1991
Earlier work this paper cites.
Explanatory coherence and the induction of properties
S. Sloman. 1997 · 1997
Earlier work this paper cites.
TRIPS: An Integrated Intelligent Problem-Solving Assistant. In AAAI/IAAI
George Ferguson and James F. Allen. 1998 · 1998
Earlier work this paper cites.
Causes and explanations: A structural-model approach. Part I: Causes
J. Halpern and J. Pearl. 2005 · 2005
Earlier work this paper cites.
Simplicity and probability in causal explanation
T. Lombrozo. 2007 · 2007
Earlier work this paper cites.
Assessing demand for intelligibility in context-aware applications. In Proceedings of the 11th International Conference on Ubiquitous Computing
Brian Y Lim and Anind K Dey. 2009 · 2009
Earlier work this paper cites.
Thinking, fast and slow
D. Kahneman. 2011 · 2011
Earlier work this paper cites.
Intelligible models for classification and regression. In KDD
Y. Lou, R. Caruana, and J. Gehrke. 2012 · 2012
Earlier work this paper cites.
A generalized taxonomy of explanations styles for traditional and social recommender systems
A. Papadimitriou, P. Symeonidis, and Y. Manolopoulos. 2012 · 2012
Cited alongside, same era.
Power to the people: The role of humans in interactive machine learning
S. Amershi, M. Cakmak, W. Knox, and T. Kulesza. 2014 · 2014
Cited alongside, same era.
Explaining and Harnessing Adversarial Examples
I. J. Goodfellow, J. Shlens, and C. Szegedy. 2014 · 2014
Cited alongside, same era.
Some Observations on Mental Models
Donald A Norman. 2014 · 2014
Cited alongside, same era.
Visualizing and understanding convolutional networks. In ECCV
M. Zeiler and R. Fergus. 2014 · 2014
Cited alongside, same era.
Neural-Symbolic Learning and Reasoning: A Survey and Interpretation
T. Besold, A. d’Avila Garcez, S. Bader, H. Bowman, P. Domingos, P. Hitzler, K. Kühnberger, L. Lamb, D. Lowd, P. Lima, L. de Penning, G. Pinkas, H. Poon, and G. Zaverucha. 2017 · 2017
Later among the works it cites.
Steps Towards Robust Artificial Intelligence
T. Dietterich. 2017 · 2017
Later among the works it cites.
Towards A Rigorous Science of Interpretable Machine Learning
F. Doshi-Velez and B. Kim. 2017 · 2017
Later among the works it cites.
Explainable Planning. In IJCAI XAI Workshop
M. Fox, D. Long, and D. Magazzeni. 2017 · 2017
Later among the works it cites.
L. A. Hendricks, R. Hu, T. Darrell, and Z. Akata. 2017 · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Intelligible models for healthcare: Predicting pneumonia risk and hospital 30-day readmission. In KDD
R. Caruana, Y. Lou, J. Gehrke, P. Koch, M. Sturm, and N. Elhadad. 2015 · 2015
Cited alongside, same era.
Principles of explanatory debugging to personalize interactive machine learning. In IUI
T. Kulesza, M. Burnett, W. Wong, and S. Stumpf. 2015 · 2015
Cited alongside, same era.
Equality of opportunity in supervised learning. In NIPS
M. Hardt, E. Price, and N. Srebro. 2016 · 2016
Cited alongside, same era.
Generating visual explanations. In ECCV
L. Hendricks, Z. Akata, M. Rohrbach, J. Donahue, B. Schiele, and T. Darrell. 2016 · 2016
Cited alongside, same era.
The Mythos of Model Interpretability. In ICML Workshop on Human Interpretability in ML
Z. Lipton. 2016 · 2016
Cited alongside, same era.
Why Should I Trust You?: Explaining the Predictions of any Classifier. In KDD
M. Ribeiro, S. Singh, and C. Guestrin. 2016 · 2016
Cited alongside, same era.
Mastering the game of Go with deep neural networks and tree search
D. Silver, A. Huang, C. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, et al · 2016
Cited alongside, same era.
Later among the works it cites.
Interpretability Beyond Feature Attribution: Quantitative Testing with Concept Activation Vectors (TCAV)
B. Kim, M. Wattenberg, J. Gilmer, C. Cai, J. Wexler, F. Viegas, and R. Sayres. 2017 · 2017
Later among the works it cites.
Understanding black-box predictions via influence functions. In ICML
P. Koh and P. Liang. 2017 · 2017
Later among the works it cites.
A Workflow for Visual Diagnostics of Binary Classifiers using Instance-Level Explanations
J. Krause, A. Dasgupta, J. Swartz, Y. Aphinyanaphongs, and E. Bertini. 2017 · 2017
Later among the works it cites.
Interpretable & Explorable Approximations of Black Box Models
H. Lakkaraju, E. Kamar, R. Caruana, and J. Leskovec. 2017 · 2017
Later among the works it cites.
A unified approach to interpreting model predictions
S. Lundberg and S. Lee. 2017 · 2017
Later among the works it cites.
Explanation in artificial intelligence: Insights from the social sciences
T. Miller. 2017 · 2017
Later among the works it cites.
Anchors: High-Precision Model-Agnostic Explanations. In AAAI
M. Ribeiro, S. Singh, and C. Guestrin. 2018 · 2018
Closest in time.
Hierarchical Expertise-Level Modeling for User Specific Robot-Behavior Explanations
S. Sreedharan, S. Srivastava, and S. Kambhampati. 2018 · 2018
Closest in time.