Fetching the paper…
Reading the bibliography…
Interpretable Machine Learning (IML) has become increasingly important in many real-world applications, such as autonomous cars and medical diagnosis, where explanations are significantly preferred to help people better understand how machine learning systems work and further enhance their trust towards systems.
Generalized linear models
J. A. Nelder and R. W. Wedderburn · 1972
Earlier work this paper cites.
Principles of rule-based expert systems
B. G. Buchanan and R. O. Duda · 1983
Earlier work this paper cites.
A survey of decision tree classifier methodology
S. R. Safavian and D. Landgrebe · 1991
Earlier work this paper cites.
An introduction to kernel and nearest-neighbor nonparametric regression
N. S. Altman · 1992
Earlier work this paper cites.
Deep neural networks for object detection
C. Szegedy, A. Toshev, and D. Erhan · 2013
Earlier work this paper cites.
Distilling knowledge from deep networks with applications to healthcare domain
Z. Che, S. Purushotham, R. Khemani, and Y. Liu · 2015
Earlier work this paper cites.
Deep learning
Y. LeCun, Y. Bengio, and G. Hinton · 2015
Earlier work this paper cites.
Interpretable classifiers using rules and bayesian analysis: Building a better stroke prediction model
B. Letham, C. Rudin, T. H. McCormick, and D. Madigan · 2015
Earlier work this paper cites.
Fully convolutional networks for semantic segmentation
J. Long, E. Shelhamer, and T. Darrell · 2015
Earlier work this paper cites.
Deep neural networks are easily fooled: High confidence predictions for unrecognizable images
A. Nguyen, J. Yosinski, and J. Clune · 2015
Earlier work this paper cites.
User trust in intelligent systems: A journey over time
D. Holliday, S. Wilson, and S. Stumpf · 2016
Earlier work this paper cites.
Interpretable decision sets: A joint framework for description and prediction
H. Lakkaraju, S. H. Bach, and J. Leskovec · 2016
Earlier work this paper cites.
Rationalizing neural predictions
T. Lei, R. Barzilay, and T. Jaakkola · 2016
Earlier work this paper cites.
Why should i trust you?: Explaining the predictions of any classifier
M. T. Ribeiro, S. Singh, and C. Guestrin · 2016
Earlier work this paper cites.
Towards a rigorous science of interpretable machine learning
F. Doshi-Velez and B. Kim · 2017
Cited alongside, same era.
Interpretable explanations of black boxes by meaningful perturbation
R. C. Fong and A. Vedaldi · 2017
Cited alongside, same era.
Distilling a neural network into a soft decision tree
N. Frosst and G. Hinton · 2017
Cited alongside, same era.
Interpretation of neural networks is fragile
A. Ghorbani, A. Abid, and J. Zou · 2017
Cited alongside, same era.
The promise and peril of human evaluation for model interpretability
B. Herman · 2017
Cited alongside, same era.
Adversarial detection with model interpretation
N. Liu, H. Yang, and X. Hu · 2018
Later among the works it cites.
Explaining deep learning models using causal inference
T. Narendra, A. Sankaran, D. Vijaykeerthy, and S. Mani · 2018
Later among the works it cites.
Interpreting neural networks with nearest neighbors
E. Wallace, S. Feng, and J. Boyd-Graber · 2018
Later among the works it cites.
Towards interpretation of recommender systems with sorted explanation paths
F. Yang, N. Liu, S. Wang, and X. Hu · 2018
Later among the works it cites.
Interpreting deep visual representations via network dissection
B. Zhou, D. Bau, A. Oliva, and A. Torralba · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
B. Kim, M. Wattenberg, J. Gilmer, C. Cai, J. Wexler, F. Viegas, and R. Sayres · 2017
Cited alongside, same era.
Understanding black-box predictions via influence functions
P. W. Koh and P. Liang · 2017
Cited alongside, same era.
Interpretable active learning
R. L. Phillips, K. H. Chang, and S. A. Friedler · 2017
Cited alongside, same era.
Grad-cam: Visual explanations from deep networks via gradient-based localization
R. R. Selvaraju, M. Cogswell, A. Das, R. Vedantam, D. Parikh, and D. Batra · 2017
Cited alongside, same era.
Transparent, explainable, and accountable ai for robotics
S. Wachter, B. Mittelstadt, and L. Floridi · 2017
Cited alongside, same era.
Trends and trajectories for explainable, accountable and intelligible systems: An hci research agenda
A. Abdul, J. Vermeulen, D. Wang, B. Y. Lim, and M. Kankanhalli · 2018
Cited alongside, same era.
What can ai do for me: Evaluating machine learning interpretations in cooperative play
S. Feng and J. Boyd-Graber · 2018
Cited alongside, same era.
A. Chattopadhyay, P. Manupriya, A. Sarkar, and V. N. Balasubramanian · 2019
Closest in time.
Techniques for interpretable machine learning
M. Du, N. Liu, and X. Hu · 2019
Closest in time.
On attribution of recurrent neural network predictions via additive decomposition
M. Du, N. Liu, F. Yang, S. Ji, and X. Hu · 2019
Closest in time.
Learning interpretable models with causal guarantees
C. Kim and O. Bastani · 2019
Closest in time.
An evaluation of the human-interpretability of explanation
I. Lage, E. Chen, J. He, M. Narayanan, B. Kim, S. Gershman, and F. Doshi-Velez · 2019
Closest in time.
Interpretable deep learning in drug discovery
K. Preuer, G. Klambauer, F. Rippmann, S. Hochreiter, and T. Unterthiner · 2019
Closest in time.
Actionable recourse in linear classification
B. Ustun, A. Spangher, and Y. Liu · 2019
Closest in time.
How sensitive are sensitivity-based explanations?
C.-K. Yeh, C.-Y. Hsieh, A. S. Suggala, D. Inouye, and P. Ravikumar · 2019
Closest in time.