Fetching the paper…
Reading the bibliography…
Local explanation frameworks aim to rationalize particular decisions made by a black-box prediction model.
Inference and missing data
Rubin, D. B. (1976) · 1976
Earlier work this paper cites.
A density-based algorithm for discovering clusters a density-based algorithm for discovering clusters in large spatial databases with noise
Ester, M., Kriegel, H.-P., Sander, J., and Xu, X. (1996) · 1996
Earlier work this paper cites.
Long short-term memory
Hochreiter, S. and Schmidhuber, J. (1997) · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
LeCun, Y., Bottou, L., Bengio, Y., and Haffner, P. (1998) · 1998
Earlier work this paper cites.
Gradient-based learning applied to document recognition
LeCun, Y., Bottou, L., Bengio, Y., and Haffner, P. (1998) · 1998
Earlier work this paper cites.
Kernel methods for missing variables
Smola, A. J., Vishwanathan, S., and Hofmann, T. (2005) · 2005
Earlier work this paper cites.
How to explain individual classification decisions
Baehrens, D., Schroeter, T., Harmeling, S., Kawanabe, M., Hansen, K., and Müller, K.-R. (2010) · 2010
Earlier work this paper cites.
Learning attitudes and attributes from multi-aspect reviews
McAuley, J., Leskovec, J., and Jurafsky, D. (2012) · 2012
Earlier work this paper cites.
An integrated encyclopedia of dna elements in the human genome
Consortium, E. P. et al. (2012) · 2012
Earlier work this paper cites.
Learning attitudes and attributes from multi-aspect reviews
McAuley, J., Leskovec, J., and Jurafsky, D. (2012) · 2012
Earlier work this paper cites.
Adadelta: An adaptive learning rate method
Zeiler, M. D. (2012) · 2012
Earlier work this paper cites.
Deep inside convolutional networks: Visualising image classification models and saliency maps
Simonyan, K., Vedaldi, A., and Zisserman, A. (2014) · 2014
Earlier work this paper cites.
Visualizing and understanding convolutional networks
Zeiler, M. D. and Fergus, R. (2014) · 2014
Earlier work this paper cites.
On pixel-wise explanations for non-linear classifier decisions by layer-wise relevance propagation
Bach, S., Binder, A., Montavon, G., Klauschen, F., Müller, K.-R., and Samek, W. (2015) · 2015
Earlier work this paper cites.
Intelligible models for healthcare: Predicting pneumonia risk and hospital 30-day readmission
Caruana, R., Lou, Y., Gehrke, J., Koch, P., Sturm, M., and Elhadad, N. (2015) · 2015
Earlier work this paper cites.
Striving for simplicity: The all convolutional net
Springenberg, J. T., Dosovitskiy, A., Brox, T., and Riedmiller, M. (2015) · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J. (2015) · 2015
Earlier work this paper cites.
Rationalizing neural predictions
Lei, T., Barzilay, R., and Jaakkola, T. (2016) · 2016
Cited alongside, same era.
The mythos of model interpretability
Lipton, Z. C. (2016) · 2016
Cited alongside, same era.
Jaspar 2016: a major expansion and update of the open-access database of transcription factor binding profiles
Mathelier, A., Fornes, O., Arenillas, D. J., Chen, C.-y., Denay, G., Lee, J., Shi, W., Shyr, C., Tan, G., Worsley-Hunt, R., et al. (2015) · 2016
Cited alongside, same era.
"Why should I trust you?": Explaining the predictions of any classifier
Ribeiro, M. T., Singh, S., and Guestrin, C. (2016) · 2016
Cited alongside, same era.
Energy distance
Rizzo, M. L. and Székely, G. J. (2016) · 2016
Cited alongside, same era.
Stealing machine learning models via prediction APIs
Tramer, F., Zhang, F., Juels, A., Reiter, M. K., and Ristenpart, T. (2016) · 2016
Cited alongside, same era.
The (un) reliability of saliency methods
Kindermans, P.-J., Hooker, S., Adebayo, J., Alber, M., Schütt, K. T., Dähne, S., Erhan, D., and Kim, B. (2017) · 2017
Later among the works it cites.
Understanding neural networks through representation erasure
Li, J., Monroe, W., and Jurafsky, D. (2017) · 2017
Later among the works it cites.
Feature visualization
Olah, C., Mordvintsev, A., and Schubert, L. (2017) · 2017
Later among the works it cites.
Interpretable predictions of clinical outcomes with an attention-based recurrent neural network
Sha, Y. and Wang, M. D. (2017) · 2017
Later among the works it cites.
Learning important features through propagating activation differences
Shrikumar, A., Greenside, P., and Kundaje, A. (2017) · 2017
Later among the works it cites.
Axiomatic attribution for deep networks
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Convolutional neural network architectures for predicting dna–protein binding
Zeng, H., Edwards, M. D., Liu, G., and Gifford, D. K. (2016) · 2016
Cited alongside, same era.
Rationalizing neural predictions
Lei, T., Barzilay, R., and Jaakkola, T. (2016) · 2016
Cited alongside, same era.
"Why should I trust you?": Explaining the predictions of any classifier
Ribeiro, M. T., Singh, S., and Guestrin, C. (2016) · 2016
Cited alongside, same era.
Energy distance
Rizzo, M. L. and Székely, G. J. (2016) · 2016
Cited alongside, same era.
Attention-based lstm for aspect-level sentiment classification
Wang, Y., Huang, M., Zhao, L., et al. (2016) · 2016
Cited alongside, same era.
Convolutional neural network architectures for predicting dna–protein binding
Zeng, H., Edwards, M. D., Liu, G., and Gifford, D. K. (2016) · 2016
Cited alongside, same era.
Sundararajan, M., Taly, A., and Yan, Q. (2017) · 2017
Later among the works it cites.
cleverhans v2.0.0: an adversarial machine learning library
Papernot, N., Carlini, N., Goodfellow, I., Feinman, R., Faghri, F., Matyasko, A., Hambardzumyan, K., Juang, Y.-L., Kurakin, A., Sheatsley, R., Garg, A., and Lin, Y.-C. (2017) · 2017
Later among the works it cites.
Learning to generate reviews and discovering sentiment
Radford, A., Jozefowicz, R., and Sutskever, I. (2017) · 2017
Later among the works it cites.
Axiomatic attribution for deep networks
Sundararajan, M., Taly, A., and Yan, Q. (2017) · 2017
Later among the works it cites.
Learning to explain: An information-theoretic perspective on model interpretation
Chen, J., Song, L., Wainwright, M. J., and Jordan, M. I. (2018) · 2018
Closest in time.
Interpretability beyond feature attribution: Quantitative testing with concept activation vectors (TCAV)
Kim, B., Wattenberg, M., Gilmer, J., Cai, C., Wexler, J., Viegas, F., and Sayres, R. (2018) · 2018
Closest in time.
Learning how to explain neural networks: PatternNet and PatternAttribution
Kindermans, P.-J., Schütt, K. T., Alber, M., Müller, K.-R., Erhan, D., Kim, B., and Dähne, S. (2018) · 2018
Closest in time.
Human decisions and machine predictions
Kleinberg, J., Lakkaraju, H., Leskovec, J., Ludwig, J., and Mullainathan, S. (2018) · 2018
Closest in time.
Beyond word importance: Contextual decomposition to extract interactions from LSTMs
Murdoch, W. J., Liu, P. J., and Yu, B. (2018) · 2018
Closest in time.
The building blocks of interpretability
Olah, C., Satyanarayan, A., Johnson, I., Carter, S., Schubert, L., Ye, K., and Mordvintsev, A. (2018) · 2018
Closest in time.
Deep learning for mortgage risk
Sirignano, J. A., Sadhwani, A., and Giesecke, K. (2018) · 2018
Closest in time.