Fetching the paper…
Reading the bibliography…
In many modern image-classification applications, understanding the cause of model's prediction can be as critical as the prediction's accuracy itself.
Y. Lecun, L. Bottou, Y. Bengio, and P. Haffner, “Gradient-based learning applied to document recognition,” Proceedings of the IEEE , vol. 86, no. 11, pp. 2278–2324, 1998
1998
Earlier work this paper cites.
L. Fei-Fei, R. Fergus, and P. Perona, “Learning generative visual models from few training examples: An incremental bayesian approach tested on 101 object categories,” 2004 Conference on Computer Vision and Pattern Recognition Workshop , pp. 178–178, 2004
2004
Earlier work this paper cites.
M. Robnik-Šikonja and I. Kononenko, “Explaining classifications for individual instances,” IEEE Transactions on Knowledge and Data Engineering , vol. 20, no. 5, pp. 589–600, May 2008
2008
Earlier work this paper cites.
E. Štrumbelj, I. Kononenko, and M. Robnik Šikonja, “Explaining instance classifications with interactions of subsets of feature values,” Data Knowl. Eng. , vol. 68, no. 10, pp. 886–904, Oct. 2009. [Online]. Available: http://dx.doi.org/10.1016/j.datak.2009.01.004
2009
Earlier work this paper cites.
Y. LeCun and C. Cortes, “MNIST handwritten digit database,” 2010. [Online]. Available: http://yann.lecun.com/exdb/mnist/
2010
Earlier work this paper cites.
A. Krizhevsky, “Learning multiple layers of features from tiny images,” University of Toronto , 05 2012
2012
Earlier work this paper cites.
K. Simonyan, A. Vedaldi, and A. Zisserman, “Deep inside convolutional networks: Visualising image classification models and saliency maps,” in Workshop at International Conference on Learning Representations , 2014
2014
Earlier work this paper cites.
D. Martens and F. Provost, “Explaining data-driven document classifications,” MIS Q. , vol. 38, no. 1, pp. 73–100, Mar. 2014. [Online]. Available: https://doi.org/10.25300/MISQ/2014/38.1.04
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
D. Kingma and J. Ba, “Adam: A method for stochastic optimization,” International Conference on Learning Representations , 12 2014
2014
Earlier work this paper cites.
S. Bach, A. Binder, G. Montavon, F. Klauschen, K.-R. Müller, and W. Samek, “On pixel-wise explanations for non-linear classifier decisions by layer-wise relevance propagation,” PLOS ONE , vol. 10, no. 7, pp. 1–46, 07 2015. [Online]. Available: https://doi.org/10.1371/journal.pone.0130140
2015
Earlier work this paper cites.
J. Springenberg, A. Dosovitskiy, T. Brox, and M. Riedmiller, “Striving for simplicity: The all convolutional net,” in ICLR (workshop track) , 2015. [Online]. Available: http://lmb.informatik.uni-freiburg.de/Publications/2015/DB15a
2015
Cited alongside, same era.
B. Kim, E. Glassman, B. Johnson, and J. Shah, “iBCM : Interactive Bayesian case model empowering humans via intuitive interaction,” 2015
2015
Cited alongside, same era.
2015
Cited alongside, same era.
K. Simonyan and A. Zisserman, “Very deep convolutional networks for large-scale image recognition,” in International Conference on Learning Representations , 2015
2015
Cited alongside, same era.
R. R. Selvaraju, M. Cogswell, A. Das, R. Vedantam, D. Parikh, and D. Batra, “Grad-cam: Visual explanations from deep networks via gradient-based localization,” in 2017 IEEE International Conference on Computer Vision (ICCV) , Oct 2017, pp. 618–626
2017
Later among the works it cites.
A. Shrikumar, P. Greenside, and A. Kundaje, “Learning important features through propagating activation differences,” in Proceedings of the 34th International Conference on Machine Learning , ser. Proceedings of Machine Learning Research, D. Precup and Y. W. Teh, Eds., vol. 70. International Convention Centre, Sydney, Australia: PMLR, 06–11 Aug 2017, pp. 3145–3153. [Online]. Available: http://proceedings.mlr.press/v70/shrikumar17a.html
2017
Later among the works it cites.
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. Liu and W. Deng, “Very deep convolutional neural network based image classification using small training sample size,” 2015 3rd IAPR Asian Conference on Pattern Recognition (ACPR) , pp. 730–734, 2015
2015
Cited alongside, same era.
M. T. Ribeiro, S. Singh, and C. Guestrin, ““why should i trust you?”: Explaining the predictions of any classifier,” in Proceedings of the 22nd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining , ser. KDD ’16. New York, NY, USA: Association for Computing Machinery, 2016, p. 1135–1144. [Online]. Available: https://doi.org/10.1145/2939672.2939778
2016
Cited alongside, same era.
C. Szegedy, V. Vanhoucke, S. Ioffe, J. Shlens, and Z. Wojna, “Rethinking the inception architecture for computer vision,” in 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , June 2016, pp. 2818–2826
2016
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in The IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , June 2016
2016
Cited alongside, same era.
2016
Cited alongside, same era.
A. Esteva, B. Kuprel, R. A. Novoa, J. Ko, S. M. Swetter, H. M. Blau, and S. Thrun, “Dermatologist-level classification of skin cancer with deep neural networks,” Nature , vol. 542, p. 115, Jan 2017. [Online]. Available: https://doi.org/10.1038/nature21056
2017
Cited alongside, same era.
S. M. Lundberg and S.-I. Lee, “A unified approach to interpreting model predictions,” in Advances in Neural Information Processing Systems 30 , I. Guyon, U. V. Luxburg, S. Bengio, H. Wallach, R. Fergus, S. Vishwanathan, and R. Garnett, Eds. Curran Associates, Inc., 2017, pp. 4765–4774. [Online]. Available: http://papers.nips.cc/paper/7062-a-unified-approach-to-interpreting-model-predictions.pdf
2017
Cited alongside, same era.
M. Sundararajan, A. Taly, and Q. Yan, “Axiomatic attribution for deep networks,” in Proceedings of the 34th International Conference on Machine Learning, ICML 2017, Sydney, NSW, Australia, 6-11 August 2017 , ser. Proceedings of Machine Learning Research, D. Precup and Y. W. Teh, Eds., vol. 70. PMLR, 2017, pp. 3319–3328. [Online]. Available: http://proceedings.mlr.press/v70/sundararajan17a.html
2017
Later among the works it cites.
W. Samek, A. Binder, G. Montavon, S. Lapuschkin, and K. Müller, “Evaluating the visualization of what a deep neural network has learned,” IEEE Transactions on Neural Networks and Learning Systems , vol. 28, no. 11, pp. 2660–2673, Nov 2017
2017
Later among the works it cites.
N. Carlini and D. Wagner, “Towards evaluating the robustness of neural networks,” in 2017 IEEE Symposium on Security and Privacy (SP) , May 2017, pp. 39–57
2017
Later among the works it cites.
2017
Later among the works it cites.
2017
Later among the works it cites.
Z. C. Lipton, “The mythos of model interpretability,” Queue , vol. 16, no. 3, p. 31–57, Jun. 2018. [Online]. Available: https://doi.org/10.1145/3236386.3241340
2018
Later among the works it cites.
M. T. Ribeiro, S. Singh, and C. Guestrin, “Anchors: High-precision model-agnostic explanations,” in AAAI , 2018
2018
Later among the works it cites.