Fetching the paper…
Reading the bibliography…
While the evaluation of explanations is an important step towards trustworthy models, it needs to be done carefully, and the employed metrics need to be well-understood.
On the ratio of two correlated normal random variables
D. V. Hinkley · 1969
Earlier work this paper cites.
Kendall’s advanced theory of statistics
A. Stuart and K. Ord · 1991
Earlier work this paper cites.
Image quality assessment: from error visibility to structural similarity
Zhou Wang, A.C. Bovik, H.R. Sheikh, and E.P. Simoncelli · 2004
Earlier work this paper cites.
How to explain individual classification decisions
David Baehrens, Timon Schroeter, Stefan Harmeling, Motoaki Kawanabe, Katja Hansen, and Klaus-Robert Müller · 2010
Earlier work this paper cites.
Deep inside convolutional networks: Visualising image classification models and saliency maps
Karen Simonyan, Andrea Vedaldi, and Andrew Zisserman · 2013
Earlier work this paper cites.
On pixel-wise explanations for non-linear classifier decisions by layer-wise relevance propagation
Sebastian Bach, Alexander Binder, Grégoire Montavon, Frederick Klauschen, Klaus-Robert Müller, and Wojciech Samek · 2015
Earlier work this paper cites.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2015
Earlier work this paper cites.
Striving for simplicity: The all convolutional net
J Springenberg, Alexey Dosovitskiy, Thomas Brox, and M Riedmiller · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
The shattered gradients problem: If resnets are the answer, then what is the question?
David Balduzzi, Marcus Frean, Lennox Leary, J. P. Lewis, Kurt Wan-Duo Ma, and Brian McWilliams · 2017
Earlier work this paper cites.
Evaluating the visualization of what a deep neural network has learned
Wojciech Samek, Alexander Binder, Grégoire Montavon, Sebastian Lapuschkin, and Klaus-Robert Müller · 2017
Earlier work this paper cites.
Smoothgrad: removing noise by adding noise
Daniel Smilkov, Nikhil Thorat, Been Kim, Fernanda B. Viégas, and Martin Wattenberg · 2017
Earlier work this paper cites.
Axiomatic attribution for deep networks
Mukund Sundararajan, Ankur Taly, and Qiqi Yan · 2017
Cited alongside, same era.
Sanity checks for saliency maps
Julius Adebayo, Justin Gilmer, Michael Muelly, Ian J. Goodfellow, Moritz Hardt, and Been Kim · 2018
Cited alongside, same era.
Towards robust interpretability with self-explaining neural networks
David Alvarez-Melis and Tommi S. Jaakkola · 2018
Cited alongside, same era.
Methods for interpreting and understanding deep neural networks
Grégoire Montavon, Wojciech Samek, and Klaus-Robert Müller · 2018
Cited alongside, same era.
Mukund Sundararajan and Ankur Taly · 2018
Cited alongside, same era.
Top-down neural attention by excitation backprop
Jianming Zhang, Sarah Adel Bargal, Zhe Lin, Jonathan Brandt, Xiaohui Shen, and Stan Sclaroff · 2018
Concise explanations of neural networks using adversarial training
Prasad Chalasani, Jiefeng Chen, Amrita Roy Chowdhury, Xi Wu, and Somesh Jha · 2020
Later among the works it cites.
A survey of safety and trustworthiness of deep neural networks: Verification, testing, adversarial attack and defence, and interpretability
Xiaowei Huang, Daniel Kroening, Wenjie Ruan, James Sharp, Youcheng Sun, Emese Thamo, Min Wu, and Xinping Yi · 2020
Later among the works it cites.
On quantitative aspects of model interpretability
An-phi Nguyen and María Rodríguez Martínez · 2020
Later among the works it cites.
Jim Nilsson and Tomas Akenine-Möller · 2020
Later among the works it cites.
When explanations lie: Why many modified BP attributions fail
Leon Sixt, Maximilian Granz, and Tim Landgraf · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
The (Un)reliability of Saliency Methods
Pieter-Jan Kindermans, Sara Hooker, Julius Adebayo, Maximilian Alber, Kristof T. Schütt, Sven Dähne, Dumitru Erhan, and Been Kim · 2019
Cited alongside, same era.
Unmasking clever hans predictors and assessing what machines really learn
Sebastian Lapuschkin, Stephan Wäldchen, Alexander Binder, Grégoire Montavon, Wojciech Samek, and Klaus-Robert Müller · 2019
Cited alongside, same era.
Gradient-Based Vs. Propagation-Based Explanations: An Axiomatic Comparison
Grégoire Montavon · 2019
Cited alongside, same era.
Pytorch: An imperative style, high-performance deep learning library
Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, et al · 2019
Cited alongside, same era.
Explaining image classifiers by removing input features using generative models
Chirag Agarwal and Anh Nguyen · 2020
Cited alongside, same era.
Evaluating and aggregating feature-based model explanations
Umang Bhatt, Adrian Weller, and José M. F. Moura · 2020
Cited alongside, same era.
Investigating sanity checks for saliency maps with image and text classification
Narine Kokhlikyan, Vivek Miglani, Bilal Alsallakh, Miguel Martin, and Orion Reblitz-Richardson · 2021
Later among the works it cites.
Interpretable deep learning: Interpretations, interpretability, trustworthiness, and beyond
Xuhong Li, Haoyi Xiong, Xingjian Li, Xuanyu Wu, Xiao Zhang, Ji Liu, Jiang Bian, and Dejing Dou · 2021
Later among the works it cites.
Explaining deep neural networks and beyond: A review of methods and applications
Wojciech Samek, Grégoire Montavon, Sebastian Lapuschkin, Christopher J. Anders, and Klaus-Robert Müller · 2021
Later among the works it cites.
Revisiting sanity checks for saliency maps
Gal Yona and Daniel Greenfeld · 2021
Later among the works it cites.
Focus! rating XAI methods and finding biases
Anna Arias-Duart, Ferran Parés, Dario Garcia-Gasulla, and Victor Gimenez-Abalos · 2022
Closest in time.
CLEVR-XAI: A benchmark dataset for the ground truth evaluation of neural network explanations
Leila Arras, Ahmed Osman, and Wojciech Samek · 2022
Closest in time.