Fetching the paper…
Reading the bibliography…
As counterfactual examples become increasingly popular for explaining decisions of deep learning models, it is essential to understand what properties quantitative evaluation metrics do capture and equally important what they do not capture.
Model agnostic contrastive explanations for structured data
Amit Dhurandhar, Tejaswini Pedapati, Avinash Balakrishnan, Pin-Yu Chen, Karthikeyan Shanmugam, and Ruchir Puri · 1906
Earlier work this paper cites.
Interpretable counterfactual explanations guided by prototypes
Arnaud Van Looveren and Janis Klaise · 1907
Earlier work this paper cites.
Preserving causal constraints in counterfactual explanations for machine learning classifiers
Divyat Mahajan, Chenhao Tan, and Amit Sharma · 1912
Earlier work this paper cites.
Counterfactual explanation based on gradual construction for deep networks
Sin-Han Kang, Honggyu Jung, Dong-Ok Won, and Seong-Whan Lee · 2008
Earlier work this paper cites.
MNIST handwritten digit database
Yann LeCun and Corinna Cortes · 2010
Earlier work this paper cites.
Intriguing properties of neural networks
Christian Szegedy, Wojciech Zaremba, Ilya Sutskever, Joan Bruna, Dumitru Erhan, Ian J. Goodfellow, and Rob Fergus · 2014
Earlier work this paper cites.
On pixel-wise explanations for non-linear classifier decisions by layer-wise relevance propagation
Sebastian Bach, Alexander Binder, Grégoire Montavon, Frederick Klauschen, Klaus Robert Müller, and Wojciech Samek · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba · 2015
Earlier work this paper cites.
Explaining nonlinear classification decisions with deep taylor decomposition
Grégoire Montavon, Sebastian Lapuschkin, Alexander Binder, Wojciech Samek, and Klaus-Robert Müller · 2016
Earlier work this paper cites.
"why should I trust you?": Explaining the predictions of any classifier
Marco Túlio Ribeiro, Sameer Singh, and Carlos Guestrin · 2016
Earlier work this paper cites.
Rethinking the inception architecture for computer vision
Christian Szegedy, Vincent Vanhoucke, Sergey Ioffe, Jon Shlens, and Zbigniew Wojna · 2016
Cited alongside, same era.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter · 2017
Cited alongside, same era.
Delving into transferable adversarial examples and black-box attacks
Yanpei Liu, Xinyun Chen, Chang Liu, and Dawn Song · 2017
Cited alongside, same era.
Counterfactual explanations without opening the black box: Automated decisions and the GDPR
Sandra Wachter, Brent Mittelstadt, and Chris Russell · 2017
Cited alongside, same era.
Towards robust interpretability with self-explaining neural networks
David Alvarez Melis and Tommi Jaakkola · 2018
Cited alongside, same era.
Explanations based on the Missing: Towards Contrastive Explanations with Pertinent Negatives
Factual and counterfactual explanations for black box decision making
Riccardo Guidotti, Anna Monreale, Fosca Giannotti, Dino Pedreschi, Salvatore Ruggieri, and Franco Turini · 2019
Later among the works it cites.
Training normalizing flows with the information bottleneck for competitive generative classification
Lynton Ardizzone, Radek Mackowiak, Carsten Rother, and Ullrich Köthe · 2020
Later among the works it cites.
Explaining machine learning classifiers through diverse counterfactual explanations
Ramaravind K Mothilal, Amit Sharma, and Chenhao Tan · 2020
Later among the works it cites.
Learning model-agnostic counterfactual explanations for tabular data
Martin Pawelczyk, Klaus Broelemann, and Gjergji Kasneci · 2020
Later among the works it cites.
FACE: feasible and actionable counterfactual explanations
Rafael Poyiadzi, Kacper Sokol, Raúl Santos-Rodríguez, Tijl De Bie, and Peter A. Flach · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Amit Dhurandhar, Pin Yu Chen, Ronny Luss, Chun Chen Tu, Paishun Ting, Karthikeyan Shanmugam, and Payel Das · 2018
Cited alongside, same era.
Progressive Growing of GANs for Improved Quality, Stability, and Variation
Tero Karras, Timo Aila, Samuli Laine, and Jaakko Lehtinen · 2018
Cited alongside, same era.
Contrastive explanations with local foil trees
Jasper van der Waa, Marcel Robeer, Jurriaan van Diggelen, Matthieu Brinkhuis, and Mark A. Neerincx · 2018
Cited alongside, same era.
Explaining image classifiers by counterfactual generation
Chun Hao Chang, Elliot Creager, Anna Goldenberg, and David Duvenaud · 2019
Cited alongside, same era.
Counterfactual Visual Explanations
Yash Goyal, Ziyan Wu, Jan Ernst, Dhruv Batra, Devi Parikh, and Stefan Lee · 2019
Cited alongside, same era.
Sumedha Singla, Brian Pollack, Junxiang Chen, and Kayhan Batmanghelich · 2020
Later among the works it cites.
ECINN: efficient counterfactuals from invertible neural networks
Frederik Hvilshøj, Alexandros Iosifidis, and Ira Assent · 2021
Closest in time.
Beyond trivial counterfactual explanations with diverse valuable explanations
Pau Rodríguez, Massimo Caccia, Alexandre Lacoste, Lee Zamparo, Issam H. Laradji, Laurent Charlin, and David Vázquez · 2021
Closest in time.
Generating interpretable counterfactual explanations by implicit minimisation of epistemic and aleatoric uncertainties
Lisa Schut, Oscar Key, Rory McGrath, Luca Costabello, Bogdan Sacaleanu, Medb Corcoran, and Yarin Gal · 2021
Closest in time.
A survey of contrastive and counterfactual explanation generation methods for explainable artificial intelligence
Ilia Stepin, José Maria Alonso, Alejandro Catalá, and Martin Pereira-Fariña · 2021
Closest in time.