Fetching the paper…
Reading the bibliography…
Counterfactuals, serving as one of the emerging type of model interpretations, have recently received attention from both researchers and practitioners.
The need for biases in learning generalizations
Tom M Mitchell. 1980 · 1980
Earlier work this paper cites.
Example-based reasoning
Edwina L Rissland. 1991 · 1991
Earlier work this paper cites.
The interaction of nature and nurture in development: A parallel distributed processing perspective (Parallel Distributed Processing and Cognitive Neuroscience PDP. CNS. 92.6)
JL McClelland. 1992 · 1992
Earlier work this paper cites.
Pattern recognition and machine learning
Christopher M Bishop. 2006 · 2006
Earlier work this paper cites.
Structural equations and causation
Ned Hall. 2007 · 2007
Earlier work this paper cites.
Nonlinear causal discovery with additive noise models
Patrik Hoyer, Dominik Janzing, Joris M Mooij, Jonas Peters, et al · 2008
Earlier work this paper cites.
Probabilistic models of cognition: Exploring representations and inductive biases
Thomas L Griffiths, Nick Chater, Charles Kemp, Amy Perfors, and Joshua B Tenenbaum. 2010 · 2010
Earlier work this paper cites.
Umbrella sampling
Johannes Kästner. 2011 · 2011
Earlier work this paper cites.
Emergence of machine learning techniques in criminology: implications of complexity in our data and in research questions
Tim Brennan and William L Oliver. 2013 · 2013
Earlier work this paper cites.
Generative adversarial nets. In NeurIPS . 2672–2680
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, et al · 2014
Earlier work this paper cites.
Conditional generative adversarial nets
Mehdi Mirza and Simon Osindero. 2014 · 2014
Earlier work this paper cites.
Categorical reparameterization with gumbel-softmax
Eric Jang, Shixiang Gu, and Ben Poole. 2016 · 2016
Earlier work this paper cites.
Examples are not enough, learn to criticize! criticism for interpretability. In NeurIPS . 2280–2288
Been Kim, Rajiv Khanna, and Oluwasanmi O Koyejo. 2016 · 2016
Earlier work this paper cites.
" Why should i trust you?" Explaining the predictions of any classifier. In Proceedings of the 22nd ACM SIGKDD conference on knowledge discovery and data mining . 1135–1144
Marco Tulio Ribeiro, Sameer Singh, and Carlos Guestrin. 2016 · 2016
Earlier work this paper cites.
Understanding black-box predictions via influence functions. In ICML . 1885–1894
Pang Wei Koh and Percy Liang. 2017 · 2017
Earlier work this paper cites.
A survey on deep learning in medical image analysis
Geert Litjens, Thijs Kooi, Babak Ehteshami Bejnordi, Arnaud Arindra Adiyoso Setio, Francesco Ciompi, et al · 2017
Cited alongside, same era.
A unified approach to interpreting model predictions. In NeurIPS . 4765–4774
Scott M Lundberg and Su-In Lee. 2017 · 2017
Cited alongside, same era.
Grad-Cam: Visual explanations from deep networks via gradient-based localization. In Proceedings of the IEEE international conference on computer vision (ICCV) . 618–626
Ramprasaath R Selvaraju, Michael Cogswell, et al · 2017
Cited alongside, same era.
Axiomatic attribution for deep networks. In Proceedings of the 34th International Conference on Machine Learning-Volume 70 . JMLR. org, 3319–3328
Mukund Sundararajan, Ankur Taly, and Qiqi Yan. 2017 · 2017
Cited alongside, same era.
Counterfactual explanations without opening the black box: Automated decisions and the GDPR
Sandra Wachter et al · 2017
Cited alongside, same era.
Towards automatic concept-based explanations
Amirata Ghorbani, James Wexler, James Zou, and Been Kim. 2019 · 2019
Later among the works it cites.
Explaining classifiers with causal concept effect (cace)
Yash Goyal, Amir Feder, Uri Shalit, and Been Kim. 2019a · 2019
Later among the works it cites.
Preserving causal constraints in counterfactual explanations for machine learning classifiers
Divyat Mahajan, Chenhao Tan, and Amit Sharma. 2019 · 2019
Later among the works it cites.
Explaining deep learning models with constrained adversarial examples. In Pacific Rim International Conference on Artificial Intelligence . Springer, 43–56
Jonathan Moore, Nils Hammerla, and Chris Watkins. 2019 · 2019
Later among the works it cites.
Measurable counterfactual local explanations for any classifier
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Relational inductive biases, deep learning, and graph networks
Peter W Battaglia, Jessica B Hamrick, Victor Bapst, et al · 2018
Cited alongside, same era.
Explanations based on the missing: Towards contrastive explanations with pertinent negatives. In NeurIPS . 592–603
Amit Dhurandhar, Pin-Yu Chen, et al · 2018
Cited alongside, same era.
xGEMs: Generating examplars to explain black-box models
Shalmali Joshi, Oluwasanmi Koyejo, Been Kim, et al · 2018
Cited alongside, same era.
Interpretability Beyond Feature Attribution: Quantitative Testing with Concept Activation Vectors (TCAV). In ICML . 2668–2677
Been Kim, Martin Wattenberg, Justin Gilmer, Carrie Cai, et al · 2018
Cited alongside, same era.
Data synthesis based on generative adversarial networks
Noseong Park, Mahmoud Mohammadi, Kshitij Gorde, Sushil Jajodia, Hongkyu Park, and Youngmin Kim. 2018 · 2018
Cited alongside, same era.
Anchors: High-precision model-agnostic explanations. In 32nd AAAI on Artificial Intelligence
Marco Tulio Ribeiro, Sameer Singh, and Carlos Guestrin. 2018 · 2018
Cited alongside, same era.
Interpretable basis decomposition for visual explanation. In ECCV . 119–134
Bolei Zhou, Yiyou Sun, David Bau, and Antonio Torralba. 2018 · 2018
Cited alongside, same era.
Adam White and Artur d’Avila Garcez. 2019 · 2019
Later among the works it cites.
Modeling tabular data using conditional gan. In NeurIPS . 7335–7345
Lei Xu, Maria Skoularidou, Alfredo Cuesta-Infante, and Kalyan Veeramachaneni. 2019 · 2019
Later among the works it cites.
Causal Discovery Toolbox: Uncovering causal relationships in Python
Diviyan Kalainathan, Olivier Goudet, and Ritik Dutta. 2020 · 2020
Later among the works it cites.
A survey of algorithmic recourse: definitions, formulations, solutions, and prospects
Amir-Hossein Karimi, Gilles Barthe, Bernhard Schölkopf, and Isabel Valera. 2020 · 2020
Later among the works it cites.
Explaining machine learning classifiers through diverse counterfactual explanations. In Proceedings on Fairness, Accountability, and Transparency (FAccT) . 607–617
Ramaravind K Mothilal, Amit Sharma, and Chenhao Tan. 2020 · 2020
Later among the works it cites.
Learning Model-Agnostic Counterfactual Explanations for Tabular Data. In Proceedings of The Web Conference 2020 . 3126–3132
Martin Pawelczyk, Klaus Broelemann, and Gjergji Kasneci. 2020 · 2020
Later among the works it cites.
Explainable Image Classification with Evidence Counterfactual
Tom Vermeire and David Martens. 2020 · 2020
Later among the works it cites.
Deep Neural Networks with Knowledge Instillation. In SDM . 370–378
Fan Yang, Ninghao Liu, Mengnan Du, Kaixiong Zhou, Shuiwang Ji, and Xia Hu. 2020 · 2020
Later among the works it cites.
On Completeness-aware Concept-Based Explanations in Deep Neural Networks
Chih-Kuan Yeh, Been Kim, et al · 2020
Later among the works it cites.
Generative Counterfactuals for Neural Networks via Attribute-Informed Perturbation
Fan Yang, Ninghao Liu, Mengnan Du, et al · 2021
Closest in time.