Fetching the paper…
Reading the bibliography…
Stalnaker RC (1968) A theory of conditionals. In: Ifs, Springer, pp 41–55
1968
Earlier work this paper cites.
Lewis DK (1973) Counterfactuals. Blackwell
1973
Earlier work this paper cites.
Lewis D (1979) Counterfactual dependence and time’s arrow. Noûs pp 455–476
1979
Earlier work this paper cites.
Lewis D (1983) Philosophical Papers Volume I. Oxford university press New York
1983
Earlier work this paper cites.
Friedman JH, et al. (1991) Multivariate adaptive regression splines. The Annals of Statistics 19(1):1–67, DOI 10.1214/aos/1176347963
1991
Earlier work this paper cites.
Fernández-Loría C, Provost F, Han X (2020) Explaining data-driven decisions made by ai systems: The counterfactual approach. 2001.07417
2001
Earlier work this paper cites.
Hitchcock C (2001) The intransitivity of causation revealed in equations and graphs. The Journal of Philosophy 98(6):273–299
2001
Earlier work this paper cites.
Woodward J (2002) What is a mechanism? a counterfactual account. Philosophy of science 69(S3):S366–S377
2002
Earlier work this paper cites.
Dalvi N, Domingos P, Sanghai S, Verma D (2004) Adversarial classification. In: Proceedings of the tenth ACM SIGKDD international conference on Knowledge discovery and data mining, pp 99–108
2004
Earlier work this paper cites.
Fernandez JC, Mounier L, Pachon C (2005) A model-based approach for robustness testing. In: IFIP international conference on testing of communicating systems, Springer, pp 333–348
2005
Earlier work this paper cites.
Bishop CM (2006) Pattern recognition and machine learning. Springer
2006
Earlier work this paper cites.
Molnar C, König G, Herbinger J, Freiesleben T, Dandl S, Scholbeck CA, Casalicchio G, Grosse-Wentrup M, Bischl B (2020) Pitfalls to avoid when interpreting machine learning models. 2007.04131
2007
Earlier work this paper cites.
Claeskens G, Hjort NL, et al. (2008) Model selection and model averaging. Cambridge Books DOI 10.1017/CBO9780511790485
2008
Earlier work this paper cites.
D’silva V, Kroening D, Weissenbacher G (2008) A survey of automated techniques for formal software verification. IEEE Transactions on Computer-Aided Design of Integrated Circuits and Systems 27(7):1165–1178
2008
Earlier work this paper cites.
Hashemi M, Fathi A (2020) Permuteattack: Counterfactual explanation of machine learning credit scorecards. 2008.10138
2008
Earlier work this paper cites.
Pearl J (2009) Causality. Cambridge University Press
2009
Earlier work this paper cites.
Browne K, Swift B (2020) Semantics and explanation: why counterfactual explanations produce adversarial examples in deep neural networks. 2012.10076
2012
Earlier work this paper cites.
Good PI, Hardin JW (2012) Common errors in statistics (and how to avoid them). John Wiley & Sons, DOI 10.1002/9781118360125
2012
Earlier work this paper cites.
Kizza JM, Kizza, Wheeler (2013) Guide to computer network security. Springer
2013
Earlier work this paper cites.
Vapnik V (2013) The nature of statistical learning theory. Springer science & business media
2013
Earlier work this paper cites.
Goodfellow IJ, Pouget-Abadie J, Mirza M, Xu B, Warde-Farley D, Ozair S, Courville A, Bengio Y (2014) Generative adversarial networks. arXiv preprint arXiv:14062661
2014
Earlier work this paper cites.
Štrumbelj E, Kononenko I (2014) Explaining prediction models and individual predictions with feature contributions. Knowledge and information systems 41(3):647–665, DOI 10.1007/s10115-013-0679-x
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
2015
Earlier work this paper cites.
Lyu C, Huang K, Liang HN (2015) A unified gradient regularization family for adversarial examples. In: 2015 IEEE international conference on data mining, IEEE, pp 301–309
2015
Earlier work this paper cites.
Bastani O, Ioannou Y, Lampropoulos L, Vytiniotis D, Nori AV, Criminisi A (2016) Measuring neural net robustness with constraints. In: Proceedings of the 30th International Conference on Neural Information Processing Systems, pp 2621–2629
2016
Earlier work this paper cites.
Burrell J (2016) How the machine ‘thinks’: Understanding opacity in machine learning algorithms. Big Data & Society 3(1):2053951715622512, DOI 10.1177/2053951715622512
2016
Earlier work this paper cites.
Byrne RM (2016) Counterfactual thought. Annual review of psychology 67:135–157
2016
Earlier work this paper cites.
Carlini N, Mishra P, Vaidya T, Zhang Y, Sherr M, Shields C, Wagner D, Zhou W (2016) Hidden voice commands. In: 25th { \{ USENIX } \} Security Symposium ( { \{ USENIX } \} Security 16), pp 513–530
2016
Earlier work this paper cites.
Goodfellow I, Bengio Y, Courville A (2016) Deep learning. MIT press
2016
Earlier work this paper cites.
Kurakin A, Goodfellow I, Bengio S, et al. (2016) Adversarial examples in the physical world
2016
Earlier work this paper cites.
Moosavi-Dezfooli SM, Fawzi A, Frossard P (2016) Deepfool: a simple and accurate method to fool deep neural networks. In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp 2574–2582
2016
Earlier work this paper cites.
Papernot N, McDaniel P, Jha S, Fredrikson M, Celik ZB, Swami A (2016b) The limitations of deep learning in adversarial settings. In: 2016 IEEE European symposium on security and privacy (EuroS&P), IEEE, pp 372–387
2016
Earlier work this paper cites.
Ribeiro MT, Singh S, Guestrin C (2016) Why should i trust you?: Explaining the predictions of any classifier. In: Proceedings of the 22nd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining, ACM, pp 1135–1144, DOI 10.1145/2939672.2939778
2016
Earlier work this paper cites.
Rozsa A, Rudd EM, Boult TE (2016) Adversarial diversity and hard positive generation. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops, pp 25–32
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
Tanay T, Griffin L (2016) A boundary tilting persepective on the phenomenon of adversarial examples. arXiv preprint arXiv:160807690
2016
Earlier work this paper cites.
Bau D, Zhou B, Khosla A, Oliva A, Torralba A (2017) Network dissection: Quantifying interpretability of deep visual representations. In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp 6541–6549
2017
Earlier work this paper cites.
Behzadan V, Munir A (2017) Vulnerability of deep reinforcement learning to policy induction attacks. In: International Conference on Machine Learning and Data Mining in Pattern Recognition, Springer, pp 262–275
2017
Earlier work this paper cites.
Brown TB, Mané D, Roy A, Abadi M, Gilmer J (2017) Adversarial patch. arXiv preprint arXiv:171209665
2017
Earlier work this paper cites.
Carlini N, Wagner D (2017) Towards evaluating the robustness of neural networks. In: 2017 IEEE Symposium on Security and Privacy, IEEE, pp 39–57
2017
Earlier work this paper cites.
Chen PY, Zhang H, Sharma Y, Yi J, Hsieh CJ (2017) Zoo: Zeroth order optimization based black-box attacks to deep neural networks without training substitute models. In: Proceedings of the 10th ACM Workshop on Artificial Intelligence and Security, pp 15–26
2017
Earlier work this paper cites.
Dong Y, Su H, Zhu J, Bao F (2017) Towards interpretable deep neural networks by leveraging adversarial examples. arXiv preprint arXiv:170805493
2017
Cited alongside, same era.
Doshi-Velez F, Kim B (2017) Towards a rigorous science of interpretable machine learning. arXiv preprint arXiv:170208608
2017
Cited alongside, same era.
Huang S, Papernot N, Goodfellow I, Duan Y, Abbeel P (2017) Adversarial attacks on neural network policies. arXiv preprint arXiv:170202284
2017
Cited alongside, same era.
Kusner MJ, Loftus J, Russell C, Silva R (2017) Counterfactual fairness. In: Advances in Neural Information Processing Systems, pp 4066–4076
2017
Cited alongside, same era.
Moosavi-Dezfooli SM, Fawzi A, Fawzi O, Frossard P (2017) Universal adversarial perturbations. In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp 1765–1773
2017
Laugel T, Lesot MJ, Marsala C, Renard X, Detyniecki M (2019a) The dangers of post-hoc interpretability: Unjustified counterfactual explanations. In: Proceedings of the Twenty-Eighth International Joint Conference on Artificial Intelligence, IJCAI-19, International Joint Conferences on Artificial Intelligence Organization, pp 2801–2807, DOI 10.24963/ijcai.2019/388
2019
Later among the works it cites.
Mahajan D, Tan C, Sharma A (2019) Preserving causal constraints in counterfactual explanations for machine learning classifiers. arXiv preprint arXiv:191203277
2019
Later among the works it cites.
Menzies P, Beebee H (2019) Counterfactual theories of causation. In: Zalta EN (ed) The Stanford Encyclopedia of Philosophy, winter 2019 edn, Metaphysics Research Lab, Stanford University
2019
Later among the works it cites.
Miller T (2019) Explanation in artificial intelligence: Insights from the social sciences. Artificial Intelligence 267:1–38
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Papernot N, McDaniel P, Goodfellow I, Jha S, Celik ZB, Swami A (2017) Practical black-box attacks against machine learning. In: Proceedings of the 2017 ACM on Asia conference on computer and communications security, pp 506–519
2017
Cited alongside, same era.
Tolomei G, Silvestri F, Haines A, Lalmas M (2017) Interpretable predictions of tree-based ensembles via actionable feature tweaking. In: Proceedings of the 23rd ACM SIGKDD international conference on knowledge discovery and data mining, pp 465–474
2017
Cited alongside, same era.
Voigt P, Von dem Bussche A (2017) The eu general data protection regulation (gdpr). A Practical Guide, 1st Ed, Cham: Springer International Publishing 10:3152676
2017
Cited alongside, same era.
Wachter S, Mittelstadt B, Russell C (2017) Counterfactual explanations without opening the black box: Automated decisions and the gdpr. Harv JL & Tech 31:841
2017
Cited alongside, same era.
2017
Cited alongside, same era.
Adadi A, Berrada M (2018) Peeking inside the black-box: a survey on explainable artificial intelligence (xai). IEEE access 6:52138–52160
2018
Cited alongside, same era.
Athalye A, Engstrom L, Ilyas A, Kwok K (2018) Synthesizing robust adversarial examples. In: International conference on machine learning, PMLR, pp 284–293
2018
Cited alongside, same era.
Molnar C (2019) Interpretable Machine Learning. https://christophm.github.io/interpretable-ml-book/
2019
Later among the works it cites.
Moore J, Hammerla N, Watkins C (2019) Explaining deep learning models with constrained adversarial examples. In: Pacific Rim International Conference on Artificial Intelligence, Springer, pp 43–56
2019
Later among the works it cites.
Páez A (2019) The pragmatic turn in explainable artificial intelligence (xai). Minds and Machines 29(3):441–459
2019
Later among the works it cites.
Rudin C (2019) Stop explaining black box machine learning models for high stakes decisions and use interpretable models instead. Nature Machine Intelligence 1(5):206–215
2019
Later among the works it cites.
Russell C (2019) Efficient search for diverse coherent explanations. In: Proceedings of the Conference on Fairness, Accountability, and Transparency, Association for Computing Machinery, New York, NY, USA, FAT* ’19, p 20–28, DOI 10.1145/3287560.3287569
2019
Later among the works it cites.
Schölkopf B (2019) Causality for machine learning. arXiv preprint arXiv:191110500
2019
Later among the works it cites.
Sokol K, Flach PA (2019) Counterfactual explanations of machine learning predictions: Opportunities and challenges for ai safety. In: Proceedings of the AAAI Workshop on Artificial Intelligence Safety
2019
Later among the works it cites.
Starr W (2019) Counterfactuals. In: Zalta EN (ed) The Stanford Encyclopedia of Philosophy, fall 2019 edn, Metaphysics Research Lab, Stanford University
2019
Later among the works it cites.
Stutz D, Hein M, Schiele B (2019) Confidence-calibrated adversarial training: Generalizing to unseen attacks. arXiv preprint arXiv:191006259
2019
Later among the works it cites.
Su J, Vargas DV, Sakurai K (2019) One pixel attack for fooling deep neural networks. IEEE Transactions on Evolutionary Computation 23(5):828–841
2019
Later among the works it cites.
Tramer F, Boneh D (2019) Adversarial training and robustness for multiple perturbations. arXiv preprint arXiv:190413000
2019
Later among the works it cites.
Ustun B, Spangher A, Liu Y (2019) Actionable recourse in linear classification. In: Proceedings of the Conference on Fairness, Accountability, and Transparency, pp 10–19
2019
Later among the works it cites.
Van Looveren A, Klaise J (2019) Interpretable counterfactual explanations guided by prototypes. arXiv preprint arXiv:190702584
2019
Later among the works it cites.
Wang X, He K, Hopcroft JE (2019) At-gan: A generative attack model for adversarial transferring on generative adversarial nets. CoRR, abs/190407793
2019
Later among the works it cites.
Wong E, Schmidt F, Kolter Z (2019) Wasserstein adversarial examples via projected sinkhorn iterations. In: International Conference on Machine Learning, PMLR, pp 6808–6817
2019
Later among the works it cites.
Yuan X, He P, Zhu Q, Li X (2019) Adversarial examples: Attacks and defenses for deep learning. IEEE transactions on neural networks and learning systems 30(9):2805–2824
2019
Later among the works it cites.
Zhang H, Chen H, Song Z, Boning D, Dhillon IS, Hsieh CJ (2019) The limitations of adversarial training and the blind-spot attack. arXiv preprint arXiv:190104684
2019
Later among the works it cites.
Asher N, Paul S, Russell C (2020) Adequate and fair explanations. arXiv preprint arXiv:200107578
2020
Closest in time.
Barocas S, Selbst AD, Raghavan M (2020) The hidden assumptions behind counterfactual explanations and principal reasons. In: Proceedings of the 2020 Conference on Fairness, Accountability, and Transparency, Association for Computing Machinery, New York, NY, USA, FAT* ’20, p 80–89, DOI 10.1145/3351095.3372830
2020
Closest in time.
Brown TB, Mann B, Ryder N, Subbiah M, Kaplan J, Dhariwal P, Neelakantan A, Shyam P, Sastry G, Askell A, et al. (2020) Language models are few-shot learners. arXiv preprint arXiv:200514165
2020
Closest in time.
Dandl S, Molnar C, Binder M, Bischl B (2020) Multi-objective counterfactual explanations. In: Bäck T, Preuss M, Deutz A, Wang H, Doerr C, Emmerich M, Trautmann H (eds) Parallel Problem Solving from Nature – PPSN XVI, Springer International Publishing, Cham, pp 448–469
2020
Closest in time.
Das A, Rad P (2020) Opportunities and challenges in explainable artificial intelligence (xai): A survey. arXiv preprint arXiv:200611371
2020
Closest in time.
Kanamori K, Takagi T, Kobayashi K, Arimura H (2020) Dace: Distribution-aware counterfactual explanation by mixed-integer linear optimization. In: Proceedings of the Twenty-Ninth International Joint Conference on Artificial Intelligence, IJCAI-20, Christian Bessiere (Ed.). International Joint Conferences on Artificial Intelligence Organization, pp 2855–2862
2020
Closest in time.
Mothilal RK, Sharma A, Tan C (2020) Explaining machine learning classifiers through diverse counterfactual explanations. In: Proceedings of the ACM Conference on Fairness, Accountability, and Transparency
2020
Closest in time.
Olah C, Cammarata N, Schubert L, Goh G, Petrov M, Carter S (2020) Zoom in: An introduction to circuits. Distill 5(3):e00024–001
2020
Closest in time.
Pawelczyk M, Broelemann K, Kasneci G (2020) Learning model-agnostic counterfactual explanations for tabular data. In: Proceedings of The Web Conference 2020, pp 3126–3132
2020
Closest in time.
Poyiadzi R, Sokol K, Santos-Rodriguez R, De Bie T, Flach P (2020) Face: Feasible and actionable counterfactual explanations. In: Proceedings of the AAAI/ACM Conference on AI, Ethics, and Society, pp 344–350
2020
Closest in time.
Senior AW, Evans R, Jumper J, Kirkpatrick J, Sifre L, Green T, Qin C, Žídek A, Nelson AW, Bridgland A, et al. (2020) Improved protein structure prediction using potentials from deep learning. Nature 577(7792):706–710
2020
Closest in time.
Serban A, Poll E, Visser J (2020) Adversarial examples on object recognition: A comprehensive survey. ACM Computing Surveys (CSUR) 53(3):1–38
2020
Closest in time.
Sharma S, Henderson J, Ghosh J (2020) Certifai:: Counterfactual explanations for robustness, transparency, interpretability, and fairness of artificial intelligence models. Proceedings of the AAAI/ACM Conference on AI, Ethics, and Society DOI 10.1145/3375627.3375812
2020
Closest in time.
Toreini E, Aitken M, Coopamootoo K, Elliott K, Zelaya CG, van Moorsel A (2020) The relationship between trust in ai and trustworthy machine learning technologies. In: Proceedings of the 2020 Conference on Fairness, Accountability, and Transparency, pp 272–283
2020
Closest in time.
Venkatasubramanian S, Alfano M (2020) The philosophical basis of algorithmic recourse. In: Proceedings of the 2020 conference on fairness, accountability, and transparency, pp 284–293
2020
Closest in time.
Verma S, Dickerson J, Hines K (2020) Counterfactual explanations for machine learning: A review. arXiv preprint arXiv:201010596
2020
Closest in time.
Cartella F, Anunciacao O, Funabiki Y, Yamaguchi D, Akishita T, Elshocht O (2021) Adversarial attacks for tabular data: Application to fraud detection and imbalanced data. arXiv preprint arXiv:210108030
2021
Closest in time.
Olson ML, Khanna R, Neal L, Li F, Wong WK (2021) Counterfactual state explanations for reinforcement learning agents via generative deep learning. Artificial Intelligence 295:103455
2021
Closest in time.
Shin D (2021) The effects of explainability and causability on perception, trust, and acceptance: Implications for explainable ai. International Journal of Human-Computer Studies 146:102551
2021
Closest in time.
Stepin I, Alonso JM, Catala A, Pereira-Fariña M (2021) A survey of contrastive and counterfactual explanation generation methods for explainable artificial intelligence. IEEE Access 9:11974–12001, DOI 10.1109/ACCESS.2021.3051315
2021
Closest in time.