Fetching the paper…
Reading the bibliography…
Many applications of data-driven models demand transparency of decisions, especially in health care, criminal justice, and other high-stakes environments.
doi:10.48550/arXiv.1910.02065
O.-M. Camburu, E. Giunchiglia, J. Foerster, T. Lukasiewicz, P. Blunsom, Can I Trust the Explainer? Verifying Post-hoc Explanatory Methods, NeurIPS 2019 Workshop on Safety and Robustness in Decision Making (1) (2019) 1–13 · 1910
Earlier work this paper cites.
doi:https://doi.org/10.1016/0095-0696(78)90006-2
D. Harrison, Jr., D. L. Rubinfeld, Hedonic housing prices and the demand for clean air, Journal of Environmental Economics and Management 5 (1) (1978) 81 – 102 · 1978
Earlier work this paper cites.
G. Van Rossum, F. L. Drake Jr, Python reference manual, Centrum voor Wiskunde en Informatica Amsterdam, 1995
1995
Earlier work this paper cites.
doi:10.1109/5.726791
Y. LeCun, L. Bottou, Y. Bengio, P. Haffner, Gradient-based learning applied to document recognition, Proceedings of the IEEE 86 (11) (1998) 2278–2324 · 1998
Earlier work this paper cites.
J. H. Friedman, Greedy function approximation: A gradient boosting machine , The Annals of Statistics 29 (5) (2001) 1189–1232. URL http://www.jstor.org/stable/2699986
2001
Earlier work this paper cites.
doi:10.1109/MCSE.2007.55
J. D. Hunter, Matplotlib: A 2d graphics environment, Computing in Science Engineering 9 (3) (2007) 90–95 · 2007
Earlier work this paper cites.
L. Gautier, rpy2: A simple and efficient access to R from Python, URL http://rpy.sourceforge.net/rpy2.html 3 (2008) 1
2008
Earlier work this paper cites.
J. Quionero-Candela, M. Sugiyama, A. Schwaighofer, N. D. Lawrence, Dataset Shift in Machine Learning, The MIT Press, 2009
2009
Earlier work this paper cites.
doi:10.1007/978-0-387-84858-7
T. Hastie, R. Tibshirani, J. H. Friedman, The Elements of Statistical Learning: Data Mining, Inference, and Prediction, 2nd Edition, Springer Series in Statistics, Springer, 2009 · 2009
Earlier work this paper cites.
2010
Earlier work this paper cites.
doi:10.25080/Majora-92bf1922-00a
Wes McKinney, Data Structures for Statistical Computing in Python, in: Stéfan van der Walt, Jarrod Millman (Eds.), Proceedings of the 9th Python in Science Conference, 2010, pp. 56 – 61 · 2010
Earlier work this paper cites.
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V. Dubourg, J. Vanderplas, A. Passos, D. Cournapeau, M. Brucher, M. Perrot, Édouard Duchesnay, Scikit-learn: Machine learning in python , Journal of Machine Learning Research 12 (85) (2011) 2825–2830. URL http://jmlr.org/papers/v12/pedregosa11a.html
2011
Earlier work this paper cites.
doi:10.1007/s11257-011-9117-5
N. Tintarev, J. Masthoff, Evaluating the effectiveness of explanations for recommender systems, User Modeling and User-Adapted Interaction 22 (4-5) (2012) 399–439 · 2012
Earlier work this paper cites.
F. Johansson, et al., mpmath: a Python library for arbitrary-precision floating-point arithmetic (version 0.18), http://mpmath.org/
2013
Earlier work this paper cites.
C. Szegedy, W. Zaremba, I. Sutskever, J. Bruna, D. Erhan, I. J. Goodfellow, R. Fergus, Intriguing properties of neural networks, in: Y. Bengio, Y. LeCun (Eds.), International Conference on Learning Representations, 2014, pp. 1–10
2014
Earlier work this paper cites.
doi:10.1371/journal.pone.0130140
S. Bach, A. Binder, G. Montavon, F. Klauschen, K.-R. Müller, W. Samek, On pixel-wise explanations for non-linear classifier decisions by layer-wise relevance propagation, PLOS ONE 10 (7) (2015) e0130140 · 2015
Earlier work this paper cites.
M. Abadi, A. Agarwal, P. Barham, E. Brevdo, Z. Chen, C. Citro, G. S. Corrado, A. Davis, J. Dean, M. Devin, S. Ghemawat, I. Goodfellow, A. Harp, G. Irving, M. Isard, Y. Jia, R. Jozefowicz, L. Kaiser, M. Kudlur, J. Levenberg, D. Mané, R. Monga, S. Moore, D. Murray, C. Olah, M. Schuster, J. Shlens, B. Steiner, I. Sutskever, K. Talwar, P. Tucker, V. Vanhoucke, V. Vasudevan, F. Viégas, O. Vinyals, P. Warden, M. Wattenberg, M. Wicke, Y. Yu, X. Zheng, TensorFlow: Large-scale machine learning on heterogeneous systems , software available from tensorflow.org (2015). URL http://tensorflow.org/
2015
Earlier work this paper cites.
C. O’Neil, Weapons of Math Destruction: How Big Data Increases Inequality and Threatens Democracy, Crown, 2016
2016
Earlier work this paper cites.
Regulation (EU) 2016/679 of the European Parliament and of the Council of 27 April 2016 on the protection of natural persons with regard to the processing of personal data and on the free movement of such data, and repealing Directive 95/46/EC (General Data Protection Regulation), Official Journal of the European Union L 119 (2016) 1–88
2016
Earlier work this paper cites.
doi:10.1145/2939672.2939778
M. T. Ribeiro, S. Singh, C. Guestrin, “Why should I trust you?”: Explaining the predictions of any classifier , in: SIGKDD International Conference on Knowledge Discovery & Data Mining, ACM, 2016, pp. 1135–1144 · 2016
Earlier work this paper cites.
doi:10.18653/v1/d16-1011
T. Lei, R. Barzilay, T. S. Jaakkola, Rationalizing neural predictions, in: J. Su, X. Carreras, K. Duh (Eds.), Conference on Empirical Methods in Natural Language Processing, The Association for Computational Linguistics, 2016, pp. 107–117 · 2016
Earlier work this paper cites.
J. Angwin, J. Larson, S. Mattu, L. Kirchner, Machine bias: there’s software used across the country to predict future criminals. And it’s biased against blacks., https://www.propublica.org/article/machine-bias-risk-assessments-in-criminal-sentencing (2016)
2016
Earlier work this paper cites.
doi:10.1093/idpl/ipx005
S. Wachter, B. Mittelstadt, L. Floridi, Why a Right to Explanation of Automated Decision-Making Does Not Exist in the General Data Protection Regulation, International Data Privacy Law 7 (2) (2017) 76–99 · 2017
Earlier work this paper cites.
doi:10.48550/arXiv.1711.07414
B. Herman, The promise and peril of human evaluation for model interpretability, in: Conference on Neural Information Processing Systems Symposium on Interpretable Machine Learning, 2017, pp. 1–6 · 2017
Earlier work this paper cites.
S. M. Lundberg, S.-I. Lee, A unified approach to interpreting model predictions , in: I. Guyon, U. von Luxburg, S. Bengio, H. M. Wallach, R. Fergus, S. V. N. Vishwanathan, R. Garnett (Eds.), Conference on Neural Information Processing Systems, 2017, pp. 4765–4774. URL https://proceedings.neurips.cc/paper/2017/hash/8a20a8621978632d76c43dfd28b67767-Abstract.html
2017
Earlier work this paper cites.
doi:10.7717/peerj-cs.103
A. Meurer, C. P. Smith, M. Paprocki, O. Čertík, S. B. Kirpichev, M. Rocklin, A. Kumar, S. Ivanov, J. K. Moore, S. Singh, T. Rathnayake, S. Vig, B. E. Granger, R. P. Muller, F. Bonazzi, H. Gupta, S. Vats, F. Johansson, F. Pedregosa, M. J. Curry, A. R. Terrel, v. Roučka, A. Saboo, I. Fernando, S. Kulal, R. Cimrman, A. Scopatz, SymPy: symbolic computing in Python, PeerJ Computer Science 3 (2017) e103 · 2017
Cited alongside, same era.
J. Buolamwini, T. Gebru, Gender Shades: Intersectional Accuracy Disparities in Commercial Gender Classification, in: Conference on Fairness, Accountability and Transparency, PMLR, 2018, pp. 77–91
2018
Cited alongside, same era.
doi:10.1007/978-3-319-98131-4_1
F. Doshi-Velez, B. Kim, Considerations for evaluation and generalization in interpretable machine learning, in: H. J. Escalante, S. Escalera, I. Guyon, X. Baró, Y. Güçlütürk, U. Güçlü, M. van Gerven (Eds.), Explainable and Interpretable Models in Computer Vision and Machine Learning, The Springer Series on Challenges in Machine Learning, Springer International Publishing, Cham, 2018, pp. 3–17 · 2018
Cited alongside, same era.
doi:10.1109/DSAA.2018.00018
L. H. Gilpin, D. Bau, B. Z. Yuan, A. Bajwa, M. A. Specter, L. Kagal, Explaining explanations: An overview of interpretability of machine learning, in: F. Bonchi, F. J. Provost, T. Eliassi-Rad, W. Wang, C. Cattuto, R. Ghani (Eds.), International Conference on Data Science and Advanced Analytics, IEEE, 2018, pp. 80–89 · 2018
Cited alongside, same era.
doi:10.1145/3375627.3375830
D. Slack, S. Hilgard, E. Jia, S. Singh, H. Lakkaraju, Fooling LIME and SHAP: Adversarial attacks on post hoc explanation methods, in: A. N. Markham, J. Powles, T. Walsh, A. L. Washington (Eds.), Conference on AI, Ethics, and Society, AAAI/ACM, 2020, pp. 180–186 · 2020
Later among the works it cites.
J. DeYoung, S. Jain, N. F. Rajani, E. Lehman, C. Xiong, R. Socher, B. C. Wallace, ERASER: A benchmark to evaluate rationalized NLP models , Transactions of the Association for Computational Linguistics (2020) 1–16. URL https://par.nsf.gov/biblio/10156029
2020
Later among the works it cites.
2020
Later among the works it cites.
doi:10.1007/s11634-020-00418-3
Y. Ramon, D. Martens, F. J. Provost, T. Evgeniou, A comparison of instance-level counterfactual explanation algorithms for behavioral and textual data: Sedc, LIME-C and SHAP-C, Advances in Data Analysis and Classification 14 (4) (2020) 801–819 · 2020
Later among the works it cites.
doi:10.1016/j.artint.2020.103428
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
doi:10.1145/3236386.3241340
Z. C. Lipton, The mythos of model interpretability, ACM Queue 16 (3) (2018) 30 · 2018
Cited alongside, same era.
G. Plumb, D. Molitor, A. S. Talwalkar, Model agnostic supervised local explanations , in: S. Bengio, H. M. Wallach, H. Larochelle, K. Grauman, N. Cesa-Bianchi, R. Garnett (Eds.), Conference on Neural Information Processing Systems, 2018, pp. 2520–2529. URL https://proceedings.neurips.cc/paper/2018/hash/b495ce63ede0f4efc9eec62cb947c162-Abstract.html
2018
Cited alongside, same era.
doi:10.1145/3236009
R. Guidotti, A. Monreale, S. Ruggieri, F. Turini, F. Giannotti, D. Pedreschi, A survey of methods for explaining black box models, ACM Computing Surveys 51 (5) (2018) 93:1–93:42 · 2018
Cited alongside, same era.
2018
Cited alongside, same era.
D. Alvarez-Melis, T. S. Jaakkola, On the robustness of interpretability methods, ICML Workshop on Human Interpretability in Machine Learning (2018) 66–71
2018
Cited alongside, same era.
doi:10.5281/zenodo.1476122
D. Servén, C. Brummitt, H. Abedi, hlink, dswah/pygam: v0.8.0 (Oct. 2018) · 2018
Cited alongside, same era.
F. I. C. (FICO), FICO explainable machine learning challenge: Home equity line of credit (HELOC) dataset, https://community.fico.com/s/explainable-machine-learning-challenge (2018)
2018
Cited alongside, same era.
doi:10.1186/s12911-019-0874-0
R. Elshawi, M. H. Al-Mallah, S. Sakr, On the interpretability of machine learning-based model for predicting hypertension, BMC Medical Informatics and Decision Making 19 (1) (2019) 1–32 · 2019
Cited alongside, same era.
R. Guidotti, Evaluating local explanation methods on ground truth, Artificial Intelligence 291 (2021) 1–16 · 2020
Later among the works it cites.
doi:10.1038/s41586-020-2649-2
C. R. Harris, K. J. Millman, S. J. van der Walt, R. Gommers, P. Virtanen, D. Cournapeau, E. Wieser, J. Taylor, S. Berg, N. J. Smith, R. Kern, M. Picus, S. Hoyer, M. H. van Kerkwijk, M. Brett, A. Haldane, J. Fernández del Río, M. Wiebe, P. Peterson, P. Gérard-Marchant, K. Sheppard, T. Reddy, W. Weckesser, H. Abbasi, C. Gohlke, T. E. Oliphant, Array programming with NumPy, Nature 585 (2020) 357–362 · 2020
Later among the works it cites.
doi:10.1038/s41592-019-0686-2
P. Virtanen, R. Gommers, T. E. Oliphant, M. Haberland, T. Reddy, D. Cournapeau, E. Burovski, P. Peterson, W. Weckesser, J. Bright, S. J. van der Walt, M. Brett, J. Wilson, K. J. Millman, N. Mayorov, A. R. J. Nelson, E. Jones, R. Kern, E. Larson, C. J. Carey, İ. Polat, Y. Feng, E. W. Moore, J. VanderPlas, D. Laxalde, J. Perktold, R. Cimrman, I. Henriksen, E. A. Quintero, C. R. Harris, A. M. Archibald, A. H. Ribeiro, F. Pedregosa, P. van Mulbregt, SciPy 1.0 Contributors, SciPy 1.0: Fundamental Algorithms for Scientific Computing in Python, Nature Methods 17 (2020) 261–272 · 2020
Later among the works it cites.
Joblib Development Team, Joblib: running python functions as pipeline jobs (2020). URL https://joblib.readthedocs.io/
2020
Later among the works it cites.
doi:10.1145/3313831.3376219
H. Kaur, H. Nori, S. Jenkins, R. Caruana, H. Wallach, J. Wortman Vaughan, Interpreting interpretability: Understanding data scientists’ use of interpretability tools for machine learning, in: R. Bernhaupt, F. F. Mueller, D. Verweij, J. Andres, J. McGrenere, A. Cockburn, I. Avellino, A. Goguey, P. Bjøn, S. Zhao, B. P. Samson, R. Kocielnik (Eds.), Conference on Human Factors in Computing Systems, ACM, 2020, pp. 1–14 · 2020
Later among the works it cites.
D. Hendrycks, K. Zhao, S. Basart, J. Steinhardt, D. Song, Natural adversarial examples , in: Conference on Computer Vision and Pattern Recognition, IEEE/Computer Vision Foundation, 2021, pp. 15262–15271. URL https://openaccess.thecvf.com/content/CVPR2021/html/Hendrycks_Natural_Adversarial_Examples_CVPR_2021_paper.html
2021
Closest in time.
L. Hogenhout, A framework for ethical AI at the United Nations , Unite Paper 2021(1), UN Office for Information and Communications Technology, https://unite.un.org/news/unite-paper-framework-ethical-ai-united-nations (Mar. 2021). URL https://unite.un.org/sites/unite.un.org/files/unite_paper_-_ethical_ai_at_the_un.pdf
2021
Closest in time.
doi:10.1016/j.artint.2021.103502
K. Aas, M. Jullum, A. Løland, Explaining individual predictions when features are dependent: More accurate approximations to shapley values, Artificial Intelligence 298 (2021) 1–24 · 2021
Closest in time.
doi:10.3390/electronics10050593
J. Zhou, A. H. Gandomi, F. Chen, A. Holzinger, Evaluating the Quality of Machine Learning Explanations: A Survey on Methods and Metrics, Electronics 10 (5) (2021) 593 · 2021
Closest in time.
2021
Closest in time.
Y. Zhou, S. Booth, M. T. Ribeiro, J. Shah, Do Feature Attribution Methods Correctly Attribute Features?, NeurIPS 1st Workshop on eXplainable AI approaches for debugging and diagnosis (XAI4Debugging) (2021) 1–22
2021
Closest in time.
R. Agarwal, L. Melnick, N. Frosst, X. Zhang, B. Lengerich, R. Caruana, G. Hinton, Neural additive models: Interpretable machine learning with neural nets , in: A. Beygelzimer, Y. Dauphin, P. Liang, J. W. Vaughan (Eds.), Advances in Neural Information Processing Systems, 2021. URL https://openreview.net/forum?id=wHkKTW2wrmm
2021
Closest in time.
L. Jiangchun, C. D. C. Santos, M. Kuhlen, A. Singh, Pdpbox: v0.2.1 (Mar. 2021). URL https://github.com/SauceCat/PDPbox
2021
Closest in time.
doi:10.21105/joss.03021
M. L. Waskom, seaborn: statistical data visualization , Journal of Open Source Software 6 (60) (2021) 3021 · 2021
Closest in time.
R Core Team, R: A Language and Environment for Statistical Computing , R Foundation for Statistical Computing, Vienna, Austria (2021). URL https://www.R-project.org/
2021
Closest in time.
N. Sellereite, M. Jullum, A. Redelmeier, shapr: Prediction Explanation with Dependence-Aware Shapley Values , r package version 0.2.0 (2021). URL https://CRAN.R-project.org/package=shapr
2021
Closest in time.
G. V. den Broeck, A. Lykov, M. Schleich, D. Suciu, On the tractability of SHAP explanations , in: Thirty-Fifth AAAI Conference on Artificial Intelligence, AAAI 2021, Thirty-Third Conference on Innovative Applications of Artificial Intelligence, IAAI 2021, The Eleventh Symposium on Educational Advances in Artificial Intelligence, EAAI 2021, Virtual Event, February 2-9, 2021, AAAI Press, 2021, pp. 6505–6513. URL https://ojs.aaai.org/index.php/AAAI/article/view/16806
2021
Closest in time.
2022
Closest in time.
2022
Closest in time.
doi:10.1007/s10044-021-01055-y
T. Vermeire, D. Brughmans, S. Goethals, R. M. B. de Oliveira, D. Martens, Explainable image classification with evidence counterfactual, Pattern Analysis and Applications 26 (2022) 1–21 · 2022
Closest in time.