One explanation does not fit all: A toolkit and taxonomy of AI explainability techniques
Original
Arya, Vijay, Rachel K. E. Bellamy, Pin-Yu Chen, Amit Dhurandhar, Michael Hind, Samuel C. Hoffman, Stephanie Houde, Q. Vera Liao, Ronny Luss, Aleksandra Mojsilovic, Sami Mourad, Pablo Pedemonte, Ramya Raghavendra, John T. Richards, Prasanna Sattigeri, Karthikeyan Shanmugam, Moninder Singh, Kush R. Varshney, Dennis Wei, and Yunfeng Zhang. 2019a · 1909
Earlier work this paper cites.
One explanation does not fit all: A toolkit and taxonomy of ai explainability techniques
Original
Arya, Vijay, Rachel KE Bellamy, Pin-Yu Chen, Amit Dhurandhar, Michael Hind, Samuel C Hoffman, Stephanie Houde, Q Vera Liao, Ronny Luss, Aleksandra Mojsilović, et al. 2019b · 1909
Earlier work this paper cites.
Better summarization evaluation with word embeddings for ROUGE
Ng, Jun-Ping and Viktoria Abrecht. 2015 · 1930
Earlier work this paper cites.
A difficulty in the concept of social welfare
Arrow, Kenneth J. 1950 · 1950
Earlier work this paper cites.
Rank analysis of incomplete block designs: I. the method of paired comparisons
Bradley, Ralph Allan and Milton E. Terry. 1952 · 1952
Earlier work this paper cites.
Automatic evaluation of machine translation quality using n-gram co-occurrence statistics
Doddington, George. 2002 · 2002
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Papineni, Kishore, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
A Rule-Based Style and Grammar Checker
Naber, Daniel. 2003 · 2003
Earlier work this paper cites.
The significance of recall in automatic metrics for mt evaluation
Lavie, Alon, Kenji Sagae, and Shyamsundar Jayaraman. 2004 · 2004
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Lin, Chin-Yew. 2004 · 2004
Earlier work this paper cites.
Wt5?! training text-to-text models to explain their predictions
Original
Narang, Sharan, Colin Raffel, Katherine Lee, Adam Roberts, Noah Fiedel, and Karishma Malkan. 2020 · 2004
Earlier work this paper cites.
METEOR: An automatic metric for MT evaluation with improved correlation with human judgments
Banerjee, Satanjeev and Alon Lavie. 2005 · 2005
Earlier work this paper cites.
A survey on tree edit distance and related problems
Bille, Philip. 2005 · 2005
Earlier work this paper cites.
Trust building with explanation interfaces
Pu, Pearl and Li Chen. 2006 · 2006
Earlier work this paper cites.
A study of translation edit rate with targeted human annotation
Snover, Matthew, Bonnie Dorr, Richard Schwartz, Linnea Micciulla, and John Makhoul. 2006 · 2006
Earlier work this paper cites.
Some issues in automatic evaluation of english-hindi mt : More blues for bleu
Ananthakrishnan, R., P. Bhattacharyya, M. Sasikumar, and R. M. Shah. 2006 · 2007
Earlier work this paper cites.
Subplex: Towards a better understanding of black box model explanations at the subpopulation level
Original
Chan, Gromit Yeuk-Yin, Jun Yuan, Kyle Overton, Brian Barr, Kim Rees, Luis Gustavo Nonato, Enrico Bertini, and Claudio T Silva. 2020 · 2007
Earlier work this paper cites.
MLQE-PE: A multilingual quality estimation and post-editing dataset
Original
Fomicheva, Marina, Shuo Sun, Erick Fonseca, Frédéric Blain, Vishrav Chaudhary, Francisco Guzmán, Nina Lopatina, Lucia Specia, and André F. T. Martins. 2020a · 2010
Earlier work this paper cites.
Efficient estimation of word representations in vector space
Mikolov, Tomás, Kai Chen, Greg Corrado, and Jeffrey Dean. 2013 · 2013
Earlier work this paper cites.
A systematic comparison of smoothing techniques for sentence-level BLEU
Chen, Boxing and Colin Cherry. 2014 · 2014
Earlier work this paper cites.
Comprehensible classification models: a position paper
Freitas, Alex A. 2014 · 2014
Earlier work this paper cites.
Multidimensional quality metrics (mqm): A framework for declaring and describing translation quality metrics
Lommel, Arle, Aljoscha Burchardt, and Hans Uszkoreit. 2014 · 2014
Earlier work this paper cites.
Understanding deep image representations by inverting them
Mahendran, Aravindh and Andrea Vedaldi. 2014 · 2014
Earlier work this paper cites.
Glove: Global vectors for word representation
Pennington, Jeffrey, Richard Socher, and Christopher D. Manning. 2014 · 2014
Earlier work this paper cites.
Intriguing properties of neural networks
Szegedy, Christian, Wojciech Zaremba, Ilya Sutskever, Joan Bruna, Dumitru Erhan, Ian Goodfellow, and Rob Fergus. 2014 · 2014
Earlier work this paper cites.
Adaptive quality estimation for machine translation
Turchi, Marco, Antonios Anastasopoulos, José G. C. de Souza, and Matteo Negri. 2014 · 2014
Earlier work this paper cites.
SemEval-2015 task 2: Semantic textual similarity, English, Spanish and pilot on interpretability
Agirre, Eneko, Carmen Banea, Claire Cardie, Daniel Cer, Mona Diab, Aitor Gonzalez-Agirre, Weiwei Guo, Inigo Lopez-Gazpio, Montse Maritxalar, Rada Mihalcea, German Rigau, Larraitz Uria, and Janyce Wiebe. 2015 · 2015
Earlier work this paper cites.
On pixel-wise explanations for non-linear classifier decisions by layer-wise relevance propagation
Bach, Sebastian, Alexander Binder, Grégoire Montavon, Frederick Klauschen, Klaus-Robert Müller, and Wojciech Samek. 2015 · 2015
Earlier work this paper cites.
Retrofitting word vectors to semantic lexicons
Faruqui, Manaal, Jesse Dodge, Sujay Kumar Jauhar, Chris Dyer, Eduard Hovy, and Noah A. Smith. 2015 · 2015
Earlier work this paper cites.
Explaining and harnessing adversarial examples
Goodfellow, Ian J., Jonathon Shlens, and Christian Szegedy. 2015 · 2015
Earlier work this paper cites.
Principles of explanatory debugging to personalize interactive machine learning
Kulesza, Todd, Margaret Burnett, Weng-Keen Wong, and Simone Stumpf. 2015 · 2015
Earlier work this paper cites.
Multidimensional quality metrics (mqm) definition
Lommel, Arle, Aljoscha Burchardt, and Hans Uszkoreit. 2015 · 2015
Earlier work this paper cites.
Results of the WMT15 metrics shared task
Stanojević, Miloš, Amir Kamran, Philipp Koehn, and Ondřej Bojar. 2015 · 2015
Earlier work this paper cites.
SemEval-2016 task 2: Interpretable semantic textual similarity
Agirre, Eneko, Aitor Gonzalez-Agirre, Iñigo Lopez-Gazpio, Montse Maritxalar, German Rigau, and Larraitz Uria. 2016 · 2016
Earlier work this paper cites.
Explaining predictions of non-linear classifiers in NLP
Arras, Leila, Franziska Horn, Grégoire Montavon, Klaus-Robert Müller, and Wojciech Samek. 2016 · 2016
Earlier work this paper cites.
Results of the WMT16 metrics shared task
Bojar, Ondřej, Yvette Graham, Amir Kamran, and Miloš Stanojević. 2016 · 2016
Earlier work this paper cites.
Can machine translation systems be evaluated by the crowd alone
Graham, Yvette, Timothy Baldwin, Alistair Moffat, and Justin Zobel. 2016 · 2016
Earlier work this paper cites.
Explainable artificial intelligence (xai)
Gunning, David. 2016 · 2016
Earlier work this paper cites.
Interacting with predictions: Visual inspection of black-box machine learning models
Krause, Josua, Adam Perer, and Kenney Ng. 2016 · 2016
Earlier work this paper cites.
Rationalizing neural predictions
Lei, Tao, Regina Barzilay, and Tommi Jaakkola. 2016 · 2016
Earlier work this paper cites.
Visualizing and understanding neural models in NLP
Li, Jiwei, Xinlei Chen, Eduard Hovy, and Dan Jurafsky. 2016 · 2016
Earlier work this paper cites.
The mythos of model interpretability
Lipton, Zachary. 2016 · 2016
Earlier work this paper cites.
FBK-HLT-NLP at SemEval-2016 task 2: A multitask, deep learning approach for interpretable semantic textual similarity
Magnolini, Simone, Anna Feltracco, and Bernardo Magnini. 2016 · 2016
Earlier work this paper cites.
“why should I trust you?”: Explaining the predictions of any classifier
Ribeiro, Marco, Sameer Singh, and Carlos Guestrin. 2016 · 2016
Earlier work this paper cites.
Results of the WMT17 metrics shared task
Bojar, Ondřej, Yvette Graham, and Amir Kamran. 2017 · 2017
Earlier work this paper cites.
SemEval-2017 task 1: Semantic textual similarity multilingual and crosslingual focused evaluation
Cer, Daniel, Mona Diab, Eneko Agirre, Iñigo Lopez-Gazpio, and Lucia Specia. 2017a · 2017
Earlier work this paper cites.
SemEval-2017 task 1: Semantic textual similarity multilingual and crosslingual focused evaluation
Cer, Daniel, Mona Diab, Eneko Agirre, Iñigo Lopez-Gazpio, and Lucia Specia. 2017b · 2017
Earlier work this paper cites.
Towards a rigorous science of interpretable machine learning
Doshi-Velez, Finale and Been Kim. 2017 · 2017
Earlier work this paper cites.
European union regulations on algorithmic decision-making and a “right to explanation”
Goodman, Bryce and Seth Flaxman. 2017 · 2017
Earlier work this paper cites.
Discourse structure in machine translation evaluation
Joty, Shafiq, Francisco Guzmán, Lluís Màrquez, and Preslav Nakov. 2017 · 2017
Earlier work this paper cites.
Best-worst scaling more reliable than rating scales: A case study on sentiment intensity annotation
Kiritchenko, Svetlana and Saif Mohammad. 2017 · 2017
Earlier work this paper cites.
Understanding black-box predictions via influence functions
Koh, Pang Wei and Percy Liang. 2017 · 2017
Earlier work this paper cites.
A unified approach to interpreting model predictions
Lundberg, Scott M and Su-In Lee. 2017 · 2017
Earlier work this paper cites.
Evaluating the visualization of what a deep neural network has learned
Samek, Wojciech, Alexander Binder, Grégoire Montavon, Sebastian Lapuschkin, and Klaus-Robert Müller. 2017 · 2017
Earlier work this paper cites.
Axiomatic attribution for deep networks
Sundararajan, Mukund, Ankur Taly, and Qiqi Yan. 2017 · 2017
Earlier work this paper cites.
bleu2vec: the painfully familiar metric on continuous vector space steroids
Tättar, Andre and Mark Fishel. 2017 · 2017
Earlier work this paper cites.
Generating natural language adversarial examples
Alzantot, Moustafa, Yash Sharma, Ahmed Elgohary, Bo-Jhang Ho, Mani Srivastava, and Kai-Wei Chang. 2018 · 2018
Earlier work this paper cites.
e-snli: natural language inference with natural language explanations
Camburu, Oana-Maria, Tim Rocktäschel, Thomas Lukasiewicz, and Phil Blunsom. 2018 · 2018
Earlier work this paper cites.
Universal sentence encoder
Cer, Daniel, Yang. Yinfei, Sheng-yi Kong, Nan Hua, Nicole Limtiaco, Rhomni St. John, Noah Constant, Mario Guajardo-Cespedes, Steve Yuan, Chris Tar, Yun-Hsuan Sung, Brian Strope, and Ray Kurzweil. 2018 · 2018
Earlier work this paper cites.
BERT: pre-training of deep bidirectional transformers for language understanding
Original
Devlin, Jacob, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Earlier work this paper cites.
HotFlip: White-box adversarial examples for text classification
Ebrahimi, Javid, Anyi Rao, Daniel Lowd, and Dejing Dou. 2018 · 2018
Earlier work this paper cites.