Fetching the paper…
Reading the bibliography…
Machine learning is an important tool for decision making, but its ethical and responsible application requires rigorous vetting of its interpretability and utility: an understudied problem, particularly for natural language processing models.
Xplain: A system for creating and explaining expert consulting programs
William R Swartout. 1983 · 1983
Earlier work this paper cites.
Building a large annotated corpus of English: The Penn Treebank
Mitchell P Marcus, Mary Ann Marcinkiewicz, and Beatrice Santorini. 1993 · 1993
Earlier work this paper cites.
Introduction to reinforcement learning
Richard S Sutton and Andrew G Barto. 1998 · 1998
Earlier work this paper cites.
Mixed-initiative interaction
JE Allen, Curry I Guinn, and E Horvtz. 1999 · 1999
Earlier work this paper cites.
Principles of mixed-initiative user interfaces. In International Conference on Human Factors in Computing Systems
Eric Horvitz. 1999 · 1999
Earlier work this paper cites.
Towards improving trust in context-aware systems by displaying system confidence. In Proceedings of the international conference on Human-computer interaction with mobile devices and services
Stavros Antifakos, Nicky Kern, Bernt Schiele, and Adrian Schwaninger. 2005 · 2005
Earlier work this paper cites.
Numeracy and decision making
Ellen Peters, Daniel Västfjäll, Paul Slovic, CK Mertz, Ketti Mazzocco, and Stephan Dickert. 2006 · 2006
Earlier work this paper cites.
Visualization of uncertainty in context aware mobile applications. In Proceedings of the international conference on Human-computer interaction with mobile devices and services
Enrico Rukzio, John Hamard, Chie Noda, and Alexander De Luca. 2006 · 2006
Earlier work this paper cites.
The design of implicit interactions: Making interactive systems less obnoxious
Wendy Ju and Larry Leifer. 2008 · 2008
Earlier work this paper cites.
Numeracy, ratio bias, and denominator neglect in judgments of risk and probability
Valerie F Reyna and Charles J Brainerd. 2008 · 2008
Earlier work this paper cites.
ImageNet: A large-scale hierarchical image database. In Computer Vision and Pattern Recognition
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei. 2009 · 2009
Earlier work this paper cites.
A survey of software learnability: metrics, methodologies and guidelines. In International Conference on Human Factors in Computing Systems
Tovi Grossman, George Fitzmaurice, and Ramtin Attar. 2009 · 2009
Earlier work this paper cites.
How to Explain Individual Classification Decisions
David Baehrens, Timon Schroeter, Stefan Harmeling, Motoaki Kawanabe, Katja Hansen, and Klaus-Robert Müller. 2010 · 2010
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning. In Proceedings of Artificial Intelligence and Statistics
Stéphane Ross, Geoffrey Gordon, and Drew Bagnell. 2011 · 2011
Earlier work this paper cites.
Besting the Quiz Master: Crowdsourcing Incremental Classification Games. In Proceedings of Empirical Methods in Natural Language Processing
Jordan L. Boyd-Graber, Brianna Satinoff, He He, and Hal Daumé III. 2012 · 2012
Earlier work this paper cites.
New Potentials for Data-Driven Intelligent Tutoring System Development and Optimization
Kenneth R. Koedinger, Emma Brunskill, Ryan S.J.d. Baker, Elizabeth A. McLaughlin, and John Stamper. 2013 · 2013
Earlier work this paper cites.
Smarter Than You Think: How Technology is Changing Our Minds for the Better
Clive Thompson. 2013 · 2013
Earlier work this paper cites.
Don’t until the final verb wait: Reinforcement learning for simultaneous machine translation. In Proceedings of Empirical Methods in Natural Language Processing
Alvin Grissom II, He He, Jordan Boyd-Graber, John Morgan, and Hal Daumé III. 2014 · 2014
Earlier work this paper cites.
A Neural Network for Factoid Question Answering over Paragraphs. In Proceedings of Empirical Methods in Natural Language Processing
Mohit Iyyer, Jordan Boyd-Graber, Leonardo Max Batista Claudino, Richard Socher, and Hal Daumé III. 2014 · 2014
Earlier work this paper cites.
Deep Inside Convolutional Networks: Visualising Image Classification Models and Saliency Maps. In Proceedings of the International Conference on Learning Representations
Karen Simonyan, Andrea Vedaldi, and Andrew Zisserman. 2014 · 2014
Earlier work this paper cites.
Explaining and Harnessing Adversarial Examples. In Proceedings of the International Conference on Learning Representations
Ian J. Goodfellow, Jonathon Shlens, and Christian Szegedy. 2015 · 2015
Earlier work this paper cites.
Elasticsearch: The Definitive Guide
Clinton Gormley and Zachary Tong. 2015 · 2015
Earlier work this paper cites.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification. In International Conference on Computer Vision
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2015 · 2015
Earlier work this paper cites.
Deep learning
Yann LeCun, Yoshua Bengio, and Geoffrey Hinton. 2015 · 2015
Cited alongside, same era.
Interpretable classifiers using rules and bayesian analysis: Building a better stroke prediction model
Benjamin Letham, Cynthia Rudin, Tyler H McCormick, David Madigan, et al · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Cited alongside, same era.
General data protection regulation
European Parliament and Council of the European Union. 2016 · 2016
Cited alongside, same era.
Opponent Modeling in Deep Reinforcement Learning. In Proceedings of the International Conference of Machine Learning
He He, Jordan L. Boyd-Graber, Kevin Kwok, and Hal Daumé III. 2016 · 2016
Cited alongside, same era.
Intervention user interfaces: a new interaction paradigm for automated systems
Albrecht Schmidt and Thomas Herrmann. 2017 · 2017
Later among the works it cites.
Mastering the game of Go without human knowledge
David Silver, Julian Schrittwieser, Karen Simonyan, Ioannis Antonoglou, Aja Huang, Arthur Guez, Thomas Hubert, Lucas Baker, Matthew Lai, Adrian Bolton, et al · 2017
Later among the works it cites.
Evaluating visual representations for topic understanding and their effects on manually generated labels
Alison Smith, Tak Yeon Lee, Forough Poursabzi-Sangdeh, Jordan Boyd-Graber, Niklas Elmqvist, and Leah Findlater. 2017 · 2017
Later among the works it cites.
Axiomatic Attribution for Deep Networks. In Proceedings of the International Conference of Machine Learning
Mukund Sundararajan, Ankur Taly, and Qiqi Yan. 2017 · 2017
Later among the works it cites.
Statement on algorithmic transparency and accountability
USACM. 2017 · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Interacting with predictions: Visual inspection of black-box machine learning models. In International Conference on Human Factors in Computing Systems
Josua Krause, Adam Perer, and Kenney Ng. 2016 · 2016
Cited alongside, same era.
Interpretable decision sets: A joint framework for description and prediction. In Knowledge Discovery and Data Mining
Himabindu Lakkaraju, Stephen H Bach, and Jure Leskovec. 2016 · 2016
Cited alongside, same era.
Understanding Neural Networks through Representation Erasure
Jiwei Li, Will Monroe, and Daniel Jurafsky. 2016 · 2016
Cited alongside, same era.
The Mythos of Model Interpretability
Zachary Chase Lipton. 2016 · 2016
Cited alongside, same era.
Why Should I Trust You?": Explaining the Predictions of Any Classifier. In Knowledge Discovery and Data Mining
Marco Túlio Ribeiro, Sameer Singh, and Carlos Guestrin. 2016 · 2016
Cited alongside, same era.
AI challenges in human-robot cognitive teaming
Tathagata Chakraborti, Subbarao Kambhampati, Matthias Scheutz, and Yu Zhang. 2017 · 2017
Cited alongside, same era.
Interpretable explanations of black boxes by meaningful perturbation. In International Conference on Computer Vision
Ruth C Fong and Andrea Vedaldi. 2017 · 2017
Cited alongside, same era.
Sanity Checks for Saliency Maps. In Proceedings of Advances in Neural Information Processing Systems
Julius Adebayo, Been Kim, Ian Goodfellow, Justin Gilmer, and Moritz Hardt. 2018 · 2018
Closest in time.
Human-Computer Question Answering: The Case for Quizbowl
Jordan Boyd-Graber, Shi Feng, and Pedro Rodriguez. 2018 · 2018
Closest in time.
Creative Writing with a Machine in the Loop: Case Studies on Slogans and Stories. In International Conference on Intelligent User Interfaces
Elizabeth Clark, Anne Spencer Ross, Chenhao Tan, Yangfeng Ji, and Noah A Smith. 2018 · 2018
Closest in time.
Towards A Rigorous Science of Interpretable Machine Learning
Finale Doshi-Velez and Been Kim. 2018 · 2018
Closest in time.
Pathologies of Neural Models Make Interpretations Difficult. In Proceedings of Empirical Methods in Natural Language Processing
Shi Feng, Eric Wallace, Alvin Grissom II, Mohit Iyyer, Pedro Rodriguez, and Jordan Boyd-Graber. 2018 · 2018
Closest in time.
Interpretation of Neural Networks is Fragile. In Association for the Advancement of Artificial Intelligence
Amirata Ghorbani, Abubakar Abid, and James Y. Zou. 2018 · 2018
Closest in time.
Evaluating Feature Importance Estimates. In ICML Workshop on Human Interpretability in Machine Learning
Sara Hooker, Dumitru Erhan, Pieter-Jan Kindermans, and Been Kim. 2018 · 2018
Closest in time.
To Trust Or Not To Trust A Classifier. In Proceedings of Advances in Neural Information Processing Systems
Heinrich Jiang, Been Kim, and Maya R. Gupta. 2018 · 2018
Closest in time.
Beyond Word Importance: Contextual Decomposition to Extract Interactions from LSTMs. In Proceedings of the International Conference on Learning Representations
W. James Murdoch, Peter J. Liu, and Bin Yu. 2018 · 2018
Closest in time.
How do Humans Understand Explanations from Machine Learning Systems? An Evaluation of the Human-Interpretability of Explanation
Menaka Narayanan, Emily Chen, Jeffrey He, Been Kim, Sam Gershman, and Finale Doshi-Velez. 2018 · 2018
Closest in time.
Deep k-Nearest Neighbors: Towards Confident, Interpretable and Robust Deep Learning
Nicolas Papernot and Patrick D. McDaniel. 2018 · 2018
Closest in time.
Semantically Equivalent Adversarial Rules for Debugging NLP Models. In Proceedings of the Association for Computational Linguistics
Marco Tulio Ribeiro, Sameer Singh, and Carlos Guestrin. 2018 · 2018
Closest in time.
Improving the Adversarial Robustness and Interpretability of Deep Neural Networks by Regularizing their Input Gradients. In Association for the Advancement of Artificial Intelligence
Andrew Slavin Ross and Finale Doshi-Velez. 2018 · 2018
Closest in time.
Please Stop Explaining Black Box Models for High Stakes Decisions. In NIPS 2018 Workshop on Critiquing and Correcting Trends in Machine Learning
Cynthia Rudin. 2018 · 2018
Closest in time.
Human-Robot Teaming. In Extended Abstracts of the 2018 CHI Conference on Human Factors in Computing Systems
David W Vinson, Leila Takayama, Jodi Forlizzi, Wendy Ju, Maya Cakmak, and Hideaki Kuzuoka. 2018 · 2018
Closest in time.
Trick Me If You Can: Adversarial Writing of Trivia Challenge Questions. In Proceedings of ACL 2018 Student Research Workshop
Eric Wallace and Jordan Boyd-Graber. 2018 · 2018
Closest in time.
On Human Predictions with Explanations and Predictions of Machine Learning Models: A Case Study on Deception Detection. In Proceedings of ACM FAT*
Vivian Lai and Chenhao Tan. 2019 · 2019
Closest in time.