Fetching the paper…
Reading the bibliography…
Despite a surge collection of XAI methods, users still struggle to obtain required AI explanations.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 1901
Earlier work this paper cites.
A technique for computer detection and correction of spelling errors
Fred J Damerau. 1964 · 1964
Earlier work this paper cites.
Binary codes capable of correcting deletions, insertions, and reversals. In Soviet physics doklady , Vol. 10. Soviet Union, 707–710
Vladimir I Levenshtein et al · 1966
Earlier work this paper cites.
Perplexity—a measure of the difficulty of speech recognition tasks
Fred Jelinek, Robert L Mercer, Lalit R Bahl, and James K Baker. 1977 · 1977
Earlier work this paper cites.
Speech & language processing
Dan Jurafsky. 2000 · 2000
Earlier work this paper cites.
Relational agents: a model and implementation of building user trust. In Proceedings of the SIGCHI conference on Human factors in computing systems . 396–403
Timothy Bickmore and Justine Cassell. 2001 · 2001
Earlier work this paper cites.
Bhavya Ghai, Q Vera Liao, Yunfeng Zhang, Rachel Bellamy, and Klaus Mueller. 2020 · 2001
Earlier work this paper cites.
The eyes have it: A task by data type taxonomy for information visualizations
Ben Shneiderman. 2003 · 2003
Earlier work this paper cites.
Are visual explanations useful? a case study in model-in-the-loop prediction
Eric Chu, Deb Roy, and Jacob Andreas. 2020 · 2007
Earlier work this paper cites.
Dynamic time warping
Meinard Müller. 2007 · 2007
Earlier work this paper cites.
A normalized Levenshtein distance metric
Li Yujian and Liu Bo. 2007 · 2007
Earlier work this paper cites.
Conversation analysis
Ian Hutchby and Robin Wooffitt. 2008 · 2008
Earlier work this paper cites.
Creative help: A story writing assistant. In International Conference on Interactive Digital Storytelling . Springer, 81–92
Melissa Roemmele and Andrew S Gordon. 2015 · 2015
Earlier work this paper cites.
Rationalizing Neural Predictions. In Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, Austin, Texas, 107–117
Tao Lei, Regina Barzilay, and Tommi Jaakkola. 2016 · 2016
Earlier work this paper cites.
"Why Should I Trust You?": Explaining the Predictions of Any Classifier. In Proceedings of the 22nd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining, San Francisco, CA, USA, August 13-17, 2016 , Balaji Krishnapuram, Mohak Shah, Alexander J. Smola, Charu C. Aggarwal, Dou Shen, and Rajeev Rastogi (Eds.). ACM, 1135–1144
Marco Túlio Ribeiro, Sameer Singh, and Carlos Guestrin. 2016 · 2016
Earlier work this paper cites.
Does String-Based Neural MT Learn Source Syntax?. In Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, Austin, Texas, 1526–1534
Xing Shi, Inkit Padhi, and Kevin Knight. 2016 · 2016
Earlier work this paper cites.
Fine-grained Analysis of Sentence Embeddings Using Auxiliary Prediction Tasks. In 5th International Conference on Learning Representations, ICLR 2017, Toulon, France, April 24-26, 2017, Conference Track Proceedings . OpenReview.net
Yossi Adi, Einat Kermany, Yonatan Belinkov, Ofer Lavi, and Yoav Goldberg. 2017 · 2017
Earlier work this paper cites.
Towards a rigorous science of interpretable machine learning
Finale Doshi-Velez and Been Kim. 2017 · 2017
Earlier work this paper cites.
The Promise and Peril of Human Evaluation for Model Interpretability
Bernease Herman. 2017 · 2017
Earlier work this paper cites.
A Unified Approach to Interpreting Model Predictions
Scott M Lundberg and Su-In Lee. 2017 · 2017
Earlier work this paper cites.
ParlAI: A Dialog Research Software Platform
A. H. Miller, W. Feng, A. Fisch, J. Lu, D. Batra, A. Bordes, D. Parikh, and J. Weston. 2017 · 2017
Earlier work this paper cites.
Explainable AI: Beware of inmates running the asylum or: How I learnt to stop worrying and love the social and behavioural sciences
Tim Miller, Piers Howe, and Liz Sonenberg. 2017 · 2017
Earlier work this paper cites.
e-SNLI: Natural Language Inference with Natural Language Explanations. In Advances in Neural Information Processing Systems 31: Annual Conference on Neural Information Processing Systems 2018, NeurIPS 2018, December 3-8, 2018, Montréal, Canada , Samy Bengio, Hanna M. Wallach, Hugo Larochelle, Kristen Grauman, Nicolò Cesa-Bianchi, and Roman Garnett (Eds.). 9560–9572
Oana-Maria Camburu, Tim Rocktäschel, Thomas Lukasiewicz, and Phil Blunsom. 2018 · 2018
Earlier work this paper cites.
Iris: A conversational agent for complex tasks. In Proceedings of the 2018 CHI conference on human factors in computing systems . 1–12
Ethan Fast, Binbin Chen, Julia Mendelsohn, Jonathan Bassen, and Michael S Bernstein. 2018 · 2018
Earlier work this paper cites.
Feedback Orchestration: Structuring Feedback for Facilitating Reflection and Revision in Writing. In Companion of the 2018 ACM Conference on Computer Supported Cooperative Work and Social Computing . 257–260
Yi-Ching Huang, Hao-Chuan Wang, and Jane Yung-jen Hsu. 2018 · 2018
Earlier work this paper cites.
Did the model understand the question?
Pramod Kaushik Mudrakarta, Ankur Taly, Mukund Sundararajan, and Kedar Dhamdhere. 2018 · 2018
Earlier work this paper cites.
Anchors: High-Precision Model-Agnostic Explanations. In Proceedings of the Thirty-Second AAAI Conference on Artificial Intelligence, (AAAI-18), the 30th innovative Applications of Artificial Intelligence (IAAI-18), and the 8th AAAI Symposium on Educational Advances in Artificial Intelligence (EAAI-18), New Orleans, Louisiana, USA, February 2-7, 2018 , Sheila A. McIlraith and Kilian Q. Weinberger (Eds.). AAAI Press, 1527–1535
Marco Túlio Ribeiro, Sameer Singh, and Carlos Guestrin. 2018 · 2018
Earlier work this paper cites.
Glass-Box: Explaining AI Decisions With Counterfactual Statements Through Conversation With a Voice-enabled Virtual Assistant.. In IJCAI . 5868–5870
Kacper Sokol and Peter A Flach. 2018 · 2018
Earlier work this paper cites.
SciBERT: A pretrained language model for scientific text
Iz Beltagy, Kyle Lo, and Arman Cohan. 2019 · 2019
Cited alongside, same era.
Chains-of-Reasoning at TextGraphs 2019 Shared Task: Reasoning over Chains of Facts for Explainable Multi-hop Inference. In Proceedings of the Thirteenth Workshop on Graph-Based Methods for Natural Language Processing (TextGraphs-13) . Association for Computational Linguistics, Hong Kong, 101–117
Rajarshi Das, Ameya Godbole, Manzil Zaheer, Shehzaad Dhuliawala, and Andrew McCallum. 2019 · 2019
Cited alongside, same era.
What can AI do for me? evaluating machine learning interpretations in cooperative play. In Proceedings of the 24th International Conference on Intelligent User Interfaces . 229–239
Shi Feng and Jordan Boyd-Graber. 2019 · 2019
Cited alongside, same era.
GEval: Tool for Debugging NLP Datasets and Models. In Proceedings of the 2019 ACL Workshop BlackboxNLP: Analyzing and Interpreting Neural Networks for NLP . Association for Computational Linguistics, Florence, Italy, 254–262
How Useful Are the Machine-Generated Interpretations to General Users? A Human Evaluation on Guessing the Incorrectly Predicted Labels. In Proceedings of the AAAI Conference on Human Computation and Crowdsourcing , Vol. 8. 168–172
Hua Shen and Ting-Hao Huang. 2020 · 2020
Later among the works it cites.
No Explainability without Accountability: An Empirical Study of Explanations and Feedback in Interactive ML. In CHI ’20: CHI Conference on Human Factors in Computing Systems, Honolulu, HI, USA, April 25-30, 2020 , Regina Bernhaupt, Florian ’Floyd’ Mueller, David Verweij, Josh Andres, Joanna McGrenere, Andy Cockburn, Ignacio Avellino, Alix Goguey, Pernille Bjøn, Shengdong Zhao, Briane Paul Samson, and Rafal Kocielnik (Eds.). ACM, 1–13
Alison Smith-Renner, Ron Fan, Melissa Birchfield, Tongshuang Wu, Jordan L. Boyd-Graber, Daniel S. Weld, and Leah Findlater. 2020 · 2020
Later among the works it cites.
Does the whole exceed its parts? the effect of ai explanations on complementary team performance. In Proceedings of the 2021 CHI Conference on Human Factors in Computing Systems . 1–16
Gagan Bansal, Tongshuang Wu, Joyce Zhou, Raymond Fok, Besmira Nushi, Ece Kamar, Marco Tulio Ribeiro, and Daniel Weld. 2021a · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Filip Graliński, Anna Wróblewska, Tomasz Stanisławek, Kamil Grabowski, and Tomasz Górecki. 2019 · 2019
Cited alongside, same era.
Acute-eval: Improved dialogue evaluation with optimized questions and multi-turn comparisons
Margaret Li, Jason Weston, and Stephen Roller. 2019 · 2019
Cited alongside, same era.
Explanation in artificial intelligence: Insights from the social sciences
Tim Miller. 2019 · 2019
Cited alongside, same era.
Model cards for model reporting. In Proceedings of the conference on fairness, accountability, and transparency . 220–229
Margaret Mitchell, Simone Wu, Andrew Zaldivar, Parker Barnes, Lucy Vasserman, Ben Hutchinson, Elena Spitzer, Inioluwa Deborah Raji, and Timnit Gebru. 2019 · 2019
Cited alongside, same era.
A slow algorithm improves users’ assessments of the algorithm’s accuracy
Joon Sung Park, Rick Barber, Alex Kirlik, and Karrie Karahalios. 2019 · 2019
Cited alongside, same era.
Explain Yourself! Leveraging Language Models for Commonsense Reasoning. In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics . Association for Computational Linguistics, Florence, Italy, 4932–4942
Nazneen Fatema Rajani, Bryan McCann, Caiming Xiong, and Richard Socher. 2019 · 2019
Cited alongside, same era.
CoQA: A Conversational Question Answering Challenge
Siva Reddy, Danqi Chen, and Christopher D. Manning. 2019 · 2019
Cited alongside, same era.
Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks. In Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP) . Association for Computational Linguistics, Hong Kong, China, 3982–3992
Nils Reimers and Iryna Gurevych. 2019 · 2019
Cited alongside, same era.
Do Neural Dialog Systems Use the Conversation History Effectively? An Empirical Study. In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics . Association for Computational Linguistics, Florence, Italy, 32–37
Chinnadhurai Sankar, Sandeep Subramanian, Chris Pal, Sarath Chandar, and Yoshua Bengio. 2019 · 2019
Cited alongside, same era.
Does the Whole Exceed its Parts? The Effect of AI Explanations on Complementary Team Performance
Gagan Bansal, Tongshuang Wu, Joyce Zhou, Raymond Fok, Besmira Nushi, Ece Kamar, Marco Túlio Ribeiro, and Daniel S. Weld. 2021b · 2021
Later among the works it cites.
AI-assisted peer review
Alessandro Checco, Lorenzo Bracciale, Pierpaolo Loreti, Stephen Pinfield, and Giuseppe Bianchi. 2021 · 2021
Later among the works it cites.
Wordcraft: A human-AI collaborative editor for story writing
Andy Coenen, Luke Davis, Daphne Ippolito, Emily Reif, and Ann Yuan. 2021 · 2021
Later among the works it cites.
Datasheets for datasets
Timnit Gebru, Jamie Morgenstern, Briana Vecchione, Jennifer Wortman Vaughan, Hanna Wallach, Hal Daumé Iii, and Kate Crawford. 2021 · 2021
Later among the works it cites.
Explanation-based human debugging of nlp models: A survey
Piyawat Lertvittayakumjorn and Francesca Toni. 2021 · 2021
Later among the works it cites.
Explaining by Conversing: The Argument for Conversational Xai Systems
Wassim Marrakchi. 2021 · 2021
Later among the works it cites.
Towards an Explainer-agnostic Conversational XAI.. In IJCAI . 4909–4910
Navid Nobani, Fabio Mercorio, and Mario Mezzanzanica. 2021 · 2021
Later among the works it cites.
An empirical comparison of instance attribution methods for NLP
Pouya Pezeshkpour, Sarthak Jain, Byron C Wallace, and Sameer Singh. 2021 · 2021
Later among the works it cites.
Manipulating and Measuring Model Interpretability
Forough Poursabzi-Sangdeh, D. Goldstein, J. Hofman, Jennifer Wortman Vaughan, and H. Wallach. 2021 · 2021
Later among the works it cites.
Explaining the Road Not Taken
Hua Shen and Ting-Hao’Kenneth’ Huang. 2021 · 2021
Later among the works it cites.
Exploring and promoting diagnostic transparency and explainability in online symptom checkers. In Proceedings of the 2021 CHI Conference on Human Factors in Computing Systems . 1–17
Chun-Hua Tsai, Yue You, Xinning Gui, Yubo Kou, and John M Carroll. 2021 · 2021
Later among the works it cites.
Polyjuice: Automated, General-purpose Counterfactual Generation
Tongshuang Wu, Marco Tulio Ribeiro, Jeffrey Heer, and Daniel S Weld. 2021 · 2021
Later among the works it cites.
Can we automate scientific reviewing?
Weizhe Yuan, Pengfei Liu, and Graham Neubig. 2021 · 2021
Later among the works it cites.
Read, Revise, Repeat: A System Demonstration for Human-in-the-loop Iterative Text Revision
Wanyu Du, Zae Myung Kim, Vipul Raheja, Dhruv Kumar, and Dongyeop Kang. 2022 · 2022
Later among the works it cites.
A Design Space for Writing Support Tools Using a Cognitive Process Model of Writing. In Proceedings of the First Workshop on Intelligent and Interactive Writing Assistants (In2Writing 2022) . 11–24
Katy Gero, Alex Calderwood, Charlotte Li, and Lydia Chilton. 2022 · 2022
Later among the works it cites.
Rethinking Explainability as a Dialogue: A Practitioner’s Perspective
Himabindu Lakkaraju, Dylan Slack, Yuxin Chen, Chenhao Tan, and Sameer Singh. 2022 · 2022
Later among the works it cites.
Coauthor: Designing a human-ai collaborative writing dataset for exploring language model capabilities. In CHI Conference on Human Factors in Computing Systems . 1–19
Mina Lee, Percy Liang, and Qian Yang. 2022 · 2022
Later among the works it cites.
How do you Converse with an Analytical Chatbot? Revisiting Gricean Maxims for Designing Analytical Conversational Behavior. In CHI Conference on Human Factors in Computing Systems . 1–17
Vidya Setlur and Melanie Tory. 2022 · 2022
Later among the works it cites.
Are Shortest Rationales the Best Explanations for Human Understanding?. In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers) . Association for Computational Linguistics, Dublin, Ireland, 10–19
Hua Shen, Tongshuang Wu, Wenbo Guo, and Ting-Hao Huang. 2022 · 2022
Later among the works it cites.
TalkToModel: Explaining Machine Learning Models with Interactive Natural Language Conversations
Dylan Slack, Satyapriya Krishna, Himabindu Lakkaraju, and Sameer Singh. 2022 · 2022
Later among the works it cites.
Exploring the Effects of Interactive Dialogue in Improving User Control for Explainable Online Symptom Checkers. In CHI Conference on Human Factors in Computing Systems Extended Abstracts . 1–7
Yuan Sun and S Shyam Sundar. 2022 · 2022
Later among the works it cites.
Interpretable Directed Diversity: Leveraging Model Explanations for Iterative Crowd Ideation. In CHI Conference on Human Factors in Computing Systems . 1–28
Yunlong Wang, Priyadarshini Venkatesh, and Brian Y Lim. 2022 · 2022
Later among the works it cites.
How to Guide Task-oriented Chatbot Users, and When: A Mixed-methods Study of Combinations of Chatbot Guidance Types and Timings. In CHI Conference on Human Factors in Computing Systems . 1–16
Su-Fang Yeh, Meng-Hsin Wu, Tze-Yu Chen, Yen-Chun Lin, XiJing Chang, You-Hsuan Chiang, and Yung-Ju Chang. 2022 · 2022
Later among the works it cites.
Parachute: Evaluating interactive human-lm co-writing systems
Hua Shen and Tongshuang Wu. 2023 · 2023
Closest in time.
Wikum: Bridging discussion forums and wikis using recursive summarization. In Proceedings of the 2017 ACM Conference on Computer Supported Cooperative Work and Social Computing . 2082–2096
Amy X Zhang, Lea Verou, and David Karger. 2017 · 2096
Closest in time.