Fetching the paper…
Reading the bibliography…
In a human-AI collaboration, users build a mental model of the AI system based on its reliability and how it presents its decision, e.g.
Quizbowl: The case for incremental question answering
Pedro Rodriguez, Shi Feng, Mohit Iyyer, He He, and Jordan Boyd-Graber. 2019 · 1904
Earlier work this paper cites.
How model accuracy and explanation fidelity influence user trust
Andrea Papenmeier, Gwenn Englebienne, and Christin Seifert. 2019 · 1907
Earlier work this paper cites.
Variabilità e mutabilità: Contributo allo studio delle distribuzioni e delle relazioni statistiche.[Fasc. I.]
Corrado Gini. 1912 · 1912
Earlier work this paper cites.
An optimum character recognition system using decision functions
Chi-Keung Chow. 1957 · 1957
Earlier work this paper cites.
Advances in prospect theory: Cumulative representation of uncertainty
Amos Tversky and Daniel Kahneman. 1992 · 1992
Earlier work this paper cites.
Trust, self-confidence, and operators’ adaptation to automation
John D Lee and Neville Moray. 1994 · 1994
Earlier work this paper cites.
Methodology matters: Doing research in the behavioral and social sciences
Joseph E McGrath. 1995 · 1995
Earlier work this paper cites.
A meta-analytic review of experiments examining the effects of extrinsic rewards on intrinsic motivation
Edward L Deci, Richard Koestner, and Richard M Ryan. 1999 · 1999
Earlier work this paper cites.
Selective question answering under domain shift
Amita Kamath, Robin Jia, and Percy Liang. 2020 · 2006
Earlier work this paper cites.
On the foundations of noise-free selective classification
Ran El-Yaniv et al. 2010 · 2010
Earlier work this paper cites.
The psychology of risk
Paul Slovic. 2010 · 2010
Earlier work this paper cites.
Machine translation evaluation versus quality estimation
Lucia Specia, Dhwaj Raj, and Marco Turchi. 2010 · 2010
Earlier work this paper cites.
Human evaluation of spoken vs. visual explanations for open-domain qa
Ana Valeria Gonzalez, Gagan Bansal, Angela Fan, Robin Jia, Yashar Mehdad, and Srinivasan Iyer. 2020 · 2012
Earlier work this paper cites.
The UX Book: Process and guidelines for ensuring a quality user experience
Rex Hartson and Pardha S Pyla. 2012 · 2012
Earlier work this paper cites.
Interrupted time series design
John Ferron and Gianna Rendina-Gobioff. 2014 · 2014
Earlier work this paper cites.
Cognition is a matter of trust: Distrust tunes cognitive processes
Ruth Mayo. 2015 · 2015
Earlier work this paper cites.
" why should i trust you?" explaining the predictions of any classifier
Marco Tulio Ribeiro, Sameer Singh, and Carlos Guestrin. 2016 · 2016
Earlier work this paper cites.
Under-versus overconfidence: An experiment on how others perceive a biased self-assessment
Carmen Thoma. 2016 · 2016
Earlier work this paper cites.
Diagnostic assessment of deep learning algorithms for detection of lymph node metastases in women with breast cancer
Babak Ehteshami Bejnordi, Mitko Veta, Paul Johannes Van Diest, Bram Van Ginneken, Nico Karssemeijer, Geert Litjens, Jeroen AWM Van Der Laak, Meyke Hermsen, Quirine F Manson, Maschenka Balkenhol, et al. 2017 · 2017
Earlier work this paper cites.
On calibration of modern neural networks
Chuan Guo, Geoff Pleiss, Yu Sun, and Kilian Q Weinberger. 2017 · 2017
Cited alongside, same era.
Triviaqa: A large scale distantly supervised challenge dataset for reading comprehension
Mandar Joshi, Eunsol Choi, Daniel S Weld, and Luke Zettlemoyer. 2017 · 2017
Cited alongside, same era.
Lukasz Kaiser, Aidan N Gomez, Noam Shazeer, Ashish Vaswani, Niki Parmar, Llion Jones, and Jakob Uszkoreit. 2017 · 2017
Cited alongside, same era.
An introduction to neural information retrieval
Bhaskar Mitra, Nick Craswell, et al. 2018 · 2018
Cited alongside, same era.
Predictive model to assess user trust: A psycho-physiological approach
Ighoyota Ben Ajenaghughrure, Sonia C Sousa, Ilkka Johannes Kosunen, and David Lamas. 2019 · 2019
Cited alongside, same era.
Calibration of machine reading systems at scale
Shehzaad Dhuliawala, Leonard Adolphs, Rajarshi Das, and Mrinmaya Sachan. 2022 · 2022
Later among the works it cites.
A review on human–machine trust evaluation: Human-centric and machine-centric perspectives
Biniam Gebru, Lydia Zeleke, Daniel Blankson, Mahmoud Nabil, Shamila Nateghi, Abdollah Homaifar, and Edward Tunstel. 2022 · 2022
Later among the works it cites.
Designing for trust: A set of design principles to increase trust in chatbot
Yunsan Guo, Jian Wang, Runfan Wu, Zeyu Li, and Lingyun Sun. 2022 · 2022
Later among the works it cites.
Teaching humans when to defer to a classifier via exemplars
Hussein Mozannar, Arvind Satyanarayan, and David Sontag. 2022 · 2022
Later among the works it cites.
It’s complicated: The relationship between user trust, model accuracy and explanations in AI
Andrea Papenmeier, Dagmar Kern, Gwenn Englebienne, and Christin Seifert. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Updates in human-AI teams: Understanding and addressing the performance/compatibility tradeoff
Gagan Bansal, Besmira Nushi, Ece Kamar, Daniel S Weld, Walter S Lasecki, and Eric Horvitz. 2019 · 2019
Cited alongside, same era.
Natural questions: A benchmark for question answering research
Tom Kwiatkowski, Jennimaria Palomaki, Olivia Redfield, Michael Collins, Ankur Parikh, Chris Alberti, Danielle Epstein, Illia Polosukhin, Matthew Kelcey, Jacob Devlin, Kenton Lee, Kristina N. Toutanova, Llion Jones, Ming-Wei Chang, Andrew Dai, Jakob Uszkoreit, Quoc Le, and Slav Petrov. 2019 · 2019
Cited alongside, same era.
An initial model of trust in chatbots for customer service—findings from a questionnaire study
Cecilie Bertinussen Nordheim, Asbjørn Følstad, and Cato Alexander Bjørkli. 2019 · 2019
Cited alongside, same era.
Understanding the effect of accuracy on trust in machine learning models
Ming Yin, Jennifer Wortman Vaughan, and Hanna Wallach. 2019 · 2019
Cited alongside, same era.
Physiological indicators for user trust in machine learning with influence enhanced fact-checking
Jianlong Zhou, Huaiwen Hu, Zhidong Li, Kun Yu, and Fang Chen. 2019 · 2019
Cited alongside, same era.
Modeling trust in human-robot interaction: A survey
Zahra Rezaei Khavas, S Reza Ahmadzadeh, and Paul Robinette. 2020 · 2020
Cited alongside, same era.
Effect of confidence and explanation on accuracy and trust calibration in AI-assisted decision making
Yunfeng Zhang, Q Vera Liao, and Rachel KE Bellamy. 2020 · 2020
Cited alongside, same era.
When confidence meets accuracy: Exploring the effects of multiple performance indicators on trust in machine learning models
Amy Rechkemmer and Ming Yin. 2022 · 2022
Later among the works it cites.
Calibrating trust in AI-assisted decision making
Amy Turner, Meena Kaushik, Mu-Ti Huang, and Srikar Varanasi. 2022 · 2022
Later among the works it cites.
Uncalibrated Models Can Improve Human-AI Collaboration
Kailas Vodrahalli, Tobias Gerstenberg, and James Zou. 2022 · 2022
Later among the works it cites.
Effects of explanations in AI-assisted decision making: Principles and comparisons
Xinru Wang and Ming Yin. 2022 · 2022
Later among the works it cites.
Human-aligned calibration for AI-Assisted decision making
Nina L Corvelo Benz and Manuel Gomez Rodriguez. 2023 · 2023
Closest in time.
Artificial intelligence and the future of teaching and learning
Miguel A. Cardona, Roberto J. Rodríguez, and Kristina Ishmael. 2023 · 2023
Closest in time.
Sabrina Chiesurin, Dimitris Dimakopoulos, Marco Antonio Sobrevilla Cabezudo, Arash Eshghi, Ioannis Papaioannou, Verena Rieser, and Ioannis Konstas. 2023 · 2023
Closest in time.
Modeling human trust and reliance in AI-assisted decision making: A markovian approach
Zhuoyan Li, Zhuoran Lu, and Ming Yin. 2023 · 2023
Closest in time.
Evaluating verifiability in generative search engines
Nelson F. Liu, Tianyi Zhang, and Percy Liang. 2023 · 2023
Closest in time.
Shuai Ma, Ying Lei, Xinru Wang, Chengbo Zheng, Chuhan Shi, Ming Yin, and Xiaojuan Ma. 2023 · 2023
Closest in time.
OpenAI. 2023 · 2023
Closest in time.
Background explanations reduce users’ over-reliance on AI: A case study on multi-hop question answering
Alicia Vikander. 2023 · 2023
Closest in time.
Poor man’s quality estimation: Predicting reference-based MT metrics without the reference
Vilém Zouhar, Shehzaad Dhuliawala, Wangchunshu Zhou, Nico Daheim, Tom Kocmi, Yuchen Eleanor Jiang, and Mrinmaya Sachan. 2023 · 2023
Closest in time.