Fetching the paper…
Reading the bibliography…
Background: Large language models (LLMs) such as OpenAI's GPT-4 or Google's PaLM 2 are proposed as viable diagnostic support tools or even spoken of as replacements for "curbside consults".
“Group discussion improves lie detection”
Nadav Klein and Nicholas Epley · 2015
Earlier work this paper cites.
“Comparative accuracy of diagnosis by collective intelligence of multiple physicians vs individual physicians”
Michael Barnett, Dhruv Boddupalli, Shantanu Nundy and David Bates · 2019
Earlier work this paper cites.
“Frameworks for collective intelligence: A systematic literature review”
Shweta Suran, Vishwajeet Pattanaik and Dirk Draheim · 2020
Earlier work this paper cites.
“Large language model (ChatGPT) as a support tool for breast tumor board”
Vera Sorin, Eyal Klang, Miri Sklair-Levy, Israel Cohen, Douglas Zippel, Nora Balint, Eli Konen and Yiftach Barash · 2023
Earlier work this paper cites.
“A Large Language Model Screening Tool to Target Patients for Best Practice Alerts: Development and Validation”
Thomas Savage, John Wang and Lisa Shieh · 2023
Earlier work this paper cites.
“Artificial intelligence and machine learning in clinical medicine, 2023”
Charlotte Haug and Jeffrey Drazen · 2023
Earlier work this paper cites.
“ChatGPT: the future of discharge summaries?”
Sajan Patel and Kyle Lam · 2023
Cited alongside, same era.
“Benefits, limits, and risks of GPT-4 as an AI chatbot for medicine”
Peter Lee, Sebastien Bubeck and Joseph Petro · 2023
Cited alongside, same era.
“Black box warning: large language models and the future of infectious diseases consultation”
Ilan Schwartz, Katherine Link, Roxana Daneshjou and Nicolás Cortés-Penfield · 2023
Cited alongside, same era.
“Worldwide AI ethics: A review of 200 guidelines and recommendations for AI governance”
Nicholas Corrêa, Camila Galvão, James Santos, Carolina Del, Edson Pinto, Camila Barbosa, Diogo Massmann, Rodrigo Mambrini, Luiza Galvão and Edmund Terem · 2023
Cited alongside, same era.
“Use of GPT-4 to diagnose complex clinical cases”
Alexander Eriksen, Sören Möller and Jesper Ryg · 2023
Cited alongside, same era.
“Automating hybrid collective intelligence in open-ended medical diagnostics”
Ralf Kurvers, Andrea Nuzzolese, Alessandro Russo, Gioele Barabucci, Stefan Herzog and Vito Trianni · 2023
Later among the works it cites.
“LLM-Blender: Ensembling Large Language Models with Pairwise Ranking and Generative Fusion”
Dongfu Jiang, Xiang Ren and Bill Lin · 2023
Later among the works it cites.
“One LLM is not Enough: Harnessing the Power of Ensemble Learning for Medical Question Answering”
Han Yang, Mingchen Li, Yongkang Xiao, Huixue Zhou, Rui Zhang and Qian Fang · 2023
Later among the works it cites.
“Chateval: Towards better llm-based evaluators through multi-agent debate”
Chi-Min Chan, Weize Chen, Yusheng Su, Jianxuan Yu, Wei Xue, Shanghang Zhang, Jie Fu and Zhiyuan Liu · 2023
Later among the works it cites.
“Diagnostic Accuracy of a Large Language Model in Pediatric Case Studies”
Joseph Barile, Alex Margolis, Grace Cason, Rachel Kim, Saia Kalash, Alexis Tchaconas and Ruth Milanaik · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Closest in time.