Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have demonstrated great potential for conducting diagnostic conversations but evaluation has been largely limited to language-only interactions, deviating from the real-world requirements of remote care delivery.
“Language models are few-shot learners”
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry and Amanda Askell · 1901
Earlier work this paper cites.
“Assessment of clinical competence using objective structured examination”
RM Harden, M Stevenson, WW Downie and GM Wilson · 1975
Earlier work this paper cites.
In Annals of Internal Medicine 85.6
“Computer-Based Medical Consultations: MYCIN.” · 1976
Earlier work this paper cites.
“INTERNIST-1, an experimental computer-based diagnostic consultant for general internal medicine”
Randolph Miller, Harry Pople and Jack Myers · 1982
Earlier work this paper cites.
“The Objective Structured Clinical Examination. The new gold standard for evaluating postgraduate clinical performance.”
David Sloan, Michael Donnelly, Richard Schwartz and William Strodel · 1995
Earlier work this paper cites.
“The prevalence and prognostic significance of electrocardiographic abnormalities”
Euan Ashley, Vinod Raxwal and Victor Froelicher · 2000
Earlier work this paper cites.
“The big five personality factors and personal values”
Sonia Roccas, Lilach Sagiv, Shalom Schwartz and Ariel Knafo · 2002
Earlier work this paper cites.
“MRCP (UK) PART 2 Clinical Examination (PACES): a review of the first four examination sessions (June 2001–July 2002)”
Jane Dacre, Mike Besser and Patricia White · 2003
Earlier work this paper cites.
“Techniques for measuring clinical competence: objective structured clinical examinations”
David Newble · 2004
Earlier work this paper cites.
Nick. Phillips, Pranav Rajpurkar, Mark Sabini, Rayan Krishnan, Sharon Zhou, Anuj Pareek, Nguyet Phu, Chris Wang, Mudit Jain, Nguyen Du, Steven Truong, Andrew. Ng and Matthew. Lungren · 2007
Earlier work this paper cites.
““Best practice” for patient-centered communication: a narrative review”
Ann King and Ruth Hoppe · 2013
Earlier work this paper cites.
“Mobile telephone text messaging for medication adherence in chronic disease: a meta-analysis”
Jitesh Thakkar, Rahul Kurup, Tracey-Lea Laba, Karla Santo, Aravinda Thiagalingam, Anthony Rodgers, Mark Woodward, Julie Redfern and Clara Chow · 2015
Earlier work this paper cites.
“WhatsApp messenger as an adjunctive tool for telemedicine: an overview”
Vincenzo Giordano, Hilton Koch, Alexandre Godoy-Santos, William Belangero, Robinson Pires and Pedro Labronici · 2017
Earlier work this paper cites.
“Clinical reasoning: defining it, teaching it, assessing it, studying it”
Larry Gruppen · 2017
Earlier work this paper cites.
“Real-world use of telemedicine-a picture is worth a thousand words”
Richard Wong and Ken Dunn · 2018
Earlier work this paper cites.
“Effectiveness of text messaging interventions for the management of depression: A systematic review and meta-analysis”
Buddhika Senanayake, Sumudu Wickramasinghe, Mark Chatfield, Julie Hansen, Sisira Edirippulige and Anthony Smith · 2019
Earlier work this paper cites.
“Acceptability, benefits, and challenges of video consulting: a qualitative study in primary care”
Eddie Donaghy, Helen Atherton, Victoria Hammersley, Hannah McNeilly, Annemieke Bikker, Lucy Robbins, John Campbell and Brian McKinstry · 2019
Earlier work this paper cites.
“PAD-UFES-20: A skin lesion dataset composed of patient data and clinical images collected from smartphones”
Andre.C. Pacheco, Gustavo. Lima, Amanda. Salomão, Breno Krohling, Igor. Biral, Gabriel. de Angelo, Fábio.R. Alves Jr, José.M. Esgario, Alana. Simora, Pedro.C. Castro, Felipe. Rodrigues, Patricia.L. Frasson, Renato. Krohling, Helder Knidel, Maria.S. Santos, Rachel. do Espírito Santo, Telma.S.G. Macedo, Tania.P. Canuto and Luíz.S. de Barros · 2020
Earlier work this paper cites.
“PTB-XL, a large publicly available electrocardiography dataset”
Patrick Wagner, Nils Strodthoff, Ralf-Dieter Bousseljot, Dieter Kreiseler, Fatima Lunze, Wojciech Samek and Tobias Schaeffter · 2020
Earlier work this paper cites.
“A deep learning system for differential diagnosis of skin diseases”
Yuan Liu, Ayush Jain, Clara Eng, David Way, Kang Lee, Peggy Bui, Kimberly Kanada, Guilherme de Oliveira, Jessica Gallegos and Sara Gabriele · 2020
Earlier work this paper cites.
“Utility of WhatsApp in healthcare provision and sharing of medical information with caregivers of children with neurodisabilties: experience from Sudan”, 2021
IN Mohamed and MA Elseed · 2021
Earlier work this paper cites.
“The state of telehealth before and after the COVID-19 pandemic”
Julia Shaver · 2022
Earlier work this paper cites.
“Smartphone technology for communications between clinicians–A scoping review”
Bernadette John, Christine McCreary and Anthony Roberts · 2022
Earlier work this paper cites.
“Using the apple watch to record multiple-lead electrocardiograms in detecting myocardial infarction: where are we now?”
Ke Li, Abdelmotagaly Elgalad, Cristiano Cardoso and Emerson Perin · 2022
Earlier work this paper cites.
“Training language models to follow instructions with human feedback”
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama and Alex Ray · 2022
Cited alongside, same era.
“Constitutional AI: Harmlessness from AI feedback”
Yuntao Bai, Saurav Kadavath, Sandipan Kundu, Amanda Askell, Jackson Kernion, Andy Jones, Anna Chen, Anna Goldie, Azalia Mirhoseini and Cameron McKinnon · 2022
Cited alongside, same era.
“Chain-of-thought prompting elicits reasoning in large language models”
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc Le and Denny Zhou · 2022
Cited alongside, same era.
“Self-consistency improves chain of thought reasoning in language models”
Xuezhi Wang, Jason Wei, Dale Schuurmans, Quoc Le, Ed Chi and Denny Zhou · 2022
Cited alongside, same era.
“The global primary care crisis”
Euan Lawson · 2023
Cited alongside, same era.
“Multimodal Healthcare AI: Identifying and Designing Clinically Relevant Vision-Language Applications for Radiology”
Nur Yildirim, Hannah Richardson, Maria Wetscherek, Junaid Bajwa, Joseph Jacob, Mark Pinnock, Stephen Harris, Daniel Coelho, Shruthi Bannur, Stephanie Hyland, Pratik Ghosh, Mercy Ranjit, Kenza Bouzid, Anton Schwaighofer, Fernando Pérez-García, Harshita Sharma, Ozan Oktay, Matthew Lungren, Javier Alvarez-Valle, Aditya Nori and Anja Thieme · 2024
Later among the works it cites.
“The future of multimodal artificial intelligence models for integrating imaging and clinical metadata: a narrative review”, 2024
BD Simon, KB Ozyoruk, DG Gelikman, SA Harmon and B Türkbey · 2024
Later among the works it cites.
Amnon Bleich, Antje Linnemann, Bjoern. Diem and Tim Conrad · 2024
Later among the works it cites.
“GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI”, 2024
Pengcheng Chen, Jin Ye, Guoan Wang, Yanjun Li, Zhongying Deng, Wei Li, Tianbin Li, Haodong Duan, Ziyan Huang, Yanzhou Su, Benyou Wang, Shaoting Zhang, Bin Fu, Jianfei Cai, Bohan Zhuang, Eric Seibel, Junjun He and Yu Qiao · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“The layered crisis of the primary care medical workforce in the European region: what evidence do we need to identify causes and solutions?”
Giuliano Russo, Julian Perelman, Tomas Zapata and Milena Šantrić-Milićević · 2023
Cited alongside, same era.
“Diagnostic accuracy of artificial intelligence in virtual primary care”
Dan Zeltzer, Lee Herzog, Yishai Pickman, Yael Steuerman, Ran Ber, Zehavi Kugler, Ran Shaul and Jon Ebbert · 2023
Cited alongside, same era.
“The role of digital literacy in achieving health equity in the third millennium society: A literature review”
Laura Campanozzi, Filippo Gibelli, Paolo Bailo, Giulio Nittari, Ascanio Sirignano and Giovanna Ricci · 2023
Cited alongside, same era.
Jungwoo Oh, Gyubok Lee, Seongsu Bae, Joon-myoung Kwon and Edward Choi · 2023
Cited alongside, same era.
“LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day”, 2023
Chunyuan Li, Cliff Wong, Sheng Zhang, Naoto Usuyama, Haotian Liu, Jianwei Yang, Tristan Naumann, Hoifung Poon and Jianfeng Gao · 2023
Cited alongside, same era.
“Med-Flamingo: a Multimodal Medical Few-shot Learner”, 2023
Michael Moor, Qian Huang, Shirley Wu, Michihiro Yasunaga, Cyril Zakka, Yash Dalmia, Eduardo Reis, Pranav Rajpurkar and Jure Leskovec · 2023
Cited alongside, same era.
“Active Acquisition for Multimodal Temporal Data: A Challenging Decision-Making Task”, 2023
Jannik Kossen, Cătălina Cangea, Eszter Vértes, Andrew Jaegle, Viorica Patraucean, Ira Ktena, Nenad Tomasev and Danielle Belgrave · 2023
Cited alongside, same era.
Later among the works it cites.
“AgentClinic: a multimodal agent benchmark to evaluate AI in simulated clinical environments”, 2024
Samuel Schmidgall, Rojin Ziaei, Carl Harris, Eduardo Reis, Jeffrey Jopling and Michael Moor · 2024
Later among the works it cites.
“Pre-trained multimodal large language model enhances dermatological diagnosis using SkinGPT-4”
Juexiao Zhou, Xiaonan He, Liyuan Sun, Jiannan Xu, Xiuying Chen, Yuetan Chu, Longxi Zhou, Xingyu Liao, Bin Zhang and Shawn Afvari · 2024
Later among the works it cites.
“MINT: A wrapper to make multi-modal and multi-image AI models interactive”, 2024
Jan Freyberg, Abhijit Roy, Terry Spitz, Beverly Freeman, Mike Schaekermann, Patricia Strachan, Eva Schnider, Renee Wong, Dale Webster, Alan Karthikesalingam, Yun Liu, Krishnamurthy Dvijotham and Umesh Telang · 2024
Later among the works it cites.
“Hospitals should check staff use of WhatsApp to ensure patient safety, UK watchdog says” URL: https://www.ft.com/content/f428e4ff-4dd9-4613-8485-dfc77fa154ec , Financial Times, 2024
Laura Hughes · 2024
Later among the works it cites.
“’What’s so wrong with Whatsapp?”
Tom Dobbs · 2024
Later among the works it cites.
“Text Messaging to Extend School-Based Suicide Prevention: Pilot Randomized Controlled Trial”
Anthony Pisani, Peter Wyman, Ian Cero, Caroline Kelberman, Kunali Gurditta, Emily Judd, Karen Schmeelk-Cone, David Mohr, David Goldston and Ashkan Ertefaie · 2024
Later among the works it cites.
“Towards conversational diagnostic ai”
Tao Tu, Shekoofeh Azizi, Danny Driess, Mike Schaekermann, Mohamed Amin, Pi-Chuan Chang, Andrew Carroll, Chuck Lau, Ryutaro Tanno and Ira Ktena · 2024
Later among the works it cites.
“Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context”, 2024
Gemini Team et al · 2024
Later among the works it cites.
“Capabilities of gemini models in medicine”
Khaled Saab, Tao Tu, Wei-Hung Weng, Ryutaro Tanno, David Stutz, Ellery Wulczyn, Fan Zhang, Tim Strother, Chunjong Park and Elahe Vedadi · 2024
Later among the works it cites.
“Challenges and recommendations in the implementation of audiovisual telemedicine communication: a systematic review”
Imelda Ritunga, Mora Claramita, Sandra Widaty and Hardyanto Soebono · 2024
Later among the works it cites.
“Towards conversational diagnostic artificial intelligence”
Tao Tu, Mike Schaekermann, Anil Palepu, Khaled Saab, Jan Freyberg, Ryutaro Tanno, Amy Wang, Brenna Li, Mohamed Amin and Yong Cheng · 2025
Closest in time.
“Towards conversational AI for disease management”
Anil Palepu, Valentin Liévin, Wei-Hung Weng, Khaled Saab, David Stutz, Yong Cheng, Kavita Kulkarni, S Mahdavi, Joëlle Barral, Dale Webster, Katherine Chou, Avinatan Hassidim, Yossi Matias, James Manyika, Ryutaro Tanno, Vivek Natarajan, Adam Rodman, Tao Tu, Alan Karthikesalingam and Mike Schaekermann · 2025
Closest in time.
“Comparison of initial artificial intelligence (AI) and final physician recommendations in AI-assisted virtual urgent care visits”
Dan Zeltzer, Zehavi Kugler, Lior Hayat, Tamar Brufman, Ran Ilan, Keren Leibovich, Tom Beer, Ilan Frank, Ran Shaul, Caroline Goldzweig and Joshua Pevnick · 2025
Closest in time.
“Gemini 2.0 Flash | Generative AI on Vertex AI | Google Cloud — cloud.google.com” [Accessed 30-04-2025], https://cloud.google.com/vertex-ai/generative-ai/docs/models/gemini/2-0-flash
2025
Closest in time.
“Conversational Medical AI: Ready for Practice”, 2025
Antoine Lizée, Pierre-Auguste Beaucoté, James Whitbeck, Marion Doumeingts, Anaël Beaugnon and Isabelle Feldhaus · 2025
Closest in time.
“Dialogue is Better Than Monologue: Instructing Medical LLMs via Strategical Conversations”, 2025
Zijie Liu, Xinyu Zhao, Jie Peng, Zhuangdi Zhu, Qingyu Chen, Xia Hu and Tianlong Chen · 2025
Closest in time.
Md Islam, Khondokar Hasan, Hasibul Shajeeb, Humayan Rana, Md Rahmand, Md Hasan, AKM Azad, Ibrahim Abdullah and Mohammad Moni · 2025
Closest in time.
“An evaluation framework for clinical use of large language models in patient interaction tasks”
Shreya Johri, Jaehwan Jeong, Benjamin Tran, Daniel Schlessinger, Shannon Wongvibulsin, Leandra Barnes, Hong-Yu Zhou, Zhuo Cai, Eliezer Van and David Kim · 2025
Closest in time.
“From Texts to Triage: The WhatsApp Clinic Experience - BMJ Global Health blog — blogs.bmj.com”, https://blogs.bmj.com/bmjgh/2025/01/14/from-texts-to-triage-the-whatsapp-clinic-experience/ , 2025
Kamal Sharma, Saranya Manickaraj and Navsangeet Saini · 2025
Closest in time.