Fetching the paper…
Reading the bibliography…
Recent large language models (LLMs) have demonstrated significant advancements, particularly in their ability to serve as agents thereby surpassing their traditional role as chatbots.
Allocation of physician time in ambulatory practice: a time and motion study in 4 specialties
Christine Sinsky, Lacey Colligan, Ling Li, Mirela Prgomet, Sam Reynolds, Lindsey Goeders, Johanna Westbrook, Michael Tutty, and George Blike · 2016
Earlier work this paper cites.
Application of artificial intelligence in the health care safety context: opportunities and challenges
Samer Ellahham, Nour Ellahham, and Mecit Can Emre Simsekler · 2020
Earlier work this paper cites.
A new paradigm for accelerating clinical data science at stanford medicine
Somalee Datta, Jose Posada, Garrick Olson, Wencheng Li, Ciaran O’Reilly, Deepa Balraj, Joseph Mesterhazy, Joseph Pallas, Priyamvada Desai, and Nigam Shah · 2020
Earlier work this paper cites.
Trust and medical ai: the challenges we face and the expertise needed to overcome them
Thomas P Quinn, Manisha Senadeera, Stephan Jacobs, Simon Coghlan, and Vuong Le · 2021
Earlier work this paper cites.
Medmcqa: A large-scale multi-subject multi-choice dataset for medical domain question answering
Ankit Pal, Logesh Kumar Umapathi, and Malaikannan Sankarasubbu · 2022
Earlier work this paper cites.
Artificial intelligence in us health care delivery
Nikhil R Sahni and Brandon Carrus · 2023
Earlier work this paper cites.
Agentbench: Evaluating llms as agents
Xiao Liu, Hao Yu, Hanchen Zhang, Yifan Xu, Xuanyu Lei, Hanyu Lai, Yu Gu, Hangliang Ding, Kaiwen Men, Kejuan Yang, et al · 2023
Earlier work this paper cites.
Gorilla: Large language model connected with massive apis
Shishir G. Patil, Tianjun Zhang, Xin Wang, and Joseph E. Gonzalez · 2023
Earlier work this paper cites.
Survey on large language model-enhanced reinforcement learning: Concept, taxonomy, and methods
Yuji Cao, Huan Zhao, Yuheng Cheng, Ting Shu, Yue Chen, Guolong Liu, Gaoqi Liang, Junhua Zhao, Jinyue Yan, and Yun Li · 2024
Earlier work this paper cites.
Llm-based agentic systems in medicine and healthcare
Jianing Qiu, Kyle Lam, Guohao Li, Amish Acharya, Tien Yin Wong, Ara Darzi, Wu Yuan, and Eric J Topol · 2024
Earlier work this paper cites.
Implications of large language models for quality and efficiency of neurologic care: emerging issues in neurology
Lidia Moura, David T Jones, Irfan S Sheikh, Shawn Murphy, Michael Kalfin, Benjamin R Kummer, Allison L Weathers, Zachary M Grinspan, Heather M Silsbee, Lyell K Jones Jr, et al · 2024
Cited alongside, same era.
Large language models and artificial intelligence: a primer for plastic surgeons on the demonstrated and potential applications, promises, and limitations of chatgpt
Jad Abi-Rafeh, Hong Hao Xu, Roy Kazan, Ruth Tevlin, and Heather Furnas · 2024
Cited alongside, same era.
How artificial intelligence could transform emergency care
Marika M Kachman, Irina Brennan, Jonathan J Oskvarek, Tayab Waseem, and Jesse M Pines · 2024
Cited alongside, same era.
The elastic ehr: A five-tiered framework for applying ai to electronic health record maintenance, configuration, and use
Colby Uptegraft, Kameron C Black, Jonathan Gale, Andrew Marshall, and Shuhan He · 2024
Cited alongside, same era.
Efficient healthcare with large language models: optimizing clinical workflow and enhancing patient care
Craft-md: A conversational evaluation framework for comprehensive assessment of clinical llms
Shreya Johri, Jaehwan Jeong, Benjamin A Tran, Daniel I Schlessinger, Shannon Wongvibulsin, Zhuo Ran Cai, Roxana Daneshjou, and Pranav Rajpurkar · 2024
Later among the works it cites.
Agentclinic: a multimodal agent benchmark to evaluate ai in simulated clinical environments
Samuel Schmidgall, Rojin Ziaei, Carl Harris, Eduardo Reis, Jeffrey Jopling, and Michael Moor · 2024
Later among the works it cites.
Mmedagent: Learning to use medical tools with multi-modal agent
Binxu Li, Tiankai Yan, Yuanting Pan, Jie Luo, Ruiyang Ji, Jiayuan Ding, Zhe Xu, Shilong Liu, Haoyu Dong, Zihao Lin, et al · 2024
Later among the works it cites.
Autobencher: Creating salient, novel, difficult datasets for language models
Xiang Lisa Li, Evan Zheran Liu, Percy Liang, and Tatsunori Hashimoto · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Satvik Tripathi, Rithvik Sukumaran, and Tessa S Cook · 2024
Cited alongside, same era.
Agentboard: An analytical evaluation board of multi-turn llm agents
Chang Ma, Junlei Zhang, Zhihao Zhu, Cheng Yang, Yujiu Yang, Yaohui Jin, Zhenzhong Lan, Lingpeng Kong, and Junxian He · 2024
Cited alongside, same era.
t a u tau -bench: A benchmark for tool-agent-user interaction in real-world domains
Shunyu Yao, Noah Shinn, Pedram Razavi, and Karthik Narasimhan · 2024
Cited alongside, same era.
Cyber insecurity in healthcare: The cost and impact on patient safety and care, 2024
Ponemon Institute · 2024
Cited alongside, same era.
Ethical and regulatory challenges of ai technologies in healthcare: A narrative review
Ciro Mennella, Umberto Maniscalco, Giuseppe De Pietro, and Massimo Esposito · 2024
Cited alongside, same era.
Superhuman performance of a large language model on the reasoning tasks of a physician
Peter G Brodeur, Thomas A Buckley, Zahir Kanjee, Ethan Goh, Evelyn Bin Ling, Priyank Jain, Stephanie Cabral, Raja-Elie Abdulnour, Adrian Haimovich, Jason A Freed, et al · 2024
Cited alongside, same era.
Balancing act: the complex role of artificial intelligence in addressing burnout and healthcare workforce dynamics
Suresh Pavuluri, Rohit Sangal, John Sather, and R Andrew Taylor · 2024
Later among the works it cites.
Many-shot in-context learning in multimodal foundation models
Yixing Jiang, Jeremy Irvin, Ji Hun Wang, Muhammad Ahmed Chaudhry, Jonathan H Chen, and Andrew Y Ng · 2024
Later among the works it cites.
Meta-prompting: Enhancing language models with task-agnostic scaffolding
Mirac Suzgun and Adam Tauman Kalai · 2024
Later among the works it cites.
The rise of agentic ai teammates in medicine
James Zou and Eric J Topol · 2025
Closest in time.
Large language models lack essential metacognition for reliable medical reasoning
Maxime Griot, Coralie Hemptinne, Jean Vanderdonckt, and Demet Yuksel · 2025
Closest in time.