Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) are transforming healthcare through the development of LLM-based agents that can understand, reason about, and assist with medical tasks.
Bertscore: Evaluating text generation with bert
Tianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger, and Yoav Artzi. 2020 · 1904
Earlier work this paper cites.
Health insurance portability and accountability act of 1996
Accountability Act. 1996 · 1996
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Di Jin, Eileen Pan, Nassim Oufattole, Wei-Hung Weng, Hanyi Fang, and Peter Szolovits. 2020 · 2009
Earlier work this paper cites.
General data protection regulation
GDPR GDPR. 2016 · 2016
Earlier work this paper cites.
Artificial intelligence: a modern approach
Stuart J Russell and Peter Norvig. 2016 · 2016
Earlier work this paper cites.
Pubmedqa: A dataset for biomedical research question answering
Qiao Jin, Bhuwan Dhingra, Zhengping Liu, William Cohen, and Xinghua Lu. 2019 · 2019
Earlier work this paper cites.
Extracting training data from large language models
Nicholas Carlini, Florian Tramer, Eric Wallace, Matthew Jagielski, Ariel Herbert-Voss, Katherine Lee, Adam Roberts, Tom Brown, Dawn Song, Ulfar Erlingsson, et al. 2021 · 2021
Earlier work this paper cites.
Transparency of ai in healthcare as a multilayered system of accountabilities: Between legal requirements and technical limitations
Anastasiya Kiseleva, Dimitris Kotzinos, and Paul De Hert. 2022 · 2022
Earlier work this paper cites.
Medmcqa: A large-scale multi-subject multi-choice dataset for medical domain question answering
Ankit Pal, Logesh Kumar Umapathi, and Malaikannan Sankarasubbu. 2022 · 2022
Earlier work this paper cites.
Wordcraft: Story writing with large language models
Ann Yuan, Andy Coenen, Emily Reif, and Daphne Ippolito. 2022 · 2022
Earlier work this paper cites.
Artificial intelligence and automation in endoscopy and surgery
F Chadebecq, L.B. Lovat, and D. Stoyanov. 2023 · 2023
Earlier work this paper cites.
User inference attacks on large language models
Nikhil Kandpal, Krishna Pillutla, Alina Oprea, Peter Kairouz, Christopher A Choquette-Choo, and Zheng Xu. 2023 · 2023
Earlier work this paper cites.
Halueval: A large-scale hallucination evaluation benchmark for large language models
Junyi Li, Xiaoxue Cheng, Wayne Xin Zhao, Jian-Yun Nie, and Ji-Rong Wen. 2023 · 2023
Earlier work this paper cites.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed Chi, Quoc Le, and Denny Zhou. 2023 · 2023
Earlier work this paper cites.
The rise and potential of large language model based agents: A survey
Zhiheng Xi, Wenxiang Chen, Xin Guo, Wei He, Yiwen Ding, Boyang Hong, Ming Zhang, Junzhe Wang, Senjie Jin, Enyu Zhou, Rui Zheng, Xiaoran Fan, Xiao Wang, Limao Xiong, Yuhao Zhou, Weiran Wang, Changhao Jiang, Yicheng Zou, Xiangyang Liu, Zhangyue Yin, Shihan Dou, Rongxiang Weng, Wensen Cheng, Qi Zhang, Wenjuan Qin, Yongyan Zheng, Xipeng Qiu, Xuanjing Huang, and Tao Gui. 2023 · 2023
Earlier work this paper cites.
Magda: Multi-agent guideline-driven diagnostic assistance
David Bani-Harouni, Nassir Navab, and Matthias Keicher. 2024 · 2024
Earlier work this paper cites.
Large language models improve clinical decision making of medical students through patient simulation and structured feedback: a randomized controlled trial
Emilia Brügge, Sarah Ricchizzi, Malin Arenbeck, Marius Niklas Keller, Lina Schur, Walter Stummer, Markus Holling, Max Hao Lu, and Dogus Darici. 2024 · 2024
Earlier work this paper cites.
Exploring large language model based intelligent agents: Definitions, methods, and prospects
Yuheng Cheng, Ceyao Zhang, Zhengwen Zhang, Xiangrui Meng, Sirui Hong, Wenhao Li, Zihao Wang, Zekai Wang, Feng Yin, Junhua Zhao, and Xiuqiang He. 2024 · 2024
Cited alongside, same era.
Exploring llm-based data annotation strategies for medical dialogue preference alignment
Chengfeng Dou, Ying Zhang, Zhi Jin, Wenpin Jiao, Haiyan Zhao, Yongqiang Zhao, and Zhengwei Tao. 2024 · 2024
Cited alongside, same era.
Llms can simulate standardized patients via agent coevolution
Zhuoyun Du, Lujie Zheng, Renjun Hu, Yuyang Xu, Xiawei Li, Ying Sun, Wei Chen, Jian Wu, Haolei Cai, and Haohao Ying. 2024 · 2024
Cited alongside, same era.
Adaptive reasoning and acting in medical language agents
Abhishek Dutta and Yen-Che Hsiao. 2024 · 2024
Cited alongside, same era.
Rx strategist: Prescription verification using llm agents system
Phuc Phan Van, Dat Nguyen Minh, An Dinh Ngoc, and Huy Phan Thanh. 2024 · 2024
Later among the works it cites.
Surgbox: Agent-driven operating room sandbox with surgery copilot
Jinlin Wu, Xusheng Liang, Xuexue Bai, and Zhen Chen. 2024 · 2024
Later among the works it cites.
Towards next-generation medical agent: How o1 is reshaping decision-making in medical scenarios
Shaochen Xu, Yifan Zhou, Zhengliang Liu, Zihao Wu, Tianyang Zhong, Huaqin Zhao, Yiwei Li, Hanqi Jiang, Yi Pan, Junhao Chen, Jin Lu, Wei Zhang, Tuo Zhang, Lu Zhang, Dajiang Zhu, Xiang Li, Wei Liu, Quanzheng Li, Andrea Sikora, Xiaoming Zhai, Zhen Xiang, and Tianming Liu. 2024 · 2024
Later among the works it cites.
Clinicallab: Aligning agents for multi-departmental clinical diagnostics in the real world
Weixiang Yan, Haitian Liu, Tengxiao Wu, Qian Chen, Wen Wang, Haoyuan Chai, Jiayi Wang, Weishan Zhao, Yixin Zhang, Renjun Zhang, Li Zhu, and Xuandong Zhao. 2024 · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zhihao Fan, Jialong Tang, Wei Chen, Siyuan Wang, Zhongyu Wei, Jun Xi, Fei Huang, and Jingren Zhou. 2024 · 2024
Cited alongside, same era.
Long term memory: The foundation of ai self-evolution
Xun Jiang, Feng Li, Han Zhao, Jiaying Wang, Jun Shao, Shihao Xu, Shu Zhang, Weiling Chen, Xavier Tang, Yize Chen, Mengyue Wu, Weizhi Ma, Mengdi Wang, and Tianqiao Chen. 2024 · 2024
Cited alongside, same era.
Mitigating cognitive biases in clinical decision-making through multi-agent conversations using large language models: Simulation study
Yuhe Ke, Rui Yang, Sui An Lie, Taylor Xin Yi Lim, Yilin Ning, Irene Li, Hairil Rizal Abdullah, Daniel Shu Wei Ting, and Nan Liu. 2024 · 2024
Cited alongside, same era.
Roles, users, benefits, and limitations of chatbots in health care: Rapid review
M Laymouna, Y Ma, D Lessard, T Schuster, K Engler, and B Lebouché. 2024 · 2024
Cited alongside, same era.
Improving clinical documentation with ai: A comparative study of sporo ai scribe and gpt-4o mini
Chanseo Lee, Sonu Kumar, Kimon A. Vogt, and Sam Meraj. 2024 · 2024
Cited alongside, same era.
Evaluating large language models as agents in the clinic
Nikita Mehandru, Brenda Y Miao, Eduardo Rodriguez Almaraz, Madhumita Sushil, Atul Janardhan Butte, and Ahmed Alaa. 2024 · 2024
Cited alongside, same era.
Polaris: A safety-focused llm constellation architecture for healthcare
Subhabrata Mukherjee, Paul Gamble, Markel Sanz Ausin, et al. 2024 · 2024
Cited alongside, same era.
Llm-based agentic systems in medicine and healthcare
Jianing Qiu, Kyle Lam, Guohao Li, Amish Acharya, Tien Yin Wong, Ara Darzi, Wu Yuan, and Eric J. Topol. 2024 · 2024
Cited alongside, same era.
Later among the works it cites.
Aipatient: Simulating patients with ehrs and llm powered agentic workflow
Huizi Yu, Jiayan Zhou, Lingyao Li, et al. 2024 · 2024
Later among the works it cites.
R-judge: Benchmarking safety risk awareness for LLM agents
Tongxin Yuan, Zhiwei He, Lingzhong Dong, Yiming Wang, Ruijie Zhao, Tian Xia, Lizhen Xu, Binglin Zhou, Fangqi Li, Zhuosheng Zhang, Rui Wang, and Gongshen Liu. 2024 · 2024
Later among the works it cites.
Clinicalagent: Clinical trial multi-agent system with large language model-based reasoning
Ling Yue, Sixue Xing, Jintai Chen, and Tianfan Fu. 2024 · 2024
Later among the works it cites.
Exploring the potential of large language models in radiological imaging systems: Improving user interface design and functional capabilities
Luyao Zhang, Jianhua Shu, Jili Hu, Fangfang Li, Junjun He, Peng Wang, and Yiqing Shen. 2024 · 2024
Later among the works it cites.
Menti: Bridging medical calculator and llm agent with nested tool calling
Yakun Zhu, Shaohang Wei, Xu Wang, Kui Xue, Xiaofan Zhang, and Shaoting Zhang. 2024 · 2024
Later among the works it cites.
Medhallbench: A new benchmark for assessing hallucination in medical large language models
Kaiwen Zuo and Yirui Jiang. 2024 · 2024
Later among the works it cites.
Empowering LLM agents with zero-shot optimal decision-making through q-learning
Jiajun Chai, Sicheng Li, Yuqian Fu, Dongbin Zhao, and Yuanheng Zhu. 2025 · 2025
Closest in time.
How does deepseek-r1 perform on usmle?
Lisle Faray de Paiva, Gijs Luijten, Behrus Puladi, and Jan Egger. 2025 · 2025
Closest in time.
Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning
Daya Guo, Dejian Yang, Haowei Zhang, Junxiao Song, Ruoyu Zhang, Runxin Xu, Qihao Zhu, Shirong Ma, Peiyi Wang, Xiao Bi, et al. 2025 · 2025
Closest in time.
O1 replication journey – part 3: Inference-time scaling for medical reasoning
Zhongzhen Huang, Gui Geng, Shengyi Hua, Zhen Huang, Haoyang Zou, Shaoting Zhang, Pengfei Liu, and Xiaofan Zhang. 2025 · 2025
Closest in time.
Unlocking efficiency and cost savings in healthcare: How swarms of llm agents can revolutionize medical operations and save millions
Swarms. 2025 · 2025
Closest in time.
The application of large language models in primary healthcare services and the challenges
Wenxin YAN, Jian HU, Huatang ZENG, Min LIU, and Wannian LIANG. 2025 · 2025
Closest in time.
A dual-agent collaboration framework based on llms for nursing robots to perform bimanual coordination tasks
Zhendong Zhao, Xiaotian Yue, Jiexin Xie, Chuanhong Fang, Zhenzhou Shao, and Shijie Guo. 2025 · 2025
Closest in time.
Kaiwen Zuo, Yirui Jiang, Fan Mo, and Pietro Lio. 2025 · 2025
Closest in time.