Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) hold great promise to revolutionize current clinical systems for their superior capacities on medical text processing tasks and medical licensing exams.
Mimic-iii, a freely accessible critical care database
Alistair EW Johnson, Tom J Pollard, Lu Shen, Li-wei H Lehman, Mengling Feng, Mohammad Ghassemi, Benjamin Moody, Peter Szolovits, Leo Anthony Celi, and Roger G Mark · 2016
Earlier work this paper cites.
Predictive models for hospital readmission risk: A systematic review of methods
Arkaitz Artetxe, Andoni Beristain, and Manuel Grana · 2018
Earlier work this paper cites.
Benchmarking deep learning models on large healthcare datasets
Sanjay Purushotham, Chuizheng Meng, Zhengping Che, and Yan Liu · 2018
Earlier work this paper cites.
Machine learning in medicine
Alvin Rajkomar, Jeffrey Dean, and Isaac Kohane · 2019
Earlier work this paper cites.
Mimic-iv
Alistair Johnson, Lucas Bulgarelli, Tom Pollard, Steven Horng, Leo Anthony Celi, and Roger Mark · 2020
Earlier work this paper cites.
Democratizing ehr analyses with fiddle: a flexible data-driven preprocessing pipeline for structured clinical data
Shengpu Tang, Parmida Davarmanesh, Yanmeng Song, Danai Koutra, Michael W Sjoding, and Jenna Wiens · 2020
Earlier work this paper cites.
Mimic-extract: A data extraction, preprocessing, and representation pipeline for mimic-iii
Shirly Wang, Matthew BA McDermott, Geeticka Chauhan, Marzyeh Ghassemi, Michael C Hughes, and Tristan Naumann · 2020
Earlier work this paper cites.
Measuring massive multitask language understanding
Dan Hendrycks, Collin Burns, Steven Basart, Andy Zou, Mantas Mazeika, Dawn Song, and Jacob Steinhardt · 2021
Earlier work this paper cites.
Mortality risk stratification using artificial intelligence-augmented electrocardiogram in cardiac intensive care unit patients
Jacob C Jentzer, Anthony H Kashou, Francisco Lopez-Jimenez, Zachi I Attia, Suraj Kapa, Paul A Friedman, and Peter A Noseworthy · 2021
Earlier work this paper cites.
What disease does this patient have? a large-scale open domain question answering dataset from medical exams
Di Jin, Eileen Pan, Nassim Oufattole, Wei-Hung Weng, Hanyi Fang, and Peter Szolovits · 2021
Earlier work this paper cites.
Lora: Low-rank adaptation of large language models
Edward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen · 2022
Earlier work this paper cites.
Large language models are zero-shot reasoners
Takeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo, and Yusuke Iwasawa · 2022
Earlier work this paper cites.
Medmcqa: A large-scale multi-subject multi-choice dataset for medical domain question answering
Ankit Pal, Logesh Kumar Umapathi, and Malaikannan Sankarasubbu · 2022
Earlier work this paper cites.
A systematic review of the prediction of hospital length of stay: Towards a unified framework
Kieran Stone, Reyer Zwiggelaar, Phil Jones, and Neil Mac Parthaláin · 2022
Earlier work this paper cites.
Twin: Personalized clinical trial digital twin generation
Trisha Das, Zifeng Wang, and Jimeng Sun · 2023
Earlier work this paper cites.
A survey on in-context learning
Qingxiu Dong, Lei Li, Damai Dai, Ce Zheng, Zhiyong Wu, Baobao Chang, Xu Sun, Jingjing Xu, and Zhifang Sui · 2023
Earlier work this paper cites.
How does chatgpt perform on the united states medical licensing examination (usmle)? the implications of large language models for medical education and knowledge assessment
Aidan Gilson, Conrad W Safranek, Thomas Huang, Vimig Socrates, Ling Chi, Richard Andrew Taylor, David Chartash, et al · 2023
Earlier work this paper cites.
Albert Q Jiang, Alexandre Sablayrolles, Arthur Mensch, Chris Bamford, Devendra Singh Chaplot, Diego de las Casas, Florian Bressand, Gianna Lengyel, Guillaume Lample, Lucile Saulnier, et al · 2023
Earlier work this paper cites.
Accuracy of a generative artificial intelligence model in a complex diagnostic challenge
Zahir Kanjee, Byron Crowe, and Adam Rodman · 2023
Earlier work this paper cites.
Evaluating gpt-4 and chatgpt on japanese medical licensing examinations
Jungo Kasai, Yuhei Kasai, Keisuke Sakaguchi, Yutaro Yamada, and Dragomir Radev · 2023
Earlier work this paper cites.
Karolina Korgul, Andrew M Bean, Felix Krones, Robert McCraith, and Adam Mahdi · 2023
Earlier work this paper cites.
Biomedgpt: Open multimodal generative pre-trained transformer for biomedicine
Yizhen Luo, Jiahuan Zhang, Siqi Fan, Kai Yang, Yushuai Wu, Mu Qiao, and Zaiqing Nie · 2023
Earlier work this paper cites.
Towards accurate differential diagnosis with large language models
Daniel McDuff, Mike Schaekermann, Tao Tu, Anil Palepu, Amy Wang, Jake Garrison, Karan Singhal, Yash Sharma, Shekoofeh Azizi, Kavita Kulkarni, et al · 2023
Cited alongside, same era.
Artificial intelligence for clinical decision support for monitoring patients in cardiovascular icus: a systematic review
Sobhan Moazemi, Sahar Vahdati, Jason Li, Sebastian Kalkhoff, Luis JV Castano, Bastian Dewitz, Roman Bibo, Parisa Sabouniaghdam, Mohammad S Tootooni, Ralph A Bundschuh, et al · 2023
Cited alongside, same era.
Liangming Pan, Michael Saxon, Wenda Xu, Deepak Nathani, Xinyi Wang, and William Yang Wang · 2023
Cited alongside, same era.
Evaluation of the performance of gpt-3.5 and gpt-4 on the polish medical final examination
Maciej Rosoł, Jakub S Gąsior, Jonasz Łaba, Kacper Korzeniewski, and Marcel Młyńczak · 2023
Cited alongside, same era.
Evaluating large language models for public health classification and extraction tasks
Joshua Harris, Timothy Laurence, Leo Loman, Fan Grayson, Toby Nonnenmacher, Harry Long, Loes WalsGriffith, Amy Douglas, Holly Fountain, Stelios Georgiou, et al · 2024
Closest in time.
Minicpm: Unveiling the potential of small language models with scalable training strategies
Shengding Hu, Yuge Tu, Xu Han, Chaoqun He, Ganqu Cui, Xiang Long, Zhi Zheng, Yewei Fang, Yuxiang Huang, Weilin Zhao, et al · 2024
Closest in time.
A comprehensive evaluation of large language models on benchmark biomedical text processing tasks
Israt Jahan, Md Tahmid Rahman Laskar, Chun Peng, and Jimmy Xiangji Huang · 2024
Closest in time.
Graphcare: Enhancing healthcare predictions with personalized knowledge graphs
Pengcheng Jiang, Cao Xiao, Adam Richard Cross, and Jimeng Sun · 2024
Closest in time.
Hidden flaws behind expert-level accuracy of gpt-4 vision in medicine
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Digital twin in healthcare: Recent updates and challenges
Tianze Sun, Xiwang He, and Zhonghai Li · 2023
Cited alongside, same era.
Hypergraph transformers for ehr-based clinical predictions
Ran Xu, Mohammed K Ali, Joyce C Ho, and Carl Yang · 2023
Cited alongside, same era.
PyHealth: A deep learning toolkit for healthcare predictive modeling
Chaoqi Yang, Zhenbang Wu, Patrick Jiang, Zhen Lin, Junyi Gao, Benjamin Danek, and Jimeng Sun · 2023
Cited alongside, same era.
Instruction tuning for large language models: A survey
Shengyu Zhang, Linfeng Dong, Xiaoya Li, Sen Zhang, Xiaofei Sun, Shuhe Wang, Jiwei Li, Runyi Hu, Tianwei Zhang, Fei Wu, et al · 2023
Cited alongside, same era.
Judging llm-as-a-judge with mt-bench and chatbot arena
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric Xing, et al · 2023
Cited alongside, same era.
Agieval: A human-centric benchmark for evaluating foundation models
Wanjun Zhong, Ruixiang Cui, Yiduo Guo, Yaobo Liang, Shuai Lu, Yanlin Wang, Amin Saied, Weizhu Chen, and Nan Duan · 2023
Cited alongside, same era.
A survey of large language models in medicine: Progress, application, and challenge
Hongjian Zhou, Boyang Gu, Xinyu Zou, Yiru Li, Sam S Chen, Peilin Zhou, Junling Liu, Yining Hua, Chengfeng Mao, Xian Wu, et al · 2023
Cited alongside, same era.
Phi-3 technical report: A highly capable language model locally on your phone
Marah Abdin, Sam Ade Jacobs, Ammar Ahmad Awan, Jyoti Aneja, Ahmed Awadallah, Hany Awadalla, Nguyen Bach, Amit Bahree, Arash Bakhtiari, Harkirat Behl, et al · 2024
Cited alongside, same era.
Qiao Jin, Fangyuan Chen, Yiliang Zhou, Ziyang Xu, Justin M Cheung, Robert Chen, Ronald M Summers, Justin F Rousseau, Peiyun Ni, Marc J Landsman, et al · 2024
Closest in time.
Digital twins for health: a scoping review
Evangelia Katsoulakis, Qi Wang, Huanmei Wu, Leili Shahriyari, Richard Fletcher, Jinwei Liu, Luke Achenie, Hongfang Liu, Pamela Jackson, Ying Xiao, et al · 2024
Closest in time.
Sunjun Kweon, Byungjin Choi, Minkyu Kim, Rae Woong Park, and Edward Choi · 2024
Closest in time.
Biomistral: A collection of open-source pretrained large language models for medical domains
Yanis Labrak, Adrien Bazoge, Emmanuel Morin, Pierre-Antoine Gourraud, Mickael Rouvier, and Richard Dufour · 2024
Closest in time.
Large language model instruction following: A survey of progresses and challenges
Renze Lou, Kai Zhang, and Wenpeng Yin · 2024
Closest in time.
Pre-trained language models in medicine: A survey
Xudong Luo, Zhiqi Deng, Binxia Yang, and Michael Y Luo · 2024
Closest in time.
To cool or not to cool? temperature network meets large foundation models via dro
Zi-Hao Qiu, Siqi Guo, Mao Xu, Tuo Zhao, Lijun Zhang, and Tianbao Yang · 2024
Closest in time.
The effect of sampling temperature on problem solving in large language models
Matthew Renze and Erhan Guven · 2024
Closest in time.
A systematic survey of prompt engineering in large language models: Techniques and applications
Pranab Sahoo, Ayush Kumar Singh, Sriparna Saha, Vinija Jain, Samrat Mondal, and Aman Chadha · 2024
Closest in time.
The prompt report: A systematic survey of prompting techniques
Sander Schulhoff, Michael Ilie, Nishant Balepur, Konstantine Kahadze, Amanda Liu, Chenglei Si, Yinheng Li, Aayush Gupta, HyoJung Han, Sevien Schulhoff, et al · 2024
Closest in time.
Medconceptsqa–open source medical concepts qa benchmark
Ofir Ben Shoham and Nadav Rappoport · 2024
Closest in time.
Large language models for data annotation: A survey, 2024
Zhen Tan, Alimohammad Beigi, Song Wang, Ruocheng Guo, Amrita Bhattacharjee, Bohan Jiang, Mansooreh Karami, Jundong Li, Lu Cheng, and Huan Liu · 2024
Closest in time.
Gemma 2: Improving open language models at a practical size
Gemma Team, Morgane Riviere, Shreya Pathak, Pier Giuseppe Sessa, Cassidy Hardin, Surya Bhupatiraju, Léonard Hussenot, Thomas Mesnard, Bobak Shahriari, Alexandre Ramé, et al · 2024
Closest in time.
Yet another ICU benchmark: A flexible multi-center framework for clinical ML
Robin van de Water, Hendrik Nils Aurel Schmidt, Paul Elbers, Patrick Thoral, Bert Arnrich, and Patrick Rockenschaub · 2024
Closest in time.
Medreqal: Examining medical knowledge recall of large language models via question answering
Juraj Vladika, Phillip Schneider, and Florian Matthes · 2024
Closest in time.
An Yang, Baosong Yang, Binyuan Hui, Bo Zheng, Bowen Yu, Chang Zhou, Chengpeng Li, Chengyuan Li, Dayiheng Liu, Fei Huang, et al · 2024
Closest in time.
Yi: Open foundation models by 01. ai
Alex Young, Bei Chen, Chao Li, Chengen Huang, Ge Zhang, Guanwei Zhang, Heng Li, Jiangcheng Zhu, Jianqun Chen, Jing Chang, et al · 2024
Closest in time.
Large language models for disease diagnosis: A scoping review
Shuang Zhou, Zidu Xu, Mian Zhang, Chunpu Xu, Yawen Guo, Zaifu Zhan, Sirui Ding, Jiashuo Wang, Kaishuai Xu, Yi Fang, et al · 2024
Closest in time.