Fetching the paper…
Reading the bibliography…
Recent advancements in artificial intelligence (AI), particularly in large language models (LLMs), have unlocked significant potential to enhance the quality and efficiency of medical care.
Archives of internal medicine 162 , 1897–1903
Barker, K.N., Flynn, E.A., Pepper, G.A., Bates, D.W., and Mikeal, R.L. (2002). Medication errors observed in 36 health care facilities · 1903
Earlier work this paper cites.
arXiv preprint arXiv:1910.04597
Glocker, B., Robinson, R., Castro, D.C., Dou, Q., and Konukoglu, E. (2019). Machine learning with multi-site imaging data: An empirical study on the impact of scanner effects · 1910
Earlier work this paper cites.
Science 237 , 1317–1323
Shepard, R.N. (1987). Toward a universal law of generalization for psychological science · 1987
Earlier work this paper cites.
Annals of internal medicine 162 , 55–63
Collins, G.S., Reitsma, J.B., Altman, D.G., and Moons, K.G. (2015). Transparent reporting of a multivariable prediction model for individual prognosis or diagnosis (tripod): the tripod statement · 2015
Earlier work this paper cites.
Journal of community hospital internal medicine perspectives 6 , 31758
Da Silva, B.A., and Krishnamurthy, M. (2016). The alarming reality of medication error: a patient case and review of pennsylvania and national data · 2016
Earlier work this paper cites.
AI alignment [Internet] 19
Christiano, P. (2016). Prosaic ai alignment · 2016
Earlier work this paper cites.
Kidney research and clinical practice 36 , 3
Lee, C.H., and Yoon, H.J. (2017). Medical big data: promise and challenges · 2017
Earlier work this paper cites.
Acm Sigcas Computers and Society 47 , 54–64
Wolf, M.J., Miller, K., and Grodzinsky, F.S. (2017). Why we should have seen that coming: comments on microsoft’s tay” experiment,” and wider implications · 2017
Earlier work this paper cites.
JAMA internal medicine 178 , 1544–1547
Gianfrancesco, M.A., Tamang, S., Yazdany, J., and Schmajuk, G. (2018). Potential biases in machine learning algorithms using electronic health record data · 2018
Earlier work this paper cites.
Nature medicine 24 , 1342–1350
De Fauw, J., Ledsam, J.R., Romera-Paredes, B., Nikolov, S., Tomasev, N., Blackwell, S., Askham, H., Glorot, X., O’Donoghue, B., Visentin, D. et al. (2018). Clinically applicable deep learning for diagnosis and referral in retinal disease · 2018
Earlier work this paper cites.
Annals of internal medicine 169 , 866–872
Rajkomar, A., Hardt, M., Howell, M.D., Corrado, G., and Chin, M.H. (2018). Ensuring fairness in machine learning to advance health equity · 2018
Earlier work this paper cites.
Nature biomedical engineering 2 , 719–731
Yu, K.H., Beam, A.L., and Kohane, I.S. (2018). Artificial intelligence in healthcare · 2018
Earlier work this paper cites.
In ISPIM Innovation Symposium. The International Society for Professional Innovation Management (ISPIM) pp. 1–15
Salla, E., Pikkarainen, M., Leväsluoto, J., Blackbright, H., and Johansson, P.E. (2018). Ai innovations and their impact on healthcare and medical expertise · 2018
Earlier work this paper cites.
New England Journal of Medicine 380 , 1347–1358
Rajkomar, A., Dean, J., and Kohane, I. (2019). Machine learning in medicine · 2019
Earlier work this paper cites.
BMJ quality & safety 28 , 238–241
Yu, K.H., and Kohane, I.S. (2019). Framing the challenges of artificial intelligence in medicine · 2019
Earlier work this paper cites.
Journal of biomedical informatics 94 , 103188
Saripalle, R., Runyan, C., and Russell, M. (2019). Using hl7 fhir to achieve interoperability in patient health record · 2019
Earlier work this paper cites.
In AMIA Annual Symposium Proceedings vol. 2019. American Medical Informatics Association pp. 592
Liu, D., Sahu, R., Ignatov, V., Gottlieb, D., and Mandl, K.D. (2019). High performance computing on flat fhir files created with the new smart/hl7 bulk data access standard · 2019
Earlier work this paper cites.
Nature medicine 25 , 37–43
Price, W.N., and Cohen, I.G. (2019). Privacy in the age of medical big data · 2019
Earlier work this paper cites.
Science 366 , 447–453
Obermeyer, Z., Powers, B., Vogeli, C., and Mullainathan, S. (2019). Dissecting racial bias in an algorithm used to manage the health of populations · 2019
Earlier work this paper cites.
Nature biomedical engineering 3 , 173–182
Lee, H., Yune, S., Mansouri, M., Kim, M., Tajmir, S.H., Guerrier, C.E., Ebert, S.A., Pomerantz, S.R., Romero, J.M., Kamalian, S. et al. (2019). An explainable deep-learning algorithm for the detection of acute intracranial haemorrhage from small datasets · 2019
Earlier work this paper cites.
JAMA network open 3 , e1919396–e1919396
Tiwari, P., Colborn, K.L., Smith, D.E., Xing, F., Ghosh, D., and Rosenberg, M.A. (2020). Assessment of a machine learning model applied to harmonized electronic health record data for the prediction of incident atrial fibrillation · 2020
Earlier work this paper cites.
Nature Machine Intelligence 2 , 305–311
Kaissis, G.A., Makowski, M.R., Rückert, D., and Braren, R.F. (2020). Secure, privacy-preserving and federated machine learning in medical imaging · 2020
Earlier work this paper cites.
JAMIA open 3 , 146–150
Laparra, E., Bethard, S., and Miller, T.A. (2020). Rethinking domain adaptation for machine learning over clinical language · 2020
Earlier work this paper cites.
Politique Etrangere pp. 202–203
Noël, J.C. (2020). Human compatible: Ai and the problem of control · 2020
Earlier work this paper cites.
In ICASSP 2021-2021 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). IEEE pp. 5879–5883
Masumura, R., Makishima, N., Ihori, M., Takashima, A., Tanaka, T., and Orihashi, S. (2021). Hierarchical transformer-based large-context end-to-end asr with large-context knowledge distillation · 2021
Earlier work this paper cites.
Applied Sciences 11 , 8275
Kumar, G., Basri, S., Imam, A.A., Khowaja, S.A., Capretz, L.F., and Balogun, A.O. (2021). Data harmonization for heterogeneous datasets: A systematic literature review · 2021
Earlier work this paper cites.
Physica Medica 83 , 108–121
Papadimitroulas, P., Brocki, L., Chung, N.C., Marchadour, W., Vermet, F., Gaubert, L., Eleftheriadis, V., Plachouris, D., Visvikis, D., Kagadis, G.C. et al. (2021). Artificial intelligence: Deep learning in oncological radiomics and challenges of interpretability and data harmonization · 2021
Earlier work this paper cites.
In AMIA Annual Symposium Proceedings vol. 2021. American Medical Informatics Association pp. 1099
Selim, M., Zhang, J., Fei, B., Zhang, G.Q., Ge, G.Y., and Chen, J. (2021). Cross-vendor ct image data harmonization using cvh-ct · 2021
Earlier work this paper cites.
Journal of Biomedical Informatics 117 , 103735
Cui, J., Zhu, H., Deng, H., Chen, Z., and Liu, D. (2021). Fearh: Federated machine learning with anonymous random hybridization on electronic medical records · 2021
Earlier work this paper cites.
Communications medicine 1 , 25
Vokinger, K.N., Feuerriegel, S., and Kesselheim, A.S. (2021). Mitigating bias in machine learning for medicine · 2021
Earlier work this paper cites.
Addressing fairness, bias, and appropriate use of artificial intelligence and machine learning in global health. Frontiers Media SA
Fletcher, R.R., Nakeshimana, A., and Olubeko, O. (2021) · 2021
Earlier work this paper cites.
Proposal for a regulation of the european parliament and of the council laying down harmonised rules on artificial intelligence (artificial intelligence act) and amending certain union legislative acts.
DOWN, L., and ACT, I. (2021) · 2021
Earlier work this paper cites.
NPJ Digital Medicine 4 , 4
Kompa, B., Snoek, J., and Beam, A.L. (2021). Second opinion needed: communicating uncertainty in medical machine learning · 2021
Earlier work this paper cites.
BMC Medical Ethics 22 , 1–5
Murdoch, B. (2021). Privacy and artificial intelligence: challenges for protecting health information in a new era · 2021
Earlier work this paper cites.
Computers in biology and medicine 129 , 104130
Thapa, C., and Camtepe, S. (2021). Precision health data: Requirements, challenges and existing techniques for data security and privacy · 2021
Earlier work this paper cites.
Physica Medica 83 , 72–78
Balagurunathan, Y., Mitchell, R., and El Naqa, I. (2021). Requirements and reliability of ai in the medical context · 2021
Cited alongside, same era.
Radiation 1 , 261–276
Pesapane, F., Bracchi, D.A., Mulligan, J.F., Linnikov, A., Maslennikov, O., Lanzavecchia, M.B., Tantrige, P., Stasolla, A., Biondetti, P., Giuggioli, P.F. et al. (2021). Legal and regulatory framework for ai solutions in healthcare in eu, us, china, and russia: new scenarios after a pandemic · 2021
Cited alongside, same era.
Journal of the American Medical Informatics Association 28 , 890–894
Quinn, T.P., Senadeera, M., Jacobs, S., Coghlan, S., and Le, V. (2021). Trust and medical ai: the challenges we face and the expertise needed to overcome them · 2021
Cited alongside, same era.
arXiv preprint arXiv:2205.12689
Agrawal, M., Hegselmann, S., Lang, H., Kim, Y., and Sontag, D. (2022). Large language models are few-shot clinical information extractors · 2022
Cited alongside, same era.
Advances in Neural Information Processing Systems 35 , 38274–38290
Tirumala, K., Markosyan, A., Zettlemoyer, L., and Aghajanyan, A. (2022). Memorization without overfitting: Analyzing the training dynamics of large language models · 2022
Cell Reports Medicine 4
Sanchez, M., Alford, K., Krishna, V., Huynh, T.M., Nguyen, C.D., Lungren, M.P., Truong, S.Q., and Rajpurkar, P. (2023). Ai-clinician collaboration via disagreement prediction: A decision pipeline and retrospective analysis of real-world radiologist-ai interactions · 2023
Later among the works it cites.
Nature Biomedical Engineering pp. 1–10
Chen, E., Prakash, S., Janapa Reddi, V., Kim, D., and Rajpurkar, P. (2023). A framework for integrating artificial intelligence for clinical care with continuous therapeutic monitoring · 2023
Later among the works it cites.
npj Digital Medicine 6 , 196
Barnett, M., Wang, D., Beadnall, H., Bischof, A., Brunacci, D., Butzkueven, H., Brown, J.W.L., Cabezas, M., Das, T., Dugal, T. et al. (2023). A real-world clinical validation for ai-based mri monitoring in multiple sclerosis · 2023
Later among the works it cites.
In Workshop on Clinical Image-Based Procedures. Springer pp. 132–141
Ricci Lara, M.A., Mosquera, C., Ferrante, E., and Echeveste, R. (2023). Towards unraveling calibration biases in medical image analysis · 2023
Later among the works it cites.
IEEE Transactions on Artificial Intelligence 4 , 383–397
Karimi, D., and Gholipour, A. (2023). Improving calibration and out-of-distribution detection in deep models for medical image segmentation · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
International Journal of Human-Computer Studies 160 , 102772
Pandey, R., Purohit, H., Castillo, C., and Shalin, V.L. (2022). Modeling and mitigating human annotation errors to design efficient stream processing systems with human-in-the-loop machine learning · 2022
Cited alongside, same era.
Computers in biology and medicine 141 , 105144
Li, C., Lin, X., Mao, Y., Lin, W., Qi, Q., Ding, X., Huang, Y., Liang, D., and Yu, Y. (2022). Domain generalization on medical imaging classification using episodic training with task augmentation · 2022
Cited alongside, same era.
Computers in Human Behavior 133 , 107296
Formosa, P., Rogers, W., Griep, Y., Bankins, S., and Richards, D. (2022). Medical ai and human dignity: Contrasting perceptions of human and artificially intelligent (ai) decision making in diagnostic and medical resource allocation contexts · 2022
Cited alongside, same era.
In Seminars in Nuclear Medicine vol. 52. Elsevier pp. 498–503
Currie, G., and Rohren, E. (2022). Social asymmetry, artificial intelligence and the medical imaging landscape · 2022
Cited alongside, same era.
PLOS Digital Health 1 , e0000022
Celi, L.A., Cellini, J., Charpignon, M.L., Dee, E.C., Dernoncourt, F., Eber, R., Mitchell, W.G., Moukheiber, L., Schirmer, J., Situ, J. et al. (2022). Sources of bias in artificial intelligence that perpetuate healthcare disparities—a global review · 2022
Cited alongside, same era.
Research Square
Van Veen, D., Van Uden, C., Blankemeier, L., Delbrouck, J.B., Aali, A., Bluethgen, C., Pareek, A., Polacin, M., Reis, E.P., Seehofnerova, A. et al. (2023). Clinical text summarization: adapting large language models can outperform human experts · 2023
Cited alongside, same era.
Nature 620 , 172–180
Singhal, K., Azizi, S., Tu, T., Mahdavi, S.S., Wei, J., Chung, H.W., Scales, N., Tanwani, A., Cole-Lewis, H., Pfohl, S. et al. (2023). Large language models encode clinical knowledge · 2023
Cited alongside, same era.
Later among the works it cites.
BMJ global health 8 , e010435
Federspiel, F., Mitchell, R., Asokan, A., Umana, C., and McCoy, D. (2023). Threats by artificial intelligence to human health and human existence · 2023
Later among the works it cites.
Artificial intelligence and machine learning (ai/ml)-enabled medical devices.
Food, U., and Administration, D. (2024) · 2024
Closest in time.
60% of americans would be uncomfortable with provider relying on ai in their own health care.
Center, P.R. (2023) · 2024
Closest in time.
Nature pp. 1–3
Abramson, J., Adler, J., Dunger, J., Evans, R., Green, T., Pritzel, A., Ronneberger, O., Willmore, L., Ballard, A.J., Bambrick, J. et al. (2024). Accurate structure prediction of biomolecular interactions with alphafold 3 · 2024
Closest in time.
arXiv preprint arXiv:2412.12767
Xie, L., Liu, H., Zeng, J., Tang, X., Han, Y., Luo, C., Huang, J., Li, Z., Wang, S., and He, Q. (2024). A survey of calibration process for black-box llms · 2024
Closest in time.
arXiv preprint arXiv:2405.11613
Bi, B., Liu, S., Mei, L., Wang, Y., Ji, P., and Cheng, X. (2024). Decoding by contrasting knowledge: Enhancing llms’ confidence on edited facts · 2024
Closest in time.
arXiv preprint arXiv:2406.08391
Kapoor, S., Gruver, N., Roberts, M., Collins, K., Pal, A., Bhatt, U., Weller, A., Dooley, S., Goldblum, M., and Wilson, A.G. (2024). Large language models must be taught to know what they don’t know · 2024
Closest in time.
arXiv preprint arXiv:2402.12563
Li, L., Chen, Z., Chen, G., Zhang, Y., Su, Y., Xing, E., and Zhang, K. (2024). Confidence matters: Revisiting intrinsic self-correction capabilities of large language models · 2024
Closest in time.
arXiv preprint arXiv:2403.05973
Ulmer, D., Gubri, M., Lee, H., Yun, S., and Oh, S.J. (2024). Calibrating large language models using their generations only · 2024
Closest in time.
arXiv preprint arXiv:2403.08819
Shen, M., Das, S., Greenewald, K., Sattigeri, P., Wornell, G., and Ghosh, S. (2024). Thermometer: Towards universal calibration for large language models · 2024
Closest in time.
arXiv preprint arXiv:2402.00367
Feng, S., Shi, W., Wang, Y., Ding, W., Balachandran, V., and Tsvetkov, Y. (2024). Don’t hallucinate, abstain: Identifying llm knowledge gaps via multi-llm collaboration · 2024
Closest in time.
Data drift in llms—causes, challenges, and strategies.
Nexla (2024) · 2024
Closest in time.
In Findings of the Association for Computational Linguistics: EMNLP 2024. pp. 8014–8029
Mousavi, S.M., Alghisi, S., and Riccardi, G. (2024). Dyknow: dynamically verifying time-sensitive factual knowledge in llms · 2024
Closest in time.
arXiv preprint arXiv:2407.04108
Price, S., Panickssery, A., Bowman, S., and Stickland, A.C. (2024). Future events as backdoor triggers: Investigating temporal vulnerabilities in llms · 2024
Closest in time.
arXiv preprint arXiv:2405.08460
Zhu, C., Chen, N., Gao, Y., and Wang, B. (2024). Is your llm outdated? evaluating llms at temporal generalization · 2024
Closest in time.
arXiv preprint arXiv:2402.08113
Schmidgall, S., Harris, C., Essien, I., Olshvang, D., Rahman, T., Kim, J.W., Ziaei, R., Eshraghian, J., Abadir, P., and Chellappa, R. (2024). Addressing cognitive bias in medical language models · 2024
Closest in time.
arXiv preprint arXiv:2409.08087
Peng, B., Chen, K., Li, M., Feng, P., Bi, Z., Liu, J., and Niu, Q. (2024). Securing large language models: Addressing bias, misinformation, and prompt attacks · 2024
Closest in time.
arXiv preprint arXiv:2404.15149
Poulain, R., Fayyaz, H., and Beheshti, R. (2024). Bias patterns in the application of llms for clinical decision support: A comprehensive study · 2024
Closest in time.
The Lancet Digital Health 6 , e12–e22
Zack, T., Lehman, E., Suzgun, M., Rodriguez, J.A., Celi, L.A., Gichoya, J., Jurafsky, D., Szolovits, P., Bates, D.W., Abdulnour, R.E.E. et al. (2024). Assessing the potential of gpt-4 to perpetuate racial and gender biases in health care: a model evaluation study · 2024
Closest in time.
The Lancet Digital Health 6 , e428–e432
Ong, J.C.L., Chang, S.Y.H., William, W., Butte, A.J., Shah, N.H., Chew, L.S.T., Liu, N., Doshi-Velez, F., Lu, W., Savulescu, J. et al. (2024). Ethical and regulatory challenges of large language models in medicine · 2024
Closest in time.
The Lancet Digital Health 6 , e662–e672
Freyer, O., Wiest, I.C., Kather, J.N., and Gilbert, S. (2024). A future role for health applications of large language models depends on regulators enforcing safety standards · 2024
Closest in time.
NPJ digital medicine 7 , 183
Haltaufderheide, J., and Ranisch, R. (2024). The ethics of chatgpt in medicine and healthcare: a systematic review on large language models (llms) · 2024
Closest in time.
In The Thirty-eight Conference on Neural Information Processing Systems Datasets and Benchmarks Track
Han, T., Kumar, A., Agarwal, C., and Lakkaraju, H. (2024). Medsafetybench: Evaluating and improving the medical safety of large language models · 2024
Closest in time.
arXiv preprint arXiv:2404.09932
Anwar, U., Saparov, A., Rando, J., Paleka, D., Turpin, M., Hase, P., Lubana, E.S., Jenner, E., Casper, S., Sourbut, O. et al. (2024). Foundational challenges in assuring alignment and safety of large language models · 2024
Closest in time.
Code of medical ethics. a
Association, A.M. (2001a) · 2024
Closest in time.
Principles of medical ethics. b
Association, A.M. (2001b) · 2024
Closest in time.
Transactions of the Association for Computational Linguistics 12 , 933–949
Mizrahi, M., Kaplan, G., Malkin, D., Dror, R., Shahaf, D., and Stanovsky, G. (2024). State of what art? a call for multi-prompt llm evaluation · 2024
Closest in time.
European Archives of Oto-Rhino-Laryngology pp. 1–13
Ostrowska, M., Kacała, P., Onolememen, D., Vaughan-Lane, K., Sisily Joseph, A., Ostrowski, A., Pietruszewska, W., Banaszewski, J., and Wróbel, M.J. (2024). To trust or not to trust: evaluating the reliability and safety of ai responses to laryngeal cancer queries · 2024
Closest in time.
arXiv preprint arXiv:2401.05654
Tu, T., Palepu, A., Schaekermann, M., Saab, K., Freyberg, J., Tanno, R., Wang, A., Li, B., Amin, M., Tomasev, N. et al. (2024). Towards conversational diagnostic ai · 2024
Closest in time.