Fetching the paper…
Reading the bibliography…
With the advent of Large Language Models (LLMs), medical artificial intelligence (AI) has experienced substantial technological progress and paradigm shifts, highlighting the potential of LLMs to streamline healthcare delivery and improve patient outcomes.
C. Mattingly, “What is clinical reasoning?” The American Journal of Occupational Therapy , vol. 45, no. 11, pp. 979–986, 1991
1991
Earlier work this paper cites.
A. Act, “Health insurance portability and accountability act of 1996,” Public law , vol. 104, p. 191, 1996
1996
Earlier work this paper cites.
J. R. Kogan, L. M. Bellini, and J. A. Shea, “Implementation of the mini-cex to evaluate medical students’ clinical skills,” Academic Medicine , vol. 77, no. 11, pp. 1156–1157, 2002
2002
Earlier work this paper cites.
J.-D. Kim, T. Ohta, Y. Tateisi, and J. Tsujii, “Genia corpus—a semantically annotated corpus for bio-textmining,” Bioinformatics , vol. 19, no. suppl_1, pp. i180–i182, 2003
2003
Earlier work this paper cites.
J. J. Norcini, L. L. Blank, F. D. Duffy, and G. S. Fortna, “The mini-cex: a method for assessing clinical skills,” Annals of internal medicine , vol. 138, no. 6, pp. 476–481, 2003
2003
Earlier work this paper cites.
J. Ramos et al. , “Using tf-idf to determine word relevance in document queries,” in Proceedings of the first instructional conference on machine learning , vol. 242, no. 1, 2003, pp. 29–48
2003
Earlier work this paper cites.
G. Zhou, J. Su, J. Zhang, and M. Zhang, “Exploring various knowledge in relation extraction,” in Proceedings of the 43rd annual meeting of the association for computational linguistics (acl’05) , 2005, pp. 427–434
2005
Earlier work this paper cites.
P. Sen, G. Namata, M. Bilgic, L. Getoor, B. Galligher, and T. Eliassi-Rad, “Collective classification in network data,” AI magazine , vol. 29, no. 3, pp. 93–93, 2008
2008
Earlier work this paper cites.
A. B. Abacha and P. Zweigenbaum, “Medical entity recognition: a comparaison of semantic and statistical methods,” in Proceedings of BioNLP 2011 workshop , 2011, pp. 56–64
2011
Earlier work this paper cites.
J.-D. Kim, Y. Wang, T. Takagi, and A. Yonezawa, “Overview of genia event task in bionlp shared task 2011,” in Proceedings of BioNLP shared task 2011 workshop , 2011, pp. 7–15
2011
Earlier work this paper cites.
H. Gurulingappa, A. M. Rajput, A. Roberts, J. Fluck, M. Hofmann-Apitius, and L. Toldo, “Development of a benchmark corpus to support the automatic extraction of drug-related adverse effects from medical case reports,” Journal of biomedical informatics , vol. 45, no. 5, pp. 885–892, 2012
2012
Earlier work this paper cites.
S. Pradhan, N. Elhadad, B. R. South, D. Martinez, L. M. Christensen, A. Vogel, H. Suominen, W. W. Chapman, and G. K. Savova, “Task 1: Share/clef ehealth evaluation lab 2013.” CLEF (working notes) , vol. 1179, 2013
2013
Earlier work this paper cites.
J.-D. Kim, Y. Wang, and Y. Yasunori, “The genia event extraction shared task, 2013 edition-overview,” in Proceedings of the BioNLP Shared Task 2013 Workshop , 2013, pp. 8–15
2013
Earlier work this paper cites.
R. I. Doğan, R. Leaman, and Z. Lu, “Ncbi disease corpus: a resource for disease name recognition and concept normalization,” Journal of biomedical informatics , vol. 47, pp. 1–10, 2014
2014
Earlier work this paper cites.
D. L. Mowery, S. Velupillai, B. R. South, L. Christensen, D. Martinez, L. Kelly, L. Goeuriot, N. Elhadad, S. Pradhan, G. Savova et al. , “Task 2: Share/clef ehealth evaluation lab 2014,” in Proceedings of CLEF 2014 , 2014
2014
Earlier work this paper cites.
S. Moon, S. Pakhomov, N. Liu, J. O. Ryan, and G. B. Melton, “A sense inventory for clinical abbreviations and acronyms created using clinical notes and medical dictionary resources,” Journal of the American Medical Informatics Association , vol. 21, no. 2, pp. 299–307, 2014
2014
Earlier work this paper cites.
A. Trotman, A. Puurula, and B. Burgess, “Improvements to bm25 and language models examined,” in Proceedings of the 19th Australasian Document Computing Symposium , 2014, pp. 58–65
2014
Earlier work this paper cites.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft coco: Common objects in context,” in Computer Vision–ECCV 2014: 13th European Conference, Zurich, Switzerland, September 6-12, 2014, Proceedings, Part V 13 , 2014, pp. 740–755
2014
Earlier work this paper cites.
S. Karimi, A. Metke-Jimenez, M. Kemp, and C. Wang, “Cadec: A corpus of adverse drug event annotations,” Journal of biomedical informatics , vol. 55, pp. 73–81, 2015
2015
Earlier work this paper cites.
J. Li, Y. Sun, R. J. Johnson, D. Sciaky, C.-H. Wei, R. Leaman, A. P. Davis, C. J. Mattingly, T. C. Wiegers, and Z. Lu, “Biocreative v cdr task corpus: a resource for chemical disease relation extraction,” Database , vol. 2016, 2016
2016
Earlier work this paper cites.
intersoft consulting, “General data protection regulation,” 2016. [Online]. Available: https://gdpr-info.eu/
2016
Earlier work this paper cites.
A. E. Johnson, T. J. Pollard, L. Shen, L.-w. H. Lehman, M. Feng, M. Ghassemi, B. Moody, P. Szolovits, L. Anthony Celi, and R. G. Mark, “Mimic-iii, a freely accessible critical care database,” Scientific data , vol. 3, no. 1, pp. 1–9, 2016
2016
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
A. Pampari, P. Raghavan, J. Liang, and J. Peng, “Emrqa: A large corpus for question answering on electronic medical records,” in 2018 Conference on Empirical Methods in Natural Language Processing, EMNLP 2018 , 2018, pp. 2357–2368
2018
Earlier work this paper cites.
Z. Wei, Q. Liu, B. Peng, H. Tou, T. Chen, X.-J. Huang, K.-F. Wong, and X. Dai, “Task-oriented dialogue system for automatic diagnosis,” in Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers) , 2018, pp. 201–207
2018
Earlier work this paper cites.
B. Nye, J. J. Li, R. Patel, Y. Yang, I. J. Marshall, A. Nenkova, and B. C. Wallace, “A corpus with multi-level annotations of patients, interventions and outcomes to support language processing for medical literature,” in Proceedings of the conference. Association for Computational Linguistics. Meeting , vol. 2018. NIH Public Access, 2018, p. 197
2018
Earlier work this paper cites.
Z. Chen, “Qasystemonmedicalgraph,” 2018. [Online]. Available: https://github.com/zhihao-chen/QASystemOnMedicalGraph
2018
Earlier work this paper cites.
2019
Earlier work this paper cites.
Q. Jin, B. Dhingra, Z. Liu, W. Cohen, and X. Lu, “Pubmedqa: A dataset for biomedical research question answering,” in Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP) , 2019, pp. 2567–2577
2019
Earlier work this paper cites.
A. E. Johnson, T. J. Pollard, S. J. Berkowitz, N. R. Greenbaum, M. P. Lungren, C.-y. Deng, R. G. Mark, and S. Horng, “Mimic-cxr, a de-identified publicly available database of chest radiographs with free-text reports,” Scientific data , vol. 6, no. 1, p. 317, 2019
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
N. S. Tawfik and M. R. Spruit, “Towards recognition of textual entailment in the biomedical domain,” in Natural Language Processing and Information Systems . Springer International Publishing, 2019, pp. 368–375
2019
Earlier work this paper cites.
A. B. Abacha, Y. Mrabet, M. Sharp, T. R. Goodwin, S. E. Shooshan, and D. Demner-Fushman, “Bridging the gap between consumers’ medication questions and trusted answers.” in MedInfo , 2019, pp. 25–29
2019
Earlier work this paper cites.
A. Ben Abacha and D. Demner-Fushman, “A question-entailment approach to question answering,” BMC Bioinform. , vol. 20, no. 1, pp. 511:1–511:23, 2019
2019
Earlier work this paper cites.
J. He, M. Fu, and M. Tu, “Applying deep matching networks to chinese medical question answering: a study and a dataset,” BMC medical informatics and decision making , vol. 19, no. 2, pp. 91–100, 2019
2019
Earlier work this paper cites.
A. B. Abacha and D. Demner-Fushman, “On the summarization of consumer health questions,” in Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics , 2019, pp. 2228–2234
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Y. Bai, “chatbot-base-on-knowledgegraph,” 2019. [Online]. Available: https://github.com/baiyang2464/chatbot-base-on-Knowledge-Graph
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
C. Raffel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, and P. J. Liu, “Exploring the limits of transfer learning with a unified text-to-text transformer,” Journal of Machine Learning Research , vol. 21, no. 140, pp. 1–67, 2020
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
X. Qiu, T. Sun, Y. Xu, Y. Shao, N. Dai, and X. Huang, “Pre-trained models for natural language processing: A survey,” Science China Technological Sciences , vol. 63, no. 10, pp. 1872–1897, 2020
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
J. Li, A. Sun, J. Han, and C. Li, “A survey on deep learning for named entity recognition,” IEEE transactions on knowledge and data engineering , vol. 34, no. 1, pp. 50–70, 2020
2020
Earlier work this paper cites.
J. Liu, Y. Chen, K. Liu, W. Bi, and X. Liu, “Event extraction as machine reading comprehension,” in Proceedings of the 2020 conference on empirical methods in natural language processing (EMNLP) , 2020, pp. 1641–1651
2020
Earlier work this paper cites.
X. Yang, J. Bian, W. R. Hogan, and Y. Wu, “Clinical concept extraction using transformers,” Journal of the American Medical Informatics Association , vol. 27, no. 12, pp. 1935–1942, 2020
2020
Earlier work this paper cites.
B. Fan, W. Fan, C. Smith et al. , “Adverse drug event detection and extraction from open data: A deep learning approach,” Information Processing & Management , vol. 57, no. 1, p. 102131, 2020
2020
Earlier work this paper cites.
S. Yadav, V. Pallagani, and A. Sheth, “Medical knowledge-enriched textual entailment framework,” 2020
2020
Earlier work this paper cites.
Q. Liu, S. Jiang, Y. Wang, and S. Li, “Liveqa: A question answering dataset over sports live,” in Chinese Computational Linguistics: 19th China National Conference, CCL 2020, Hainan, China, October 30–November 1, 2020, Proceedings 19 . Springer, 2020, pp. 316–328
2020
Earlier work this paper cites.
G. Zeng, Q. Wu, Y. Zhang, Z. Yu, E. Xing, and P. Xie, “Develop medical dialogue systems for covid-19,” https://github.com/UCSD-AI4H/COVID-Dialogue , 2020
2020
Earlier work this paper cites.
M. Savery, A. B. Abacha, S. Gayen, and D. Demner-Fushman, “Question-driven summarization of answers to consumer health questions,” Scientific Data , vol. 7, no. 1, p. 322, 2020
2020
Earlier work this paper cites.
L. L. Wang, K. Lo, Y. Chandrasekhar, R. Reas, J. Yang, D. Burdick, D. Eide, K. Funk, Y. Katsis, R. Kinney et al. , “Cord-19: The covid-19 open research dataset,” ArXiv , 2020
2020
Earlier work this paper cites.
J. Lee, W. Yoon, S. Kim, D. Kim, S. Kim, C. H. So, and J. Kang, “Biobert: a pre-trained biomedical language representation model for biomedical text mining,” Bioinformatics , vol. 36, no. 4, pp. 1234–1240, 2020
2020
Earlier work this paper cites.
H.-C. Shin, Y. Zhang, E. Bakhturina, R. Puri, M. Patwary, M. Shoeybi, and R. Mani, “Biomegatron: larger biomedical domain language model,” in Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP) , 2020, pp. 4700–4706
2020
Earlier work this paper cites.
P. Lewis, E. Perez, A. Piktus, F. Petroni, V. Karpukhin, N. Goyal, H. Küttler, M. Lewis, W.-t. Yih, T. Rocktäschel et al. , “Retrieval-augmented generation for knowledge-intensive nlp tasks,” Advances in Neural Information Processing Systems , vol. 33, pp. 9459–9474, 2020
2020
Earlier work this paper cites.
Z. Chu, S. L. Rathbun, and S. Li, “Matching in selective and balanced representation space for treatment effects estimation,” in Proceedings of the 29th ACM International Conference on Information & Knowledge Management , 2020, pp. 205–214
2020
Earlier work this paper cites.
R. Wheeler, Clinical law for clinical practice . CRC Press, 2020
2020
Earlier work this paper cites.
X. Pan, M. Zhang, S. Ji, and M. Yang, “Privacy risks of general-purpose language models,” in 2020 IEEE Symposium on Security and Privacy (SP) . IEEE, 2020, pp. 1314–1331
2020
Earlier work this paper cites.
D. Jin, E. Pan, N. Oufattole, W.-H. Weng, H. Fang, and P. Szolovits, “What disease does this patient have? a large-scale open domain question answering dataset from medical exams,” Applied Sciences , vol. 11, no. 14, p. 6421, 2021
2021
Earlier work this paper cites.
Y. Gu, R. Tinn, H. Cheng, M. Lucas, N. Usuyama, X. Liu, T. Naumann, J. Gao, and H. Poon, “Domain-specific language model pretraining for biomedical natural language processing,” ACM Transactions on Computing for Healthcare (HEALTH) , vol. 3, no. 1, pp. 1–23, 2021
2021
Earlier work this paper cites.
F. S. Yazi, W.-T. Vong, V. Raman, P. H. H. Then, and M. J. Lunia, “An experimental evaluation of deep neural network model performance for the recognition of contradictory medical research claims using small and medium-sized corpora,” pp. 68–77, 2021
2021
Earlier work this paper cites.
E. J. Hu, P. Wallis, Z. Allen-Zhu, Y. Li, S. Wang, L. Wang, W. Chen et al. , “Lora: Low-rank adaptation of large language models,” in International Conference on Learning Representations , 2021
2021
Earlier work this paper cites.
D. Hendrycks, C. Burns, S. Basart, A. Zou, M. Mazeika, D. Song, and J. Steinhardt, “Measuring massive multitask language understanding,” Proceedings of the International Conference on Learning Representations (ICLR) , 2021
2021
Earlier work this paper cites.
H.-S. Sheu, Z. Chu, D. Qi, and S. Li, “Knowledge-guided article embedding refinement for session-based news recommendation,” IEEE Transactions on Neural Networks and Learning Systems , vol. 33, no. 12, pp. 7921–7927, 2021
2021
Earlier work this paper cites.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark et al. , “Learning transferable visual models from natural language supervision,” in International conference on machine learning . PMLR, 2021, pp. 8748–8763
2021
Earlier work this paper cites.
Z. Chu, S. L. Rathbun, and S. Li, “Graph infomax adversarial learning for treatment effect estimation with networked observational data,” in Proceedings of the 27th ACM SIGKDD Conference on Knowledge Discovery & Data Mining , 2021, pp. 176–184
2021
Earlier work this paper cites.
I. Garrido-Muñoz, A. Montejo-Ráez, F. Martínez-Santiago, and L. A. Ureña-López, “A survey on bias in deep nlp,” Applied Sciences , vol. 11, no. 7, p. 3184, 2021
2021
Earlier work this paper cites.
E. M. Bender, T. Gebru, A. McMillan-Major, and S. Shmitchell, “On the dangers of stochastic parrots: Can language models be too big?” in Proceedings of the 2021 ACM conference on fairness, accountability, and transparency , 2021, pp. 610–623
2021
Earlier work this paper cites.
2022
Earlier work this paper cites.
J. Wei, X. Wang, D. Schuurmans, M. Bosma, brian ichter, F. Xia, E. H. Chi, Q. V. Le, and D. Zhou, “Chain of thought prompting elicits reasoning in large language models,” in Advances in Neural Information Processing Systems , A. H. Oh, A. Agarwal, D. Belgrave, and K. Cho, Eds., 2022
2022
Earlier work this paper cites.
A. Pal, L. K. Umapathi, and M. Sankarasubbu, “Medmcqa: A large-scale multi-subject multi-choice dataset for medical domain question answering,” in Proceedings of the Conference on Health, Inference, and Learning , vol. 174, 2022, pp. 248–260
2022
Earlier work this paper cites.
Q. Lu, D. Dou, and T. Nguyen, “ClinicalT5: A generative language model for clinical text,” in Findings of the Association for Computational Linguistics , 2022, pp. 5436–5443
2022
Earlier work this paper cites.
R. Luo, L. Sun, Y. Xia, T. Qin, S. Zhang, H. Poon, and T.-Y. Liu, “Biogpt: generative pre-trained transformer for biomedical text generation and mining,” Briefings in bioinformatics , vol. 23, no. 6, p. bbac409, 2022
2022
Earlier work this paper cites.
X. Yang, A. Chen, N. PourNejatian, H. C. Shin, K. E. Smith, C. Parisien, C. Compas, C. Martin, A. B. Costa, M. G. Flores et al. , “A large language model for electronic health records,” NPJ digital medicine , vol. 5, no. 1, p. 194, 2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
J. Yang, S. C. Han, and J. Poon, “A survey on extraction of causal relations from natural language text,” vol. 64, no. 5, pp. 1161–1186, 2022-05-01
2022
Earlier work this paper cites.
L. Ouyang, J. Wu, X. Jiang, D. Almeida, C. Wainwright, P. Mishkin, C. Zhang, S. Agarwal, K. Slama, A. Ray et al. , “Training language models to follow instructions with human feedback,” Advances in neural information processing systems , vol. 35, pp. 27 730–27 744, 2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
F. Xia, B. Li, Y. Weng, S. He, K. Liu, B. Sun, S. Li, and J. Zhao, “MedConQA: Medical conversational question answering system based on knowledge graphs,” in Proceedings of the The 2022 Conference on Empirical Methods in Natural Language Processing: System Demonstrations , 2022, pp. 148–158
2022
Earlier work this paper cites.
S. Sotudeh, N. Goharian, and Z. Young, “Mentsum: A resource for exploring summarization of mental health online posts,” in Proceedings of the thirteenth language resources and evaluation conference , 2022, pp. 2682–2692
2022
Earlier work this paper cites.
M. Agrawal, S. Hegselmann, H. Lang, Y. Kim, and D. Sontag, “Large language models are few-shot clinical information extractors,” in Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing , 2022, pp. 1998–2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Cited alongside, same era.
S. Xue, Y. Wang, Z. Chu, X. Shi, C. Jiang, H. Hao, G. Jiang, X. Feng, J. Zhang, and J. Zhou, “Prompt-augmented temporal point process for streaming event sequence,” Advances in Neural Information Processing Systems , vol. 36, pp. 18 885–18 905, 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Y. Anand, Z. Nussbaum, B. Duderstadt, B. Schmidt, and A. Mulyar, “Gpt4all: Training an assistant-style chatbot with large scale data distillation from gpt-3.5-turbo,” https://github.com/nomic-ai/gpt4all , 2023
2023
2023
Later among the works it cites.
W.-L. Chiang, Z. Li, Z. Lin, Y. Sheng, Z. Wu, H. Zhang, L. Zheng, S. Zhuang, Y. Zhuang, J. E. Gonzalez et al. , “Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality,” See https://vicuna. lmsys. org (accessed 14 April 2023) , vol. 2, no. 3, p. 6, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
2023
Cited alongside, same era.
M. Xu, “Medicalgpt: Training medical gpt model,” https://github.com/shibing624/MedicalGPT , 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
W. Y. Wei Zhu and X. Wang, “Shennong-tcm: A traditional chinese medicine large language model,” https://github.com/michael-wzhu/ShenNong-TCM-LLM , 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
B. Min, H. Ross, E. Sulem, A. P. B. Veyseh, T. H. Nguyen, O. Sainz, E. Agirre, I. Heintz, and D. Roth, “Recent advances in natural language processing via large pre-trained language models: A survey,” ACM Computing Surveys , vol. 56, no. 2, pp. 1–40, 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
V. Sorin, Y. Barash, E. Konen, and E. Klang, “Large language models for oncological applications,” Journal of Cancer Research and Clinical Oncology , vol. 149, no. 11, pp. 9505–9508, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
C. W. Safranek, A. E. Sidamon-Eristoff, A. Gilson, and D. Chartash, “The role of large language models in medical education: applications and implications,” p. e50945, 2023
2023
Later among the works it cites.
Z. Chu, R. Li, S. Rathbun, and S. Li, “Continual causal inference with incremental observational data,” in 2023 IEEE 39th International Conference on Data Engineering . IEEE, 2023, pp. 3430–3439
2023
Later among the works it cites.
J. Clusmann, F. R. Kolbinger, H. S. Muti, Z. I. Carrero, J.-N. Eckardt, N. G. Laleh, C. M. L. Löffler, S.-C. Schwarzkopf, M. Unger, G. P. Veldhuizen et al. , “The future landscape of large language models in medicine,” Communications medicine , vol. 3, no. 1, p. 141, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
J. Zhang, K. Bao, Y. Zhang, W. Wang, F. Feng, and X. He, “Is chatgpt fair for recommendation? evaluating fairness in large language model recommendation,” in Proceedings of the 17th ACM Conference on Recommender Systems , 2023, pp. 993–999
2023
Later among the works it cites.
S. Dwivedi, S. Ghosh, and S. Dwivedi, “Breaking the bias: Gender fairness in llms using prompt engineering and in-context learning,” Rupkatha Journal on Interdisciplinary Studies in Humanities , vol. 15, no. 4, 2023
2023
Later among the works it cites.
J. A. Omiye, J. C. Lester, S. Spichak, V. Rotemberg, and R. Daneshjou, “Large language models propagate race-based medicine,” NPJ Digital Medicine , vol. 6, no. 1, p. 195, 2023
2023
Later among the works it cites.
M. Karabacak and K. Margetis, “Embracing large language models for medical applications: opportunities and challenges,” Cureus , vol. 15, no. 5, 2023
2023
Later among the works it cites.
C. Blease and J. Torous, “Chatgpt and mental healthcare: balancing benefits with risks of harms,” BMJ Ment Health , vol. 26, no. 1, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
Z. Chu and S. Li, “Causal effect estimation: Recent progress, challenges, and opportunities,” Machine Learning for Causal Inference , pp. 79–100, 2023
2023
Later among the works it cites.
A. A. Parray, Z. M. Inam, D. Ramonfaur, S. S. Haider, S. K. Mistry, and A. K. Pandya, “Chatgpt and global public health: applications, challenges, ethical considerations and mitigation strategies,” 2023
2023
Later among the works it cites.
S. Mathavan, “Mitigating linguistic bias in bert-based medical diagnosis models,” Ph.D. dissertation, 2023
2023
Later among the works it cites.
J. Elsborg and M. Salvatore, “Using llms and explainable ml to analyze biomarkers at single-cell level for improved understanding of diseases,” Biomolecules , vol. 13, no. 10, p. 1516, 2023
2023
Later among the works it cites.
M. A. Rahman, “A survey on security and privacy of multimodal llms-connected healthcare perspective,” in 2023 IEEE Globecom Workshops (GC Wkshps) . IEEE, 2023, pp. 1807–1812
2023
Later among the works it cites.
2023
Later among the works it cites.
J. Liu, P. Zhou, Y. Hua, D. Chong, Z. Tian, A. Liu, H. Wang, C. You, Z. Guo, L. Zhu et al. , “Benchmarking large language models on cmexam-a comprehensive chinese medical exam dataset,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.
Y. Cai, L. Wang, Y. Wang, G. de Melo, Y. Zhang, Y. Wang, and L. He, “Medbench: A large-scale chinese benchmark for evaluating medical large language models,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 38, no. 16, 2024, pp. 17 709–17 717
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
C. Zakka, R. Shad, A. Chaurasia, A. R. Dalal, J. L. Kim, M. Moor, R. Fong, C. Phillips, K. Alexander, E. Ashley et al. , “Almanac—retrieval-augmented language models for clinical medicine,” NEJM AI , vol. 1, no. 2, p. AIoa2300068, 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
Z. Sun, Y. Shen, Q. Zhou, H. Zhang, Z. Chen, D. Cox, Y. Yang, and C. Gan, “Principle-driven self-alignment of language models from scratch with minimal human supervision,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
J. A. Omiye, H. Gui, S. J. Rezaei, J. Zou, and R. Daneshjou, “Large language models in medicine: The potentials and pitfalls: A narrative review,” Annals of Internal Medicine , vol. 177, no. 2, pp. 210–220, 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
Y. Zhuang, Y. Yu, K. Wang, H. Sun, and C. Zhang, “Toolqa: A dataset for llm question answering with external tools,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.
2024
Closest in time.
V. Liévin, C. E. Hother, A. G. Motzfeldt, and O. Winther, “Can large language models reason about medical questions?” Patterns , vol. 5, no. 3, 2024
2024
Closest in time.
L. Luo, J. Ning, Y. Zhao, Z. Wang, Z. Ding, P. Chen, W. Fu, Q. Han, G. Xu, Y. Qiu, D. Pan, J. Li, H. Li, W. Feng, S. Tu, Y. Liu, Z. Yang, J. Wang, Y. Sun, and H. Lin, “Taiyi: A bilingual fine-tuned large language model for diverse biomedical tasks,” Journal of the American Medical Informatics Association , 2024
2024
Closest in time.
S. Yang, H. Zhao, S. Zhu, G. Zhou, H. Xu, Y. Jia, and H. Zan, “Zhongjing: Enhancing the chinese medical capabilities of large language model through expert feedback and real-world multi-turn dialogue,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 38, no. 17, 2024, pp. 19 368–19 376
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
Y. Wang, Z. Chu, X. Ouyang, S. Wang, H. Hao, Y. Shen, J. Gu, S. Xue, J. Zhang, Q. Cui et al. , “Llmrg: Improving recommendations through large language model reasoning graphs,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 38, no. 17, 2024, pp. 19 189–19 196
2024
Closest in time.
2024
Closest in time.
S. Pan, L. Luo, Y. Wang, C. Chen, J. Wang, and X. Wu, “Unifying large language models and knowledge graphs: A roadmap,” IEEE Transactions on Knowledge and Data Engineering , 2024
2024
Closest in time.
Y. Tian, H. Song, Z. Wang, H. Wang, Z. Hu, F. Wang, N. V. Chawla, and P. Xu, “Graph neural prompting with large language models,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 38, no. 17, 2024, pp. 19 080–19 088
2024
Closest in time.
W. Wei, X. Ren, J. Tang, Q. Wang, L. Su, S. Cheng, J. Wang, D. Yin, and C. Huang, “Llmrec: Large language models with graph augmentation for recommendation,” in Proceedings of the 17th ACM International Conference on Web Search and Data Mining , 2024, pp. 806–815
2024
Closest in time.
X. Guan, Y. Liu, H. Lin, Y. Lu, B. He, X. Han, and L. Sun, “Mitigating large language model hallucinations via autonomous knowledge graph-based retrofitting,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 38, no. 16, 2024, pp. 18 126–18 134
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
N. Mehandru, B. Y. Miao, E. R. Almaraz, M. Sushil, A. J. Butte, and A. Alaa, “Evaluating large language models as agents in the clinic,” npj Digital Medicine , vol. 7, no. 1, p. 84, 2024
2024
Closest in time.
H. Liu, C. Li, Q. Wu, and Y. J. Lee, “Visual instruction tuning,” Advances in neural information processing systems , vol. 36, 2024
2024
Closest in time.
2024
Closest in time.
Z. Chu, M. Hu, Q. Cui, L. Li, and S. Li, “Task-driven causal feature distillation: Towards trustworthy risk prediction,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 38, no. 10, 2024, pp. 11 642–11 650
2024
Closest in time.
2024
Closest in time.
D. Van Veen, C. Van Uden, L. Blankemeier, J.-B. Delbrouck, A. Aali, C. Bluethgen, A. Pareek, M. Polacin, E. P. Reis, A. Seehofnerová et al. , “Adapted large language models can outperform medical experts in clinical text summarization,” Nature Medicine , pp. 1–9, 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
L. A. Jawad, “Security and privacy in digital healthcare systems: Challenges and mitigation strategies,” Abhigyan , no. 1, pp. 23–31, 2024
2024
Closest in time.
S. Arvisais-Anhalt, S. L. Gonias, and S. G. Murray, “Establishing priorities for implementation of large language models in pathology and laboratory medicine,” Academic Pathology , vol. 11, no. 1, p. 100101, 2024
2024
Closest in time.
A. Lawson McLean, Y. Wu, A. C. Lawson McLean, and V. Hristidis, “Large language models as decision aids in neuro-oncology: a review of shared decision-making applications,” Journal of Cancer Research and Clinical Oncology , vol. 150, no. 3, p. 139, 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
T. Zack, E. Lehman, M. Suzgun, J. A. Rodriguez, L. A. Celi, J. Gichoya, D. Jurafsky, P. Szolovits, D. W. Bates, R.-E. E. Abdulnour et al. , “Assessing the potential of gpt-4 to perpetuate racial and gender biases in health care: a model evaluation study,” The Lancet Digital Health , vol. 6, no. 1, pp. e12–e22, 2024
2024
Closest in time.
A. H. Shariatmadari, S. Guo, S. Srinivasan, and A. Zhang, “Harnessing the power of knowledge graphs to enhance llm explainability in the biomedical domain,” 2024
2024
Closest in time.
2024
Closest in time.
T. Singh, H. Aditya, V. K. Madisetti, and A. Bahga, “Whispered tuning: Data privacy preservation in fine-tuning llms through differential privacy,” Journal of Software Engineering and Applications , vol. 17, no. 1, pp. 1–22, 2024
2024
Closest in time.
L. Yuan, Y. Chen, G. Cui, H. Gao, F. Zou, X. Cheng, H. Ji, Z. Liu, and M. Sun, “Revisiting out-of-distribution robustness in nlp: Benchmarks, analysis, and llms evaluations,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.
2024
Closest in time.