Fetching the paper…
Reading the bibliography…
Medical Large Language Models (MLLMs) have demonstrated potential in healthcare applications, yet their propensity for hallucinations -- generating medically implausible or inaccurate information -- presents substantial risks to patient care.
Generating biomedical question answering corpora from Q&A forums
Lamurias, A.; Sousa, D.; and Couto, F. M. 2020 · 2020
Earlier work this paper cites.
MLEC-QA: A Chinese multi-choice biomedical question answering dataset
Li, J.; Zhong, S.; and Chen, K. 2021 · 2021
Earlier work this paper cites.
Datasetgan: Efficient labeled data factory with minimal human effort
Zhang, Y.; Ling, H.; Gao, J.; Yin, K.; Lafleche, J.-F.; Barriuso, A.; Torralba, A.; and Fidler, S. 2021 · 2021
Earlier work this paper cites.
Few-shot learning for medical text: A systematic review
Ge, Y.; Guo, Y.; Yang, Y.-C.; Al-Garadi, M. A.; and Sarker, A. 2022 · 2022
Earlier work this paper cites.
InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning
Dai, W.; Li, J.; Li, D.; Tiong, A. M. H.; Zhao, J.; Wang, W.; Li, B.; Fung, P.; and Hoi, S. 2023 · 2023
Earlier work this paper cites.
MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Fu, C.; Chen, P.; Shen, Y.; Qin, Y.; Zhang, M.; Lin, X.; Qiu, Z.; Lin, W.; Yang, J.; Zheng, X.; Li, K.; Sun, X.; and Ji, R. 2023 · 2023
Earlier work this paper cites.
MedAlpaca–an open-source collection of medical conversational AI models and training data
Han, T.; Adams, L. C.; Papaioannou, J.-M.; Grundmann, P.; Oberhauser, T.; Löser, A.; Truhn, D.; and Bressem, K. K. 2023 · 2023
Earlier work this paper cites.
Performance of ChatGPT on USMLE: potential for AI-assisted medical education using large language models
Kung, T. H.; Cheatham, M.; Medenilla, A.; Sillos, C.; De Leon, L.; Elepaño, C.; Madriaga, M.; Aggabao, R.; Diaz-Candido, G.; Maningo, J.; et al. 2023 · 2023
Earlier work this paper cites.
Med-HALT: Medical Domain Hallucination Test for Large Language Models
Pal, A.; Umapathi, L. K.; and Sankarasubbu, M. 2023 · 2023
Earlier work this paper cites.
Mededit: Model editing for medical question answering with external knowledge bases
Shi, Y.; Xu, S.; Liu, Z.; Liu, T.; Li, X.; and Liu, N. 2023 · 2023
Earlier work this paper cites.
Large language models encode clinical knowledge
Singhal, K.; Azizi, S.; Tu, T.; Mahdavi, S.; Wei, J.; Chung, H.; Scales, N.; Tanwani, A.; Cole-Lewis, H.; Pfohl, S.; Payne, P.; Seneviratne, M.; Gamble, P.; Kelly, C.; Babiker, A.; Schärli, N.; Chowdhery, A.; Mansfield, P.; Demner-Fushman, D.; and Natarajan, V. 2023 · 2023
Cited alongside, same era.
Aligning Large Multimodal Models with Factually Augmented RLHF
Sun, Z.; Shen, S.; Cao, S.; Liu, H.; Li, C.; Shen, Y.; Gan, C.; Gui, L.; Wang, Y.-X.; Yang, Y.; Keutzer, K.; and Darrell, T. 2023 · 2023
Cited alongside, same era.
Xraygpt: Chest radiographs summarization using medical vision-language models
Thawkar, O.; Shaker, A.; Mullappilly, S. S.; Cholakkal, H.; Anwer, R. M.; Khan, S.; Laaksonen, J.; and Khan, F. S. 2023 · 2023
Cited alongside, same era.
Improving Automated Data Annotation with Self-Supervised Learning: A Pathway to Robust AI Models
Thirunagalingam, A. 2023 · 2023
Cited alongside, same era.
Reinforcement Learning from Human Feedback: Aligning AI Systems with Human Preferences
Alabi, M.; and Wick, L. 2024 · 2024
Closest in time.
MIMIC-Ext-MIMIC-CXR-VQA: A Complex, Diverse, And Large-Scale Visual Question Answering Dataset for Chest X-ray Images
Bae, S.; Kyung, D.; Ryu, J.; Cho, E.; Lee, G.; Kweon, S.; Oh, J.; JI, L.; Chang, E.; Kim, T.; et al. 2024 · 2024
Closest in time.
Data Collection and Labeling Techniques for Machine Learning
Huang, Q.; and Zhao, T. 2024 · 2024
Closest in time.
MedExQA: Medical Question Answering Benchmark with Multiple Explanations
Kim, Y.; Wu, J.; Abdulle, Y.; and Wu, H. 2024 · 2024
Closest in time.
A Liver Cancer Question-Answering System Based on Next-Generation Intelligence and the Large Model Med-PaLM 2
Qian, J.; Jin, Z.; Zhang, Q.; Cai, G.; and Liu, B. 2024 · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Tian, Y.; Gan, R.; Song, Y.; Zhang, J.; and Zhang, Y. 2023 · 2023
Cited alongside, same era.
Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data
Wu, C.; Zhang, X.; Zhang, Y.; Wang, Y.; and Xie, W. 2023 · 2023
Cited alongside, same era.
LVLM-eHub: A Comprehensive Evaluation Benchmark for Large Vision-Language Models
Xu, P.; Shao, W.; Zhang, K.; Gao, P.; Liu, S.; Lei, M.; Meng, F.; Huang, S.; Qiao, Y. J.; and Luo, P. 2023 · 2023
Cited alongside, same era.
mPLUG-Owl2: Revolutionizing Multi-modal Large Language Model with Modality Collaboration
Ye, Q.; Xu, H.; Ye, J.; Yan, M.; Hu, A.; Liu, H.; Qian, Q.; Zhang, J.; Huang, F.; and Zhou, J. 2023 · 2023
Cited alongside, same era.
A Survey of Large Language Models
Zhao, W. X.; Zhou, K.; Li, J.; Tang, T.; Wang, X.; Hou, Y.; Min, Y.; Zhang, B.; Zhang, J.; Dong, Z.; Du, Y.; Yang, C.; Chen, Y.; Chen, Z.; Jiang, J.; Ren, R.; Li, Y.; Tang, X.; Liu, Z.; Liu, P.; Nie, J.; and rong Wen, J. 2023 · 2023
Cited alongside, same era.
MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Zhu, D.; Chen, J.; Shen, X.; Li, X.; and Elhoseiny, M. 2023 · 2023
Cited alongside, same era.
LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day
Li, C.; Wong, C.; Zhang, S.; Usuyama, N.; Liu, H.; Yang, J.; Naumann, T.; Poon, H.; and Gao, J. 2023a
Cited in the paper.
HaluEval: A Large-Scale Hallucination Evaluation Benchmark for Large Language Models
Li, J.; Cheng, X.; Zhao, W. X.; Nie, J.; and rong Wen, J. 2023b
Cited in the paper.
Closest in time.
Large language models for data annotation: A survey
Tan, Z.; Li, D.; Wang, S.; Beigi, A.; Jiang, B.; Bhattacharjee, A.; Karami, M.; Li, J.; Cheng, L.; and Liu, H. 2024 · 2024
Closest in time.
Safety challenges of AI in medicine
Wang, X.; Zhang, N. X.; He, H.; Nguyen, T.; Yu, K.-H.; Deng, H.; Brandt, C.; Bitterman, D. S.; Pan, L.; Cheng, C.-Y.; et al. 2024 · 2024
Closest in time.
Zuo, K.; Jiang, Y.; Mo, F.; and Lio, P. 2024 · 2024
Closest in time.
Zuo, K.; Tang, J.; Qin, H.; Luo, B.; He, L.; and Tang, S. 2025 · 2025
Closest in time.