Fetching the paper…
Reading the bibliography…
The application of the Multi-modal Large Language Models (MLLMs) in medical clinical scenarios remains underexplored.
Rouge: A package for automatic evaluation of summaries
Lin, C.-Y. 2004 · 2004
Earlier work this paper cites.
A dataset of clinically generated visual questions and answers about radiology images
Lau, J. J.; Gayen, S.; Ben Abacha, A.; and Demner-Fushman, D. 2018 · 2018
Earlier work this paper cites.
Diagnostics, 9th ed
PMPH. 2018 · 2018
Earlier work this paper cites.
spaCy: Industrial-strength Natural Language Processing in Python
Honnibal, M.; Montani, I.; Van Landeghem, S.; and Boyd, A. 2020 · 2020
Earlier work this paper cites.
Towards Visual Question Answering on Pathology Images
He, X.; Cai, Z.; Wei, W.; Zhang, Y.; Mou, L.; Xing, E.; and Xie, P. 2021 · 2021
Earlier work this paper cites.
What disease does this patient have? a large-scale open domain question answering dataset from medical exams
Jin, D.; Pan, E.; Oufattole, N.; Weng, W.-H.; Fang, H.; and Szolovits, P. 2021 · 2021
Earlier work this paper cites.
Slake: A semantically-labeled knowledge-enhanced dataset for medical visual question answering
Liu, B.; Zhan, L.-M.; Xu, L.; Ma, L.; Yang, Y.; and Wu, X.-M. 2021 · 2021
Earlier work this paper cites.
MedMCQA: A Large-scale Multi-Subject Multi-Choice Dataset for Medical domain Question Answering
Pal, A.; Umapathi, L. K.; and Sankarasubbu, M. 2022 · 2022
Earlier work this paper cites.
Large language models encode clinical knowledge
Singhal, K.; Azizi, S.; Tu, T.; Mahdavi, S. S.; Wei, J.; Chung, H. W.; Scales, N.; Tanwani, A.; Cole-Lewis, H.; Pfohl, S.; et al. 2022 · 2022
Earlier work this paper cites.
Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Bai, J.; Bai, S.; Yang, S.; and et al. 2023 · 2023
Cited alongside, same era.
Disc-medllm: Bridging general large language models and real-world medical consultation
Bao, Z.; Chen, W.; Xiao, S.; Ren, K.; Wu, J.; Zhong, C.; Peng, J.; Huang, X.; and Wei, Z. 2023 · 2023
Cited alongside, same era.
Leveraging large language models for decision support in personalized oncology
Benary, M.; Wang, X. D.; Schmidt, M.; Soll, D.; Hilfenhaus, G.; Nassir, M.; Sigler, C.; Knödler, M.; Keller, U.; Beule, D.; et al. 2023 · 2023
Cited alongside, same era.
Liao, Y.; Meng, Y.; Liu, H.; Wang, Y.; and Wang, Y. 2023 · 2023
Cited alongside, same era.
ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools
GLM, T.; Zeng, A.; Xu, B.; and et al. 2024 · 2024
Closest in time.
Evaluation and mitigation of the limitations of large language models in clinical decision-making
Hager, P.; Jungmann, F.; Holland, R.; Bhagat, K.; Hubrecht, I.; Knauer, M.; Vielhauer, J.; Makowski, M.; Braren, R.; Kaissis, G.; et al. 2024 · 2024
Closest in time.
Ct2rep: Automated radiology report generation for 3d medical imaging
Hamamci, I. E.; Er, S.; and Menze, B. 2024 · 2024
Closest in time.
CRAFT-MD: A Conversational Evaluation Framework for Comprehensive Assessment of Clinical LLMs
Johri, S.; Jeong, J.; Tran, B. A.; Schlessinger, D. I.; Wongvibulsin, S.; Cai, Z. R.; Daneshjou, R.; and Rajpurkar, P. 2024 · 2024
Closest in time.
Llava-med: Training a large language-and-vision assistant for biomedicine in one day
Li, C.; Wong, C.; Zhang, S.; Usuyama, N.; Liu, H.; Yang, J.; Naumann, T.; Poon, H.; and Gao, J. 2024 · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
NHC. 2023 · 2023
Cited alongside, same era.
Xraygpt: Chest radiographs summarization using medical vision-language models
Thawkar, O.; Shaker, A.; Mullappilly, S. S.; Cholakkal, H.; Anwer, R. M.; Khan, S.; Laaksonen, J.; and Khan, F. S. 2023 · 2023
Cited alongside, same era.
Towards generalist foundation model for radiology
Wu, C.; Zhang, X.; Zhang, Y.; Wang, Y.; and Xie, W. 2023 · 2023
Cited alongside, same era.
Huatuogpt, towards taming language model to be a doctor
Zhang, H.; Chen, J.; Jiang, F.; Yu, F.; Chen, Z.; Li, J.; Chen, G.; Wu, X.; Zhang, Z.; Xiao, Q.; et al. 2023 · 2023
Cited alongside, same era.
Peking University Health Science Center Clinical Medicine Program Objective Structured Clinical Examination (OSCE) Instructions
BJMU. 2024 · 2024
Cited alongside, same era.
HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale
Chen, J.; Ouyang, R.; Gao, A.; Chen, S.; Chen, G. H.; Wang, X.; Zhang, R.; Cai, Z.; Ji, K.; Yu, G.; et al. 2024a
Cited in the paper.
How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
Chen, Z.; Wang, W.; Tian, H.; and et al. 2024b
Cited in the paper.
Mini-InternVL 1.5: A Powerful Pocket Multimodal Model with 8% Parameters for 80% Performance
Chen, Z.; Zhangwei, G.; Erfei, C.; and et al. 2024c
Cited in the paper.
Closest in time.
Automatic Interactive Evaluation for Large Language Models with State Aware Patient Simulator
Liao, Y.; Meng, Y.; Wang, Y.; Liu, H.; Wang, Y.; and Wang, Y. 2024 · 2024
Closest in time.
Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reid, M.; Savinov, N.; Teplyashin, D.; and et al. 2024 · 2024
Closest in time.
AgentClinic: a multimodal agent benchmark to evaluate AI in simulated clinical environments
Schmidgall, S.; Ziaei, R.; Harris, C.; Reis, E.; Jopling, J.; and Moor, M. 2024 · 2024
Closest in time.
Towards conversational diagnostic ai
Tu, T.; Palepu, A.; Schaekermann, M.; Saab, K.; Freyberg, J.; Tanno, R.; Wang, A.; Li, B.; Amin, M.; Tomasev, N.; et al. 2024 · 2024
Closest in time.