Fetching the paper…
Reading the bibliography…
Security concerns related to Large Language Models (LLMs) have been extensively explored, yet the safety implications for Multimodal Large Language Models (MLLMs), particularly in medical contexts (MedMLLMs), remain insufficiently studied.
Overconfidence as a cause of diagnostic error in medicine
Eta S Berner and Mark L Graber · 2008
Earlier work this paper cites.
Diagnostic error in medicine: analysis of 583 physician-reported errors
Gordon D Schiff, Omar Hasan, Seijeoung Kim, Richard Abrams, Karen Cosby, Bruce L Lambert, Arthur S Elstein, Scott Hasler, Martin L Kabongo, Nela Krosnjar, et al · 2009
Earlier work this paper cites.
The incidence of diagnostic error in medicine
Mark L Graber · 2013
Earlier work this paper cites.
Evasion attacks against machine learning at test time
Battista Biggio, Igino Corona, Davide Maiorca, Blaine Nelson, Nedim Šrndić, Pavel Laskov, Giorgio Giacinto, and Fabio Roli · 2013
Earlier work this paper cites.
Explaining and harnessing adversarial examples
I. J. Goodfellow, J. Shlens, and C. Szegedy · 2014
Earlier work this paper cites.
Towards evaluating the robustness of neural networks
N. Carlini and D. Wagner · 2017
Earlier work this paper cites.
Towards deep learning models resistant to adversarial attacks
A. Madry, A. Makelov, L. Schmidt, D. Tsipras, and A. Vladu · 2017
Earlier work this paper cites.
The space of transferable adversarial examples
Florian Tramèr, Alexey Kurakin, Nicolas Papernot, Ian Goodfellow, Dan Boneh, and Patrick McDaniel · 2017
Earlier work this paper cites.
Labeled Optical Coherence Tomography (OCT) for Classification
Daniel Kermany · 2017
Earlier work this paper cites.
A dataset of clinically generated visual questions and answers about radiology images
Jason J Lau, Soumya Gayen, Asma Ben Abacha, and Dina Demner-Fushman · 2018
Earlier work this paper cites.
Automated measurement of fetal head circumference using 2D ultrasound images
Thomas L. A. van den Heuvel, Dagmar de Bruijn, Chris L. de Korte, and Bram van Ginneken · 2018
Earlier work this paper cites.
Study of Segmentation Technique and Stereology to Detect PCO Follicles on USG Images
Untari Novia Wisesty, Irba Fairuz Thufailah, Ria May Dewi, Adiwijaya, and Jondri · 2018
Earlier work this paper cites.
Identifying Medical Diagnoses and Treatable Diseases by Image-Based Deep Learning
Daniel S. Kermany, Michael Goldbaum, Wenjia Cai, Carolina C. S. Valentim, Huiying Liang, Sally L. Baxter, Alex McKeown, Ge Yang, Xiaokang Wu, Fangbing Yan, Justin Dong, Made K. Prasadha, Jacqueline Pei, Magdalene Y. L. Ting, Jie Zhu, Christina Li, Sierra Hewett, Jason Dong, Ian Ziyar, Alexander Shi, Runze Zhang, Lianghong Zheng, Rui Hou, William Shi, Xin Fu, Yaou Duan, Viet A. N. Huu, Cindy Wen, Edward D. Zhang, Charlotte L. Zhang, Oulan Li, Xiaobo Wang, Michael A. Singer, Xiaodong Sun, Jie Xu, Ali Tafreshi, M. Anthony Lewis, Huimin Xia, and Kang Zhang · 2018
Earlier work this paper cites.
Chexpert: A large chest radiograph dataset with uncertainty labels and expert comparison
Jeremy Irvin, Pranav Rajpurkar, Michael Ko, Yifan Yu, Silviana Ciurea-Ilcus, Chris Chute, Henrik Marklund, Behzad Haghgoo, Robyn Ball, Katie Shpanskaya, et al · 2019
Earlier work this paper cites.
Mimic-cxr, a de-identified publicly available database of chest radiographs with free-text reports
Alistair EW Johnson, Tom J Pollard, Seth J Berkowitz, Nathaniel R Greenbaum, Matthew P Lungren, Chih-ying Deng, Roger G Mark, and Steven Horng · 2019
Earlier work this paper cites.
Lumbar Spine MRI Dataset, April 2019
Sud Sudirman, Ala Al Kafri, Friska Natalia, Hira Meidia, Nunik Afriliana, Wasfi Al-Rashdan, Mohammad Bashtawi, and Mohammed Al-Jumaily · 2019
Earlier work this paper cites.
Gpt-3: Its nature, scope, limits, and consequences
Luciano Floridi and Massimo Chiriatti · 2020
Earlier work this paper cites.
Autoprompt: Eliciting knowledge from language models with automatically generated prompts
T. Shin, Y. Razeghi, R. L. Logan IV, E. Wallace, and S. Singh · 2020
Earlier work this paper cites.
Medicat: A dataset of medical images, captions, and textual references
Sanjay Subramanian, Lucy Lu Wang, Sachin Mehta, Ben Bogin, Madeleine van Zuylen, Sravanthi Parasa, Sameer Singh, Matt Gardner, and Hannaneh Hajishirzi · 2020
Earlier work this paper cites.
COVID-19 and common pneumonia chest CT dataset
Jackie Yan · 2020
Earlier work this paper cites.
Performance of a deep neural network algorithm based on a small medical image dataset: incremental impact of 3d-to-2d reformation combined with novel data augmentation, photometric conversion, or transfer learning
Vikash Gupta, Mutlu Demirer, Matthew Bigelow, Kevin J Little, Sema Candemir, Luciano M Prevedello, Richard D White, Thomas P O’Donnell, Michael Wels, and Barbaros S Erdal · 2020
Earlier work this paper cites.
PadChest: A large chest x-ray image dataset with multi-label annotated reports
Aurelia Bustos, Antonio Pertusa, Jose-Maria Salinas, and Maria de la Iglesia-Vayá · 2020
Earlier work this paper cites.
Dataset of breast ultrasound images
Walid Al-Dhabyani, Mohammed Gomaa, Hussien Khaled, and Aly Fahmy · 2020
Earlier work this paper cites.
Dataset of Breast mammography images with Masses
Ting-Yu Lin and Mei-Ling Huang · 2020
Earlier work this paper cites.
Automatic segmentation of multiple cardiovascular structures from cardiac computed tomography angiography images using deep learning
Lohendran Baskaran, Subhi J Al’Aref, Gabriel Maliakal, Benjamin C Lee, Zhuoran Xu, Jeong W Choi, Sang-Eun Lee, Ji Min Sung, Fay Y Lin, Simon Dunham, et al · 2020
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Earlier work this paper cites.
Zero-shot text-to-image generation
Aditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray, Chelsea Voss, Alec Radford, Mark Chen, and Ilya Sutskever · 2021
Earlier work this paper cites.
Gpt-3: What’s it good for?
Robert Dale · 2021
Earlier work this paper cites.
Extracting training data from large language models
N. Carlini, F. Tramer, E. Wallace, M. Jagielski, A. Herbert-Voss, K. Lee, A. Roberts, T. Brown, D. Song, and U. Erlingsson · 2021
Earlier work this paper cites.
On the dangers of stochastic parrots: Can language models be too big?
Emily M. Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell · 2021
Earlier work this paper cites.
Preventing language models from learning sensitive information
Sean McGregor et al · 2021
Earlier work this paper cites.
Towards visual question answering on pathology images
Xuehai He · 2021
Earlier work this paper cites.
Slake: A semantically-labeled knowledge-enhanced dataset for medical visual question answering
Bo Liu, Li-Ming Zhan, Li Xu, Lin Ma, Yan Yang, and Xiao-Ming Wu · 2021
Earlier work this paper cites.
Brain tumor mri dataset
Msoud Nickparvar · 2021
Earlier work this paper cites.
Automatic detection of 39 fundus diseases and conditions in retinal photographs using deep neural networks
Ling-Ping Cen, Jie Ji, Jian-Wei Lin, Si-Tong Ju, Hong-Jie Lin, Tai-Ping Li, Yun Wang, Jian-Feng Yang, Yu-Fen Liu, Shaoying Tan, Li Tan, Dongjie Li, Yifan Wang, Dezhi Zheng, Yongqun Xiong, Hanfu Wu, Jingjing Jiang, Zhenggen Wu, Dingguo Huang, Tingkun Shi, Binyao Chen, Jianling Yang, Xiaoling Zhang, Li Luo, Chukai Huang, Guihua Zhang, Yuqiang Huang, Tsz Kin Ng, Haoyu Chen, Weiqi Chen, Chi Pui Pang, and Mingzhi Zhang · 2021
Earlier work this paper cites.
Data contamination: From memorization to exploitation
Inbal Magar and Roy Schwartz · 2022
Earlier work this paper cites.
Tufts Dental Database: A Multimodal Panoramic X-Ray Dataset for Benchmarking Diagnostic Systems
Karen Panetta, Rahul Rajendran, Aruna Ramesh, Shishir Paramathma Rao, and Sos Agaian · 2022
Earlier work this paper cites.
Fracture Detection in Wrist X-ray Images Using Deep Learning-Based Object Detection Models
Fırat Hardalaç, Fatih Uysal, Ozan Peker, Murat Çiçeklidağ, Tolga Tolunay, Nil Tokgöz, Uğurhan Kutbay, Boran Demirciler, and Fatih Mert · 2022
Cited alongside, same era.
Dataset for fetus framework, 9 2022
Chen Cui and Fajin Dong · 2022
Cited alongside, same era.
Common carotid artery ultrasound images, 2022
Agata Momot · 2022
Cited alongside, same era.
Alzheimer mri preprocessed dataset
Sachin Kumar and Sourabh Shastri · 2022
Cited alongside, same era.
Foundation models for generalist medical artificial intelligence
Michael Moor, Oishi Banerjee, Zahra Shakeri Hossein Abad, Harlan M Krumholz, Jure Leskovec, Eric J Topol, and Pranav Rajpurkar · 2023
Cited alongside, same era.
Large language models in medicine
Arun James Thirunavukarasu, Darren Shu Jeng Ting, Kabilan Elangovan, Laura Gutierrez, Ting Fang Tan, and Daniel Shu Wei Ting · 2023
Pixel-wise wireless capsule endoscopy image annotated dataset for clear and contaminated region segmentation, 12 2023
Vahid Sadeghi, Alireza Mehridehnavi, Yasaman Sanahmadi, and Mohsen Sharifi · 2023
Later among the works it cites.
Imagecas: A large-scale dataset and benchmark for coronary artery segmentation based on computed tomography angiography images
An Zeng, Chunbiao Wu, Guisen Lin, Wen Xie, Jin Hong, Meiping Huang, Jian Zhuang, Shanshan Bi, Dan Pan, Najeeb Ullah, Kaleem Nawaz Khan, Tianchen Wang, Yiyu Shi, Xiaomeng Li, and Xiaowei Xu · 2023
Later among the works it cites.
Non-contrast cardiac ct images dataset with coronary artery calcium scoring, 1 2023
Ali Kazemi, Ahmad Keshtkar, Saeid Rashidi, Naser Aslanabadi, Behrouz Khodadad, and Mahdad Esmaeili · 2023
Later among the works it cites.
Towards generalist biomedical ai
Tao Tu, Shekoofeh Azizi, Danny Driess, Mike Schaekermann, Mohamed Amin, Pi-Chuan Chang, Andrew Carroll, Charles Lau, Ryutaro Tanno, Ira Ktena, et al · 2024
Closest in time.
A liver cancer question-answering system based on next-generation intelligence and the large model med-palm 2
Jili Qian, Zhengyu Jin, Quan Zhang, Guoqing Cai, and Beichang Liu · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Large language models encode clinical knowledge
Karan Singhal, Shekoofeh Azizi, Tao Tu, S Sara Mahdavi, Jason Wei, Hyung Won Chung, Nathan Scales, Ajay Tanwani, Heather Cole-Lewis, Stephen Pfohl, et al · 2023
Cited alongside, same era.
Towards expert-level medical question answering with large language models
Karan Singhal, Tao Tu, Juraj Gottweis, Rory Sayres, Ellery Wulczyn, Le Hou, Kevin Clark, Stephen Pfohl, Heather Cole-Lewis, Darlene Neal, et al · 2023
Cited alongside, same era.
Benchmarking medical large language models
Sadra Bakhshandeh · 2023
Cited alongside, same era.
A medical multimodal large language model for future pandemics
Fenglin Liu, Tingting Zhu, Xian Wu, Bang Yang, Chenyu You, Chenyang Wang, Lei Lu, Zhangdaihong Liu, Yefeng Zheng, Xu Sun, et al · 2023
Cited alongside, same era.
Cxr-llava: Multimodal large language model for interpreting chest x-ray images
Seowoo Lee, Jiwon Youn, Mansu Kim, and Soon Ho Yoon · 2023
Cited alongside, same era.
Pmc-vqa: Visual instruction tuning for medical visual question answering
Xiaoman Zhang, Chaoyi Wu, Ziheng Zhao, Weixiong Lin, Ya Zhang, Yanfeng Wang, and Weidi Xie · 2023
Cited alongside, same era.
Closest in time.
M3d: Advancing 3d medical image analysis with multi-modal large language models
Fan Bai, Yuxin Du, Tiejun Huang, Max Q-H Meng, and Bo Zhao · 2024
Closest in time.
A survey on large language model (llm) security and privacy: The good, the bad, and the ugly
Yifan Yao, Jinhao Duan, Kaidi Xu, Yuanfang Cai, Zhibo Sun, and Yue Zhang · 2024
Closest in time.
Mm-safetybench: A benchmark for safety evaluation of multimodal large language models, 2024
Xin Liu, Yichen Zhu, Jindong Gu, Yunshi Lan, Chao Yang, and Yu Qiao · 2024
Closest in time.
Benchmarking trustworthiness of multimodal large language models: A comprehensive study
Yichi Zhang, Yao Huang, Yitong Sun, Chang Liu, Zhe Zhao, Zhengwei Fang, Yifan Wang, Huanran Chen, Xiao Yang, Xingxing Wei, et al · 2024
Closest in time.
Gpt-4v (ision) is a human-aligned evaluator for text-to-3d generation
Tong Wu, Guandao Yang, Zhibing Li, Kai Zhang, Ziwei Liu, Leonidas Guibas, Dahua Lin, and Gordon Wetzstein · 2024
Closest in time.
Cogvlm: Visual expert for pretrained language models, 2024
Weihan Wang, Qingsong Lv, Wenmeng Yu, Wenyi Hong, Ji Qi, Yan Wang, Junhui Ji, Zhuoyi Yang, Lei Zhao, Xixuan Song, Jiazheng Xu, Bin Xu, Juanzi Li, Yuxiao Dong, Ming Ding, and Jie Tang · 2024
Closest in time.
Llava-phi: Efficient multi-modal assistant with small language model
Yichen Zhu, Minjie Zhu, Ning Liu, Zhicai Ou, Xiaofeng Mou, and Jian Tang · 2024
Closest in time.
Instructblip: Towards general-purpose vision-language models with instruction tuning
Wenliang Dai, Junnan Li, Dongxu Li, Anthony Meng Huat Tiong, Junqi Zhao, Weisheng Wang, Boyang Li, Pascale N Fung, and Steven Hoi · 2024
Closest in time.
What matters when building vision-language models?
Hugo Laurençon, Léo Tronchon, Matthieu Cord, and Victor Sanh · 2024
Closest in time.
Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality, March 2023
The Vicuna Team · 2024
Closest in time.
Llava-med: Training a large language-and-vision assistant for biomedicine in one day
Chunyuan Li, Cliff Wong, Sheng Zhang, Naoto Usuyama, Haotian Liu, Jianwei Yang, Tristan Naumann, Hoifung Poon, and Jianfeng Gao · 2024
Closest in time.
Chexagent: Towards a foundation model for chest x-ray interpretation
Zhihong Chen, Maya Varma, Jean-Benoit Delbrouck, Magdalini Paschali, Louis Blankemeier, Dave Van Veen, Jeya Maria Jose Valanarasu, Alaa Youssef, Joseph Paul Cohen, Eduardo Pontes Reis, et al · 2024
Closest in time.
Jailbroken: How does llm safety training fail?
Alexander Wei, Nika Haghtalab, and Jacob Steinhardt · 2024
Closest in time.
Jailbreaking attack against multimodal large language model
Zhenxing Niu, Haodong Ren, Xinbo Gao, Gang Hua, and Rong Jin · 2024
Closest in time.
Visual adversarial examples jailbreak aligned large language models
Xiangyu Qi, Kaixuan Huang, Ashwinee Panda, Peter Henderson, Mengdi Wang, and Prateek Mittal · 2024
Closest in time.
Images are achilles’ heel of alignment: Exploiting visual vulnerabilities for jailbreaking multimodal large language models, 2024
Yifan Li, Hangyu Guo, Kun Zhou, Wayne Xin Zhao, and Ji-Rong Wen · 2024
Closest in time.
Weidi Luo, Siyuan Ma, Xiaogeng Liu, Xiaoyu Guo, and Chaowei Xiao · 2024
Closest in time.
Synth-empathy: Towards high-quality synthetic empathy data
Hao Liang, Linzhuang Sun, Jingxuan Wei, Xijie Huang, Linkun Sun, Bihui Yu, Conghui He, and Wentao Zhang · 2024
Closest in time.
Synthvlm: High-efficiency and high-quality synthetic data for vision language models
Zheng Liu, Hao Liang, Xijie Huang, Wentao Xiong, Qinhan Yu, Conghui He, Bin Cui, and Wentao Zhang · 2024
Closest in time.
Agfsync: Leveraging ai-generated feedback for preference optimization in text-to-image generation
Jingkun An, Yinghao Zhu, Zongjian Li, Haoran Feng, Xijie Huang, Bohua Chen, Yemin Shi, and Chengwei Pan · 2024
Closest in time.
Keyvideollm: Towards large-scale video keyframe selection
Hao Liang, Jiapeng Li, Tianyi Bai, Xijie Huang, Chong Chen, Conghui He, Bin Cui, and Wentao Zhang · 2024
Closest in time.
Bge m3-embedding: Multi-lingual, multi-functionality, multi-granularity text embeddings through self-knowledge distillation, 2024
Jianlv Chen, Shitao Xiao, Peitian Zhang, Kun Luo, Defu Lian, and Zheng Liu · 2024
Closest in time.
Strengthening multimodal large language model with bootstrapped preference optimization
Renjie Pi, Tianyang Han, Wei Xiong, Jipeng Zhang, Runtao Liu, Rui Pan, and Tong Zhang · 2024
Closest in time.
Siyuan Ma, Weidi Luo, Yu Wang, Xiaogeng Liu, Muhao Chen, Bo Li, and Chaowei Xiao · 2024
Closest in time.
An image is worth 1000 lies: Adversarial transferability across prompts on vision-language models
Haochen Luo, Jindong Gu, Fengyuan Liu, and Philip Torr · 2024
Closest in time.
On prompt-driven safeguarding for large language models
Chujie Zheng, Fan Yin, Hao Zhou, Fandong Meng, Jie Zhou, Kai-Wei Chang, Minlie Huang, and Nanyun Peng · 2024
Closest in time.
Brain stroke prediction ct scan image dataset
Noshin Tasnia · 2024
Closest in time.
Dermoscopy images
Sergio Tascón and Esteban Vaca · 2024
Closest in time.
OCTDL: Optical Coherence Tomography Dataset for Image-Based Deep Learning Methods, March 2024
Mikhail Kulyabin, Aleksei Zhdanov, Anastasia Nikiforova, Andrey Stepichev, Anna Kuznetsova, Vasilii Borisov, Mikhail Ronkin, Alexander Bogachev, Sergey Korotkich, and Andreas Maier · 2024
Closest in time.
Mmotu dataset
Lang Li · 2024
Closest in time.
Visualization tools for high resolution fundus dataset
Dataset Ninja · 2024
Closest in time.
Brain cancer mri object detection & segmentation dataset
TrainingData · 2024
Closest in time.