Fetching the paper…
Reading the bibliography…
Large Vision-Language Models (LVLMs) are capable of handling diverse data types such as imaging, text, and physiological signals, and can be applied in various fields.
The mammographic images analysis society digital mammogram database
John Suckling · 1994
Earlier work this paper cites.
Development of a digital image database for chest radiographs with and without a lung nodule: receiver operating characteristic analysis of radiologists’ detection of pulmonary nodules
Junji Shiraishi, Shigehiko Katsuragawa, Junpei Ikezoe, Tsuneo Matsumoto, Takeshi Kobayashi, Ken-ichi Komatsu, Mitate Matsui, Hiroshi Fujita, Yoshie Kodera, and Kunio Doi · 2000
Earlier work this paper cites.
Evaluation of three-dimensional finite element-based deformable registration of pre-and intraoperative prostate imaging
Aditya Bharatha, Masanori Hirose, Nobuhiko Hata, Simon K Warfield, Matthieu Ferrant, Kelly H Zou, Eduardo Suarez-Santana, Juan Ruiz-Alzola, Anthony D’amico, Robert A Cormack, et al · 2001
Earlier work this paper cites.
Ridge-based vessel segmentation in color images of the retina
Joes Staal, Michael D Abràmoff, Meindert Niemeijer, Max A Viergever, and Bram Van Ginneken · 2004
Earlier work this paper cites.
Pap-smear benchmark data for pattern classification
Jan Jantzen, Jonas Norup, Georgios Dounias, and Beth Bjerregaard · 2005
Earlier work this paper cites.
Evaluation and benchmark for biological image segmentation
Elisa Drelie Gelasca, Jiyun Byun, Boguslaw Obara, and BS Manjunath · 2008
Earlier work this paper cites.
Automatic recognition of corneal nerve structures in images from confocal microscopy
Fabio Scarpa, Enrico Grisan, and Alfredo Ruggeri · 2008
Earlier work this paper cites.
Comparison and evaluation of methods for liver segmentation from ct datasets
Tobias Heimann, Bram Van Ginneken, Martin A Styner, Yulia Arzhaeva, Volker Aurich, Christian Bauer, Andreas Beck, Christoph Becker, Reinhard Beichel, Gy"̈orgy Bekes, et al · 2009
Earlier work this paper cites.
Automatic classification of lymphoma images with transform-based global features
Nikita V Orlov, Wayne W Chen, David Mark Eckley, Tomasz J Macura, Lior Shamir, Elaine S Jaffe, and Ilya G Goldberg · 2010
Earlier work this paper cites.
Rim-one: An open retinal image database for optic nerve evaluation
Francisco Fumero, Silvia Alayón, José L Sanchez, Jose Sigut, and M Gonzalez-Hernandez · 2011
Earlier work this paper cites.
Automatic evaluation of corneal nerve tortuosity in images from in vivo confocal microscopy
Fabio Scarpa, Xiaodong Zheng, Yuichi Ohashi, and Alfredo Ruggeri · 2011
Earlier work this paper cites.
Automatic segmentation of closed-contour features in ophthalmic images using graph theory and dynamic programming
Stephanie J Chiu, Cynthia A Toth, Catherine Bowes Rickman, Joseph A Izatt, and Sina Farsiu · 2012
Earlier work this paper cites.
Clavicle segmentation in chest radiographs
Laurens Hogeweg, Clara I Sánchez, Pim A de Jong, Pragnya Maduskar, and Bram van Ginneken · 2012
Earlier work this paper cites.
Automated analysis of retinal images for detection of referable diabetic retinopathy
Michael D Abràmoff, James C Folk, Dennis P Han, Jonathan D Walker, David F Williams, Stephen R Russell, Pascale Massin, Beatrice Cochener, Philippe Gain, Li Tang, et al · 2013
Earlier work this paper cites.
Robust vessel segmentation in fundus images
Attila Budai, R"̈udiger Bock, Andreas Maier, Joachim Hornegger, and Georg Michelson · 2013
Earlier work this paper cites.
Automatic cone photoreceptor segmentation using graph theory and dynamic programming
Stephanie J Chiu, Yuliya Lokhnygina, Adam M Dubis, Alfredo Dubra, Joseph Carroll, Joseph A Izatt, and Sina Farsiu · 2013
Earlier work this paper cites.
Automated separation of binary overlapping trees in low-contrast color retinal images
Qiao Hu, Michael D Abràmoff, and Mona K Garvin · 2013
Earlier work this paper cites.
Ph 2-a dermoscopic image database for research and benchmarking
Teresa Mendonça, Pedro M Ferreira, Jorge S Marques, André RS Marcal, and Jorge Rozeira · 2013
Earlier work this paper cites.
An automated method for retinal arteriovenous nicking quantification from color fundus images
Uyen TV Nguyen, Alauddin Bhuiyan, Laurence AF Park, Ryo Kawasaki, Tien Y Wong, Jie Jin Wang, Paul Mitchell, and Kotagiri Ramamohanarao · 2013
Earlier work this paper cites.
Decoding tumour phenotype by noninvasive imaging using a quantitative radiomics approach
Hugo JWL Aerts, Emmanuel Rios Velazquez, Ralph TH Leijenaar, Chintan Parmar, Patrick Grossmann, Sara Carvalho, Johan Bussink, René Monshouwer, Benjamin Haibe-Kains, Derek Rietveld, et al · 2014
Earlier work this paper cites.
Tree topology estimation
Rolando Estrada, Carlo Tomasi, Scott C Schmidler, and Sina Farsiu · 2014
Earlier work this paper cites.
Texture characterization of bone radiograph images. application to osteoporosis diagnosis
Khaled Harrar · 2014
Earlier work this paper cites.
Two public chest x-ray datasets for computer-aided screening of pulmonary diseases
Stefan Jaeger, Sema Candemir, Sameer Antani, Yì-Xiáng J Wáng, Pu-Xuan Lu, and George Thoma · 2014
Earlier work this paper cites.
Evaluation of prostate segmentation algorithms for mri: the promise12 challenge
Geert Litjens, Robert Toth, Wendy Van De Ven, Caroline Hoeks, Sjoerd Kerkstra, Bram Van Ginneken, Graham Vincent, Gwenael Guillard, Neil Birbeck, Jindang Zhang, et al · 2014
Earlier work this paper cites.
Comparison of macular octs in right and left eyes of normal people
Tahereh Mahmudi, Rahele Kafieh, Hossein Rabbani, Mohammadreza Akhlagi, et al · 2014
Earlier work this paper cites.
Identification of suitable fundus images using automated quality assessment methods
Uğur Şevik, Cemal K"̈ose, Tolga Berber, and Hidayet Erd"̈ol · 2014
Earlier work this paper cites.
Wm-dova maps for accurate polyp highlighting in colonoscopy: Validation vs. saliency maps from physicians
Jorge Bernal, F Javier Sánchez, Gloria Fernández-Esparrach, Debora Gil, Cristina Rodríguez, and Fernando Vilariño · 2015
Earlier work this paper cites.
The multimodal brain tumor image segmentation benchmark (brats)
H Menze Bjoern, Jakab Andras, Bauer Stefan, Kalpathy-Cramer Jayashree, Farahani Keyvan, Kirby Justin, et al · 2015
Earlier work this paper cites.
Diabetic retinopathy detection, 2015
Emma Dugas, Jared, Jorge, and Will Cukierski · 2015
Earlier work this paper cites.
Med-node: A computer-assisted melanoma diagnosis system using non-dermoscopic images
Ioannis Giotis, Nynke Molders, Sander Land, Michael Biehl, Marcel F Jonkman, and Nicolai Petkov · 2015
Earlier work this paper cites.
Multi-site collection of lung ct data with nodule segmentations
K Jayashree and N Sandy · 2015
Earlier work this paper cites.
Miccai multi-atlas labeling beyond the cranial vault–workshop and challenge
Bennett Landman, Zhoubing Xu, J Igelsias, Martin Styner, T Langerak, and Arno Klein · 2015
Earlier work this paper cites.
An improved joint optimization of multiple level set functions for the segmentation of overlapping cervical cells
Zhi Lu, Gustavo Carneiro, and Andrew P Bradley · 2015
Earlier work this paper cites.
Interactive whole-heart segmentation in congenital heart disease
Danielle F Pace, Adrian V Dalca, Tal Geva, Andrew J Powell, Mehdi H Moghari, and Polina Golland · 2015
Earlier work this paper cites.
An open access thyroid ultrasound image database
Lina Pedraza, Carlos Vargas, Fabián Narváez, Oscar Durán, Emma Muñoz, and Eduardo Romero · 2015
Earlier work this paper cites.
Fully automatic segmentation of fluorescein leakage in subjects with diabetic macular edema
Hossein Rabbani, Michael J Allingham, Priyatham S Mettu, Scott W Cousins, and Sina Farsiu · 2015
Earlier work this paper cites.
A dataset for breast cancer histopathological image classification
Fabio A Spanhol, Luiz S Oliveira, Caroline Petitjean, and Laurent Heutte · 2015
Earlier work this paper cites.
Benchmark for algorithms segmenting the left atrium from 3d ct and mri datasets
Catalina Tobon-Gomez, Arjan J Geers, Jochen Peters, J"̈urgen Weese, Karen Pinto, Rashed Karim, Mohammed Ammar, Abdelaziz Daoudi, Jan Margeta, Zulma Sandoval, et al · 2015
Earlier work this paper cites.
David Gutman, Noel CF Codella, Emre Celebi, Brian Helba, Michael Marchetti, Nabin Mishra, and Allan Halpern · 2016
Earlier work this paper cites.
Evaluation of three algorithms for the segmentation of overlapping cervical cells
Zhi Lu, Gustavo Carneiro, Andrew P Bradley, Daniela Ushizima, Masoud S Nosrati, Andrea GC Bianchi, Claudia M Carneiro, and Ghassan Hamarneh · 2016
Earlier work this paper cites.
Ultrasound nerve segmentation, 2016
Anna Montoya, Hasnin, kaggle446, shirzad, Will Cukierski, and yffud · 2016
Earlier work this paper cites.
Texture descriptors ensembles enable image-based classification of maturation of human stem cell-derived retinal pigmented epithelium
Loris Nanni, Michelangelo Paci, Florentino Luciano Caetano dos Santos, Heli Skottman, Kati Juuti-Uusitalo, and Jari Hyttinen · 2016
Earlier work this paper cites.
Endonet: a deep architecture for recognition tasks on laparoscopic videos
Andru P Twinanda, Sherif Shehata, Didier Mutter, Jacques Marescaux, Michel De Mathelin, and Nicolas Padoy · 2016
Earlier work this paper cites.
Radbench: benchmarking image interpretation skills
Chris Wright and Pauline Reeves · 2016
Earlier work this paper cites.
Segmentation labels and radiomic features for the pre-operative scans of the tcga-lgg collection
Spyridon Bakas, Hamed Akbari, Aristeidis Sotiras, Michel Bilello, Martin Rozycki, Justin Kirby, John Freymann, Keyvan Farahani, and Christos Davatzikos · 2017
Earlier work this paper cites.
Advancing the cancer genome atlas glioma mri collections with expert segmentation labels and radiomic features
Spyridon Bakas, Hamed Akbari, Aristeidis Sotiras, Michel Bilello, Martin Rozycki, Justin S Kirby, John B Freymann, Keyvan Farahani, and Christos Davatzikos · 2017
Earlier work this paper cites.
Intel & mobileodt cervical cancer screening, 2017
jljones BenO, Kumar H, Meg Risdal, Vadim Sherman MRao, Wendy Kan Vipul, and Yau Ben-Or · 2017
Earlier work this paper cites.
Comparative validation of polyp detection methods in video colonoscopy: results from the miccai 2015 endoscopic vision challenge
Jorge Bernal, Nima Tajkbaksh, Francisco Javier Sanchez, Bogdan J Matuszewski, Hao Chen, Lequan Yu, Quentin Angermann, Olivier Romain, Bjørn Rustad, Ilangko Balasingham, et al · 2017
Earlier work this paper cites.
Longitudinal multiple sclerosis lesion segmentation: resource and challenge
Aaron Carass, Snehashis Roy, Amod Jog, Jennifer L Cuzzocreo, Elizabeth Magrath, Adrian Gherman, Julia Button, James Nguyen, Ferran Prados, Carole H Sudre, et al · 2017
Earlier work this paper cites.
Long and short survival in adenocarcinoma lung cts
HL Goldgof Dmitry, Hawkins Samuel, Schabath Matthew, Stringfield Olya, Garcia Alberto, Balagurunathan Yoganand, Kim Jongphil, Eschrich Steven, Berglund Anders, Gatenby Robert, et al · 2017
Earlier work this paper cites.
Fire: fundus image registration dataset
Carlos Hernandez-Matas, Xenophon Zabulis, Areti Triantafyllou, Panagiota Anyfanti, Stella Douma, and Antonis A Argyros · 2017
Earlier work this paper cites.
Nih clinical center provides one of the largest publicly available chest x-ray datasets to scientific community, 2017
National Institutes of Health et al · 2017
Earlier work this paper cites.
Kvasir: A multi-class image dataset for computer aided gastrointestinal disease detection
Konstantin Pogorelov, Kristin Ranheim Randel, Carsten Griwodz, Sigrun Losada Eskeland, Thomas de Lange, Dag Johansen, Concetto Spampinato, Duc-Tien Dang-Nguyen, Mathias Lux, Peter Thelin Schmidt, et al · 2017
Earlier work this paper cites.
Evaluation of segmentation methods on head and neck ct: auto-segmentation challenge 2015
Patrik F Raudaschl, Paolo Zaffino, Gregory C Sharp, Maria Francesca Spadea, Antong Chen, Benoit M Dawant, Thomas Albrecht, Tobias Gass, Christoph Langguth, Marcel L"̈uthi, et al · 2017
Earlier work this paper cites.
Segmentation labels for the pre-operative scans of the tcga-gbm collection [data set]
Bakas S, Akbari H, Sotiras A, Bilello M, Rozycki M, Kirby J, Freymann J, Farahani K, and Davatzikos C · 2017
Earlier work this paper cites.
Validation, comparison, and combination of algorithms for automatic detection of pulmonary nodules in computed tomography images: the luna16 challenge
Arnaud Arindra Adiyoso Setio, Alberto Traverso, Thomas De Bel, Moira SN Berens, Cas Van Den Bogaard, Piergiorgio Cerello, Hao Chen, Qi Dou, Maria Evelina Fantacci, Bram Geurts, et al · 2017
Earlier work this paper cites.
Gland segmentation in colon histology images: The glas challenge contest
Korsuk Sirinukunwattana, Josien PW Pluim, Hao Chen, Xiaojuan Qi, Pheng-Ann Heng, Yun Bo Guo, Li Yang Wang, Bogdan J Matuszewski, Elia Bruni, Urko Sanchez, et al · 2017
Earlier work this paper cites.
Adversarial discriminative domain adaptation
Eric Tzeng, Judy Hoffman, Kate Saenko, and Trevor Darrell · 2017
Earlier work this paper cites.
Chestx-ray8: Hospital-scale chest x-ray database and benchmarks on weakly-supervised classification and localization of common thorax diseases
Xiaosong Wang, Yifan Peng, Le Lu, Zhiyong Lu, Mohammadhadi Bagheri, and Ronald M Summers · 2017
Earlier work this paper cites.
Spyridon Bakas, Mauricio Reyes, Andras Jakab, Stefan Bauer, Markus Rempfler, Alessandro Crimi, Russell Takeshi Shinohara, Christoph Berger, Sung Min Ha, Martin Rozycki, et al · 2018
Earlier work this paper cites.
Deep learning techniques for automatic mri cardiac multi-structures segmentation and diagnosis: is the problem solved?
Olivier Bernard, Alain Lalande, Clement Zotti, Frederick Cervenansky, Xin Yang, Pheng-Ann Heng, Irem Cetin, Karim Lekadir, Oscar Camara, Miguel Angel Gonzalez Ballester, et al · 2018
Earlier work this paper cites.
Knee osteoarthritis severity grading dataset
Pingjun Chen · 2018
Earlier work this paper cites.
Data from qin-prostate-repeatability
A Fedorov, M Schwier, D Clunie, C Herz, S Pieper, R Kikinis, C Tempany, and F Fennessy · 2018
Earlier work this paper cites.
Joint optic disc and cup segmentation based on multi-label deep network and polar transformation
Huazhu Fu, Jun Cheng, Yanwu Xu, Damon Wing Kee Wong, Jiang Liu, and Xiaochun Cao · 2018
Earlier work this paper cites.
Pupil localization using geodesic distance
Radovan Fusek · 2018
Earlier work this paper cites.
Tool detection and operative skill assessment in surgical videos using region-based convolutional neural networks
Amy Jin, Serena Yeung, Jeffrey Jopling, Jonathan Krause, Dan Azagury, Arnold Milstein, and Li Fei-Fei · 2018
Earlier work this paper cites.
Algorithms for left atrial wall segmentation and thickness–evaluation on an open-source ct and mri image database
Rashed Karim, Lauren-Emma Blake, Jiro Inoue, Qian Tao, Shuman Jia, R James Housden, Pranav Bhagirath, Jean-Luc Duval, Marta Varela, Jonathan M Behar, et al · 2018
Earlier work this paper cites.
100,000 histological images of human colorectal cancer and healthy tissue
Jakob Nikolas Kather, Niels Halama, and Alexander Marx · 2018
Earlier work this paper cites.
Seven-point checklist and skin lesion classification using multitask multimodal neural nets
Jeremy Kawahara, Sara Daneshvar, Giuseppe Argenziano, and Ghassan Hamarneh · 2018
Earlier work this paper cites.
Labeled optical coherence tomography (oct) and chest x-ray images for classification
Daniel Kermany, Kang Zhang, Michael Goldbaum, et al · 2018
Earlier work this paper cites.
A dataset of clinically generated visual questions and answers about radiology images
Jason J Lau, Soumya Gayen, Asma Ben Abacha, and Dina Demner-Fushman · 2018
Earlier work this paper cites.
A new dataset of computed-tomography angiography images for computer-aided detection of pulmonary embolism
Mojtaba Masoudi, Hamid-Reza Pourreza, Mahdi Saadatmand-Tarzjan, Noushin Eftekhari, Fateme Shafiee Zargar, and Masoud Pezeshki Rad · 2018
Earlier work this paper cites.
A new cervical cytology dataset for nucleus detection and image classification (cervix93) and methods for cervical nucleus detection, 2018
Hady Ahmady Phoulady and Peter R. Mouton · 2018
Earlier work this paper cites.
Indian diabetic retinopathy image dataset (idrid): a database for diabetic retinopathy screening research
Prasanna Porwal, Samiksha Pachade, Ravi Kamble, Manesh Kokare, Girish Deshmukh, Vivek Sahasrabuddhe, and Fabrice Meriaudeau · 2018
Earlier work this paper cites.
Data from brain-tumor-progression
Prah M Schmainda KM · 2018
Earlier work this paper cites.
The ham10000 dataset, a large collection of multi-source dermatoscopic images of common pigmented skin lesions
Philipp Tschandl, Cliff Rosendahl, and Harald Kittler · 2018
Earlier work this paper cites.
You only look on lymphocytes once
Mart van Rijthoven, Zaneta Swiderska-Chadaj, Katja Seeliger, Jeroen van der Laak, and Francesco Ciompi · 2018
Earlier work this paper cites.
2017 robotic instrument segmentation challenge
Max Allan, Alex Shvets, Thomas Kurmann, Zichen Zhang, Rahul Duggal, Yun-Hsuan Su, Nicola Rieke, Iro Laina, Niveditha Kalavakonda, Sebastian Bodenstedt, et al · 2019
Earlier work this paper cites.
Structured crowdsourcing enables convolutional segmentation of histology images
Mohamed Amgad, Habiba Elfandy, Hagar Hussein, Lamees A Atteya, Mai AT Elsebaie, Lamia S Abo Elnasr, Rokia A Sakr, Hazem SE Salem, Ahmed F Ismail, Anas M Saad, et al · 2019
Earlier work this paper cites.
Bach: Grand challenge on breast cancer histology images
Guilherme Aresta, Teresa Araújo, Scotty Kwok, Sai Saketh Chennamsetty, Mohammed Safwan, Varghese Alex, Bahram Marami, Marcel Prastawa, Monica Chan, Michael Donovan, et al · 2019
Earlier work this paper cites.
Lung and colon cancer histopathological image dataset (lc25000)
Andrew A Borkowski, Marilyn M Bui, L Brannon Thomas, Catherine P Wilson, Lauren A DeLand, and Stephen M Mastorides · 2019
Earlier work this paper cites.
Clinical-grade computational pathology using weakly supervised deep learning on whole slide images
Gabriele Campanella, Matthew G Hanna, Luke Geneslaw, Allen Miraflor, Vitor Werneck Krauss Silva, Klaus J Busam, Edi Brogi, Victor E Reuter, David S Klimstra, and Thomas J Fuchs · 2019
Earlier work this paper cites.
Data from aapm rt-mac grand challenge 2019
C. Cardenas, A. Mohamed, G. Sharp, M. Gooding, H. Veeraraghavan, and J. & Yang · 2019
Earlier work this paper cites.
Cnns for automatic glaucoma assessment using fundus images: an extensive validation
Andres Diaz-Pinto, Sandra Morales, Valery Naranjo, Thomas K"̈ohler, Jose M Mossi, and Amparo Navea · 2019
Earlier work this paper cites.
Pannuke: an open pan-cancer histology dataset for nuclei instance segmentation and classification
Jevgenij Gamper, Navid Alemi Koohbanani, Ksenija Benet, Ali Khuram, and Nasir Rajpoot · 2019
Earlier work this paper cites.
Mild-net: Minimal information loss dilated network for gland instance segmentation in colon histology images
Simon Graham, Hao Chen, Jevgenij Gamper, Qi Dou, Pheng-Ann Heng, David Snead, Yee Wah Tsang, and Nasir Rajpoot · 2019
Earlier work this paper cites.
Ivdm3seg - miccai 2018 challenge intervertebral disc localization and segmentation from 3d multi-modality mr (m3) images (2018)
Daniel Belavy Guoyan Zheng, Shuo Li · 2019
Cited alongside, same era.
Isbi 2019 c-nmc challenge: Classification in cancer cell imaging
Anubha Gupta and Ritu Gupta · 2019
Cited alongside, same era.
The rsna pediatric bone age machine learning challenge
Safwan S Halabi, Luciano M Prevedello, Jayashree Kalpathy-Cramer, Artem B Mamonov, Alexander Bilbily, Mark Cicero, Ian Pan, Lucas Araújo Pereira, Rafael Teixeira Sousa, Nitamar Abdala, et al · 2019
Cited alongside, same era.
Nicholas Heller, Niranjan Sathianathen, Arveen Kalapara, Edward Walczak, Keenan Moore, Heather Kaluzniak, Joel Rosenberg, Paul Blake, Zachary Rengel, Makinna Oestreich, et al · 2019
Cited alongside, same era.
Palm: pathologic myopia challenge. comput. vis. med
F Huazhu, L Fei, and IO José · 2019
The university of california san francisco preoperative diffuse glioma mri dataset
Evan Calabrese, Javier E Villanueva-Meyer, Jeffrey D Rudie, Andreas M Rauschecker, Ujjwal Baid, Spyridon Bakas, Soonmee Cha, John T Mongan, and Christopher P Hess · 2022
Later among the works it cites.
Expert tumor annotations and radiomic features for the ispy1/acrin 6657 trial data collection
R Chitalia et al · 2022
Later among the works it cites.
Digestpath: A benchmark dataset with challenge review for the pathological detection and segmentation of digestive-system
Qian Da, Xiaodi Huang, Zhongyu Li, Yanfei Zuo, Chenbin Zhang, Jingxin Liu, Wen Chen, Jiahui Li, Dou Xu, Zhiqiang Hu, et al · 2022
Later among the works it cites.
Adam challenge: Detecting age-related macular degeneration from fundus images
Huihui Fang, Fei Li, Huazhu Fu, Xu Sun, Xingxing Cao, Fengbin Lin, Jaemin Son, Sunho Kim, Gwenole Quellec, Sarah Matta, et al · 2022
Later among the works it cites.
Dataset and evaluation algorithm design for goals challenge
Huihui Fang, Fei Li, Huazhu Fu, Junde Wu, Xiulan Zhang, and Yanwu Xu · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
A novel deep learning method for automatic assessment of human sperm images
Soroush Javadi and Seyed Abolghasem Mirroshandel · 2019
Cited alongside, same era.
Mimic-cxr-jpg, a large publicly available database of labeled chest radiographs
Alistair EW Johnson, Tom J Pollard, Nathaniel R Greenbaum, Matthew P Lungren, Chih-ying Deng, Yifan Peng, Zhiyong Lu, Roger G Mark, Seth J Berkowitz, and Steven Horng · 2019
Cited alongside, same era.
Aptos 2019 blindness detection, 2019
Sohier Dane Karthik, Maggie · 2019
Cited alongside, same era.
A self-adaptive deep learning method for automated eye laterality detection based on color fundus photography
Chi Liu, Xiaotong Han, Zhixi Li, Jason Ha, Guankai Peng, Wei Meng, and Mingguang He · 2019
Cited alongside, same era.
C_nmc_2019 dataset: All challenge dataset of isbi 2019
S. Mourya, S. Kant, P. Kumar, A. Gupta, and R Gupta · 2019
Cited alongside, same era.
Lndb: a lung nodule database on computed tomography
João Pedrosa, Guilherme Aresta, Carlos Ferreira, Márcio Rodrigues, Patrícia Leitão, André Silva Carvalho, João Rebelo, Eduardo Negrão, Isabel Ramos, António Cunha, et al · 2019
Cited alongside, same era.
Amber L Simpson, Michela Antonelli, Spyridon Bakas, Michel Bilello, Keyvan Farahani, Bram Van Ginneken, Annette Kopp-Schneider, Bennett A Landman, Geert Litjens, Bjoern Menze, et al · 2019
Cited alongside, same era.
Later among the works it cites.
Uw-madison gi tract image segmentation, 2022
happyharrycn, Maggie, Phil Culliton, Poonam Yadav, and Sangjune Laurence Lee · 2022
Later among the works it cites.
Ravir: A dataset and methodology for the semantic segmentation and quantitative analysis of retinal arteries and veins in infrared reflectance imaging
Ali Hatamizadeh, Hamid Hosseini, Niraj Patel, Jinseo Choi, Cameron C Pole, Cory M Hoeferlin, Steven D Schwartz, and Demetri Terzopoulos · 2022
Later among the works it cites.
A web-scraped skin image database of monkeypox, chickenpox, smallpox, cowpox, and measles
Towhidul Islam, Mohammad Arafat Hussain, Forhad Uddin Hasan Chowdhury, and BM Riazul Islam · 2022
Later among the works it cites.
Amos: A large-scale abdominal multi-organ benchmark for versatile medical image segmentation
Yuanfeng Ji, Haotian Bai, Chongjian Ge, Jie Yang, Ye Zhu, Ruimao Zhang, Zhen Li, Lingyan Zhanng, Wanling Ma, Xiang Wan, et al · 2022
Later among the works it cites.
Deep learning methods for automatic evaluation of delayed enhancement-mri. the results of the emidec challenge
Alain Lalande, Zhihao Chen, Thibaut Pommier, Thomas Decourselle, Abdul Qayyum, Michel Salomon, Dominique Ginhac, Youssef Skandarani, Arnaud Boucher, Khawla Brahim, et al · 2022
Later among the works it cites.
Deepdrid: Diabetic retinopathy—grading and image quality estimation challenge
Ruhan Liu, Xiangning Wang, Qiang Wu, Ling Dai, Xi Fang, Tao Yan, Jaemin Son, Shiqi Tang, Jiang Li, Zijian Gao, Adrian Galdran, J.M. Poorneshwaran, Hao Liu, Jie Wang, Yerui Chen, Prasanna Porwal, Gavin Siew Wei Tan, Xiaokang Yang, Chao Dai, Haitao Song, Mingang Chen, Huating Li, Weiping Jia, Dinggang Shen, Bin Sheng, and Ping Zhang · 2022
Later among the works it cites.
X-metric: An n-dimensional information-theoretic framework for groupwise registration and deep combined computing
Xinzhe Luo and Xiahai Zhuang · 2022
Later among the works it cites.
Fast and low-gpu-memory abdomen ct organ segmentation: the flare challenge
Jun Ma, Yao Zhang, Song Gu, Xingle An, Zhihe Wang, Cheng Ge, Congcong Wang, Fan Zhang, Yu Wang, Yinan Xu, et al · 2022
Later among the works it cites.
Radimagenet: an open radiologic deep learning research dataset for effective transfer learning
Xueyan Mei, Zelong Liu, Philip M Robson, Brett Marinelli, Mingqian Huang, Amish Doshi, Adam Jacobi, Chendi Cao, Katherine E Link, Thomas Yang, et al · 2022
Later among the works it cites.
Olives dataset: Ophthalmic labels for investigating visual eye semantics
Mohit Prabhushankar, Kiran Kokilepersaud, Yash-yee Logan, Stephanie Trejo Corona, Ghassan AlRegib, and Charles Wykoff · 2022
Later among the works it cites.
Avt: Multicenter aortic vessel tree cta dataset collection with ground truth segmentation masks
Lukas Radl, Yuan Jin, Antonio Pepe, Jianning Li, Christina Gsaxner, Fen-hua Zhao, and Jan Egger · 2022
Later among the works it cites.
Rapid artificial intelligence solutions in a pandemic—the covid-19-20 lung ct lesion segmentation challenge
Holger R Roth, Ziyue Xu, Carlos Tor-Díez, Ramon Sanchez Jacob, Jonathan Zember, Jose Molto, Wenqi Li, Sheng Xu, Baris Turkbey, Evrim Turkbey, et al · 2022
Later among the works it cites.
Surgical-vqa: Visual question answering in surgical scenes using transformer
Lalithkumar Seenivasan, Mobarakol Islam, Adithya K Krishna, and Hongliang Ren · 2022
Later among the works it cites.
The extreme cardiac mri analysis challenge under respiratory motion (cmrxmotion)
Shuo Wang, Chen Qin, Chengyan Wang, Kang Wang, Haoran Wang, Chen Chen, Cheng Ouyang, Xutong Kuang, Chengliang Dai, Yuanhan Mo, et al · 2022
Later among the works it cites.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al · 2023
Later among the works it cites.
Openflamingo: An open-source framework for training large autoregressive vision-language models
Anas Awadalla, Irena Gao, Josh Gardner, Jack Hessel, Yusuf Hanafy, Wanrong Zhu, Kalyani Marathe, Yonatan Bitton, Samir Gadre, Shiori Sagawa, Jenia Jitsev, Simon Kornblith, Pang Wei Koh, Gabriel Ilharco, Mitchell Wortsman, and Ludwig Schmidt · 2023
Later among the works it cites.
Qwen-vl: A versatile vision-language model for understanding, localization, text reading, and beyond
Jinze Bai, Shuai Bai, Shusheng Yang, Shijie Wang, Sinan Tan, Peng Wang, Junyang Lin, Chang Zhou, and Jingren Zhou · 2023
Later among the works it cites.
Qwen-vl: A versatile vision-language model for understanding, localization, text reading, and beyond, 2023
Jinze Bai, Shuai Bai, Shusheng Yang, Shijie Wang, Sinan Tan, Peng Wang, Junyang Lin, Chang Zhou, and Jingren Zhou · 2023
Later among the works it cites.
The liver tumor segmentation benchmark (lits)
Patrick Bilic, Patrick Christ, Hongwei Bran Li, Eugene Vorontsov, Avi Ben-Cohen, Georgios Kaissis, Adi Szeskin, Colin Jacobs, Gabriel Efrain Humpire Mamani, Gabriel Chartrand, et al · 2023
Later among the works it cites.
The río hortega university hospital glioblastoma dataset: A comprehensive collection of preoperative, early postoperative and recurrence mri scans (rhuh-gbm)
Santiago Cepeda, Sergio García-García, Ignacio Arrese, Francisco Herrero, Trinidad Escudero, Tomás Zamora, and Rosario Sarabia · 2023
Later among the works it cites.
Sharegpt4v: Improving large multi-modal models with better captions, 2023
Lin Chen, Jinsong Li, Xiaoyi Dong, Pan Zhang, Conghui He, Jiaqi Wang, Feng Zhao, and Dahua Lin · 2023
Later among the works it cites.
Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks
Zhe Chen, Jiannan Wu, Wenhai Wang, Weijie Su, Guo Chen, Sen Xing, Muyan Zhong, Qinglong Zhang, Xizhou Zhu, Lewei Lu, Bin Li, Ping Luo, Tong Lu, Yu Qiao, and Jifeng Dai · 2023
Later among the works it cites.
Xtuner: A toolkit for efficiently fine-tuning llm
XTuner Contributors · 2023
Later among the works it cites.
Airogs: artificial intelligence for robust glaucoma screening challenge
Coen De Vente, Koenraad A Vermeer, Nicolas Jaccard, He Wang, Hongyi Sun, Firas Khader, Daniel Truhn, Temirgali Aimyshev, Yerkebulan Zhanibekuly, Tien-Dung Le, et al · 2023
Later among the works it cites.
Expert knowledge-aware image difference graph representation learning for difference-aware medical visual question answering
Xinyue Hu, Lin Gu, Qiyuan An, Mengliang Zhang, Liangchen Liu, Kazuma Kobayashi, Tatsuya Harada, Ronald M Summers, and Yingying Zhu · 2023
Later among the works it cites.
Stress-testing pelvic autosegmentation algorithms using anatomical edge cases
Aasheesh Kanwar, Brandon Merz, Cheryl Claunch, Shushan Rana, Arthur Hung, and Reid F Thompson · 2023
Later among the works it cites.
Segment anything
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloe Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C Berg, Wan-Yen Lo, et al · 2023
Later among the works it cites.
Obelics: An open web-scale filtered dataset of interleaved image-text documents, 2023
Hugo Laurençon, Lucile Saulnier, Léo Tronchon, Stas Bekman, Amanpreet Singh, Anton Lozhkov, Thomas Wang, Siddharth Karamcheti, Alexander M. Rush, Douwe Kiela, Matthieu Cord, and Victor Sanh · 2023
Later among the works it cites.
Visual instruction tuning, 2023
Haotian Liu, Chunyuan Li, Qingyang Wu, and Yong Jae Lee · 2023
Later among the works it cites.
Qilin-med-vl: Towards chinese large vision-language model for general healthcare
Junling Liu, Ziming Wang, Qichen Ye, Dading Chong, Peilin Zhou, and Yining Hua · 2023
Later among the works it cites.
Efficient automatic segmentation for multi-level pulmonary arteries: The parse challenge
Gongning Luo, Kuanquan Wang, Jun Liu, Shuo Li, Xinjie Liang, Xiangyu Li, Shaowei Gan, Wei Wang, Suyu Dong, Wenyi Wang, et al · 2023
Later among the works it cites.
Harvard glaucoma detection and progression: A multimodal multitask dataset and generalization-reinforced semi-supervised learning, 2023
Yan Luo, Min Shi, Yu Tian, Tobias Elze, and Mengyu Wang · 2023
Later among the works it cites.
Jun Ma, Yao Zhang, Song Gu, Cheng Ge, Shihao Ma, Adamo Young, Cheng Zhu, Kangkang Meng, Xin Yang, Ziyan Huang, et al · 2023
Later among the works it cites.
Deep learning segmentation of the right ventricle in cardiac mri: The m&ms challenge
Carlos Martín-Isla, Víctor M Campello, Cristian Izquierdo, Kaisar Kushibar, Carla Sendra-Balcells, Polyxeni Gkontra, Alireza Sojoudi, Mitchell J Fulton, Tewodros Weldebirhan Arega, Kumaradevan Punithakumar, et al · 2023
Later among the works it cites.
Voxel-level segmentation of pathologically-proven adrenocortical carcinoma with ki-67 expression (adrenal-acc-ki67-seg)[data set]
AW Moawad, AA Ahmed, et al · 2023
Later among the works it cites.
Med-flamingo: a multimodal medical few-shot learner
Michael Moor, Qian Huang, Shirley Wu, Michihiro Yasunaga, Yash Dalmia, Jure Leskovec, Cyril Zakka, Eduardo Pontes Reis, and Pranav Rajpurkar · 2023
Later among the works it cites.
Fuzzy attention neural network to tackle discontinuity in airway segmentation
Yang Nan, Javier Del Ser, Zeyu Tang, Peng Tang, Xiaodan Xing, Yingying Fang, Francisco Herrera, Witold Pedrycz, Simon Walsh, and Guang Yang · 2023
Later among the works it cites.
Han-seg: The head and neck organ-at-risk ct and mr segmentation dataset
Gašper Podobnik, Primož Strojan, Primož Peterlin, Bulat Ibragimov, and Tomaž Vrtovec · 2023
Later among the works it cites.
A tumour and liver automatic segmentation (atlas) dataset on contrast-enhanced magnetic resonance imaging for hepatocellular carcinoma
Félix Quinton, Romain Popoff, Benoît Presles, Sarah Leclerc, Fabrice Meriaudeau, Guillaume Nodari, Olivier Lopez, Julie Pellegrinelli, Olivier Chevallier, Dominique Ginhac, et al · 2023
Later among the works it cites.
Pandagpt: One model to instruction-follow them all, 2023
Yixuan Su, Tian Lan, Huayang Li, Jialu Xu, Yan Wang, and Deng Cai · 2023
Later among the works it cites.
Generative multimodal models are in-context learners
Quan Sun, Yufeng Cui, Xiaosong Zhang, Fan Zhang, Qiying Yu, Zhengxiong Luo, Yueze Wang, Yongming Rao, Jingjing Liu, Tiejun Huang, et al · 2023
Later among the works it cites.
Gemini: a family of highly capable multimodal models
Gemini Team, Rohan Anil, Sebastian Borgeaud, Yonghui Wu, Jean-Baptiste Alayrac, Jiahui Yu, Radu Soricut, Johan Schalkwyk, Andrew M Dai, Anja Hauth, et al · 2023
Later among the works it cites.
Cogvlm: Visual expert for pretrained language models, 2023
Weihan Wang, Qingsong Lv, Wenmeng Yu, Wenyi Hong, Ji Qi, Yan Wang, Junhui Ji, Zhuoyi Yang, Lei Zhao, Xixuan Song, Jiazheng Xu, Bin Xu, Juanzi Li, Yuxiao Dong, Ming Ding, and Jie Tang · 2023
Later among the works it cites.
Totalsegmentator: Robust segmentation of 104 anatomic structures in ct images
Jakob Wasserthal, Hanns-Christian Breit, Manfred T Meyer, Maurice Pradella, Daniel Hinck, Alexander W Sauter, Tobias Heye, Daniel T Boll, Joshy Cyriac, Shan Yang, et al · 2023
Later among the works it cites.
Towards generalist foundation model for radiology by leveraging web-scale 2d&3d medical data, 2023
Chaoyi Wu, Xiaoman Zhang, Ya Zhang, Yanfeng Wang, and Weidi Xie · 2023
Later among the works it cites.
Sa-med2d-20m dataset: Segment anything in 2d medical imaging with 20 million masks
Jin Ye, Junlong Cheng, Jianpin Chen, Zhongying Deng, Tianbin Li, Haoyu Wang, Yanzhou Su, Ziyan Huang, Jilong Chen, Lei Jiang, et al · 2023
Later among the works it cites.
mplug-owl2: Revolutionizing multi-modal large language model with modality collaboration
Qinghao Ye, Haiyang Xu, Jiabo Ye, Ming Yan, Haowei Liu, Qi Qian, Ji Zhang, Fei Huang, and Jingren Zhou · 2023
Later among the works it cites.
Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi
Xiang Yue, Yuansheng Ni, Kai Zhang, Tianyu Zheng, Ruoqi Liu, Ge Zhang, Samuel Stevens, Dongfu Jiang, Weiming Ren, Yuxuan Sun, et al · 2023
Later among the works it cites.
Multi-site, multi-domain airway tree modeling
Minghui Zhang, Yangqian Wu, Hanxiao Zhang, Yulei Qin, Hao Zheng, Wen Tang, Corey Arnold, Chenhao Pei, Pengxin Yu, Yang Nan, et al · 2023
Later among the works it cites.
Pan Zhang, Xiaoyi Dong, Bin Wang, Yuhang Cao, Chao Xu, Linke Ouyang, Zhiyuan Zhao, Shuangrui Ding, Songyang Zhang, Haodong Duan, Wenwei Zhang, Hang Yan, Xinyue Zhang, Wei Li, Jingwen Li, Kai Chen, Conghui He, Xingcheng Zhang, Yu Qiao, Dahua Lin, and Jiaqi Wang · 2023
Later among the works it cites.
On the challenges and perspectives of foundation models for medical image analysis
Shaoting Zhang and Dimitris Metaxas · 2023
Later among the works it cites.
Pmc-vqa: Visual instruction tuning for medical visual question answering, 2023
Xiaoman Zhang, Chaoyi Wu, Ziheng Zhao, Weixiong Lin, Ya Zhang, Yanfeng Wang, and Weidi Xie · 2023
Later among the works it cites.
Yi: Open foundation models by 01.ai, 2024
01. AI, :, Alex Young, Bei Chen, Chao Li, Chengen Huang, Ge Zhang, Guanwei Zhang, Heng Li, Jiangcheng Zhu, Jianqun Chen, Jing Chang, Kaidong Yu, Peng Liu, Qiang Liu, Shawn Yue, Senbin Yang, Shiming Yang, Tao Yu, Wen Xie, Wenhao Huang, Xiaohui Hu, Xiaoyi Ren, Xinyao Niu, Pengcheng Nie, Yuchi Xu, Yudong Liu, Yue Wang, Yuxuan Cai, Zhenyu Gu, Zhiyuan Liu, and Zonghong Dai · 2024
Closest in time.
The claude 3 model family: Opus, sonnet, haiku
AI Anthropic · 2024
Closest in time.
Are we on the right way for evaluating large vision-language models?
Lin Chen, Jinsong Li, Xiaoyi Dong, Pan Zhang, Yuhang Zang, Zehui Chen, Haodong Duan, Jiaqi Wang, Yu Qiao, Dahua Lin, et al · 2024
Closest in time.
How far are we to gpt-4v? closing the gap to commercial multimodal models with open-source suites
Zhe Chen, Weiyun Wang, Hao Tian, Shenglong Ye, Zhangwei Gao, Erfei Cui, Wenwen Tong, Kongzhi Hu, Jiapeng Luo, Zheng Ma, et al · 2024
Closest in time.
Instructblip: Towards general-purpose vision-language models with instruction tuning
Wenliang Dai, Junnan Li, Dongxu Li, Anthony Meng Huat Tiong, Junqi Zhao, Weisheng Wang, Boyang Li, Pascale N Fung, and Steven Hoi · 2024
Closest in time.
Xiaoyi Dong, Pan Zhang, Yuhang Zang, Yuhang Cao, Bin Wang, Linke Ouyang, Xilin Wei, Songyang Zhang, Haodong Duan, Maosong Cao, Wenwei Zhang, Yining Li, Hang Yan, Yang Gao, Xinyue Zhang, Wei Li, Jingwen Li, Kai Chen, Conghui He, Xingcheng Zhang, Yu Qiao, Dahua Lin, and Jiaqi Wang · 2024
Closest in time.
Xiaoyi Dong, Pan Zhang, Yuhang Zang, Yuhang Cao, Bin Wang, Linke Ouyang, Songyang Zhang, Haodong Duan, Wenwei Zhang, Yining Li, Hang Yan, Yang Gao, Zhe Chen, Xinyue Zhang, Wei Li, Jingwen Li, Wenhai Wang, Kai Chen, Conghui He, Xingcheng Zhang, Jifeng Dai, Yu Qiao, Dahua Lin, and Jiaqi Wang · 2024
Closest in time.
Vlmevalkit: An open-source toolkit for evaluating large multi-modality models, 2024
Haodong Duan, Junming Yang, Yuxuan Qiao, Xinyu Fang, Lin Chen, Yuan Liu, Xiaoyi Dong, Yuhang Zang, Pan Zhang, Jiaqi Wang, Dahua Lin, and Kai Chen · 2024
Closest in time.
Mme: A comprehensive evaluation benchmark for multimodal large language models, 2024
Chaoyou Fu, Peixian Chen, Yunhang Shen, Yulei Qin, Mengdan Zhang, Xu Lin, Jinrui Yang, Xiawu Zheng, Ke Li, Xing Sun, Yunsheng Wu, and Rongrong Ji · 2024
Closest in time.
Meddr: Diagnosis-guided bootstrapping for large-scale medical vision-language learning
Sunan He, Yuxiang Nie, Zhixuan Chen, Zhiyuan Cai, Hongmei Wang, Shu Yang, and Hao Chen · 2024
Closest in time.
Large multilingual models pivot zero-shot multimodal learning across languages, 2024
Jinyi Hu, Yuan Yao, Chongyi Wang, Shan Wang, Yinxu Pan, Qianyu Chen, Tianyu Yu, Hanghao Wu, Yue Zhao, Haoye Zhang, Xu Han, Yankai Lin, Jiao Xue, Dahai Li, Zhiyuan Liu, and Maosong Sun · 2024
Closest in time.
Omnimedvqa: A new large-scale comprehensive evaluation benchmark for medical lvlm
Yutao Hu, Tianbin Li, Quanfeng Lu, Wenqi Shao, Junjun He, Yu Qiao, and Ping Luo · 2024
Closest in time.
Llava-med: Training a large language-and-vision assistant for biomedicine in one day
Chunyuan Li, Cliff Wong, Sheng Zhang, Naoto Usuyama, Haotian Liu, Jianwei Yang, Tristan Naumann, Hoifung Poon, and Jianfeng Gao · 2024
Closest in time.
Mini-gemini: Mining the potential of multi-modality vision language models
Yanwei Li, Yuechen Zhang, Chengyao Wang, Zhisheng Zhong, Yixin Chen, Ruihang Chu, Shaoteng Liu, and Jiaya Jia · 2024
Closest in time.
Monkey: Image resolution and text label are important things for large multi-modal models, 2024
Zhang Li, Biao Yang, Qiang Liu, Zhiyin Ma, Shuo Zhang, Jingxu Yang, Yabo Sun, Yuliang Liu, and Xiang Bai · 2024
Closest in time.
Llava-next: Improved reasoning, ocr, and world knowledge, January 2024
Haotian Liu, Chunyuan Li, Yuheng Li, Bo Li, Yuanhan Zhang, Sheng Shen, and Yong Jae Lee · 2024
Closest in time.
Deepseek-vl: towards real-world vision-language understanding
Haoyu Lu, Wen Liu, Bo Zhang, Bingxuan Wang, Kai Dong, Bo Liu, Jingxiang Sun, Tongzheng Ren, Zhuoshu Li, Yaofeng Sun, et al · 2024
Closest in time.
Drac 2022: A public benchmark for diabetic retinopathy analysis on ultra-wide optical coherence tomography angiography images
Bo Qian, Hao Chen, Xiangning Wang, Zhouyu Guan, Tingyao Li, Yixiao Jin, Yilan Wu, Yang Wen, Haoxuan Che, Gitaek Kwon, et al · 2024
Closest in time.
Abdomenatlas-8k: Annotating 8,000 ct volumes for multi-organ segmentation in three weeks
Chongyu Qu, Tiezheng Zhang, Hualin Qiao, Yucheng Tang, Alan L Yuille, Zongwei Zhou, et al · 2024
Closest in time.
Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Machel Reid, Nikolay Savinov, Denis Teplyashin, Dmitry Lepikhin, Timothy Lillicrap, Jean-baptiste Alayrac, Radu Soricut, Angeliki Lazaridou, Orhan Firat, Julian Schrittwieser, et al · 2024
Closest in time.
Preoperative ct and survival data for patients undergoing resection of colorectal liver metastases
Amber L Simpson, Jacob Peoples, John M Creasy, Gabor Fichtinger, Natalie Gangai, Krishna N Keshavamurthy, Andras Lasso, Jinru Shia, Michael I D’Angelica, and Richard KG Do · 2024
Closest in time.
Nodule detection and generation on chest x-rays: Node21 challenge
Ecem Sogancioglu, Bram van Ginneken, Finn Behrendt, Marcel Bengs, Alexander Schlaefer, Miron Radu, Di Xu, Ke Sheng, Fabien Scalzo, Eric Marcus, et al · 2024
Closest in time.
Pathmmu: A massive multimodal expert-level benchmark for understanding and reasoning in pathology
Yuxuan Sun, Hao Wu, Chenglu Zhu, Sunyi Zheng, Qizi Chen, Kai Zhang, Yunlong Zhang, Xiaoxiao Lan, Mengyue Zheng, Jingxiong Li, et al · 2024
Closest in time.
Fuseg: The foot ulcer segmentation challenge
Chuanbo Wang, Amirreza Mahbod, Isabella Ellinger, Adrian Galdran, Sandeep Gopalakrishnan, Jeffrey Niezgoda, and Zeyun Yu · 2024
Closest in time.
LLaVA-UHD: an lmm perceiving any aspect ratio and high-resolution images
Ruyi Xu, Yuan Yao, Zonghao Guo, Junbo Cui, Zanlin Ni, Chunjiang Ge, Tat-Seng Chua, Zhiyuan Liu, and Gao Huang · 2024
Closest in time.
Kaining Ying, Fanqing Meng, Jin Wang, Zhiqian Li, Han Lin, Yue Yang, Hao Zhang, Wenbo Zhang, Yuqi Lin, Shuo Liu, et al · 2024
Closest in time.
Rlaif-v: Aligning mllms through open-source ai feedback for super gpt-4v trustworthiness
Tianyu Yu, Haoye Zhang, Yuan Yao, Yunkai Dang, Da Chen, Xiaoman Lu, Ganqu Cui, Taiwen He, Zhiyuan Liu, Tat-Seng Chua, and Maosong Sun · 2024
Closest in time.
Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi
Xiang Yue, Yuansheng Ni, Kai Zhang, Tianyu Zheng, Ruoqi Liu, Ge Zhang, Samuel Stevens, Dongfu Jiang, Weiming Ren, Yuxuan Sun, et al · 2024
Closest in time.