Fetching the paper…
Reading the bibliography…
Self-supervised learning in vision-language processing exploits semantic alignment between imaging and text modalities.
Influence of prior radiologic information on the interpretation of radiographic examinations
Uwa O. Aideyan, Kevin Berbaum, and Wilbur L. Smith · 1995
Earlier work this paper cites.
PhysioBank, PhysioToolkit, and PhysioNet: components of a new research resource for complex physiologic signals
Ary L Goldberger, Luis AN Amaral, Leon Glass, Jeffrey M Hausdorff, Plamen Ch Ivanov, Roger G Mark, Joseph E Mietus, George B Moody, Chung-Kang Peng, and H Eugene Stanley · 2000
Earlier work this paper cites.
Optimization of mutual information for multiresolution image registration
Philippe Thévenaz and Michael Unser · 2000
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu · 2002
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin · 2004
Earlier work this paper cites.
Symmetric diffeomorphic image registration with cross-correlation: Evaluating automated labeling of elderly and neurodegenerative brain
B.B. Avants, C.L. Epstein, M. Grossman, and J.C. Gee · 2006
Earlier work this paper cites.
A nonparametric-based rib suppression method for chest radiographs
Jiann-Shu Lee, Jing-Wein Wang, Hsing-Hsien Wu, and Ming-Zheng Yuan · 2012
Earlier work this paper cites.
Toward a comprehensive framework for the spatiotemporal statistical analysis of longitudinal shape data
Stanley Durrleman, Xavier Pennec, Alain Trouvé, José Braga, Guido Gerig, and Nicholas Ayache · 2013
Earlier work this paper cites.
The design of simpleitk
Bradley C Lowekamp, David T Chen, Luis Ibáñez, and Daniel Blezek · 2013
Earlier work this paper cites.
Jimmy Lei Ba, Jamie Ryan Kiros, and Geoffrey E. Hinton · 2016
Earlier work this paper cites.
Preparing a collection of radiology examinations for distribution and retrieval
Dina Demner-Fushman, Marc D Kohli, Marc B Rosenman, Sonya E Shooshan, Laritza Rodriguez, Sameer Antani, George R Thoma, and Clement J McDonald · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
ChestX-Ray8: Hospital-scale chest X-ray database and benchmarks on weakly-supervised classification and localization of common thorax diseases
Xiaosong Wang, Yifan Peng, Le Lu, Zhiyong Lu, Mohammadhadi Bagheri, and Ronald M Summers · 2017
Earlier work this paper cites.
Fully convolutional siamese networks for change detection
R. C. Daudt, B. L. Saux, and Alexandre Boulch · 2018
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Earlier work this paper cites.
Generating wikipedia by summarizing long sequences
Peter J Liu, Mohammad Saleh, Etienne Pot, Ben Goodrich, Ryan Sepassi, Lukasz Kaiser, and Noam Shazeer · 2018
Earlier work this paper cites.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter · 2018
Earlier work this paper cites.
Representation learning with contrastive predictive coding
Aaron van den Oord, Yazhe Li, and Oriol Vinyals · 2018
Earlier work this paper cites.
Longitudinal detection of radiological abnormalities with time-modulated LSTM
Ruggiero Santeramo, Samuel Joseph Withey, and G. Montana · 2018
Earlier work this paper cites.
CheXpert: A large chest radiograph dataset with uncertainty labels and expert comparison
Jeremy Irvin, Pranav Rajpurkar, Michael Ko, Yifan Yu, Silviana Ciurea-Ilcus, Chris Chute, Henrik Marklund, Behzad Haghgoo, Robyn Ball, Katie Shpanskaya, et al · 2019
Earlier work this paper cites.
MIMIC-CXR database (version 2.0.0)
A. Johnson, T. Pollard, S.J. Berkowitz, R. Mark, and S. Horng · 2019
Earlier work this paper cites.
Longitudinal change detection on chest X-rays using geometric correlation maps
Dong Yul Oh, Jihang Kim, and Kyong Joon Lee · 2019
Earlier work this paper cites.
End-to-end change detection for high resolution satellite images using improved UNet++
Daifeng Peng, Yongjun Zhang, and Haiyan Guan · 2019
Earlier work this paper cites.
Sentence-bert: Sentence embeddings using siamese bert-networks
Nils Reimers, Iryna Gurevych, Nils Reimers, Iryna Gurevych, Nandan Thakur, Nils Reimers, Johannes Daxenberger, Iryna Gurevych, Nils Reimers, Iryna Gurevych, et al · 2019
Earlier work this paper cites.
Augmenting the national institutes of health chest radiograph dataset with expert annotations of possible pneumonia
George Shih, Carol C Wu, Safwan S Halabi, Marc D Kohli, Luciano M Prevedello, Tessa S Cook, Arjun Sharma, Judith K Amorosa, Veronica Arteaga, Maya Galperin-Aizenberg, et al · 2019
Earlier work this paper cites.
Towards predicting the evolution of lung tumors during radiotherapy observed on a longitudinal MR imaging study via a deep learning algorithm
Chuang Wang, Andreas Rimner, Yu chi Hu, Neelam Tyagi, Jue Jiang, Ellen Yorke, Sadegh Riyahi, Gig S. Mageras, Joseph O. Deasy, and Pengpeng Zhang · 2019
Earlier work this paper cites.
Deep learning predicts lung cancer treatment response from serial medical imaging
Yiwen Xu, Ahmed Hosny, Roman Zeleznik, Chintan Parmar, Thibaud P. Coroller, Idalid Ivy Franco, Raymond H. Mak, and Hugo J.W.L. Aerts · 2019
Earlier work this paper cites.
Quantifying attention flow in transformers
Samira Abnar and Willem Zuidema · 2020
Cited alongside, same era.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Cited alongside, same era.
End-to-end object detection with transformers
Nicolas Carion, Francisco Massa, Gabriel Synnaeve, Nicolas Usunier, Alexander Kirillov, and Sergey Zagoruyko · 2020
Cited alongside, same era.
UNITER: UNiversal Image-TExt Representation learning
Yen-Chun Chen, Linjie Li, Licheng Yu, Ahmed El Kholy, Faisal Ahmed, Zhe Gan, Yu Cheng, and Jingjing Liu · 2020
Cited alongside, same era.
Generating radiology reports via memory-driven transformer
Zhihong Chen, Yan Song, Tsung-Hui Chang, and Xiang Wan · 2020
Cited alongside, same era.
Debiased contrastive learning
Ching-Yao Chuang, Joshua Robinson, Yen-Chen Lin, Antonio Torralba, and Stefanie Jegelka · 2020
Broaden your views for self-supervised video learning
Adria Recasens, Pauline Luc, Jean-Baptiste Alayrac, Luyu Wang, Florian Strub, Corentin Tallec, Mateusz Malinowski, Viorica Pătrăucean, Florent Altché, Michal Valko, et al · 2021
Later among the works it cites.
COVID-19 prognosis via self-supervised representation learning and multi-image prediction
Anuroop Sriram, Matthew Muckley, Koustuv Sinha, F. Shamout, Joelle Pineau, K. Geras, L. Azour, Y. Aphinyanaphongs, N. Yakubova, and William H. Moore · 2021
Later among the works it cites.
Medaug: Contrastive learning leveraging patient metadata improves representations for chest x-ray interpretation
Yen Nhi Truong Vu, Richard Wang, Niranjan Balachandar, Can Liu, Andrew Y Ng, and Pranav Rajpurkar · 2021
Later among the works it cites.
Vlmo: Unified vision-language pre-training with mixture-of-modality-experts
Wenhui Wang, Hangbo Bao, Li Dong, and Furu Wei · 2021
Later among the works it cites.
Finetuned language models are zero-shot learners
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2020
Cited alongside, same era.
Self-supervised co-training for video representation learning
Tengda Han, Weidi Xie, and Andrew Zisserman · 2020
Cited alongside, same era.
ACR practice guideline for communication of diagnostic imaging findings
American College of Radiology (ACR) · 2020
Cited alongside, same era.
Chest x-ray findings and temporal lung changes in patients with covid-19 pneumonia
Liqa A Rousan, Eyhab Elobeid, Musaab Karrar, and Yousef Khader · 2020
Cited alongside, same era.
Change detection based on artificial intelligence: State-of-the-art and challenges
Wenzhong Shi, Min Zhang, Rui Zhang, Shanxiong Chen, and Zhao Zhan · 2020
Cited alongside, same era.
Combining automatic labelers and expert annotations for accurate radiology report labeling using BERT
Akshay Smit, Saahil Jain, Pranav Rajpurkar, Anuj Pareek, Andrew Ng, and Matthew Lungren · 2020
Cited alongside, same era.
Jason Wei, Maarten Bosma, Vincent Zhao, Kelvin Guu, Adams Wei Yu, Brian Lester, Nan Du, Andrew M Dai, and Quoc V Le · 2021
Later among the works it cites.
Chest imagenome dataset (version 1.0.0)
Joy Wu, Nkechinyere Agu, Ismini Lourentzou, Arjun Sharma, Joseph Paguio, Jasper Seth Yao, Edward Christopher Dee, William Mitchell, Satyananda Kashyap, Andrea Giovannini, Leo Anthony Celi, Tanveer Syeda-Mahmood, and Mehdi Moradi · 2021
Later among the works it cites.
Rethinking and improving relative position encoding for vision transformer
Kan Wu, Houwen Peng, Minghao Chen, Jianlong Fu, and Hongyang Chao · 2021
Later among the works it cites.
Filip: Fine-grained interactive language-image pre-training
Lewei Yao, Runhui Huang, Lu Hou, Guansong Lu, Minzhe Niu, Hang Xu, Xiaodan Liang, Zhenguo Li, Xin Jiang, and Chunjing Xu · 2021
Later among the works it cites.
Contrastive learning with temporal correlated medical images: A case study using lung segmentation in chest x-rays
Dewen Zeng, John N Kheir, Peng Zeng, and Yiyu Shi · 2021
Later among the works it cites.
Contrastive learning of global and local video representations
Zhaoyang Zeng, Daniel McDuff, Yale Song, et al · 2021
Later among the works it cites.
Biomedical and clinical english model packages for the stanza python nlp library
Yuhao Zhang, Yuhui Zhang, Peng Qi, Christopher D Manning, and Curtis P Langlotz · 2021
Later among the works it cites.
Leveraging time irreversibility with order-contrastive pre-training
Monica N Agrawal, Hunter Lang, Michael Offin, Lior Gazit, and David Sontag · 2022
Later among the works it cites.
Flamingo: a visual language model for few-shot learning
Jean-Baptiste Alayrac, Jeff Donahue, Pauline Luc, Antoine Miech, Iain Barr, Yana Hasson, Karel Lenc, Arthur Mensch, Katie Millican, Malcolm Reynolds, et al · 2022
Later among the works it cites.
Making the most of text semantics to improve biomedical vision–language processing
Benedikt Boecking, Naoto Usuyama, Shruthi Bannur, Daniel C. Castro, Anton Schwaighofer, Stephanie Hyland, Maria Wetscherek, Tristan Naumann, Aditya Nori, Javier Alvarez-Valle, Hoifung Poon, and Ozan Oktay · 2022
Later among the works it cites.
MS-CXR: Making the most of text semantics to improve biomedical vision–language processing (version 0.1)
Benedikt Boecking, Naoto Usuyama, Shruthi Bannur, Daniel C. Castro, Anton Schwaighofer, Stephanie Hyland, Maria Wetscherek, Tristan Naumann, Aditya Nori, Javier Alvarez-Valle, Hoifung Poon, and Ozan Oktay · 2022
Later among the works it cites.
Robust contrastive learning against noisy views
Ching-Yao Chuang, R Devon Hjelm, Xin Wang, Vibhav Vineet, Neel Joshi, Antonio Torralba, Stefanie Jegelka, and Yale Song · 2022
Later among the works it cites.
An empirical study of training end-to-end vision-and-language transformers
Zi-Yi Dou, Yichong Xu, Zhe Gan, Jianfeng Wang, Shuohang Wang, Lijuan Wang, Chenguang Zhu, Pengchuan Zhang, Lu Yuan, Nanyun Peng, et al · 2022
Later among the works it cites.
Zero-shot out-of-distribution detection based on the pretrained model clip
Sepideh Esmaeilpour, Bing Liu, Eric Robertson, and Lei Shu · 2022
Later among the works it cites.
Chexrelnet: An anatomy-aware model for tracking longitudinal relationships between chest x-rays
Gaurang Karwande, Amarachi B Mbakwe, Joy T Wu, Leo A Celi, Mehdi Moradi, and Ismini Lourentzou · 2022
Later among the works it cites.
BLIP: Bootstrapping language-image pre-training for unified vision-language understanding and generation
Junnan Li, Dongxu Li, Caiming Xiong, and Steven Hoi · 2022
Later among the works it cites.
Multi-modal understanding and generation for medical images and text via vision-language pre-training
Jong Hak Moon, Hyungyung Lee, Woncheol Shin, Young-Hak Kim, and Edward Choi · 2022
Later among the works it cites.
How do vision transformers work?
Namuk Park and Songkuk Kim · 2022
Later among the works it cites.
Vignav Ramesh, Nathan Andrew Chi, and Pranav Rajpurkar · 2022
Later among the works it cites.
Flava: A foundational language and vision alignment model
Amanpreet Singh, Ronghang Hu, Vedanuj Goswami, Guillaume Couairon, Wojciech Galuba, Marcus Rohrbach, and Douwe Kiela · 2022
Later among the works it cites.
Clinical-BERT: Vision-language pre-training for radiograph diagnosis and reports generation
Bin Yan and Mingtao Pei · 2022
Later among the works it cites.
Coca: Contrastive captioners are image-text foundation models
Jiahui Yu, Zirui Wang, Vijay Vasudevan, Legg Yeung, Mojtaba Seyedhosseini, and Yonghui Wu · 2022
Later among the works it cites.
Time is matter: Temporal self-supervision for video transformers
Sukmin Yun, Jaehyung Kim, Dongyoon Han, Hwanjun Song, Jung-Woo Ha, and Jinwoo Shin · 2022
Later among the works it cites.