Fetching the paper…
Reading the bibliography…
3D medical images such as computed tomography are widely used in clinical practice, offering a great potential for automatic diagnosis.
E. Alsentzer, J. Murphy, W. Boag, W.-H. Weng, D. Jin, T. Naumann, and M. McDermott, “Publicly available clinical BERT embeddings,” in Proceedings of the 2nd Clinical Natural Language Processing Workshop . Minneapolis, Minnesota, USA: Association for Computational Linguistics, Jun. 2019, pp. 72–78. [Online]. Available: https://www.aclweb.org/anthology/W19-1909
1909
Earlier work this paper cites.
P. J. Rousseeuw, “Silhouettes: a graphical aid to the interpretation and validation of cluster analysis,” Journal of computational and applied mathematics , vol. 20, pp. 53–65, 1987
1987
Earlier work this paper cites.
C. P. Langlotz, “Radlex: a new method for indexing online educational materials,” pp. 1595–1597, 2006
2006
Earlier work this paper cites.
D. T. Ginat and R. Gupta, “Advances in computed tomography imaging technology,” Annual review of biomedical engineering , vol. 16, no. 1, pp. 431–453, 2014
2014
Earlier work this paper cites.
C. Deng, X. Tang, J. Yan, W. Liu, and X. Gao, “Discriminative dictionary learning with common label alignment for cross-modal retrieval,” IEEE Transactions on Multimedia , vol. 18, no. 2, pp. 208–218, 2015
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Earlier work this paper cites.
Y. Long, L. Liu, L. Shao, F. Shen, G. Ding, and J. Han, “From zero-shot learning to conventional supervised classification: Unseen visual data synthesis,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2017, pp. 1627–1636
2017
Earlier work this paper cites.
A. E. Johnson, T. J. Pollard, S. J. Berkowitz, N. R. Greenbaum, M. P. Lungren, C.-y. Deng, R. G. Mark, and S. Horng, “Mimic-cxr, a de-identified publicly available database of chest radiographs with free-text reports,” Scientific data , vol. 6, no. 1, p. 317, 2019
2019
Earlier work this paper cites.
J. Lu, D. Batra, D. Parikh, and S. Lee, “Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks,” Advances in neural information processing systems , vol. 32, 2019
2019
Earlier work this paper cites.
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova, “Bert: Pre-training of deep bidirectional transformers for language understanding,” in Proceedings of the 2019 conference of the North American chapter of the association for computational linguistics: human language technologies, volume 1 (long and short papers) , 2019, pp. 4171–4186
2019
Earlier work this paper cites.
E. Svoboda, “Artificial intelligence is improving the detection of lung cancer,” Nature , vol. 587, no. 7834, pp. S20–S20, 2020
2020
Earlier work this paper cites.
A. Bustos, A. Pertusa, J.-M. Salinas, and M. De La Iglesia-Vaya, “Padchest: A large chest x-ray image dataset with multi-label annotated reports,” Medical image analysis , vol. 66, p. 101797, 2020
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark et al. , “Learning transferable visual models from natural language supervision,” in International conference on machine learning . PMLR, 2021, pp. 8748–8763
2021
Earlier work this paper cites.
R. L. Draelos, D. Dov, M. A. Mazurowski, J. Y. Lo, R. Henao, G. D. Rubin, and L. Carin, “Machine-learning-based multiple abnormality prediction with large-scale chest computed tomography volumes,” Medical image analysis , vol. 67, p. 101857, 2021
2021
Earlier work this paper cites.
S.-C. Huang, L. Shen, M. P. Lungren, and S. Yeung, “Gloria: A multimodal global-local representation learning framework for label-efficient medical image recognition,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 3942–3951
2021
Earlier work this paper cites.
W. Kim, B. Son, and I. Kim, “Vilt: Vision-and-language transformer without convolution or region supervision,” in International conference on machine learning . PMLR, 2021, pp. 5583–5594
2021
Earlier work this paper cites.
Z. Huang, Z. Zeng, Y. Huang, B. Liu, D. Fu, and J. Fu, “Seeing out of the box: End-to-end pre-training for vision-language representation learning,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2021, pp. 12 976–12 985
2021
Earlier work this paper cites.
Z. Liu, Y. Lin, Y. Cao, H. Hu, Y. Wei, Z. Zhang, S. Lin, and B. Guo, “Swin transformer: Hierarchical vision transformer using shifted windows,” in Proceedings of the IEEE/CVF international conference on computer vision , 2021, pp. 10 012–10 022
2021
Earlier work this paper cites.
Y. Gu, R. Tinn, H. Cheng, M. Lucas, N. Usuyama, X. Liu, T. Naumann, J. Gao, and H. Poon, “Domain-specific language model pretraining for biomedical natural language processing,” ACM Transactions on Computing for Healthcare (HEALTH) , vol. 3, no. 1, pp. 1–23, 2021
2021
Earlier work this paper cites.
E. Tiu, E. Talius, P. Patel, C. P. Langlotz, A. Y. Ng, and P. Rajpurkar, “Expert-level detection of pathologies from unannotated chest x-ray images via self-supervised learning,” Nature Biomedical Engineering , vol. 6, no. 12, pp. 1399–1406, 2022
2022
Cited alongside, same era.
Z. Wang, Z. Wu, D. Agarwal, and J. Sun, “Medclip: Contrastive learning from unpaired medical images and text,” in Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing , 2022, pp. 3876–3887
2022
Cited alongside, same era.
F. Wang, Y. Zhou, S. Wang, V. Vardhanabhuti, and L. Yu, “Multi-granularity cross-modal alignment for generalized medical visual representation learning,” Advances in Neural Information Processing Systems , vol. 35, pp. 33 536–33 549, 2022
2022
Cited alongside, same era.
2022
Z. Zhang, W. Ke, Y. Zhu, X. Liang, J. Liu, Q. Ye, and T. Zhang, “Language-driven visual consensus for zero-shot semantic segmentation,” IEEE Transactions on Circuits and Systems for Video Technology , 2024
2024
Later among the works it cites.
I. E. Hamamci, S. Er, F. Almas, A. G. Simsek, S. N. Esirgun, I. Dogan, M. F. Dasdelen, B. Wittmann, E. Simsar, M. Simsar et al. , “A foundation model utilizing chest ct volumes and radiology reports for supervised-level zero-shot detection of abnormalities,” CoRR , 2024
2024
Later among the works it cites.
W. Cao, J. Zhang, Y. Xia, T. C. Mok, Z. Li, X. Ye, L. Lu, J. Zheng, Y. Tang, and L. Zhang, “Bootstrapping chest ct image understanding by distilling knowledge from x-ray expert models,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, pp. 11 238–11 247
2024
Later among the works it cites.
L. Blankemeier, J. P. Cohen, A. Kumar, D. Van Veen, S. J. S. Gardezi, M. Paschali, Z. Chen, J.-B. Delbrouck, E. Reis, C. Truyts et al. , “Merlin: A vision language foundation model for 3d computed tomography,” Research Square , pp. rs–3, 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
L. Yao, J. Han, Y. Wen, X. Liang, D. Xu, W. Zhang, Z. Li, C. Xu, and H. Xu, “Detclip: Dictionary-enriched visual-concept paralleled pre-training for open-world detection,” Advances in Neural Information Processing Systems , vol. 35, pp. 9125–9138, 2022
2022
Cited alongside, same era.
B. Boecking, N. Usuyama, S. Bannur, D. C. Castro, A. Schwaighofer, S. Hyland, M. Wetscherek, T. Naumann, A. Nori, J. Alvarez-Valle et al. , “Making the most of text semantics to improve biomedical vision–language processing,” in European conference on computer vision . Springer, 2022, pp. 1–21
2022
Cited alongside, same era.
Z. Chen, Y. Du, J. Hu, Y. Liu, G. Li, X. Wan, and T.-H. Chang, “Multi-modal masked autoencoders for medical vision-and-language pre-training,” in International Conference on Medical Image Computing and Computer-Assisted Intervention . Springer, 2022
2022
Cited alongside, same era.
R. Luo, L. Sun, Y. Xia, T. Qin, S. Zhang, H. Poon, and T.-Y. Liu, “Biogpt: generative pre-trained transformer for biomedical text generation and mining,” Briefings in bioinformatics , vol. 23, no. 6, p. bbac409, 2022
2022
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
C. Pellegrini, M. Keicher, E. Özsoy, P. Jiraskova, R. Braren, and N. Navab, “Xplainer: From x-ray observations to explainable zero-shot diagnosis,” in International Conference on Medical Image Computing and Computer-Assisted Intervention . Springer, 2023, pp. 420–429
2023
Cited alongside, same era.
C. Wu, X. Zhang, Y. Zhang, Y. Wang, and W. Xie, “Medklip: Medical knowledge enhanced language-image pre-training for x-ray diagnosis,” in Proceedings of the IEEE/CVF international conference on computer vision , 2023, pp. 21 372–21 383
2023
Cited alongside, same era.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
H. Lai, Q. Yao, Z. Jiang, R. Wang, Z. He, X. Tao, and S. K. Zhou, “Carzero: Cross-attention alignment for radiology zero-shot classification,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, pp. 11 137–11 146
2024
Later among the works it cites.
2024
Later among the works it cites.
X. Yu, L. Zhang, Z. Wu, and D. Zhu, “Core-periphery multi-modality feature alignment for zero-shot medical image analysis,” IEEE Transactions on Medical Imaging , 2024
2024
Later among the works it cites.
Q. Team, “Qwen2 technical report,” arXiv preprint arXiv:2407.10671 , 2024
2024
Later among the works it cites.
AI@Meta, “Llama 3 model card,” 2024. [Online]. Available: https://github.com/meta-llama/llama3/blob/main/MODEL_CARD.md
2024
Later among the works it cites.
2024
Later among the works it cites.
National Library of Medicine (US), “UMLS Knowledge Sources [dataset on the internet],” Bethesda (MD): National Library of Medicine (US), May 2024, release 2024AA. [cited 2024 Jul 15]. [Online]. Available: http://www.nlm.nih.gov/research/umls/licensedcontent/umlsknowledgesources.html
2024
Later among the works it cites.
L. Mei, K. Deng, Z. Cui, Y. Fang, Y. Li, H. Lai, M. S. Tonetti, and D. Shen, “Clinical knowledge-guided hybrid classification network for automatic periodontal disease diagnosis in x-ray image,” Medical Image Analysis , vol. 99, p. 103376, 2025
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.