Fetching the paper…
Reading the bibliography…
Medical artificial general intelligence (MAGI) enables one foundation model to solve different medical tasks, which is very practical in the medical domain.
K. Papineni, S. Roukos, T. Ward, and W.-J. Zhu, “Bleu: a method for automatic evaluation of machine translation,” in ACL , 2002
2002
Earlier work this paper cites.
C.-Y. Lin, “Rouge: A package for automatic evaluation of summaries,” in ACL 2004 , 2004
2004
Earlier work this paper cites.
S. Banerjee and A. Lavie, “Meteor: An automatic metric for mt evaluation with improved correlation with human judgments,” in IEEvaluation@ACL , 2005
2005
Earlier work this paper cites.
R. Vedantam, C. L. Zitnick, and D. Parikh, “Cider: Consensus-based image description evaluation,” 2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , pp. 4566–4575, 2015
2015
Earlier work this paper cites.
D. Demner-Fushman, M. D. Kohli, M. B. Rosenman, S. E. Shooshan, L. M. Rodriguez, S. Antani, G. R. Thoma, and C. J. McDonald, “Preparing a collection of radiology examinations for distribution and retrieval,” Journal of the American Medical Informatics Association : JAMIA , vol. 23 2, pp. 304–10, 2016
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , pp. 770–778, 2016
2016
Earlier work this paper cites.
Z. Yang, X. He, J. Gao, L. Deng, and A. Smola, “Stacked attention networks for image question answering,” 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , pp. 21–29, 2016
2016
Earlier work this paper cites.
X. Wang, Y. Peng, L. Lu, Z. Lu, M. Bagheri, and R. M. Summers, “Chestx-ray8: Hospital-scale chest x-ray database and benchmarks on weakly-supervised classification and localization of common thorax diseases,” 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , pp. 3462–3471, 2017
2017
Earlier work this paper cites.
G. Huang, Z. Liu, and K. Q. Weinberger, “Densely connected convolutional networks,” 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , pp. 2261–2269, 2017
2017
Earlier work this paper cites.
G. Huang, Z. Liu, L. Van Der Maaten, and K. Q. Weinberger, “Densely connected convolutional networks,” in CVPR , 2017
2017
Earlier work this paper cites.
J. J. Lau, S. Gayen, A. B. Abacha, and D. Demner-Fushman, “A dataset of clinically generated visual questions and answers about radiology images,” Scientific Data , vol. 5, 2018
2018
Earlier work this paper cites.
B. Jing, P. Xie, and E. P. Xing, “On the automatic generation of medical imaging reports,” in ACL , 2018
2018
Earlier work this paper cites.
C. Y. Li, X. Liang, Z. Hu, and E. P. Xing, “Hybrid retrieval-generation reinforced agent for medical image report generation,” in NeurIPS , 2018
2018
Earlier work this paper cites.
J.-H. Kim, J. Jun, and B.-T. Zhang, “Bilinear attention networks,” in NeurIPS , 2018
2018
Earlier work this paper cites.
B. D. Nguyen, T.-T. Do, B. X. Nguyen, T. K. Do, E. Tjiputra, and Q. D. Tran, “Overcoming data limitation in medical visual question answering,” in MICCAI , 2019
2019
Earlier work this paper cites.
H. Tan and M. Bansal, “Lxmert: Learning cross-modality encoder representations from transformers,” in Proceedings of the Conference on Empirical Methods in Natural Language Processing (EMNLP) , 2019, pp. 5099–5110
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
C. Li, Z. Li, Z. Ge, and M. Li, “Knowledge driven temporal activity localization,” Journal of Visual Communication and Image Representation , vol. 64, p. 102628, 2019
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
K. He, H. Fan, Y. Wu, S. Xie, and R. B. Girshick, “Momentum contrast for unsupervised visual representation learning,” 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , pp. 9726–9735, 2019
2019
Cited alongside, same era.
I. Beltagy, K. Lo, and A. Cohan, “Scibert: A pretrained language model for scientific text,” in EMNLP , 2019
2019
Cited alongside, same era.
2020
Cited alongside, same era.
L.-M. Zhan, B. Liu, L. Fan, J. Chen, and X.-M. Wu, “Medical visual question answering via conditional reasoning,” Proceedings of the 28th ACM International Conference on Multimedia , 2020
2020
Cited alongside, same era.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, G. Krueger, and I. Sutskever, “Learning transferable visual models from natural language supervision,” in International Conference on Machine Learning , 2021
2021
Later among the works it cites.
K. Zhou, J. Yang, C. C. Loy, and Z. Liu, “Learning to prompt for vision-language models,” International Journal of Computer Vision , vol. 130, pp. 2337 – 2348, 2021
2021
Later among the works it cites.
P. Liu, W. Yuan, J. Fu, Z. Jiang, H. Hayashi, and G. Neubig, “Pre-train, prompt, and predict: A systematic survey of prompting methods in natural language processing,” ACM Computing Surveys , vol. 55, pp. 1 – 35, 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
G. Li, N. Duan, Y. Fang, M. Gong, and D. Jiang, “Unicoder-vl: A universal encoder for vision and language by cross-modal pre-training.” in Proceedings of the Association for the Advance of Artificial Intelligence (AAAI) , vol. 34, no. 7, 2020, pp. 11 336–11 344
2020
Cited alongside, same era.
X. Li, X. Yin, C. Li, P. Zhang, X. Hu, L. Zhang, L. Wang, H. Hu, L. Dong, F. Wei, Y. Choi, and J. Gao, “Oscar: Object-semantics aligned pre-training for vision-language tasks,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2020, pp. 121–137
2020
Cited alongside, same era.
Y.-C. Chen, L. Li, L. Yu, A. E. Kholy, F. Ahmed, Z. Gan, Y. Cheng, and J. Liu, “Uniter: Universal image-text representation learning,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2020, pp. 104–120
2020
Cited alongside, same era.
P. Qi, Y. Zhang, Y. Zhang, J. Bolton, and C. D. Manning, “Stanza: A python natural language processing toolkit for many human languages,” in ACL , 2020
2020
Cited alongside, same era.
Y. Zhang, X. Wang, Z. Xu, Q. Yu, A. L. Yuille, and D. Xu, “When radiology report generation meets knowledge graph,” in AAAI , 2020
2020
Cited alongside, same era.
Q. Guan and Y. Huang, “Multi-label chest x-ray image classification via category-wise residual attention learning,” Pattern Recognition Letters , vol. 130, pp. 259–266, 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
W. Liu, P. Zhou, Z. Zhao, Z. Wang, Q. Ju, H. Deng, and P. Wang, “K-bert: Enabling language representation with knowledge graph,” in Proceedings of the AAAI Conference on Artificial Intelligence , no. 03, 2020, pp. 2901–2908
2020
Cited alongside, same era.
2021
Later among the works it cites.
2021
Later among the works it cites.
Z. Yuan, Y. Liu, C. Tan, S. Huang, and F. Huang, “Improving biomedical pretrained language models with knowledge,” in Workshop on Biomedical Natural Language Processing , 2021
2021
Later among the works it cites.
Y. Cui, Z. Yu, C. Wang, Z. Zhao, J. Zhang, M. Wang, and J. Yu, “Rosita: Enhancing vision-and-language semantic alignments via cross- and intra-modal knowledge integration,” Proceedings of the 29th ACM International Conference on Multimedia , 2021
2021
Later among the works it cites.
F. Yu, J. Tang, W. Yin, Y. Sun, H. Tian, H. Wu, and H. Wang, “Ernie-vil: Knowledge enhanced vision-language representations through scene graphs,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 35, no. 4, 2021, pp. 3208–3216
2021
Later among the works it cites.
S. Jain, A. Agrawal, A. Saporta, S. Q. Truong, D. N. Duong, T. Bui, P. Chambon, Y. Zhang, M. P. Lungren, A. Y. Ng, C. Langlotz, and P. Rajpurkar, “Radgraph: Extracting clinical entities and relations from radiology reports,” in Thirty-fifth Conference on Neural Information Processing Systems Datasets and Benchmarks Track (Round 1) , 2021. [Online]. Available: https://openreview.net/forum?id=pMWtc5NKd7V
2021
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
Z. Chen, G. Li, and X. Wan, “Align, reason and learn: Enhancing medical vision-and-language pre-training with knowledge,” Proceedings of the 30th ACM International Conference on Multimedia , 2022
2022
Later among the works it cites.
——, “Conditional prompt learning for vision-language models,” 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , pp. 16 795–16 804, 2022
2022
Later among the works it cites.
H. Song, L. Dong, W. Zhang, T. Liu, and F. Wei, “Clip models are few-shot learners: Empirical studies on vqa and visual entailment,” in Annual Meeting of the Association for Computational Linguistics , 2022
2022
Later among the works it cites.
Z. Wang, Z. Wu, D. Agarwal, and J. Sun, “Medclip: Contrastive learning from unpaired medical images and text,” in Conference on Empirical Methods in Natural Language Processing , 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
M. Li, W. Cai, K. Verspoor, S. Pan, X. Liang, and X. Chang, “Cross-modal clinical graph transformer for ophthalmic report generation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2022, pp. 20 656–20 665
2022
Later among the works it cites.