Fetching the paper…
Reading the bibliography…
Medical phrase grounding is crucial for identifying relevant regions in medical images based on phrase queries, facilitating accurate image analysis and diagnosis.
Q. Du, V. Faber, and M. Gunzburger, “Centroidal voronoi tessellations: Applications and algorithms,” SIAM review , vol. 41, no. 4, pp. 637–676, 1999
1999
Earlier work this paper cites.
R. Girshick, “Fast r-cnn,” in ICCV , 2015, pp. 1440–1448
2015
Earlier work this paper cites.
C. Rupprecht, I. Laina, R. DiPietro, M. Baust, F. Tombari, N. Navab, and G. D. Hager, “Learning in an uncertain world: Representing ambiguity through multiple hypotheses,” in ICCV , 2017, pp. 3591–3600
2017
Earlier work this paper cites.
G. Wang, W. Li, S. Ourselin, and T. Vercauteren, “Automatic brain tumor segmentation using convolutional neural networks with test-time augmentation,” in International MICCAI Brainlesion Workshop . Springer, 2018, pp. 61–72
2018
Earlier work this paper cites.
M. Sensoy, L. Kaplan, and M. Kandemir, “Evidential deep learning to quantify classification uncertainty,” in NeurIPS , 2018, pp. 3183–3193
2018
Earlier work this paper cites.
B. A. Plummer, P. Kordas, M. H. Kiapour, S. Zheng, R. Piramuthu, and S. Lazebnik, “Conditional image-text embedding networks,” in ECCV , 2018, pp. 249–264
2018
Earlier work this paper cites.
S. Kohl, B. Romera-Paredes, C. Meyer, J. De Fauw, J. R. Ledsam, K. Maier-Hein, S. Eslami, D. Jimenez Rezende, and O. Ronneberger, “A probabilistic u-net for segmentation of ambiguous images,” NeurIPS , vol. 31, 2018
2018
Earlier work this paper cites.
L. Yu, S. Wang, X. Li, C.-W. Fu, and P.-A. Heng, “Uncertainty-aware self-ensembling model for semi-supervised 3d left atrium segmentation,” in MICCAI . Springer, 2019, pp. 605–613
2019
Earlier work this paper cites.
J. Wang and L. Specia, “Phrase localization without paired training examples,” in ICCV , 2019, pp. 4663–4672
2019
Earlier work this paper cites.
H. Rezatofighi, N. Tsoi, J. Gwak, A. Sadeghian, I. Reid, and S. Savarese, “Generalized intersection over union: A metric and a loss for bounding box regression,” in CVPR , 2019, pp. 658–666
2019
Earlier work this paper cites.
A. E. Johnson, T. J. Pollard, S. J. Berkowitz, N. R. Greenbaum, M. P. Lungren, C.-y. Deng, R. G. Mark, and S. Horng, “Mimic-cxr, a de-identified publicly available database of chest radiographs with free-text reports,” Scientific Data , vol. 6, no. 1, p. 317, 2019
2019
Earlier work this paper cites.
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova, “Bert: Pre-training of deep bidirectional transformers for language understanding,” in Proceedings of the 2019 conference of the North American chapter of the association for computational linguistics: human language technologies, volume 1 (long and short papers) , 2019, pp. 4171–4186
2019
Earlier work this paper cites.
T. Nair, D. Precup, D. L. Arnold, and T. Arbel, “Exploring uncertainty measures in deep networks for multiple sclerosis lesion detection and segmentation,” Medical image analysis , vol. 59, p. 101557, 2020
2020
Earlier work this paper cites.
J. Van Amersfoort, L. Smith, Y. W. Teh, and Y. Gal, “Uncertainty estimation using a single deep deterministic neural network,” in ICML . PMLR, 2020, pp. 9690–9700
2020
Earlier work this paper cites.
A. Mehrtash, W. M. Wells, C. M. Tempany, P. Abolmaesumi, and T. Kapur, “Confidence calibration and predictive uncertainty estimation for deep medical image segmentation,” IEEE Transactions on Medical Imaging , vol. 39, no. 12, pp. 3868–3878, 2020
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
J. Deng, Z. Yang, T. Chen, W. Zhou, and H. Li, “Transvg: End-to-end visual grounding with transformers,” in ICCV , 2021, pp. 1769–1779
2021
Earlier work this paper cites.
J. Mukhoti, J. van Amersfoort, P. H. Torr, and Y. Gal, “Deep deterministic uncertainty for semantic segmentation,” in International Conference on Machine Learning Workshop on Uncertainty and Robustness in Deep Learning , 2021
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
L. Muchen and S. Leonid, “Referring transformer: A one-step approach to multi-task visual grounding,” in NeurIPS , 2021
2021
Cited alongside, same era.
Y. Du, Z. Fu, Q. Liu, and Y. Wang, “Visual grounding with transformers,” in ICME . IEEE, 2022, pp. 1–6
2022
Cited alongside, same era.
C. Zhu, Y. Zhou, Y. Shen, G. Luo, X. Pan, M. Lin, C. Chen, L. Cao, X. Sun, and R. Ji, “Seqtr: A simple yet universal network for visual grounding,” in ECCV . Springer, 2022, pp. 598–615
2022
Cited alongside, same era.
L. Huang, S. Ruan, P. Decazes, and T. Denoeux, “Lymphoma segmentation from 3d pet-ct images using a deep evidential network,” International Journal of Approximate Reasoning , vol. 149, pp. 39–60, 2022
2022
Cited alongside, same era.
K. Zou, X. Yuan, X. Shen, M. Wang, and H. Fu, “Tbrats: Trusted brain tumor segmentation,” in MICCAI . Springer, 2022, pp. 503–513
2023
Later among the works it cites.
2023
Later among the works it cites.
X. Hu, L. Gu, Q. An, M. Zhang, L. Liu, K. Kobayashi, T. Harada, R. M. Summers, and Y. Zhu, “Expert knowledge-aware image difference graph representation learning for difference-aware medical visual question answering,” in ACM KDD , 2023, pp. 4156–4165
2023
Later among the works it cites.
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2022
Cited alongside, same era.
J. H. Moon, H. Lee, W. Shin, Y.-H. Kim, and E. Choi, “Multi-modal understanding and generation for medical images and text via vision-language pre-training,” IEEE Journal of Biomedical and Health Informatics , vol. 26, no. 12, pp. 6070–6080, 2022
2022
Cited alongside, same era.
F. Wang, Y. Zhou, S. Wang, V. Vardhanabhuti, and L. Yu, “Multi-granularity cross-modal alignment for generalized medical visual representation learning,” NeurIPS , vol. 35, pp. 33 536–33 549, 2022
2022
Cited alongside, same era.
B. Boecking, N. Usuyama, S. Bannur, D. C. Castro, A. Schwaighofer, S. Hyland, M. Wetscherek, T. Naumann, A. Nori, J. Alvarez-Valle, H. Poon, and O. Oktay, “Making the most of text semantics to improve biomedical vision–language processing,” in ECCV , S. Avidan, G. Brostow, M. Cissé, G. M. Farinella, and T. Hassner, Eds. Cham: Springer Nature Switzerland, 2022, pp. 1–21
2022
Cited alongside, same era.
E. J. Hu, Y. Shen, P. Wallis, Z. Allen-Zhu, Y. Li, S. Wang, L. Wang, W. Chen et al. , “Lora: Low-rank adaptation of large language models.” ICLR , vol. 1, no. 2, p. 3, 2022
2022
Cited alongside, same era.
Z. Chen, Y. Zhou, A. Tran, J. Zhao, and et al., “Medical phrase grounding with region-phrase context contrastive alignment,” in Medical Image Computing and Computer Assisted Intervention – MICCAI 2023 . Cham: Springer Nature Switzerland, 2023, pp. 371–381
2023
Cited alongside, same era.
L. Zhou, Z. Zhou, K. Mao, and Z. He, “Joint visual grounding and tracking with natural language specification,” in CVPR , 2023, pp. 23 151–23 160
2023
Cited alongside, same era.
A. Ichinose, T. Hatsutani, K. Nakamura, Y. Kitamura, S. Iizuka, E. Simo-Serra, S. Kido, and N. Tomiyama, “Visual grounding of whole radiology reports for 3d ct images,” in MICCAI . Springer, 2023, pp. 611–621
2023
Cited alongside, same era.
2023
Later among the works it cites.
R. He, P. Cascante-Bonilla, Z. Yang, A. C. Berg, and V. Ordonez, “Improved visual grounding through self-consistent explanations,” in CVPR , 2024, pp. 13 095–13 105
2024
Closest in time.
H. Liu, C. Li, Q. Wu, and Y. J. Lee, “Visual instruction tuning,” NeurIPS , vol. 36, 2024
2024
Closest in time.
Z. Chen, J. Wu, W. Wang, W. Su, G. Chen, S. Xing, M. Zhong, Q. Zhang, X. Zhu, L. Lu et al. , “Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks,” in CVPR , 2024, pp. 24 185–24 198
2024
Closest in time.
2024
Closest in time.
K. Vilouras, P. Sanchez, A. Q. O’Neil, and S. A. Tsaftaris, “Zero-shot medical phrase grounding with off-the-shelf diffusion models,” IEEE Journal of Biomedical and Health Informatics , 2024
2024
Closest in time.
W. Zhao, Z. Deng, S. Yadav, and P. S. Yu, “Heterogeneous knowledge grounding for medical question answering with retrieval augmented large language model,” in Companion Proceedings of the ACM Web Conference 2024 , 2024, pp. 1590–1594
2024
Closest in time.
2024
Closest in time.
L. Huang, S. Ruan, Y. Xing, and M. Feng, “A review of uncertainty quantification in medical image analysis: probabilistic and non-probabilistic methods,” Medical Image Analysis , p. 103223, 2024
2024
Closest in time.
Y. Luo, Q. Yang, Y. Fan, H. Qi, and M. Xia, “Measurement guidance in diffusion models: Insight from medical image synthesis,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 2024
2024
Closest in time.
J. Ma, Y. He, F. Li, L. Han, C. You, and B. Wang, “Segment anything in medical images,” Nature Communications , vol. 15, no. 1, p. 654, 2024
2024
Closest in time.
2024
Closest in time.
L. Huang, S. Ruan, P. Decazes, and T. Denœux, “Deep evidential fusion with uncertainty quantification and reliability learning for multimodal medical image segmentation,” Information Fusion , vol. 113, p. 102648, 2025
2025
Closest in time.
M. Wang, T. Lin, A. Lin, K. Yu, Y. Peng, L. Wang, C. Chen, K. Zou, H. Liang, M. Chen et al. , “Enhancing diagnostic accuracy in rare and common fundus diseases with a knowledge-rich vision-language model,” Nature Communications , vol. 16, no. 1, p. 5528, 2025
2025
Closest in time.
X. Wang, Y. Peng, L. Lu, Z. Lu, M. Bagheri, and R. M. Summers, “Chestx-ray8: Hospital-scale chest x-ray database and benchmarks on weakly-supervised classification and localization of common thorax diseases,” in CVPR , 2017, pp. 2097–2106
2097
Closest in time.