Fetching the paper…
Reading the bibliography…
Medical Visual Question Answering (VQA) is a multi-modal challenging task widely considered by research communities of the computer vision and natural language processing.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
J. Masci, U. Meier, D. Cireşan, and J. Schmidhuber, “Stacked convolutional auto-encoders for hierarchical feature extraction,” in International conference on artificial neural networks . Springer, 2011, pp. 52–59
2011
Earlier work this paper cites.
2014
Earlier work this paper cites.
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. L. Zitnick, and D. Parikh, “Vqa: Visual question answering,” in Proceedings of the IEEE international conference on computer vision , 2015, pp. 2425–2433
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
S. Ren, K. He, R. Girshick, and J. Sun, “Faster r-cnn: towards real-time object detection with region proposal networks,” IEEE transactions on pattern analysis and machine intelligence , vol. 39, no. 6, pp. 1137–1149, 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
Q. Wu, P. Wang, C. Shen, A. Dick, and A. Van Den Hengel, “Ask me anything: Free-form visual question answering based on knowledge from external sources,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 4622–4630
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
C. Finn, P. Abbeel, and S. Levine, “Model-agnostic meta-learning for fast adaptation of deep networks,” in International Conference on Machine Learning . PMLR, 2017, pp. 1126–1135
2017
Earlier work this paper cites.
K. Saito, A. Shin, Y. Ushiku, and T. Harada, “Dualnet: Domain-invariant network for visual question answering,” in 2017 IEEE International Conference on Multimedia and Expo (ICME) . IEEE, 2017, pp. 829–834
2017
Earlier work this paper cites.
H. Ben-Younes, R. Cadene, M. Cord, and N. Thome, “Mutan: Multimodal tucker fusion for visual question answering,” in Proceedings of the IEEE international conference on computer vision , 2017, pp. 2612–2620
2017
Earlier work this paper cites.
Z. Yu, J. Yu, J. Fan, and D. Tao, “Multi-modal factorized bilinear pooling with co-attention learning for visual question answering,” in Proceedings of the IEEE international conference on computer vision , 2017, pp. 1821–1830
2017
Earlier work this paper cites.
Q. Wu, C. Shen, P. Wang, A. Dick, and A. Van Den Hengel, “Image captioning and visual question answering based on attributes and external knowledge,” IEEE transactions on pattern analysis and machine intelligence , vol. 40, no. 6, pp. 1367–1381, 2017
2017
Cited alongside, same era.
P. Wang, Q. Wu, C. Shen, and A. van den Hengel, “The vqa-machine: Learning how to use existing vision algorithms to answer new questions,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2017, pp. 1173–1182
2017
Cited alongside, same era.
2017
Cited alongside, same era.
A. Jungo, R. Meier, E. Ermis, M. Blatti-Moreno, E. Herrmann, R. Wiest, and M. Reyes, “On the effect of inter-observer variability for a reliable estimation of uncertainty of medical image segmentation,” in International Conference on Medical Image Computing and Computer-Assisted Intervention . Springer, 2018, pp. 682–690
L. Li, Z. Gan, Y. Cheng, and J. Liu, “Relation-aware graph attention network for visual question answering,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2019, pp. 10 313–10 322
2019
Later among the works it cites.
H. Abdeltawab, F. Khalifa, F. Taher, N. S. Alghamdi, M. Ghazal, G. Beache, T. Mohamed, R. Keynton, and A. El-Baz, “A deep learning-based approach for automatic segmentation and quantification of the left ventricle from cardiac cine mr images,” Computerized Medical Imaging and Graphics , vol. 81, p. 101717, 2020
2020
Later among the works it cites.
M. Alfano, B. Lenzitti, G. L. Bosco, C. Muriana, T. Piazza, and G. Vizzini, “Design, development and validation of a system for automatic help to medical text understanding,” International journal of medical informatics , vol. 138, p. 104109, 2020
2020
Later among the works it cites.
J.-J. Qiu, J. Yin, W. Qian, J.-H. Liu, Z.-X. Huang, H.-P. Yu, L. Ji, and X.-X. Zeng, “A novel multiresolution-statistical texture analysis architecture: Radiomics-aided diagnosis of pdac based on plain ct images,” IEEE Transactions on Medical Imaging , 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
H. Wang, S. N. Ahmed, and M. Mandal, “Computer-aided diagnosis of cavernous malformations in brain mr images,” Computerized Medical Imaging and Graphics , vol. 66, pp. 115–123, 2018
2018
Cited alongside, same era.
J. J. Lau, S. Gayen, A. B. Abacha, and D. Demner-Fushman, “A dataset of clinically generated visual questions and answers about radiology images,” Scientific data , vol. 5, no. 1, pp. 1–10, 2018
2018
Cited alongside, same era.
D. Teney, P. Anderson, X. He, and A. Van Den Hengel, “Tips and tricks for visual question answering: Learnings from the 2017 challenge,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2018, pp. 4223–4232
2018
Cited alongside, same era.
Z. Yu, J. Yu, C. Xiang, J. Fan, and D. Tao, “Beyond bilinear: Generalized multimodal factorized high-order pooling for visual question answering,” IEEE transactions on neural networks and learning systems , vol. 29, no. 12, pp. 5947–5959, 2018
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
P. Anderson, X. He, C. Buehler, D. Teney, M. Johnson, S. Gould, and L. Zhang, “Bottom-up and top-down attention for image captioning and visual question answering,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2018, pp. 6077–6086
2018
Cited alongside, same era.
W. A. Al and I. D. Yun, “Partial policy-based reinforcement learning for anatomical landmark localization in 3d medical images,” IEEE transactions on medical imaging , vol. 39, no. 4, pp. 1245–1255, 2019
2019
Cited alongside, same era.
2020
Later among the works it cites.
L.-M. Zhan, B. Liu, L. Fan, J. Chen, and X.-M. Wu, “Medical visual question answering via conditional reasoning,” in Proceedings of the 28th ACM International Conference on Multimedia , 2020, pp. 2345–2354
2020
Later among the works it cites.
M. H. Vu, T. Löfstedt, T. Nyholm, and R. Sznitman, “A question-centric model for visual question answering in medical imaging,” IEEE transactions on medical imaging , vol. 39, no. 9, pp. 2856–2868, 2020
2020
Later among the works it cites.
W. Guo, Y. Zhang, X. Wu, J. Yang, X. Cai, and X. Yuan, “Re-attention for visual question answering,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 34, no. 01, 2020, pp. 91–98
2020
Later among the works it cites.
2020
Later among the works it cites.
Z. Chen, X. Guo, P. Y. Woo, and Y. Yuan, “Super-resolution enhanced medical image diagnosis with sample affinity interaction,” IEEE Transactions on Medical Imaging , vol. 40, no. 5, pp. 1377–1389, 2021
2021
Closest in time.
Y. Tang, Y. Tang, Y. Zhu, J. Xiao, and R. M. Summers, “A disentangled generative model for disease decomposition in chest x-rays via normal image synthesis,” Medical Image Analysis , vol. 67, p. 101839, 2021
2021
Closest in time.
X. Li, M. Cui, J. Li, R. Bai, Z. Lu, and U. Aickelin, “A hybrid medical text classification framework: Integrating attentive rule construction and neural network,” Neurocomputing , vol. 443, pp. 345–355, 2021
2021
Closest in time.
Y. Fan, S. Zhou, Y. Li, and R. Zhang, “Deep learning approaches for extracting adverse events and indications of dietary supplements from clinical text,” Journal of the American Medical Informatics Association , vol. 28, no. 3, pp. 569–577, 2021
2021
Closest in time.
X. Xie, J. Niu, X. Liu, Z. Chen, S. Tang, and S. Yu, “A survey on incorporating domain knowledge into deep learning for medical image analysis,” Medical Image Analysis , p. 101985, 2021
2021
Closest in time.