Fetching the paper…
Reading the bibliography…
Automatically generating medical reports for retinal images is one of the promising ways to help ophthalmologists reduce their workload and improve work efficiency.
“Long short-term memory,”
Sepp Hochreiter and Jürgen Schmidhuber, · 1997
Earlier work this paper cites.
“Bleu: a method for automatic evaluation of machine translation,”
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu, · 2002
Earlier work this paper cites.
“Vision 2020: The right to sight: a global initiative to eliminate avoidable blindness,”
Louis Pizzarello, Adenike Abiose, Timothy Ffytche, Rainaldo Duerksen, R Thulasiraj, Hugh Taylor, Hannah Faal, Gullapali Rao, Ivo Kocur, and Serge Resnikoff, · 2004
Earlier work this paper cites.
“Rouge: A package for automatic evaluation of summaries,”
Chin-Yew Lin, · 2004
Earlier work this paper cites.
“Show, attend and tell: Neural image caption generation with visual attention,”
Kelvin Xu, Jimmy Ba, Ryan Kiros, Kyunghyun Cho, Aaron Courville, Ruslan Salakhutdinov, Richard Zemel, and Yoshua Bengio, · 2015
Earlier work this paper cites.
“Deep visual-semantic alignments for generating image descriptions,”
Andrej Karpathy and Li Fei-Fei, · 2015
Earlier work this paper cites.
“Show and tell: A neural image caption generator,”
Oriol Vinyals, Alexander Toshev, Samy Bengio, and Dumitru Erhan, · 2015
Earlier work this paper cites.
“From captions to visual concepts and back,”
Hao Fang, Saurabh Gupta, Forrest Iandola, Rupesh K Srivastava, Li Deng, Piotr Dollár, Jianfeng Gao, Xiaodong He, Margaret Mitchell, John C Platt, et al., · 2015
Earlier work this paper cites.
“Cider: Consensus-based image description evaluation,”
Ramakrishna Vedantam, C Lawrence Zitnick, and Devi Parikh, · 2015
Earlier work this paper cites.
“Ask your neurons: A neural-based approach to answering questions about images,”
Mateusz Malinowski, Marcus Rohrbach, and Mario Fritz, · 2015
Earlier work this paper cites.
“Generating visual explanations,”
Lisa Anne Hendricks, Zeynep Akata, Marcus Rohrbach, Jeff Donahue, Bernt Schiele, and Trevor Darrell, · 2016
Earlier work this paper cites.
“Spice: Semantic propositional image caption evaluation,”
Peter Anderson, Basura Fernando, Mark Johnson, and Stephen Gould, · 2016
Earlier work this paper cites.
“Robustness analysis of visual question answering models by basic questions,”
Jia-Hong Huang, · 2017
Earlier work this paper cites.
“Vqabq: Visual question answering by basic questions,”
Jia-Hong Huang, Modar Alfadly, and Bernard Ghanem, · 2017
Cited alongside, same era.
“Vqa: Visual question answering,”
Aishwarya Agrawal, Jiasen Lu, Stanislaw Antol, Margaret Mitchell, C Lawrence Zitnick, Devi Parikh, and Dhruv Batra, · 2017
Cited alongside, same era.
“Improved image captioning via policy gradient optimization of spider,”
Siqi Liu, Zhenhai Zhu, Ning Ye, Sergio Guadarrama, and Kevin Murphy, · 2017
Cited alongside, same era.
“Textray: Mining clinical reports to gain a broad understanding of chest x-rays,”
Jonathan Laserson, Christine Dan Lantsman, Michal Cohen-Sfady, Itamar Tamir, Eli Goz, Chen Brestel, Shir Bar, Maya Atar, and Eldad Elnekave, · 2018
Cited alongside, same era.
“On the automatic generation of medical imaging reports,”
Baoyu Jing, Pengtao Xie, Eric Xing, Baoyu Jing, Pengtao Xie, and Eric Xing, · 2018
Cited alongside, same era.
“Hybrid retrieval-generation reinforced agent for medical image report generation,”
“Assessing the robustness of visual question answering,”
Jia-Hong Huang, Modar Alfadly, Bernard Ghanem, and Marcel Worring, · 2019
Later among the works it cites.
“Deliberate attention networks for image captioning,”
Lianli Gao, Kaixuan Fan, Jingkuan Song, Xianglong Liu, Xing Xu, and Heng Tao Shen, · 2019
Later among the works it cites.
“Knowledge-driven encode, retrieve, paraphrase for medical image report generation,”
Christy Y Li, Xiaodan Liang, Zhiting Hu, and Eric P Xing, · 2019
Later among the works it cites.
“Vascular inflammation risk factors in retinal disease,”
Ileana Soto, Mark P Krebs, Alaina M Reagan, and Gareth R Howell, · 2019
Later among the works it cites.
“Characterizing speech adversarial examples using self-attention u-net enhancement,”
Chao-Han Yang, Jun Qi, Pin-Yu Chen, Xiaoli Ma, and Chin-Hui Lee, · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yuan Li, Xiaodan Liang, Zhiting Hu, and Eric P Xing, · 2018
Cited alongside, same era.
“Robustness analysis of visual qa models by basic questions,”
Jia-Hong Huang, Cuong Duc Dao, Modar Alfadly, C Huck Yang, and Bernard Ghanem, · 2018
Cited alongside, same era.
“Synthesizing new retinal symptom images by multiple generative models,”
Yi-Chieh Liu, Hao-Hsiang Yang, C-H Huck Yang, Jia-Hong Huang, Meng Tian, Hiromasa Morikawa, Yi-Chang James Tsai, and Jesper Tegner, · 2018
Cited alongside, same era.
“Auto-classification of retinal diseases in the limit of sparse data using a two-streams machine learning model,”
C-H Huck Yang, Fangyu Liu, Jia-Hong Huang, Meng Tian, MD I-Hung Lin, Yi Chieh Liu, Hiromasa Morikawa, Hao-Hsiang Yang, and Jesper Tegner, · 2018
Cited alongside, same era.
“A novel hybrid machine learning model for auto-classification of retinal diseases,”
C-H Huck Yang, Jia-Hong Huang, Fangyu Liu, Fang-Yi Chiu, Mengya Gao, Weifeng Lyu, Jesper Tegner, et al., · 2018
Cited alongside, same era.
“A novel framework for robustness analysis of visual qa models,”
Jia-Hong Huang, Cuong Duc Dao, Modar Alfadly, and Bernard Ghanem, · 2019
Cited alongside, same era.
“Silco: Show a few images, localize the common object,”
Tao Hu, Pascal Mettes, Jia-Hong Huang, and Cees GM Snoek, · 2019
Cited alongside, same era.
“Query-controllable video summarization,”
Jia-Hong Huang and Marcel Worring, · 2020
Later among the works it cites.
“X-linear attention networks for image captioning,”
Yingwei Pan, Ting Yao, Yehao Li, and Tao Mei, · 2020
Later among the works it cites.
“Deepopht: medical report generation for retinal images via deep models and visual explanation,”
Jia-Hong Huang, C-H Huck Yang, Fangyu Liu, Meng Tian, Yi-Chieh Liu, Ting-Wei Wu, I Lin, Kang Wang, Hiromasa Morikawa, Hernghua Chang, et al., · 2021
Closest in time.
Chao-Han Huck Yang, Sabato Marco Siniscalchi, and Chin-Hui Lee, · 2021
Closest in time.
“Deep context-encoding network for retinal image captioning,”
Jia-Hong Huang, Ting-Wei Wu, Chao-Han Huck Yang, and Marcel Worring, · 2021
Closest in time.
“Contextualized keyword representations for multi-modal retinal image captioning,”
Jia-Hong Huang, Ting-Wei Wu, and Marcel Worring, · 2021
Closest in time.
“Gpt2mvs: Generative pre-trained transformer-2 formulti-modal video summarization,”
Jia-Hong Huang, Luka Murn, Marta Mrak, and Marcel Worring, · 2021
Closest in time.