Fetching the paper…
Reading the bibliography…
Image captioning research achieved breakthroughs in recent years by developing neural models that can generate diverse and high-quality descriptions for images drawn from the same distribution as training images.
On a measure of divergence between two statistical populations defined by their probability distributions
Anil Bhattacharyya. 1943 · 1943
Earlier work this paper cites.
BLEU: a method for automatic evaluation of machine translation. In Proceedings of the 40th annual meeting on association for computational linguistics . Association for Computational Linguistics, 311–318
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
The Bhattacharyya space for feature selection and its application to texture segmentation
Constantino Carlos Reyes-Aldasoro and Abhir Bhalerao. 2006 · 2006
Earlier work this paper cites.
Every picture tells a story: Generating sentences from images. In European conference on computer vision . Springer, 15–29
Ali Farhadi, Mohsen Hejrati, Mohammad Amin Sadeghi, Peter Young, Cyrus Rashtchian, Julia Hockenmaier, and David Forsyth. 2010 · 2010
Earlier work this paper cites.
Collective generation of natural image descriptions. In Proceedings of the 50th Annual Meeting of the Association for Computational Linguistics: Long Papers-Volume 1 . Association for Computational Linguistics, 359–368
Polina Kuznetsova, Vicente Ordonez, Alexander C Berg, Tamara L Berg, and Yejin Choi. 2012 · 2012
Earlier work this paper cites.
Framing image description as a ranking task: Data, models and evaluation metrics
Micah Hodosh, Peter Young, and Julia Hockenmaier. 2013 · 2013
Earlier work this paper cites.
Microsoft coco: Common objects in context. In European conference on computer vision . Springer, 740–755
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick. 2014 · 2014
Earlier work this paper cites.
Deep neural networks are easily fooled: High confidence predictions for unrecognizable images. In Proceedings of the IEEE conference on computer vision and pattern recognition . 427–436
Anh Nguyen, Jason Yosinski, and Jeff Clune. 2015 · 2015
Earlier work this paper cites.
Cider: Consensus-based image description evaluation. In Proceedings of the IEEE conference on computer vision and pattern recognition . 4566–4575
Ramakrishna Vedantam, C Lawrence Zitnick, and Devi Parikh. 2015 · 2015
Earlier work this paper cites.
Show and tell: A neural image caption generator. In Proceedings of the IEEE conference on computer vision and pattern recognition . 3156–3164
Oriol Vinyals, Alexander Toshev, Samy Bengio, and Dumitru Erhan. 2015 · 2015
Earlier work this paper cites.
Rich image captioning in the wild. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops . 49–56
Kenneth Tran, Xiaodong He, Lei Zhang, Jian Sun, Cornelia Carapcea, Chris Thrasher, Chris Buehler, and Chris Sienkiewicz. 2016 · 2016
Earlier work this paper cites.
On calibration of modern neural networks. In International Conference on Machine Learning . PMLR, 1321–1330
Chuan Guo, Geoff Pleiss, Yu Sun, and Kilian Q Weinberger. 2017 · 2017
Earlier work this paper cites.
A Baseline for Detecting Misclassified and Out-of-Distribution Examples in Neural Networks
Dan Hendrycks and Kevin Gimpel. 2017 · 2017
Earlier work this paper cites.
Openimages: A public dataset for large-scale multi-label and multi-class image classification
Ivan Krasin, Tom Duerig, Neil Alldrin, Vittorio Ferrari, Sami Abu-El-Haija, Alina Kuznetsova, Hassan Rom, Jasper Uijlings, Stefan Popov, Andreas Veit, et al · 2017
Earlier work this paper cites.
Understanding blind people’s experiences with computer-generated captions of social media images. In Proceedings of the 2017 CHI Conference on Human Factors in Computing Systems . ACM, 5988–5999
Haley MacLeod, Cynthia L Bennett, Meredith Ringel Morris, and Edward Cutrell. 2017 · 2017
Cited alongside, same era.
Bottom-up and top-down attention for image captioning and visual question answering. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 6077–6086
Peter Anderson, Xiaodong He, Chris Buehler, Damien Teney, Mark Johnson, Stephen Gould, and Lei Zhang. 2018 · 2018
Cited alongside, same era.
Real-Time Self-Driving Car Navigation Using Deep Neural Network. In 2018 4th International Conference on Green Technology and Sustainable Development (GTSD) . IEEE, 7–12
Truong-Dong Do, Minh-Thien Duong, Quoc-Vu Dang, and My-Ha Le. 2018 · 2018
Cited alongside, same era.
Hierarchical Neural Story Generation. In Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Association for Computational Linguistics, Melbourne, Australia, 889–898
Angela Fan, Mike Lewis, and Yann Dauphin. 2018 · 2018
Generating Diverse and Informative Natural Language Fashion Feedback
Gil Sadeh, Lior Fritz, Gabi Shalev, and Eduard Oks. 2019a · 2019
Later among the works it cites.
Joint visual-textual embedding for multimodal style search
Gil Sadeh, Lior Fritz, Gabi Shalev, and Eduard Oks. 2019b · 2019
Later among the works it cites.
Applications of artificial neural networks in health care organizational decision-making: A scoping review
Nida Shahid, Tim Rappon, and Whitney Berta. 2019 · 2019
Later among the works it cites.
Meshed-memory transformer for image captioning. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 10578–10587
Marcella Cornia, Matteo Stefanini, Lorenzo Baraldi, and Rita Cucchiara. 2020 · 2020
Later among the works it cites.
Pretrained Transformers Improve Out-of-Distribution Robustness. In Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics . Association for Computational Linguistics, Online, 2744–2751
Dan Hendrycks, Xiaoyuan Liu, Eric Wallace, Adam Dziedzic, Rishabh Krishnan, and Dawn Song. 2020 · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Actigraphy-based sleep/wake pattern detection using convolutional neural networks
Lena Granovsky, Gabi Shalev, Nancy Yacovzada, Yotam Frank, and Shai Fine. 2018 · 2018
Cited alongside, same era.
Training confidence-calibrated classifiers for detecting out-of-distribution samples
Kimin Lee, Honglak Lee, Kibok Lee, and Jinwoo Shin. 2018a · 2018
Cited alongside, same era.
A simple unified framework for detecting out-of-distribution samples and adversarial attacks
Kimin Lee, Kibok Lee, Honglak Lee, and Jinwoo Shin. 2018b · 2018
Cited alongside, same era.
Enhancing the reliability of out-of-distribution image detection in neural networks
Shiyu Liang, Yixuan Li, and Rayadurgam Srikant. 2018 · 2018
Cited alongside, same era.
Neural baby talk. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 7219–7228
Jiasen Lu, Jianwei Yang, Dhruv Batra, and Devi Parikh. 2018 · 2018
Cited alongside, same era.
Out-of-distribution detection using multiple semantic label representations. In Advances in Neural Information Processing Systems . 7375–7385
Gabi Shalev, Yossi Adi, and Joseph Keshet. 2018 · 2018
Cited alongside, same era.
nocaps: novel object captioning at scale. In Proceedings of the IEEE International Conference on Computer Vision . 8948–8957
Harsh Agrawal, Karan Desai, Yufei Wang, Xinlei Chen, Rishabh Jain, Mark Johnson, Dhruv Batra, Devi Parikh, Stefan Lee, and Peter Anderson. 2019 · 2019
Cited alongside, same era.
Benchmarking Neural Network Robustness to Common Corruptions and Perturbations
Dan Hendrycks and Thomas Dietterich. 2019 · 2019
Cited alongside, same era.
Later among the works it cites.
The curious case of neural text degeneration
Ari Holtzman, Jan Buys, Maxwell Forbes, and Yejin Choi. 2020 · 2020
Later among the works it cites.
Generalized odin: Detecting out-of-distribution image without learning from out-of-distribution data. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 10951–10960
Yen-Chang Hsu, Yilin Shen, Hongxia Jin, and Zsolt Kira. 2020 · 2020
Later among the works it cites.
Redesigning the classification layer by randomizing the class representation vectors
Gabi Shalev, Gal-Lev Shalev, and Joseph Keshet. 2020 · 2020
Later among the works it cites.
Self-supervised out-of-distribution detection in brain CT scans
Abinav Ravi Venkatakrishnan, Seong Tae Kim, Rami Eisawy, Franz Pfister, and Nassir Navab. 2020 · 2020
Later among the works it cites.
Improving image captioning by leveraging intra-and inter-layer global representation in transformer network. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 35. 1655–1663
Jiayi Ji, Yunpeng Luo, Xiaoshuai Sun, Fuhai Chen, Gen Luo, Yongjian Wu, Yue Gao, and Rongrong Ji. 2021 · 2021
Later among the works it cites.
Oodformer: Out-of-distribution detection transformer
Rajat Koner, Poulami Sinhamahapatra, Karsten Roscher, Stephan Günnemann, and Volker Tresp. 2021 · 2021
Later among the works it cites.
k k Folden: k k -Fold Ensemble for Out-Of-Distribution Detection
Xiaoya Li, Jiwei Li, Xiaofei Sun, Chun Fan, Tianwei Zhang, Fei Wu, Yuxian Meng, and Jun Zhang. 2021 · 2021
Later among the works it cites.
On Randomized Classification Layers and Their Implications in Natural Language Generation. In Proceedings of the Third Workshop on Multimodal Artificial Intelligence . 6–11
Gal-Lev Shalev, Gabi Shalev, and Joseph Keshet. 2021 · 2021
Later among the works it cites.
Image captioning as an assistive technology: Lessons learned from vizwiz 2020 challenge
Pierre Dognin, Igor Melnyk, Youssef Mroueh, Inkit Padhi, Mattia Rigotti, Jarret Ross, Yair Schiff, Richard A Young, and Brian Belgodere. 2022 · 2022
Closest in time.