Fetching the paper…
Reading the bibliography…
Image captioning has made substantial progress with huge supporting image collections sourced from the web.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
BLEU: a method for automatic evaluation of machine translation. In Proceedings of the 40th annual meeting on association for computational linguistics . Association for Computational Linguistics, 311–318
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Meteor 1.3: Automatic metric for reliable optimization and evaluation of machine translation systems. In Proceedings of the sixth workshop on statistical machine translation . 85–91
Michael Denkowski and Alon Lavie. 2011 · 2011
Earlier work this paper cites.
Women and news: A long and winding road
Karen Ross and Cynthia Carter. 2011 · 2011
Earlier work this paper cites.
Microsoft coco: Common objects in context. In European conference on computer vision . Springer, 740–755
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick. 2014 · 2014
Earlier work this paper cites.
Deep captioning with multimodal recurrent neural networks (m-rnn)
Junhua Mao, Wei Xu, Yi Yang, Jiang Wang, Zhiheng Huang, and Alan Yuille. 2014 · 2014
Earlier work this paper cites.
Deep inside convolutional networks: visualising image classification models and saliency maps
K Simonyan, A Vedaldi, and A Zisserman. 2014 · 2014
Earlier work this paper cites.
Long-term recurrent convolutional networks for visual recognition and description. In Proceedings of the IEEE conference on computer vision and pattern recognition . 2625–2634
Jeffrey Donahue, Lisa Anne Hendricks, Sergio Guadarrama, Marcus Rohrbach, Subhashini Venugopalan, Kate Saenko, and Trevor Darrell. 2015 · 2015
Earlier work this paper cites.
From captions to visual concepts and back. In Proceedings of the IEEE conference on computer vision and pattern recognition . 1473–1482
Hao Fang, Saurabh Gupta, Forrest Iandola, Rupesh K Srivastava, Li Deng, Piotr Dollár, Jianfeng Gao, Xiaodong He, Margaret Mitchell, John C Platt, et al · 2015
Earlier work this paper cites.
Deep visual-semantic alignments for generating image descriptions. In Proceedings of the IEEE conference on computer vision and pattern recognition . 3128–3137
Andrej Karpathy and Li Fei-Fei. 2015 · 2015
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks. In Advances in neural information processing systems . 91–99
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun. 2015 · 2015
Earlier work this paper cites.
Cider: Consensus-based image description evaluation. In Proceedings of the IEEE conference on computer vision and pattern recognition . 4566–4575
Ramakrishna Vedantam, C Lawrence Zitnick, and Devi Parikh. 2015 · 2015
Earlier work this paper cites.
Show and tell: A neural image caption generator. In Proceedings of the IEEE conference on computer vision and pattern recognition . 3156–3164
Oriol Vinyals, Alexander Toshev, Samy Bengio, and Dumitru Erhan. 2015 · 2015
Earlier work this paper cites.
It’s a Man’s Wikipedia? Assessing Gender Inequality in an Online Encyclopedia. In International AAAI Conference on Weblogs and Social Media . USA, 454–463
Claudia Wagner, David Garcia, Mohsen Jadidi, and Markus Strohmaier. 2015 · 2015
Earlier work this paper cites.
Man is to computer programmer as woman is to homemaker? debiasing word embeddings. In Advances in neural information processing systems . 4349–4357
Tolga Bolukbasi, Kai-Wei Chang, James Y Zou, Venkatesh Saligrama, and Adam T Kalai. 2016 · 2016
Earlier work this paper cites.
Equality of opportunity in supervised learning. In Advances in neural information processing systems . 3315–3323
Moritz Hardt, Eric Price, and Nati Srebro. 2016 · 2016
Earlier work this paper cites.
Attention correctness in neural image captioning. In Thirty-First AAAI Conference on Artificial Intelligence
Chenxi Liu, Junhua Mao, Fei Sha, and Alan Yuille. 2017 · 2017
Cited alongside, same era.
Knowing when to look: Adaptive attention via a visual sentinel for image captioning. In Proceedings of the IEEE conference on computer vision and pattern recognition . 375–383
Jiasen Lu, Caiming Xiong, Devi Parikh, and Richard Socher. 2017 · 2017
Cited alongside, same era.
Self-critical sequence training for image captioning. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 7008–7024
Steven J Rennie, Etienne Marcheret, Youssef Mroueh, Jerret Ross, and Vaibhava Goel. 2017 · 2017
Cited alongside, same era.
Grad-cam: Visual explanations from deep networks via gradient-based localization. In Proceedings of the IEEE international conference on computer vision . 618–626
Ramprasaath R Selvaraju, Michael Cogswell, Abhishek Das, Ramakrishna Vedantam, Devi Parikh, and Dhruv Batra. 2017 · 2017
Cited alongside, same era.
Convnets and imagenet beyond accuracy: Understanding mistakes and uncovering biases. In Proceedings of the European Conference on Computer Vision (ECCV) . 498–512
Pierre Stock and Moustapha Cisse. 2018 · 2018
Later among the works it cites.
Getting Gender Right in Neural Machine Translation. In Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing . 3003–3008
Eva Vanmassenhove, Christian Hardmeier, and Andy Way. 2018 · 2018
Later among the works it cites.
Top-down neural attention by excitation backprop
Jianming Zhang, Sarah Adel Bargal, Zhe Lin, Jonathan Brandt, Xiaohui Shen, and Stan Sclaroff. 2018 · 2018
Later among the works it cites.
Exposing and correcting the gender bias in image captioning datasets and models
Shruti Bhargava. 2019 · 2019
Later among the works it cites.
Identifying and Reducing Gender Bias in Word-Level Language Models. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Student Research Workshop . 7–15
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Men Also Like Shopping: Reducing Gender Bias Amplification using Corpus-level Constraints. In Proceedings of the 2017 Conference on Empirical Methods in Natural Language Processing . 2979–2989
Jieyu Zhao, Tianlu Wang, Mark Yatskar, Vicente Ordonez, and Kai-Wei Chang. 2017 · 2017
Cited alongside, same era.
Don’t just assume; look and answer: Overcoming priors for visual question answering. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 4971–4980
Aishwarya Agrawal, Dhruv Batra, Devi Parikh, and Aniruddha Kembhavi. 2018 · 2018
Cited alongside, same era.
Bottom-Up and Top-Down Attention for Image Captioning and Visual Question Answering. In 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE, 6077–6086
Peter Anderson, Xiaodong He, Chris Buehler, Damien Teney, Mark Johnson, Stephen Gould, and Lei Zhang. 2018 · 2018
Cited alongside, same era.
Data Statements for Natural Language Processing: Toward Mitigating System Bias and Enabling Better Science
Emily M Bender and Batya Friedman. 2018 · 2018
Cited alongside, same era.
Gender shades: Intersectional accuracy disparities in commercial gender classification. In Conference on fairness, accountability and transparency . 77–91
Joy Buolamwini and Timnit Gebru. 2018 · 2018
Cited alongside, same era.
Women also snowboard: Overcoming bias in captioning models. In European Conference on Computer Vision . Springer, 793–811
Lisa Anne Hendricks, Kaylee Burns, Kate Saenko, Trevor Darrell, and Anna Rohrbach. 2018 · 2018
Cited alongside, same era.
Tell me where to look: Guided attention inference network. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 9215–9223
Kunpeng Li, Ziyan Wu, Kuan-Chuan Peng, Jan Ernst, and Yun Fu. 2018 · 2018
Cited alongside, same era.
Neural baby talk. In Proceedings of the IEEE conference on computer vision and pattern recognition . 7219–7228
Jiasen Lu, Jianwei Yang, Dhruv Batra, and Devi Parikh. 2018 · 2018
Cited alongside, same era.
Shikha Bordia and Samuel Bowman. 2019 · 2019
Later among the works it cites.
Understanding the Origins of Bias in Word Embeddings. In International Conference on Machine Learning . 803–811
Marc-Etienne Brunet, Colleen Alkalay-Houlihan, Ashton Anderson, and Richard Zemel. 2019 · 2019
Later among the works it cites.
Equalizing Gender Bias in Neural Machine Translation with Word Embeddings Techniques. In Proceedings of the First Workshop on Gender Bias in Natural Language Processing . 147–154
Joel Escudé Font and Marta R Costa-jussà. 2019 · 2019
Later among the works it cites.
A Comprehensive Survey of Deep Learning for Image Captioning
MD Hossain, Ferdous Sohel, Mohd Fairuz Shiratuddin, and Hamid Laga. 2019 · 2019
Later among the works it cites.
Algorithmic bias? an empirical study of apparent gender-based discrimination in the display of stem career ads
Anja Lambrecht and Catherine Tucker. 2019 · 2019
Later among the works it cites.
Mitigating gender bias in natural language processing: Literature review
Tony Sun, Andrew Gaut, Shirlyn Tang, Yuxin Huang, Mai ElSherief, Jieyu Zhao, Diba Mirza, Elizabeth Belding, Kai-Wei Chang, and William Yang Wang. 2019 · 2019
Later among the works it cites.
Detection of abusive language: the problem of biased datasets. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) . 602–608
Michael Wiegand, Josef Ruppenhofer, and Thomas Kleinbauer. 2019 · 2019
Later among the works it cites.
Gender Bias in Contextualized Word Embeddings. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) . 629–634
Jieyu Zhao, Tianlu Wang, Mark Yatskar, Ryan Cotterell, Vicente Ordonez, and Kai-Wei Chang. 2019 · 2019
Later among the works it cites.
Fairness in deep learning: A computational perspective
Mengnan Du, Fan Yang, Na Zou, and Xia Hu. 2020 · 2020
Closest in time.
Shortcut Learning in Deep Neural Networks
Robert Geirhos, Jörn-Henrik Jacobsen, Claudio Michaelis, Richard Zemel, Wieland Brendel, Matthias Bethge, and Felix A Wichmann. 2020 · 2020
Closest in time.
Show, attend and tell: Neural image caption generation with visual attention. In International conference on machine learning . 2048–2057
Kelvin Xu, Jimmy Ba, Ryan Kiros, Kyunghyun Cho, Aaron Courville, Ruslan Salakhudinov, Rich Zemel, and Yoshua Bengio. 2015 · 2057
Closest in time.