Fetching the paper…
Reading the bibliography…
We use coherence relations inspired by computational models of discourse to study the information needs and goals of image captioning.
A calculus of cohesion
B Phillips. 1977 · 1977
Earlier work this paper cites.
Why is discourse coherent
Jerry R Hobbs. 1978 · 1978
Earlier work this paper cites.
Meta-talk: Organizational and evaluative brackets in discourse
Deborah Schiffrin. 1980 · 1980
Earlier work this paper cites.
On the coherence and structure of discourse
Jerry R. Hobbs. 1985 · 1985
Earlier work this paper cites.
Understanding comics: The invisible art
Scott McCloud. 1993 · 1993
Earlier work this paper cites.
Discourse relations: A structural and presuppositional account using lexicalised tag
Bonnie Webber, Alistair Knott, Matthew Stone, and Aravind Joshi. 1999 · 1999
Earlier work this paper cites.
A taxonomy of relationships between images and text
Emily E Marsh and Marilyn Domas White. 2003 · 2003
Earlier work this paper cites.
The Penn discourse treebank 2.0
Rashmi Prasad, Nikhil Dinesh, Alan Lee, Eleni Miltsakaki, Livio Robaldo, Aravind K Joshi, and Bonnie L Webber. 2008 · 2008
Earlier work this paper cites.
A formal semantic analysis of gesture
Alex Lascarides and Matthew Stone. 2009 · 2009
Earlier work this paper cites.
Information structure: Towards an integrated formal theory of pragmatics
Craige Roberts. 2012 · 2012
Earlier work this paper cites.
Baselines and bigrams: Simple, good sentiment and topic classification
Sida Wang and Christopher D Manning. 2012 · 2012
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg S Corrado, and Jeff Dean. 2013 · 2013
Earlier work this paper cites.
Deixis (even without pointing)
Una Stojnic, Matthew Stone, and Ernest Lepore. 2013 · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick. 2014 · 2014
Cited alongside, same era.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher D. Manning. 2014 · 2014
Cited alongside, same era.
From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions
Peter Young, Alice Lai, Micah Hodosh, and Julia Hockenmaier. 2014 · 2014
Cited alongside, same era.
CIDEr: Consensus-based image description evaluation
Ramakrishna Vedantam, C Lawrence Zitnick, and Devi Parikh. 2015 · 2015
Cited alongside, same era.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Cited alongside, same era.
Guided open vocabulary image captioning with constrained beam search
Object hallucination in image captioning
Anna Rohrbach, Lisa Anne Hendricks, Kaylee Burns, Trevor Darrell, and Kate Saenko. 2018 · 2018
Later among the works it cites.
Conceptual captions: A cleaned, hypernymed, image alt-text dataset for automatic image captioning
Piyush Sharma, Nan Ding, Sebastian Goodman, and Radu Soricut. 2018 · 2018
Later among the works it cites.
Cite: A corpus of image-text discourse relations
Malihe Alikhani, Sreyasi Nag Chowdhury, Gerard de Melo, and Matthew Stone. 2019 · 2019
Later among the works it cites.
Show, control and tell: A framework for generating controllable and grounded captions
Marcella Cornia, Lorenzo Baraldi, and Rita Cucchiara. 2019 · 2019
Later among the works it cites.
Mscap: Multi-style image captioning with unpaired stylized text
Longteng Guo, Jing Liu, Peng Yao, Jiangwei Li, and Hanqing Lu. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Peter Anderson, Basura Fernando, Mark Johnson, and Stephen Gould. 2017 · 2017
Cited alongside, same era.
Attend to you: Personalized image captioning with context sequence memory networks
Cesc Chunseong Park, Byeongchang Kim, and Gunhee Kim. 2017 · 2017
Cited alongside, same era.
Conventions of viewpoint coherence in film
Samuel Cumming, Gabriel Greenberg, and Rory Kelly. 2017 · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
Universal sentence encoder for English
Daniel Cer, Yinfei Yang, Sheng-yi Kong, Nan Hua, Nicole Limtiaco, Rhomni St. John, Noah Constant, Mario Guajardo-Cespedes, Steve Yuan, Chris Tar, Brian Strope, and Ray Kurzweil. 2018 · 2018
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Cited alongside, same era.
A formal semantics for situated conversation
Julia Hunter, Nicholas Asher, and Alex Lascarides. 2018 · 2018
Cited alongside, same era.
Integrating text and image: Determining multimodal document intent in Instagram posts
Julia Kruk, Jonah Lubin, Karan Sikka, Xiao Lin, Dan Jurafsky, and Ajay Divakaran. 2019 · 2019
Later among the works it cites.
Best practices for the human evaluation of automatically generated text
Chris van der Lee, Albert Gatt, Emiel van Miltenburg, Sander Wubben, and Emiel Krahmer. 2019 · 2019
Later among the works it cites.
Understanding, categorizing and predicting semantic image-text relations
Christian Otto, Matthias Springstein, Avishek Anand, and Ralph Ewerth. 2019 · 2019
Later among the works it cites.
Can neural image captioning be controlled via forced attention?
Philipp Sadler, Tatjana Scheffler, and David Schlangen. 2019 · 2019
Later among the works it cites.
Engaging image captioning via personality
Kurt Shuster, Samuel Humeau, Hexiang Hu, Antoine Bordes, and Jason Weston. 2019 · 2019
Later among the works it cites.
Categorizing and inferring the relationship between the text and image of twitter posts
Alakananda Vempala and Daniel Preotiuc-Pietro. 2019 · 2019
Later among the works it cites.
Ultra fine-grained image semantic embedding
Da-Cheng Juan, Chun-Ta Lu, Zhen Li, Futang Peng, Aleksei Timofeev, Yi-Ting Chen, Yaxi Gao, Tom Duerig, Andrew Tomkins, and Sujith Ravi. 2020 · 2020
Closest in time.
The open images dataset v4
Alina Kuznetsova, Hassan Rom, Neil Alldrin, Jasper Uijlings, Ivan Krasin, Jordi Pont-Tuset, Shahab Kamali, Stefan Popov, Matteo Malloci, Alexander Kolesnikov, et al. 2020 · 2020
Closest in time.