Fetching the paper…
Reading the bibliography…
In this paper, we introduce the task of automatically generating text to describe the differences between two similar images.
Phrase localization and visual relationship detection with comprehensive image-language cues
Bryan A Plummer, Arun Mallya, Christopher M Cervantes, Julia Hockenmaier, and Svetlana Lazebnik. 2017 · 1937
Earlier work this paper cites.
The mathematics of statistical machine translation: Parameter estimation
Peter F Brown, Vincent J Della Pietra, Stephen A Della Pietra, and Robert L Mercer. 1993 · 1993
Earlier work this paper cites.
A density-based algorithm for discovering clusters in large spatial databases with noise
Martin Ester, Hans-Peter Kriegel, Jörg Sander, Xiaowei Xu, et al. 1996 · 1996
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Advances in automatic text summarization
Inderjeet Mani and Mark T Maybury. 1999 · 1999
Earlier work this paper cites.
Automatic analysis of the difference image for unsupervised change detection
Lorenzo Bruzzone and Diego F Prieto. 2000 · 2000
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Image change detection algorithms: a systematic survey
Richard J Radke, Srinivas Andra, Omar Al-Kofahi, and Badrinath Roysam. 2005 · 2005
Earlier work this paper cites.
A visual vocabulary for flower classification
M-E. Nilsback and A. Zisserman. 2006 · 2006
Earlier work this paper cites.
A survey of text summarization extractive techniques
Vishal Gupta and Gurpreet Singh Lehal. 2010 · 2010
Earlier work this paper cites.
Novel dataset for fine-grained image categorization: Stanford dogs
Aditya Khosla, Nityananda Jayadevaprakash, Bangpeng Yao, and Fei-Fei Li. 2011 · 2011
Earlier work this paper cites.
A large-scale benchmark dataset for event recognition in surveillance video
Sangmin Oh, Anthony Hoogs, Amitha Perera, Naresh Cuntoor, Chia-Chih Chen, Jong Taek Lee, Saurajit Mukherjee, JK Aggarwal, Hyungtae Lee, Larry Davis, et al. 2011 · 2011
Earlier work this paper cites.
Scikit-learn: Machine learning in Python
Fabian Pedregosa, Gaël Varoquaux, Alexandre Gramfort, Vincent Michel, Bertrand Thirion, Olivier Grisel, Mathieu Blondel, Peter Prettenhofer, Ron Weiss, Vincent Dubourg, et al. 2011 · 2011
Earlier work this paper cites.
Discovering a lexicon of parts and attributes
Subhransu Maji. 2012 · 2012
Earlier work this paper cites.
Meteor universal: Language specific translation evaluation for any target language
Michael Denkowski and Alon Lavie. 2014 · 2014
Cited alongside, same era.
Referitgame: Referring to objects in photographs of natural scenes
Sahar Kazemzadeh, Vicente Ordonez, Mark Matten, and Tamara Berg. 2014 · 2014
Cited alongside, same era.
Similarity comparisons for interactive fine-grained categorization
Catherine Wah, Grant Van Horn, Steve Branson, Subhransu Maji, Pietro Perona, and Serge Belongie. 2014 · 2014
Cited alongside, same era.
Vqa: Visual question answering
Stanislaw Antol, Aishwarya Agrawal, Jiasen Lu, Margaret Mitchell, Dhruv Batra, C Lawrence Zitnick, and Devi Parikh. 2015 · 2015
Cited alongside, same era.
Flickr30k entities: Collecting region-to-phrase correspondences for richer image-to-sentence models
Bryan A Plummer, Liwei Wang, Chris M Cervantes, Juan C Caicedo, Julia Hockenmaier, and Svetlana Lazebnik. 2015 · 2015
Cited alongside, same era.
What to talk about and how? selective generation using lstms with coarse-to-fine alignment
Hongyuan Mei, TTI UChicago, Mohit Bansal, and Matthew R Walter. 2016 · 2016
Later among the works it cites.
Grounding of textual phrases in images by reconstruction
Anna Rohrbach, Marcus Rohrbach, Ronghang Hu, Trevor Darrell, and Bernt Schiele. 2016 · 2016
Later among the works it cites.
Learning deep structure-preserving image-text embeddings
Liwei Wang, Yin Li, and Svetlana Lazebnik. 2016 · 2016
Later among the works it cites.
Ask, attend and answer: Exploring question-guided spatial attention for visual question answering
Huijuan Xu and Kate Saenko. 2016 · 2016
Later among the works it cites.
A generalized statistical model for binary change detection in multispectral images
Massimo Zanetti and Lorenzo Bruzzone. 2016 · 2016
Later among the works it cites.
Modeling relationships in referential expressions with compositional modular networks
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Alexander M Rush, Sumit Chopra, and Jason Weston. 2015 · 2015
Cited alongside, same era.
Cider: Consensus-based image description evaluation
Ramakrishna Vedantam, C Lawrence Zitnick, and Devi Parikh. 2015 · 2015
Cited alongside, same era.
Show, attend and tell: Neural image caption generation with visual attention
Kelvin Xu, Jimmy Ba, Ryan Kiros, Kyunghyun Cho, Aaron C Courville, Ruslan Salakhutdinov, Richard S Zemel, and Yoshua Bengio. 2015 · 2015
Cited alongside, same era.
Reasoning about pragmatics with neural listeners and speakers
Jacob Andreas and Dan Klein. 2016 · 2016
Cited alongside, same era.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Cited alongside, same era.
Globally coherent text generation with neural checklist models
Chloé Kiddon, Luke Zettlemoyer, and Yejin Choi. 2016 · 2016
Cited alongside, same era.
Neural text generation from structured data with application to the biography domain
Rémi Lebret, David Grangier, and Michael Auli. 2016 · 2016
Cited alongside, same era.
Ronghang Hu, Marcus Rohrbach, Jacob Andreas, Trevor Darrell, and Kate Saenko. 2017 · 2017
Later among the works it cites.
Automatically generating commit messages from diffs using neural machine translation
Siyuan Jiang, Ameer Armaly, and Collin McMillan. 2017 · 2017
Later among the works it cites.
An image captioning codebase in pytorch
Ruotian Luo. 2017 · 2017
Later among the works it cites.
Automatic differentiation in pytorch
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer. 2017 · 2017
Later among the works it cites.
Self-critical sequence training for image captioning
Steven J Rennie, Etienne Marcheret, Youssef Mroueh, Jarret Ross, and Vaibhava Goel. 2017 · 2017
Later among the works it cites.
Context-aware captions from context-agnostic supervision
Ramakrishna Vedantam, Samy Bengio, Kevin Murphy, Devi Parikh, and Gal Chechik. 2017 · 2017
Later among the works it cites.
Actor-critic sequence training for image captioning
Li Zhang, Flood Sung, Feng Liu, Tao Xiang, Shaogang Gong, Yongxin Yang, and Timothy M Hospedales. 2017 · 2017
Later among the works it cites.
Learning to generate move-by-move commentary for chess games from large-scale social forum data
Harsh Jhamtani, Varun Gangal, Eduard Hovy, Graham Neubig, and Taylor Berg-Kirkpatrick. 2018 · 2018
Closest in time.
’lighter’can still be dark: Modeling comparative color descriptions
Olivia Winn and Smaranda Muresan. 2018 · 2018
Closest in time.