Learning distributions over logical forms for referring expression generation. In Proceedings of the 2013 conference on empirical methods in natural language processing . 1914–1925
Nicholas FitzGerald, Yoav Artzi, and Luke Zettlemoyer. 2013 · 1925
Earlier work this paper cites.
Dense Captioning with Joint Inference and Visual Context. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) . 1978–1987
Linjie Yang, Kevin Tang, Jianchao Yang, and Li-Jia Li. 2016 · 1987
Earlier work this paper cites.
A training algorithm for optimal margin classifiers. In Proceedings of the fifth annual workshop on Computational learning theory . ACM, 144–152
Bernhard E Boser, Isabelle M Guyon, and Vladimir N Vapnik. 1992 · 1992
Earlier work this paper cites.
A maximum entropy approach to natural language processing
Adam L Berger, Vincent J Della Pietra, and Stephen A Della Pietra. 1996 · 1996
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann LeCun, Léon Bottou, Yoshua Bengio, and Patrick Haffner. 1998 · 1998
Earlier work this paper cites.
A framework for multiple-instance learning. In Advances in neural information processing systems . 570–576
Oded Maron and Tomás Lozano-Pérez. 1998 · 1998
Earlier work this paper cites.
Actor-critic algorithms. In Advances in neural information processing systems . 1008–1014
Vijay R Konda and John N Tsitsiklis. 2000 · 2000
Earlier work this paper cites.
Gray scale and rotation invariant texture classification with local binary patterns. In European Conference on Computer Vision . Springer, 404–420
Timo Ojala, Matti Pietikäinen, and Topi Mäenpää. 2000 · 2000
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation. In Advances in neural information processing systems . 1057–1063
Richard S Sutton, David A McAllester, Satinder P Singh, and Yishay Mansour. 2000 · 2000
Earlier work this paper cites.
BLEU: a method for automatic evaluation of machine translation. In Proceedings of the 40th annual meeting on association for computational linguistics . Association for Computational Linguistics, 311–318
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
A neural probabilistic language model
Yoshua Bengio, Réjean Ducharme, Pascal Vincent, and Christian Jauvin. 2003 · 2003
Earlier work this paper cites.
Latent dirichlet allocation
David M Blei, Andrew Y Ng, and Michael I Jordan. 2003 · 2003
Earlier work this paper cites.
Minimum error rate training in statistical machine translation. In Proceedings of the 41st Annual Meeting on Association for Computational Linguistics-Volume 1 . Association for Computational Linguistics, 160–167
Franz Josef Och. 2003 · 2003
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries. In Text summarization branches out: Proceedings of the ACL-04 workshop , Vol. 8. Barcelona, Spain
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Automatic evaluation of machine translation quality using longest common subsequence and skip-bigram statistics. In Proceedings of the 42nd Annual Meeting on Association for Computational Linguistics . Association for Computational Linguistics, 605
Chin-Yew Lin and Franz Josef Och. 2004 · 2004
Earlier work this paper cites.
Distinctive image features from scale-invariant keypoints
David G Lowe. 2004 · 2004
Earlier work this paper cites.
Understanding inverse document frequency: on theoretical arguments for IDF
Stephen Robertson. 2004 · 2004
Earlier work this paper cites.
METEOR: An automatic metric for MT evaluation with improved correlation with human judgments. In Proceedings of the acl workshop on intrinsic and extrinsic evaluation measures for machine translation and/or summarization , Vol. 29. 65–72
Satanjeev Banerjee and Alon Lavie. 2005 · 2005
Earlier work this paper cites.
Histograms of oriented gradients for human detection. In Computer Vision and Pattern Recognition, 2005. CVPR 2005. IEEE Computer Society Conference on , Vol. 1. IEEE, 886–893
Navneet Dalal and Bill Triggs. 2005 · 2005
Earlier work this paper cites.
BLEU in characters: towards automatic MT evaluation in languages without word delimiters. In Companion Volume to the Proceedings of the Second International Joint Conference on Natural Language Processing . 81–86
Etienne Denoual and Yves Lepage. 2005 · 2005
Earlier work this paper cites.
Re-evaluation the Role of Bleu in Machine Translation Research.. In EACL , Vol. 6. 249–256
Chris Callison-Burch, Miles Osborne, and Philipp Koehn. 2006 · 2006
Earlier work this paper cites.
Generating typed dependency parses from phrase structure parses. In Proceedings of LREC , Vol. 6. Genoa Italy, 449–454
Marie-Catherine De Marneffe, Bill MacCartney, and Christopher D Manning. 2006 · 2006
Earlier work this paper cites.
The iapr tc-12 benchmark: A new evaluation resource for visual information systems. In International workshop ontoImage , Vol. 5. 10
Michael Grubinger, Paul Clough, Henning Müller, and Thomas Deselaers. 2006 · 2006
Earlier work this paper cites.
Building a semantically transparent corpus for the generation of referring expressions. In Proceedings of the Fourth International Natural Language Generation Conference . Association for Computational Linguistics, 130–132
Kees van Deemter, Ielka van der Sluis, and Albert Gatt. 2006 · 2006
Earlier work this paper cites.
Meteor, m-bleu and m-ter: Evaluation metrics for high-correlation with human rankings of machine translation output. In Proceedings of the Third Workshop on Statistical Machine Translation . Association for Computational Linguistics, 115–118
Abhaya Agarwal and Alon Lavie. 2008 · 2008
Earlier work this paper cites.
Freebase: a collaboratively created graph database for structuring human knowledge. In Proceedings of the 2008 ACM SIGMOD international conference on Management of data . AcM, 1247–1250
Kurt Bollacker, Colin Evans, Praveen Paritosh, Tim Sturge, and Jamie Taylor. 2008 · 2008
Earlier work this paper cites.
The use of spatial relations in referring expression generation. In Proceedings of the Fifth International Natural Language Generation Conference . Association for Computational Linguistics, 59–67
Jette Viethen and Robert Dale. 2008 · 2008
Earlier work this paper cites.
Generating image descriptions using dependency relational patterns. In Proceedings of the 48th annual meeting of the association for computational linguistics . Association for Computational Linguistics, 1250–1258
Ahmet Aker and Robert Gaizauskas. 2010 · 2010
Earlier work this paper cites.
Every picture tells a story: Generating sentences from images. In European conference on computer vision . Springer, 15–29
Ali Farhadi, Mohsen Hejrati, Mohammad Amin Sadeghi, Peter Young, Cyrus Rashtchian, Julia Hockenmaier, and David Forsyth. 2010 · 2010
Earlier work this paper cites.
A game-theoretic approach to generating spatial descriptions. In Proceedings of the 2010 conference on empirical methods in natural language processing . Association for Computational Linguistics, 410–419
Dave Golland, Percy Liang, and Dan Klein. 2010 · 2010
Earlier work this paper cites.
Recurrent neural network based language model. In Eleventh Annual Conference of the International Speech Communication Association
Tomáš Mikolov, Martin Karafiát, Lukáš Burget, Jan Černockỳ, and Sanjeev Khudanpur. 2010 · 2010
Earlier work this paper cites.
Natural reference to objects in a visual domain. In Proceedings of the 6th international natural language generation conference . Association for Computational Linguistics, 95–104
Margaret Mitchell, Kees van Deemter, and Ehud Reiter. 2010 · 2010
Earlier work this paper cites.
Eye movement analysis for activity recognition using electrooculography
Andreas Bulling, Jamie A Ward, Hans Gellersen, and Gerhard Troster. 2011 · 2011
Earlier work this paper cites.
Learning photographic global tonal adjustment with a database of input/output image pairs. In Computer Vision and Pattern Recognition (CVPR), 2011 IEEE Conference on . IEEE, 97–104
Vladimir Bychkovsky, Sylvain Paris, Eric Chan, and Frédo Durand. 2011 · 2011
Earlier work this paper cites.
Baby talk: Understanding and generating image descriptions. In Proceedings of the 24th CVPR . Citeseer
Girish Kulkarni, Visruth Premraj, Sagnik Dhar, Siming Li, Yejin Choi, Alexander C Berg, and Tamara L Berg. 2011 · 2011
Earlier work this paper cites.
Composing simple image descriptions using web-scale n-grams. In Proceedings of the Fifteenth Conference on Computational Natural Language Learning . Association for Computational Linguistics, 220–228
Siming Li, Girish Kulkarni, Tamara L Berg, Alexander C Berg, and Yejin Choi. 2011 · 2011
Earlier work this paper cites.
Im2text: Describing images using 1 million captioned photographs. In Advances in Neural Information Processing Systems . 1143–1151
Vicente Ordonez, Girish Kulkarni, and Tamara L Berg. 2011 · 2011
Earlier work this paper cites.
Learning to recognize daily actions using gaze. In European Conference on Computer Vision . Springer, 314–327
Alireza Fathi, Yin Li, and James M Rehg. 2012 · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks. In Advances in neural information processing systems . 1097–1105
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton. 2012 · 2012
Earlier work this paper cites.
Collective generation of natural image descriptions. In Proceedings of the 50th Annual Meeting of the Association for Computational Linguistics: Long Papers-Volume 1 . Association for Computational Linguistics, 359–368
Polina Kuznetsova, Vicente Ordonez, Alexander C Berg, Tamara L Berg, and Yejin Choi. 2012 · 2012
Earlier work this paper cites.
Active visual segmentation
Ajay K Mishra, Yiannis Aloimonos, Loong Fah Cheong, and Ashraf Kassim. 2012 · 2012
Earlier work this paper cites.
Midge: Generating image descriptions from computer vision detections. In Proceedings of the 13th Conference of the European Chapter of the Association for Computational Linguistics . Association for Computational Linguistics, 747–756
Margaret Mitchell, Xufeng Han, Jesse Dodge, Alyssa Mensch, Amit Goyal, Alex Berg, Kota Yamaguchi, Tamara Berg, Karl Stratos, and Hal Daumé III. 2012 · 2012
Earlier work this paper cites.
Image description using visual dependency representations. In Proceedings of the 2013 Conference on Empirical Methods in Natural Language Processing . 1292–1302
Desmond Elliott and Frank Keller. 2013 · 2013
Earlier work this paper cites.
Devise: A deep visual-semantic embedding model. In Advances in neural information processing systems . 2121–2129
Andrea Frome, Greg S Corrado, Jon Shlens, Samy Bengio, Jeff Dean, and Tomas Mikolov. 2013 · 2013
Earlier work this paper cites.
Framing image description as a ranking task: Data, models and evaluation metrics
Micah Hodosh, Peter Young, and Julia Hockenmaier. 2013 · 2013
Earlier work this paper cites.
From where and how to what we see. In Proceedings of the IEEE International Conference on Computer Vision . 625–632
S Karthikeyan, Vignesh Jagadeesh, Renuka Shenoy, Miguel Ecksteinz, and BS Manjunath. 2013 · 2013
Earlier work this paper cites.
Efficient estimation of word representations in vector space
Original
Tomas Mikolov, Kai Chen, Greg Corrado, and Jeffrey Dean. 2013 · 2013
Earlier work this paper cites.
Generating Expressions that Refer to Visible Objects.. In HLT-NAACL . 1174–1184
Margaret Mitchell, Kees Van Deemter, and Ehud Reiter. 2013 · 2013
Earlier work this paper cites.