Z. Wu and M. Palmer, “Verbs semantics and lexical selection,” in Proc. Conf. Association for Computational Linguistics , 1994
1994
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner, “Gradient-based learning applied to document recognition,” Proc. IEEE , vol. 86, no. 11, pp. 2278–2324, 1998
1998
Earlier work this paper cites.
K. Papineni, S. Roukos, T. Ward, and W.-J. Zhu, “BLEU: a method for automatic evaluation of machine translation,” in Proc. Conf. Association for Computational Linguistics , 2002
2002
Earlier work this paper cites.
C. Zhang, J. C. Platt, and P. A. Viola, “Multiple instance boosting for object detection,” in Proc. Advances in Neural Inf. Process. Syst. , 2005
2005
Earlier work this paper cites.
S. Banerjee and A. Lavie, “METEOR: An automatic metric for MT evaluation with improved correlation with human judgments,” in Proceedings of the acl workshop on intrinsic and extrinsic evaluation measures for machine translation and/or summarization , 2005
2005
Earlier work this paper cites.
J. Vogel and B. Schiele, “Semantic modeling of natural scenes for content-based image retrieval,” Int. J. Comput. Vision , vol. 72, no. 2, pp. 133–157, 2007
2007
Earlier work this paper cites.
S. Auer, C. Bizer, G. Kobilarov, J. Lehmann, R. Cyganiak, and Z. Ives, Dbpedia: A nucleus for a web of open data . Springer, 2007
2007
Earlier work this paper cites.
K. Bollacker, C. Evans, P. Paritosh, T. Sturge, and J. Taylor, “Freebase: a collaboratively created graph database for structuring human knowledge,” in Proc. ACM SIGMOD Int. Conf. Management of Data , 2008
2008
Earlier work this paper cites.
A. Farhadi, I. Endres, D. Hoiem, and D. Forsyth, “Describing objects by their attributes,” in Proc. IEEE Conf. Comp. Vis. Patt. Recogn. , 2009
2009
Earlier work this paper cites.
L.-J. Li, H. Su, L. Fei-Fei, and E. P. Xing, “Object bank: A high-level image representation for scene classification & semantic feature sparsification,” in Proc. Advances in Neural Inf. Process. Syst. , 2010, pp. 1378–1386
2010
Earlier work this paper cites.
A. Farhadi, M. Hejrati, M. A. Sadeghi, P. Young, C. Rashtchian, J. Hockenmaier, and D. Forsyth, “Every picture tells a story: Generating sentences from images,” in Proc. Eur. Conf. Comp. Vis. , 2010
2010
Earlier work this paper cites.
B. Z. Yao, X. Yang, L. Lin, M. W. Lee, and S.-C. Zhu, “I2t: Image parsing to text description,” Proc. IEEE , vol. 98, no. 8, pp. 1485–1508, 2010
2010
Earlier work this paper cites.
D. Ferrucci, E. Brown, J. Chu-Carroll, J. Fan, D. Gondek, A. A. Kalyanpur, A. Lally, J. W. Murdock, E. Nyberg, J. Prager et al. , “Building Watson: An overview of the DeepQA project,” AI magazine , vol. 31, no. 3, pp. 59–79, 2010
2010
Earlier work this paper cites.
X. Glorot and Y. Bengio, “Understanding the difficulty of training deep feedforward neural networks,” in Proc. Int. Conf. Artificial Intell. & Stat. , 2010, pp. 249–256
2010
Earlier work this paper cites.
Y. Jia, M. Salzmann, and T. Darrell, “Learning cross-modality similarity for multinomial data,” in Proc. IEEE Int. Conf. Comp. Vis. , 2011
2011
Earlier work this paper cites.
V. Ordonez, G. Kulkarni, and T. L. Berg, “Im2text: Describing images using 1 million captioned photographs,” in Proc. Advances in Neural Inf. Process. Syst. , 2011
2011
Earlier work this paper cites.
S. Li, G. Kulkarni, T. L. Berg, A. C. Berg, and Y. Choi, “Composing simple image descriptions using web-scale n-grams,” in Proc. Conf. Computational Natural Language Learning , 2011
2011
Earlier work this paper cites.
Y. Yang, C. L. Teo, H. Daumé III, and Y. Aloimonos, “Corpus-guided sentence generation of natural images,” in Proc. Conf. Empirical Methods in Natural Language Processing , 2011
2011
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in Proc. Advances in Neural Inf. Process. Syst. , 2012
2012
Earlier work this paper cites.
Y. Su and F. Jurie, “Improving image classification using semantic attributes,” Int. J. Comput. Vision , vol. 100, no. 1, pp. 59–77, 2012
2012
Earlier work this paper cites.
L.-J. Li, H. Su, Y. Lim, and L. Fei-Fei, “Objects as attributes for scene classification,” in Trends and Topics in Computer Vision . Springer, 2012, pp. 57–69
2012
Earlier work this paper cites.
M. Hodosh, P. Young, and J. Hockenmaier, “Framing image description as a ranking task: Data, models and evaluation metrics,” J. Arti. Intell. Res. , pp. 853–899, 2013
2013
Earlier work this paper cites.
G. Kulkarni, V. Premraj, V. Ordonez, S. Dhar, S. Li, Y. Choi, A. C. Berg, and T. L. Berg, “Babytalk: Understanding and generating simple image descriptions,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 35, no. 12, pp. 2891–2903, 2013
2013
Earlier work this paper cites.
M. Rohrbach, W. Qiu, I. Titov, S. Thater, M. Pinkal, and B. Schiele, “Translating video content to natural language descriptions,” in Proc. IEEE Int. Conf. Comp. Vis. , 2013
2013
Earlier work this paper cites.
J. Berant, A. Chou, R. Frostig, and P. Liang, “Semantic Parsing on Freebase from Question-Answer Pairs.” in Proc. Conf. Empirical Methods in Natural Language Processing , 2013, pp. 1533–1544
2013
Earlier work this paper cites.
K. Cho, B. van Merrienboer, C. Gulcehre, F. Bougares, H. Schwenk, and Y. Bengio, “Learning phrase representations using rnn encoder-decoder for statistical machine translation,” in Proc. Conf. Empirical Methods in Natural Language Processing , 2014
2014
Earlier work this paper cites.
I. Sutskever, O. Vinyals, and Q. V. Le, “Sequence to sequence learning with neural networks,” in Proc. Advances in Neural Inf. Process. Syst. , 2014
2014
Earlier work this paper cites.