Fetching the paper…
Reading the bibliography…
Progress in language and image understanding by machines has sparkled the interest of the research community in more open-ended, holistic tasks, and refueled an old AI dream of building intelligent machines.
Computing machinery and intelligence
Turing, A. M. (1950) · 1950
Earlier work this paper cites.
Verbs semantics and lexical selection
Wu, Z. and Palmer, M. (1994) · 1994
Earlier work this paper cites.
Wordnet: a lexical database for english
Miller, G. A. (1995) · 1995
Earlier work this paper cites.
Space in language and cognition: Explorations in cognitive diversity
Levinson, S. C. (2003) · 2003
Earlier work this paper cites.
The pascal visual object classes (voc) challenge
Everingham, M., Van Gool, L., Williams, C. K. I., Winn, J., and Zisserman, A. (2010) · 2010
Cited alongside, same era.
Imagenet classification with deep convolutional neural networks
Krizhevsky, A., Sutskever, I., and Hinton, G. E. (2012) · 2012
Cited alongside, same era.
Jointly learning to parse and perceive: Connecting natural language to the physical world
Krishnamurthy, J. and Kollar, T. (2013) · 2013
Cited alongside, same era.
Learning dependency-based compositional semantics
Liang, P., Jordan, M. I., and Klein, D. (2013) · 2013
Cited alongside, same era.
A multi-world approach to question answering about real-world scenes based on uncertain input
Malinowski, M. and Fritz, M. (2014a)
Cited in the paper.
A pooling approach to modelling spatial relations for image retrieval and annotation
Malinowski, M. and Fritz, M. (2014b)
Cited in the paper.
Towards a visual turing challenge
Malinowski, M. and Fritz, M. (2014c)
Cited in the paper.
Long-term recurrent convolutional networks for visual recognition and description
Donahue, J., Hendricks, L. A., Guadarrama, S., Rohrbach, M., Venugopalan, S., Saenko, K., and Darrell, T. (2014) · 2014
Later among the works it cites.
Deep visual-semantic alignments for generating image descriptions
Karpathy, A. and Fei-Fei, L. (2014) · 2014
Later among the works it cites.
Deep fragment embeddings for bidirectional image sentence mapping
Karpathy, A., Joulin, A., and Fei-Fei, L. (2014) · 2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…