Fetching the paper…
Reading the bibliography…
We propose a method for automatically answering questions about images by bringing together recent advances from natural language processing and computer vision.
Fuzzy sets
Zadeh, L.A.: · 1965
Earlier work this paper cites.
Exploratory data analysis
Tukey, J.W.: · 1977
Earlier work this paper cites.
Verbs semantics and lexical selection
Wu, Z., Palmer, M.: · 1994
Earlier work this paper cites.
Wordnet: a lexical database for english
Miller, G.A.: · 1995
Earlier work this paper cites.
Grounding spatial language in perception: an empirical and computational investigation
Regier, T., Carlson, L.A.: · 2001
Earlier work this paper cites.
Online learning of relaxed ccg grammars for parsing to logical form
Zettlemoyer, L.S., Collins, M.: · 2007
Earlier work this paper cites.
Interpretation of spatial language in a map navigation task
Levit, M., Roy, D.: · 2007
Earlier work this paper cites.
Situated dialogue and spatial organization: What, where… and why
Kruijff, G.J.M., Zender, H., Jensfelt, P., Christensen, H.I.: · 2007
Earlier work this paper cites.
Learning color names from real-world images
Van De Weijer, J., Schmid, C., Verbeek, J.: · 2007
Earlier work this paper cites.
Introduction to information retrieval
Manning, C.D., Raghavan, P., Schütze, H.: · 2008
Cited alongside, same era.
Inducing probabilistic ccg grammars from logical form with higher-order unification
Kwiatkowski, T., Zettlemoyer, L., Goldwater, S., Steedman, M.: · 2010
Cited alongside, same era.
Learning to follow navigational directions
Vogel, A., Jurafsky, D.: · 2010
Cited alongside, same era.
Scalable probabilistic databases with factor graphs and mcmc
Wick, M., McCallum, A., Miklau, G.: · 2010
Cited alongside, same era.
Understanding natural language commands for robotic navigation and mobile manipulation
Tellex, S., Kollar, T., Dickerson, S., Walter, M.R., Banerjee, A.G., Teller, S.J., Roy, N.: · 2011
Cited alongside, same era.
A joint model of language and perception for grounded attribute learning
Matuszek, C., Fitzgerald, N., Zettlemoyer, L., Bo, L., Fox, D.: · 2012
Learning dependency-based compositional semantics
Liang, P., Jordan, M.I., Klein, D.: · 2013
Later among the works it cites.
Jointly learning to parse and perceive: Connecting natural language to the physical world
Krishnamurthy, J., Kollar, T.: · 2013
Later among the works it cites.
Learning to parse natural language commands to a robot control system
Matuszek, C., Herbst, E., Zettlemoyer, L., Fox, D.: · 2013
Later among the works it cites.
Perceptual organization and recognition of indoor scenes from rgb-d images
Gupta, S., Arbelaez, P., Malik, J.: · 2013
Later among the works it cites.
Grounding spatial relations for human-robot interaction
Guadarrama, S., Riano, L., Golland, D., Gouhring, D., Jia, Y., Klein, D., Abbeel, P., Darrell, T.: · 2013
Later among the works it cites.
Youtube2text: Recognizing and describing arbitrary activities using semantic hierarchies and zero-shot recognition
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Indoor segmentation and support inference from rgbd images
Silberman, N., Hoiem, D., Kohli, P., Fergus, R.: · 2012
Cited alongside, same era.
Image retrieval with structured object queries using latent ranking svm
Lan, T., Yang, W., Wang, Y., Mori, G.: · 2012
Cited alongside, same era.
Guadarrama, S., Krishnamoorthy, N., Malkarnenkar, G., Mooney, R., Darrell, T., Saenko, K.: · 2013
Later among the works it cites.
What are you talking about? text-to-image coreference
Kong, C., Lin, D., Bansal, M., Urtasun, R., Fidler, S.: · 2014
Closest in time.
Deep fragment embeddings for bidirectional image sentence mapping
Karpathy, A., Joulin, A., Fei-Fei, L.: · 2014
Closest in time.