Fetching the paper…
Reading the bibliography…
Visual Question Answering (VQA) has attracted a lot of attention in both Computer Vision and Natural Language Processing communities, not least because it offers insight into the relationships between two important sources of information.
Z. Wu and M. Palmer, “Verbs semantics and lexical selection,” in Proc. Conf. the Association for Computational Linguistics , 1994
1994
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
H. Liu and P. Singh, “ConceptNet—a practical commonsense reasoning tool-kit,” BT technology journal , vol. 22, no. 4, pp. 211–226, 2004
2004
Earlier work this paper cites.
L. S. Zettlemoyer and M. Collins, “Learning to map sentences to logical form: Structured classification with probabilistic categorial grammars,” in Proc. Uncertainty in Artificial Intell. , 2005
2005
Earlier work this paper cites.
——, “Learning context-dependent mappings from sentences to logical form,” in Proceedings of ACL-IJCNLP , 2005
2005
Earlier work this paper cites.
S. Auer, C. Bizer, G. Kobilarov, J. Lehmann, R. Cyganiak, and Z. Ives, Dbpedia: A nucleus for a web of open data . Springer, 2007
2007
Earlier work this paper cites.
M. Banko, M. J. Cafarella, S. Soderland, M. Broadhead, and O. Etzioni, “Open information extraction for the web,” in Proc. Int. Joint Conf. on Artificial Intell. , 2007
2007
Earlier work this paper cites.
K. Bollacker, C. Evans, P. Paritosh, T. Sturge, and J. Taylor, “Freebase: a collaboratively created graph database for structuring human knowledge,” in Proc. ACM SIGMOD/PODS Conf. , 2008, pp. 1247–1250
2008
Earlier work this paper cites.
E. Prud’Hommeaux, A. Seaborne et al. , “SPARQL query language for RDF,” W3C recommendation , vol. 15, 2008
2008
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in Proc. IEEE Conf. Comp. Vis. Patt. Recogn. , 2009
2009
Earlier work this paper cites.
A. Carlson, J. Betteridge, B. Kisiel, and B. Settles, “Toward an Architecture for Never-Ending Language Learning.” in Proc. National Conf. Artificial Intell. , 2010
2010
Earlier work this paper cites.
T. Mikolov, M. Karafiát, L. Burget, J. Cernockỳ, and S. Khudanpur, “Recurrent neural network based language model.” in Interspeech , vol. 2, 2010, p. 3
2010
Earlier work this paper cites.
O. Etzioni, A. Fader, J. Christensen, S. Soderland, and M. Mausam, “Open Information Extraction: The Second Generation.” in Proc. Int. Joint Conf. on Artificial Intell. , 2011
2011
Earlier work this paper cites.
A. Fader, S. Soderland, and O. Etzioni, “Identifying relations for open information extraction,” in Proc. Conf. Empirical Methods Natural Language Processing , 2011
2011
Earlier work this paper cites.
O. Kolomiyets and M.-F. Moens, “A survey on question answering technology from an information retrieval perspective,” Information Sciences , vol. 181, no. 24, pp. 5412–5434, 2011
2011
Earlier work this paper cites.
C.-C. Chang and C.-J. Lin, “LIBSVM: a library for support vector machines,” ACM Transactions on Intelligent Systems and Technology (TIST) , vol. 2, no. 3, p. 27, 2011
2011
Earlier work this paper cites.
J. Lu, J. Yang, D. Batra, and D. Parikh, “Hierarchical question-image co-attention for visual question answering,” in Proc. Adv. Neural Inf. Process. Syst. , 2016
2011
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in Proc. Adv. Neural Inf. Process. Syst. , 2012
2012
Earlier work this paper cites.
O. Erling, “Virtuoso, a Hybrid RDBMS/Graph Column Store.” IEEE Data Eng. Bull. , vol. 35, no. 1, pp. 3–8, 2012
2012
Earlier work this paper cites.
C. Unger, L. Bühmann, J. Lehmann, A.-C. Ngonga Ngomo, D. Gerber, and P. Cimiano, “Template-based question answering over RDF data,” in WWW , 2012
2012
Earlier work this paper cites.
X. Chen, A. Shrivastava, and A. Gupta, “Neil: Extracting visual knowledge from web data,” in Proc. IEEE Int. Conf. Comp. Vis. , 2013
2013
Earlier work this paper cites.
J. Hoffart, F. M. Suchanek, K. Berberich, and G. Weikum, “YAGO2: A spatially and temporally enhanced knowledge base from Wikipedia,” in Proc. Int. Joint Conf. on Artificial Intell. , 2013
2013
Earlier work this paper cites.
J. Berant, A. Chou, R. Frostig, and P. Liang, “Semantic Parsing on Freebase from Question-Answer Pairs.” in Proc. Conf. Empirical Methods Natural Language Processing , 2013, pp. 1533–1544
2013
Earlier work this paper cites.
Q. Cai and A. Yates, “Large-scale Semantic Parsing via Schema Matching and Lexicon Extension.” in Proc. Conf. the Association for Computational Linguistics , 2013
2013
Earlier work this paper cites.
P. Liang, M. I. Jordan, and D. Klein, “Learning dependency-based compositional semantics,” Computational Linguistics , vol. 39, no. 2, pp. 389–446, 2013
2013
Earlier work this paper cites.
T. Kwiatkowski, E. Choi, Y. Artzi, and L. Zettlemoyer, “Scaling semantic parsers with on-the-fly ontology matching,” in Proc. Conf. Empirical Methods Natural Language Processing , 2013
2013
Earlier work this paper cites.
J. Krishnamurthy and T. Kollar, “Jointly learning to parse and perceive: Connecting natural language to the physical world,” Transactions of the Association for Computational Linguistics , vol. 1, pp. 193–206, 2013
2013
Cited alongside, same era.
2013
Cited alongside, same era.
T. Mikolov, I. Sutskever, K. Chen, G. S. Corrado, and J. Dean, “Distributed representations of words and phrases and their compositionality,” in Proc. Adv. Neural Inf. Process. Syst. , 2013, pp. 3111–3119
2013
Cited alongside, same era.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft COCO: Common objects in context,” in Proc. Eur. Conf. Comp. Vis. , 2014
2014
Cited alongside, same era.
2015
Later among the works it cites.
2015
Later among the works it cites.
F. Mahdisoltani, J. Biega, and F. Suchanek, “YAGO3: A knowledge base from multilingual Wikipedias,” in CIDR , 2015
2015
Later among the works it cites.
S. W.-t. Yih, M.-W. Chang, X. He, and J. Gao, “Semantic parsing via staged query graph generation: Question answering with knowledge base,” in Proceedings of ACL-IJCNLP , 2015
2015
Later among the works it cites.
L. Dong, F. Wei, M. Zhou, and K. Xu, “Question answering over freebase with multi-column convolutional neural networks.” in Proceedings of ACL-IJCNLP , 2015
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2014
Cited alongside, same era.
M. Malinowski and M. Fritz, “A multi-world approach to question answering about real-world scenes based on uncertain input,” in Proc. Adv. Neural Inf. Process. Syst. , 2014, pp. 1682–1690
2014
Cited alongside, same era.
K. Tu, M. Meng, M. W. Lee, T. E. Choe, and S.-C. Zhu, “Joint video and text parsing for understanding events and answering queries,” MultiMedia, IEEE , vol. 21, no. 2, pp. 42–70, 2014
2014
Cited alongside, same era.
D. Vrandečić and M. Krötzsch, “Wikidata: a free collaborative knowledgebase,” Communications of the ACM , vol. 57, no. 10, pp. 78–85, 2014
2014
Cited alongside, same era.
R. W. Group et al. , “Resource description framework,” 2014, http://www.w3.org/standards/techs/rdf
2014
Cited alongside, same era.
N. Tandon, G. De Melo, and G. Weikum, “Acquiring Comparative Commonsense Knowledge from the Web.” in Proc. National Conf. Artificial Intell. , 2014
2014
Cited alongside, same era.
J. Berant and P. Liang, “Semantic parsing via paraphrasing.” in Proc. Conf. the Association for Computational Linguistics , 2014
2014
Cited alongside, same era.
A. Fader, L. Zettlemoyer, and O. Etzioni, “Open question answering over curated and extracted knowledge bases,” in Proc. ACM Int. Conf. Knowledge Discovery & Data Mining , 2014
2014
Cited alongside, same era.
2015
Later among the works it cites.
A. Bordes, N. Usunier, S. Chopra, and J. Weston, “Large-scale simple question answering with memory networks,” in Proc. Int. Conf. Learn. Representations , 2015
2015
Later among the works it cites.
J. Weston, S. Chopra, and A. Bordes, “Memory networks,” in Proc. Int. Conf. Learn. Representations , 2015
2015
Later among the works it cites.
S. Sukhbaatar, J. Weston, R. Fergus et al. , “End-to-end memory networks,” in Proc. Adv. Neural Inf. Process. Syst. , 2015
2015
Later among the works it cites.
2015
Later among the works it cites.
R. Girshick, “Fast R-CNN,” in Proc. IEEE Int. Conf. Comp. Vis. , 2015
2015
Later among the works it cites.
2015
Later among the works it cites.
Y. Zhu, O. Groth, M. Bernstein, and L. Fei-Fei, “Visual7W: Grounded Question Answering in Images,” in Proc. IEEE Conf. Comp. Vis. Patt. Recogn. , 2016
2016
Closest in time.
J. Andreas, M. Rohrbach, T. Darrell, and D. Klein, “Neural Module Networks,” in Proc. IEEE Conf. Comp. Vis. Patt. Recogn. , 2016
2016
Closest in time.
Z. Yang, X. He, J. Gao, L. Deng, and A. Smola, “Stacked Attention Networks for Image Question Answering,” in Proc. IEEE Conf. Comp. Vis. Patt. Recogn. , 2016
2016
Closest in time.
S. Reddy, O. Täckström, M. Collins, T. Kwiatkowski, D. Das, M. Steedman, and M. Lapata, “Transforming dependency structures to logical forms for semantic parsing,” Transactions of the Association for Computational Linguistics , vol. 4, pp. 127–140, 2016
2016
Closest in time.
C. Xiao, M. Dymetman, and C. Gardent, “Sequence-based structured prediction for semantic parsing,” in Proc. Conf. the Association for Computational Linguistics , 2016
2016
Closest in time.
A. Kumar, O. Irsoy, P. Ondruska, M. Iyyer, J. Bradbury, I. Gulrajani, V. Zhong, R. Paulus, and R. Socher, “Ask me anything: Dynamic memory networks for natural language processing,” in Proc. Int. Conf. Mach. Learn. , 2016
2016
Closest in time.
C. Xiong, S. Merity, and R. Socher, “Dynamic memory networks for visual and textual question answering,” in Proc. Int. Conf. Mach. Learn. , 2016
2016
Closest in time.
Q. Wu, P. Wang, C. Shen, A. van den Hengel, and A. Dick, “Ask Me Anything: Free-form Visual Question Answering Based on Knowledge from External Sources,” in Proc. IEEE Conf. Comp. Vis. Patt. Recogn. , 2016
2016
Closest in time.
K. Narasimhan, A. Yala, and R. Barzilay, “Improving information extraction by acquiring external evidence with reinforcement learning,” in Proc. Conf. Empirical Methods Natural Language Processing , 2016
2016
Closest in time.
Q. Wu, C. Shen, A. van den Hengel, L. Liu, and A. Dick, “What Value Do Explicit High-Level Concepts Have in Vision to Language Problems?” in Proc. IEEE Conf. Comp. Vis. Patt. Recogn. , 2016
2016
Closest in time.
P. Zhang, Y. Goyal, D. Summers-Stay, D. Batra, and D. Parikh, “Yin and yang: Balancing and answering binary visual questions,” in Proc. IEEE Conf. Comp. Vis. Patt. Recogn. , 2016
2016
Closest in time.
2016
Closest in time.
R. Krishna, Y. Zhu, O. Groth, J. Johnson, K. Hata, J. Kravitz, S. Chen, Y. Kalantidis, L.-J. Li, D. A. Shamma et al. , “Visual genome: Connecting language and vision using crowdsourced dense image annotations,” Int. J. Comp. Vis. , vol. 123, no. 1, pp. 32–73, 2017
2017
Closest in time.
P. Wang, Q. Wu, C. Shen, A. van den Hengel, and A. Dick, “Explicit Knowledge-based Reasoning for Visual Question Answering,” in Proc. Int. Joint Conf. on Artificial Intell. , 2017
2017
Closest in time.