Fetching the paper…
Reading the bibliography…
Fact-based Visual Question Answering (FVQA), a challenging variant of VQA, requires a QA-system to include facts from a diverse knowledge graph (KG) in its reasoning process to produce an answer.
Conceptnet: A practical commonsense reasoning toolkit
Hugo Liu and Push Singh. 2004 · 2004
Earlier work this paper cites.
Dbpedia: A nucleus for a web of open data
Sören Auer, Christian Bizer, Georgi Kobilarov, Jens Lehmann, Zachary Ives, and et al. 2007 · 2007
Earlier work this paper cites.
Yago: A core of semantic knowledge
Fabian M. Suchanek, Gjergji Kasneci, and Gerhard Weikum. 2007 · 2007
Earlier work this paper cites.
Freebase: a collaboratively created graph database for structuring human knowledge
Kurt Bollacker, Colin Evans, Praveen Paritosh, Tim Sturge, and Jamie Taylor. 2008 · 2008
Earlier work this paper cites.
Noise-contrastive estimation: A new estimation principle for unnormalized statistical models
Michael Gutmann and Aapo Hyvärinen. 2010 · 2010
Earlier work this paper cites.
Review relational knowledge: the foundation of higher cognition
Graeme S. Halford, William H. Wilson, and Steven Phillips. 2010 · 2010
Earlier work this paper cites.
A three-way model for collective learningon multi-relational data
M. Nickel, V. Tresp, and H.P. Kriegel. 2011 · 2011
Earlier work this paper cites.
Translating embeddings for modeling multi-relational data
Antoine Bordes, Nicolas Usunier, Jason Weston, and Oksana Yakhnenko. 2013 · 2013
Earlier work this paper cites.
Reasoning with neural tensor networks for knowledgebase completion
R. Socher, D. Chen, C.D. Manning, and A.Y. Ng. 2013 · 2013
Earlier work this paper cites.
Knowledge vault: a web-scale approach to probabilistic knowledge fusion
Xin Dong, Evgeniy Gabrilovich, Geremy Heitz, Wilko Horn, Ni Lao, Kevin Patrick Murphy, Thomas Strohmann, Shaohua Sun, and Wei Zhang. 2014 · 2014
Earlier work this paper cites.
Microsoft COCO: Common objects in context
T-Y Lin, M Maire, S Belongie, J Hays, P Perona, D Ramanan, P Dollár, and CL Zitnick. 2014 · 2014
Earlier work this paper cites.
Never-ending language learning
T. Mitchell and E. Fredkin. 2014 · 2014
Earlier work this paper cites.
GloVe: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Earlier work this paper cites.
Webchild: harvesting and organizing commonsense knowledge from the web
Niket Tandon, Gerard de Melo, Fabian M. Suchanek, and Gerhard Weikum. 2014 · 2014
Cited alongside, same era.
Vqa: Visual question answering
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. L. Zitnick, and D. Parikh. 2015 · 2015
Cited alongside, same era.
Faster r-cnn: Towards real-time object detection with region proposal networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun. 2015 · 2015
Cited alongside, same era.
ImageNet Large Scale Visual Recognition Challenge
Olga Russakovsky, Jia Deng, Hao Su, Jonathan Krause, Sanjeev Satheesh, Sean Ma, Zhiheng Huang, Andrej Karpathy, Aditya Khosla, Michael Bernstein, Alexander C. Berg, and Li Fei-Fei. 2015 · 2015
Cited alongside, same era.
Very deep convolutional networks for large-scale image classification
K Simonyan and A Zisserman. 2015 · 2015
Cited alongside, same era.
Deep residual learning for image recognition
No classification without representation: Assessing geodiversity issues in open data sets for the developing world
Shreya Shankar, Yoni Halpern, Eric Breck, James Atwood, Jimbo Wilson, and D. Sculley. 2017 · 2017
Later among the works it cites.
Places: A 10 million image database for scene recognition
Bolei Zhou, Agata Lapedriza, Aditya Khosla, Aude Oliva, and Antonio Torralba. 2017 · 2017
Later among the works it cites.
Convolutional 2D knowledge graph embeddings
T. Dettmers, P. Minervini, P. Stenetorp, and S. Riedel. 2018 · 2018
Later among the works it cites.
Debiasing knowledge graphs: Why female presidents are not like female popes
Krzysztof Janowicz, Bo Yan, Blake Regalia, Rui Zhu, and Gengchen Mai. 2018 · 2018
Later among the works it cites.
Out of the box: Reasoning with graph convolution nets for factual visual question answering
Medhini Narasimhan, Svetlana Lazebnik, and Alexander Schwing. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kaiming He, X. Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Cited alongside, same era.
Hierarchical question-image co-attention for visual question answering
Jiasen Lu, Jianwei Yang, Dhruv Batra, and Devi Parikh. 2016 · 2016
Cited alongside, same era.
FVQA: fact-based visual question answering
Peng Wang, Qi Wu, Chunhua Shen, Anton van den Hengel, and Anthony R. Dick. 2016 · 2016
Cited alongside, same era.
What value do explicit high level concepts have in vision to language problems?
Q. Wu, C. Shen, L. Liu, A. Dick, and A. Van Den Hengel. 2016 · 2016
Cited alongside, same era.
Wide residual networks
Sergey Zagoruyko and Nikos Komodakis. 2016 · 2016
Cited alongside, same era.
Making the v in vqa matter: Elevating the role of image understanding in visual question answering
Yash Goyal, Tejas Khot, Douglas Summers-Stay, Dhruv Batra, and Devi Parikh. 2017 · 2017
Cited alongside, same era.
Semi-supervised classification with graph convolutional networks
Thomas N. Kipf and Max Welling. 2017 · 2017
Cited alongside, same era.
Straight to the facts: Learning knowledge base retrieval for factual visual question answering
Medhini Narasimhan and Alexander G. Schwing. 2018 · 2018
Later among the works it cites.
Keras-retinanet for open images challenge 2018
ZFTurbo. 2018 · 2018
Later among the works it cites.
Measuring social bias in knowledge graph embeddings
Joseph Fisher, Dave Palfrey, Christos Christodoulopoulos, and Arpit Mittal. 2019 · 2019
Later among the works it cites.
Visual question answering as reading comprehension
Hui Li, Peng Wang, Chunhua Shen, and A. V. D. Hengel. 2019 · 2019
Later among the works it cites.
Rotate: Knowledge graph embedding by relational rotation in complex space
Zhiqing Sun, Zhi-Hong Deng, Jian-Yun Nie, and Jian Tang. 2019 · 2019
Later among the works it cites.
Albert: A lite bert for self-supervised learning of language representations
Zhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel, Piyush Sharma, and Radu Soricut. 2020 · 2020
Closest in time.
Improving multi-hop question answering over knowledge graphs using knowledge base embeddings
Apoorv Saxena, Aditay Tripathi, and Partha Talukdar. 2020 · 2020
Closest in time.