Fetching the paper…
Reading the bibliography…
Knowledge-based visual question answering (QA) aims to answer a question which requires visually-grounded external knowledge beyond image content itself.
Conceptnet—a practical commonsense reasoning tool-kit
Hugo Liu and Push Singh. 2004 · 2004
Earlier work this paper cites.
Dbpedia: A nucleus for a web of open data
Sören Auer, Christian Bizer, Georgi Kobilarov, Jens Lehmann, Richard Cyganiak, and Zachary Ives. 2007 · 2007
Earlier work this paper cites.
Freebase: a collaboratively created graph database for structuring human knowledge
Kurt Bollacker, Colin Evans, Praveen Paritosh, Tim Sturge, and Jamie Taylor. 2008 · 2008
Earlier work this paper cites.
Question answering with subgraph embeddings
Antoine Bordes, Sumit Chopra, and Jason Weston. 2014a · 2014
Earlier work this paper cites.
Learning phrase representations using RNN encoder–decoder for statistical machine translation
Kyunghyun Cho, Bart van Merriënboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
GloVe: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Earlier work this paper cites.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V. Le. 2014 · 2014
Earlier work this paper cites.
Webchild: harvesting and organizing commonsense knowledge from the web
Niket Tandon, Gerard de Melo, Fabian M. Suchanek, and Gerhard Weikum. 2014 · 2014
Earlier work this paper cites.
Wikidata: a free collaborative knowledgebase
Denny Vrandečić and Markus Krötzsch. 2014 · 2014
Earlier work this paper cites.
VQA: visual question answering
Stanislaw Antol, Aishwarya Agrawal, Jiasen Lu, Margaret Mitchell, Dhruv Batra, C. Lawrence Zitnick, and Devi Parikh. 2015 · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba. 2015 · 2015
Earlier work this paper cites.
End-to-end memory networks
Sainbayar Sukhbaatar, Arthur Szlam, Jason Weston, and Rob Fergus. 2015 · 2015
Earlier work this paper cites.
Jason Weston, Sumit Chopra, and Antoine Bordes. 2015 · 2015
Earlier work this paper cites.
Gated graph sequence neural networks
Yujia Li, Daniel Tarlow, Marc Brockschmidt, and Richard S. Zemel. 2016 · 2016
Earlier work this paper cites.
Key-value memory networks for directly reading documents
Alexander Miller, Adam Fisch, Jesse Dodge, Amir-Hossein Karimi, Antoine Bordes, and Jason Weston. 2016 · 2016
Earlier work this paper cites.
Visual7w: Grounded question answering in images
Yuke Zhu, Oliver Groth, Michael S. Bernstein, and Li Fei-Fei. 2016 · 2016
Earlier work this paper cites.
Semi-supervised classification with graph convolutional networks
Thomas N. Kipf and Max Welling. 2017 · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
Explicit knowledge-based reasoning for visual question answering
Peng Wang, Qi Wu, Chunhua Shen, Anthony R. Dick, and Anton van den Hengel. 2017 · 2017
Cited alongside, same era.
Go for a walk and arrive at the answer: Reasoning over paths in knowledge bases using reinforcement learning
Rajarshi Das, Shehzaad Dhuliawala, Manzil Zaheer, Luke Vilnis, Ishan Durugkar, Akshay Krishnamurthy, Alex Smola, and Andrew McCallum. 2018 · 2018
Cited alongside, same era.
Bilinear attention networks
Jin-Hwa Kim, Jaehyun Jun, and Byoung-Tak Zhang. 2018 · 2018
Cited alongside, same era.
Deeper insights into graph convolutional networks for semi-supervised learning
Multimodal transformer for unaligned multimodal language sequences
Yao-Hung Hubert Tsai, Shaojie Bai, Paul Pu Liang, J. Zico Kolter, Louis-Philippe Morency, and Ruslan Salakhutdinov. 2019 · 2019
Later among the works it cites.
Hypergcn: A new method for training graph convolutional networks on hypergraphs
Naganand Yadati, Madhav Nimishakavi, Prateek Yadav, Vikram Nitin, Anand Louis, and Partha P. Talukdar. 2019 · 2019
Later among the works it cites.
From recognition to cognition: Visual commonsense reasoning
Rowan Zellers, Yonatan Bisk, Ali Farhadi, and Yejin Choi. 2019 · 2019
Later among the works it cites.
Retinaface: Single-shot multi-level face localisation in the wild
Jiankang Deng, Jia Guo, Evangelos Ververas, Irene Kotsia, and Stefanos Zafeiriou. 2020 · 2020
Later among the works it cites.
Open domain question answering based on text enhanced knowledge graph with hyperedge infusion
Jiale Han, Bo Cheng, and Xu Wang. 2020a · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Qimai Li, Zhichao Han, and Xiao-Ming Wu. 2018 · 2018
Cited alongside, same era.
Out of the box: Reasoning with graph convolution nets for factual visual question answering
Medhini Narasimhan, Svetlana Lazebnik, and Alexander G. Schwing. 2018 · 2018
Cited alongside, same era.
Straight to the facts: Learning knowledge base retrieval for factual visual question answering
Medhini Narasimhan and Alexander G Schwing. 2018 · 2018
Cited alongside, same era.
Fvqa: Fact-based visual question answering
Peng Wang, Qi Wu, Chunhua Shen, Anthony Dick, and Anton van den Hengel. 2018 · 2018
Cited alongside, same era.
Arcface: Additive angular margin loss for deep face recognition
Jiankang Deng, Jia Guo, Niannan Xue, and Stefanos Zafeiriou. 2019 · 2019
Cited alongside, same era.
GQA: A new dataset for real-world visual reasoning and compositional question answering
Drew A. Hudson and Christopher D. Manning. 2019 · 2019
Cited alongside, same era.
KagNet: Knowledge-aware graph networks for commonsense reasoning
Bill Yuchen Lin, Xinyue Chen, Jamin Chen, and Xiang Ren. 2019 · 2019
Cited alongside, same era.
Two-phase hypergraph based reasoning with dynamic relations for multi-hop KBQA
Jiale Han, Bo Cheng, and Xu Wang. 2020b · 2020
Later among the works it cites.
Language generation with multi-hop reasoning on commonsense knowledge graph
Haozhe Ji, Pei Ke, Shaohan Huang, Furu Wei, Xiaoyan Zhu, and Minlie Huang. 2020 · 2020
Later among the works it cites.
Hypergraph attention networks for multimodal learning
Eun-Sol Kim, Woo-Young Kang, Kyoung-Woon On, Yu-Jung Heo, and Byoung-Tak Zhang. 2020 · 2020
Later among the works it cites.
Stepwise reasoning for multi-relation question answering over knowledge graph with weak supervision
Yunqi Qiu, Yuanzhuo Wang, Xiaolong Jin, and Kun Zhang. 2020 · 2020
Later among the works it cites.
Visuo-linguistic question answering (VLQA) challenge
Shailaja Keyur Sampat, Yezhou Yang, and Chitta Baral. 2020 · 2020
Later among the works it cites.
Improving multi-hop question answering over knowledge graphs using knowledge base embeddings
Apoorv Saxena, Aditay Tripathi, and Partha Talukdar. 2020 · 2020
Later among the works it cites.
Knowledge graph alignment network with gated multi-hop neighborhood aggregation
Zequn Sun, Chengming Wang, Wei Hu, Muhao Chen, Jian Dai, Wei Zhang, and Yuzhong Qu. 2020 · 2020
Later among the works it cites.
Modelling long-distance node relations for KBQA with global dynamic graph
Xu Wang, Shuai Zhao, Jiale Han, Bo Cheng, Hao Yang, Jianchang Ao, and Zhenzi Li. 2020 · 2020
Later among the works it cites.
Mucko: Multi-layer cross-modal knowledge reasoning for fact-based visual question answering
Zihao Zhu, Jing Yu, Yujing Wang, Yajing Sun, Yue Hu, and Qi Wu. 2020 · 2020
Later among the works it cites.
Knowledge base question answering through recursive hypergraphs
Naganand Yadati, Dayanidhi R S, Vaishnavi S, Indira K M, and Srinidhi G. 2021 · 2021
Later among the works it cites.
An interpretable reasoning network for multi-relation question answering
Mantong Zhou, Minlie Huang, and Xiaoyan Zhu. 2018 · 2022
Closest in time.