2021

Weakly-Supervised Visual-Retriever-Reader for Knowledge-based Question Answering

Luo, Man, Zeng, Yankai, Banerjee, Pratyay et al.

Understand

Knowledge-based visual question answering (VQA) requires answering questions with external knowledge in addition to the content of images.

  • One dataset that is mostly used in evaluating knowledge-based VQA is OK-VQA, but it lacks a gold standard knowledge corpus for retrieval.
  • Existing work leverage different knowledge bases (e.g., ConceptNet and Wikipedia) to obtain external knowledge.
  • Because of varying knowledge bases, it is hard to fairly compare models' performance.

Reading the bibliography…