Fetching the paper…
Reading the bibliography…
Visual question answering is concerned with answering free-form questions about an image.
Gqa: A new dataset for real-world visual reasoning and compositional question answering
Hudson, D. A. and Manning, C. D · 1902
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Williams, R. J · 1992
Earlier work this paper cites.
Long short-term memory
Hochreiter, S. and Schmidhuber, J · 1997
Earlier work this paper cites.
Glove: Global vectors for word representation
Pennington, J., Socher, R., and Manning, C. D · 2014
Earlier work this paper cites.
Vqa: Visual question answering
Antol, S., Agrawal, A., Lu, J., Mitchell, M., Batra, D., Lawrence Zitnick, C., and Parikh, D · 2015
Earlier work this paper cites.
Image retrieval using scene graphs
Johnson, J., Krishna, R., Stark, M., Li, L.-J., Shamma, D., Bernstein, M., and Fei-Fei, L · 2015
Earlier work this paper cites.
Semi-supervised classification with graph convolutional networks
Kipf, T. N. and Welling, M · 2016
Earlier work this paper cites.
Stacked attention networks for image question answering
Yang, Z., He, X., Gao, J., Deng, L., and Smola, A · 2016
Cited alongside, same era.
Clevr: A diagnostic dataset for compositional language and elementary visual reasoning
Johnson, J., Hariharan, B., van der Maaten, L., Fei-Fei, L., Lawrence Zitnick, C., and Girshick, R · 2017
Cited alongside, same era.
Graph-structured representations for visual question answering
Teney, D., Liu, L., and van Den Hengel, A · 2017
Cited alongside, same era.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., and Polosukhin, I · 2017
Cited alongside, same era.
Veličković, P., Cucurull, G., Casanova, A., Romero, A., Lio, P., and Bengio, Y · 2017
Cited alongside, same era.
Bottom-up and top-down attention for image captioning and visual question answering
Scene graph reasoning with prior visual relationship for visual question answering
Yang, Z., Qin, Z., Yu, J., and Hu, Y · 2018
Later among the works it cites.
Neural motifs: Scene graph parsing with global context
Zellers, R., Yatskar, M., Thomson, S., and Choi, Y · 2018
Later among the works it cites.
Explainable and explicit visual reasoning over scene graphs
Shi, J., Zhang, H., and Li, J · 2019
Later among the works it cites.
Efficientdet: Scalable and efficient object detection
Tan, M., Pang, R., and Le, Q. V · 2019
Later among the works it cites.
Graphical contrastive losses for scene graph parsing
Zhang, J., Shih, K. J., Elgammal, A., Tao, A., and Catanzaro, B · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Anderson, P., He, X., Buehler, C., Teney, D., Johnson, M., Gould, S., and Zhang, L · 2018
Cited alongside, same era.
Go for a walk and arrive at the answer: Reasoning over paths in knowledge bases using reinforcement learning
Das, R., Dhuliawala, S., Zaheer, M., Vilnis, L., Durugkar, I., Krishnamurthy, A., Smola, A., and McCallum, A · 2018
Cited alongside, same era.
Learning by abstraction: The neural state machine
Hudson, D. and Manning, C. D
Cited in the paper.
Reasoning on knowledge graphs with debate dynamics
Hildebrandt, M., Serna, J. A. Q., Ma, Y., Ringsquandl, M., Joblin, M., and Tresp, V · 2020
Closest in time.
Koner, R., Sinhamahapatra, P., and Tresp, V · 2020
Closest in time.