Fetching the paper…
Reading the bibliography…
In this paper, we exploit a memory-augmented neural network to predict accurate answers to visual questions, even when those answers occur rarely in the training set.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Bidirectional recurrent neural networks
M. Schuster and K. K. Paliwal · 1997
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
T. Mikolov, I. Sutskever, K. Chen, G. S. Corrado, and J. Dean · 2013
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
D. Bahdanau, K. Cho, and Y. Bengio · 2014
Earlier work this paper cites.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
K. Cho, B. van Merrienboer, C. Gulcehre, F. Bougares, H. Schwenk, and Y. Bengio · 2014
Earlier work this paper cites.
A. Graves, G. Wayne, and I. Danihelka · 2014
Earlier work this paper cites.
Microsoft COCO: Common objects in context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2014
Earlier work this paper cites.
Deep networks with internal selective attention through feedback connections
M. F. Stollenga, J. Masci, F. J. Gomez, and J. Schmidhuber · 2014
Earlier work this paper cites.
J. Weston, S. Chopra, and A. Bordes · 2014
Earlier work this paper cites.
VQA: Visual Question Answering
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. L. Zitnick, and D. Parikh · 2015
Earlier work this paper cites.
Are You Talking to a Machine? Dataset and Methods for Multilingual Image Question Answering
H. Gao, J. Mao, J. Zhou, Z. Huang, L. Wang, and W. Xu · 2015
Earlier work this paper cites.
Inferring algorithmic patterns with stack-augmented recurrent nets
A. Joulin and T. Mikolov · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2015
Earlier work this paper cites.
Effective approaches to attention-based neural machine translation
T. Luong, H. Pham, and C. D. Manning · 2015
Earlier work this paper cites.
Ask Your Neurons: A Neural-based Approach to Answering Questions about Images
M. Malinowski, M. Rohrbach, and M. Fritz · 2015
Earlier work this paper cites.
Recurrent neural networks with external memory for language understanding
B. Peng and K. Yao · 2015
Cited alongside, same era.
Image Question Answering: A Visual Semantic Embedding Model and a New Dataset
M. Ren, R. Kiros, and R. Zemel · 2015
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2015
Cited alongside, same era.
Weakly supervised memory networks
S. Sukhbaatar, A. Szlam, J. Weston, and R. Fergus · 2015
Cited alongside, same era.
Going deeper with convolutions
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich · 2015
Cited alongside, same era.
Hierarchical question-image co-attention for visual question answering
J. Lu, J. Yang, D. Batra, and D. Parikh · 2016
Later among the works it cites.
Adding gradient noise improves learning for very deep networks
A. Neelakantan, L. Vilnis, Q. V. Le, I. Sutskever, L. Kaiser, K. Kurach, and J. Martens · 2016
Later among the works it cites.
Training recurrent answering units with joint loss minimization for vqa
H. Noh and B. Han · 2016
Later among the works it cites.
Image Question Answering using Convolutional Neural Network with Dynamic Parameter Prediction
H. Noh, P. H. Seo, and B. Han · 2016
Later among the works it cites.
Meta-learning with memory-augmented neural networks
A. Santoro, S. Bartunov, M. Botvinick, D. Wierstra, and T. P. Lillicrap · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
K. Xu, J. Ba, R. Kiros, A. Courville, R. Salakhutdinov, R. Zemel, and Y. Bengio · 2015
Cited alongside, same era.
Simple baseline for visual question answering
B. Zhou, Y. Tian, S. Sukhbaatar, A. Szlam, and R. Fergus · 2015
Cited alongside, same era.
Neural Module Networks
J. Andreas, M. Rohrbach, T. Darrell, and D. Klein · 2016
Cited alongside, same era.
S. Chandar, S. Ahn, H. Larochelle, P. Vincent, G. Tesauro, and Y. Bengio · 2016
Cited alongside, same era.
Multimodal compact bilinear pooling for visual question answering and visual grounding
A. Fukui, D. H. Park, D. Yang, A. Rohrbach, T. Darrell, and M. Rohrbach · 2016
Cited alongside, same era.
Pointing the unknown words
Ç. Gülçehre, S. Ahn, R. Nallapati, B. Zhou, and Y. Bengio · 2016
Cited alongside, same era.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Cited alongside, same era.
K. J. Shih, S. Singh, and D. Hoiem · 2016
Later among the works it cites.
Ask Me Anything: Free-form Visual Question Answering Based on Knowledge from External Sources
Q. Wu, P. Wang, C. Shen, A. Dick, and A. v. d. Hengel · 2016
Later among the works it cites.
Dynamic memory networks for visual and textual question answering
C. Xiong, S. Merity, and R. Socher · 2016
Later among the works it cites.
Ask, Attend and Answer: Exploring Question-Guided Spatial Attention for Visual Question Answering
H. Xu and K. Saenko · 2016
Later among the works it cites.
Stacked Attention Networks for Image Question Answering
Z. Yang, X. He, J. Gao, L. Deng, and A. Smola · 2016
Later among the works it cites.
Visual7w: Grounded question answering in images
Y. Zhu, O. Groth, M. S. Bernstein, and L. Fei-Fei · 2016
Later among the works it cites.
Making the V in VQA matter: Elevating the role of image understanding in Visual Question Answering
Y. Goyal, T. Khot, D. Summers-Stay, D. Batra, and D. Parikh · 2017
Closest in time.
Learning to remember rare events
Ł. Kaiser, O. Nachum, A. Roy, and S. Bengio · 2017
Closest in time.
The vqa-machine: Learning how to use existing vision algorithms to answer new questions
P. Wang, Q. Wu, C. Shen, and A. van den Hengel · 2017
Closest in time.