Fetching the paper…
Reading the bibliography…
Neural network architectures with memory and attention mechanisms exhibit certain reasoning capabilities required for question answering.
Textrunner: Open information extraction on the web
Yates, A., Banko, M., Broadhead, M., Cafarella, M. J., Etzioni, O., and Soderland, S · 2007
Earlier work this paper cites.
Baby talk: Understanding and generating image descriptions
Kulkarni, G., Premraj, V., Dhar, S., Li, S., Choi, Y., Berg, A. C., and Berg, T. L · 2011
Earlier work this paper cites.
Joint Learning of Words and Meaning Representations for Open-Text Semantic Parsing
Bordes, A., Glorot, X., Weston, J., and Bengio, Y · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Krizhevsky, A., Sutskever, I., and Hinton, G. E · 2012
Earlier work this paper cites.
Learning a recurrent visual representation for image caption generation
Chen, X. and Zitnick, C. L · 2014
Earlier work this paper cites.
Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation
Cho, K., van Merrienboer, B., Gulcehre, C., Bahdanau, D., Bougares, F., Schwenk, H., and Bengio, Y · 2014
Earlier work this paper cites.
Empirical evaluation of gated recurrent neural networks on sequence modeling
Chung, J., Gulcehre, C., Cho, K., and Bengio, Y · 2014
Earlier work this paper cites.
A Visual Turing Test for Computer Vision Systems
Geman, D., Geman, S., Hallonquist, N., and Younes, L · 2014
Earlier work this paper cites.
Graves, A., Wayne, G., and Danihelka, I · 2014
Earlier work this paper cites.
A Neural Network for Factoid Question Answering over Paragraphs
Iyyer, M., Boyd-Graber, J., Claudino, L., Socher, R., and Daumé III, H · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, Diederik and Ba, Jimmy · 2014
Earlier work this paper cites.
Microsoft COCO: Common Objects in Context
Lin, T. Y., Maire, M., Belongie, S., Hays, J., Perona, P., Ramanan, D., Dollár, P., and Zitnick, C. L · 2014
Earlier work this paper cites.
A Multi-World Approach to Question Answering about Real-World Scenes based on Uncertain Input
Malinowski, M. and Fritz, M · 2014
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
Simonyan, K. and Zisserman, A · 2014
Cited alongside, same era.
Grounded compositional semantics for finding and describing images with sentences
Socher, R., Karpathy, A., Le, Q. V., Manning, C. D., and Ng, A. Y · 2014
Cited alongside, same era.
Deep Networks with Internal Selective Attention through Feedback Connections
Stollenga, M. F., J. Masci, F. Gomez, and Schmidhuber, J · 2014
Cited alongside, same era.
VQA: Visual Question Answering
Antol, S., Agrawal, A., Lu, J., Mitchell, M., Batra, D., Zitnick, C. L., and Parikh, D · 2015
Cited alongside, same era.
Effective approaches to attention-based neural machine translation
Luong, M. T., Pham, H., and Manning, C. D · 2015
Later among the works it cites.
Learning to Answer Questions From Image Using Convolutional Neural Network
Ma, L., Lu, Z., and Li, H · 2015
Later among the works it cites.
Ask your neurons: A neural-based approach to answering questions about images
Malinowski, M., Rohrbach, M., and Fritz, M · 2015
Later among the works it cites.
Image question answering using convolutional neural network with dynamic parameter prediction
Noh, H., Seo, P. H., and Han, B · 2015
Later among the works it cites.
Towards neural network-based reasoning
Peng, B., Lu, Z., Li, H., and Wong, K · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Neural machine translation by jointly learning to align and translate
Bahdanau, D., Cho, K., and Bengio, Y · 2015
Cited alongside, same era.
From captions to visual concepts and back
Fang, H., Gupta, S., Iandola, F., Srivastava, R., Deng, L., Dollar, P., Gao, J., He, X., Mitchell, M., and Platt, J · 2015
Cited alongside, same era.
Inferring algorithmic patterns with stack-augmented recurrent nets
Joulin, A. and Mikolov, T · 2015
Cited alongside, same era.
Kaiser, L. and Sutskever, I · 2015
Cited alongside, same era.
Deep Visual-Semantic Alignments for Generating Image Descriptions
Karpathy, A. and Fei-Fei, L · 2015
Cited alongside, same era.
Ask Me Anything: Dynamic Memory Networks for Natural Language Processing
Kumar, A., Irsoy, O., Ondruska, P., Iyyer, M., Bradbury, J., Gulrajani, I., and Socher, R · 2015
Cited alongside, same era.
A Hierarchical Neural Autoencoder for Paragraphs and Documents
Li, J., Luong, M. T., and Jurafsky, D · 2015
Cited alongside, same era.
End-to-end memory networks
Sukhbaatar, S., Szlam, A., Weston, J., and Fergus, R · 2015
Later among the works it cites.
Ask Me Anything: Free-form Visual Question Answering Based on Knowledge from External Sources
Wu, Q., Wang, P., Shen, C., Hengel, A. van den, and Dick, A · 2015
Later among the works it cites.
Show, attend and tell: Neural image caption generation with visual attention
Xu, K., Ba, J., Kiros, R., Cho, K., Courville, A. C., Salakhutdinov, R., Zemel, R. S., and Bengio, Y · 2015
Later among the works it cites.
Stacked attention networks for image question answering
Yang, Z., He, X., Gao, J., Deng, L., and Smola, A · 2015
Later among the works it cites.
Simple baseline for visual question answering
Zhou, B., Tian, Y., Sukhbaatar, S., Szlam, A., and Fergus, R · 2015
Later among the works it cites.
Learning to Compose Neural Networks for Question Answering
Andreas, J., Rohrbach, M., Darrell, T., and Klein, D · 2016
Closest in time.