Fetching the paper…
Reading the bibliography…
Bilinear models provide rich representations compared with linear models.
Long Short-Term Memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Separating style and content with bilinear models
Joshua B Tenenbaum and William T Freeman · 2000
Earlier work this paper cites.
Finding frequent items in data streams
Moses Charikar, Kevin Chen, and Martin Farach-Colton · 2002
Earlier work this paper cites.
Unsupervised learning of image transformations
Roland Memisevic and Geoffrey E Hinton · 2007
Earlier work this paper cites.
Bilinear classifiers for visual recognition
Hamed Pirsiavash, Deva Ramanan, and Charless C. Fowlkes · 2009
Earlier work this paper cites.
Feature hashing for large scale multitask learning
Kilian Weinberger, Anirban Dasgupta, John Langford, Alex Smola, and Josh Attenberg · 2009
Earlier work this paper cites.
Learning to represent spatial transformations with factored higher-order Boltzmann machines
Roland Memisevic and Geoffrey E Hinton · 2010
Earlier work this paper cites.
Lecture 6.5-rmsprop: Divide the gradient by a running average of its recent magnitude
Tijmen Tieleman and Geoffrey Hinton · 2012
Earlier work this paper cites.
Fast and scalable polynomial kernels via explicit feature maps
Ninh Pham and Rasmus Pagh · 2013
Earlier work this paper cites.
Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation
Kyunghyun Cho, Bart Van Merriënboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio · 2014
Earlier work this paper cites.
VQA: Visual Question Answering
Stanislaw Antol, Aishwarya Agrawal, Jiasen Lu, Margaret Mitchell, Dhruv Batra, C. Lawrence Zitnick, and Devi Parikh · 2015
Earlier work this paper cites.
Compressing Neural Networks with the Hashing Trick
Wenlin Chen, James T. Wilson, Stephen Tyree, Kilian Q. Weinberger, and Yixin Chen · 2015
Cited alongside, same era.
A Theoretically Grounded Application of Dropout in Recurrent Neural Networks
Yarin Gal · 2015
Cited alongside, same era.
Spatial Transformer Networks
Max Jaderberg, Karen Simonyan, Andrew Zisserman, and Koray Kavukcuoglu · 2015
Cited alongside, same era.
Skip-Thought Vectors
Ryan Kiros, Yukun Zhu, Ruslan Salakhutdinov, Richard S. Zemel, Antonio Torralba, Raquel Urtasun, and Sanja Fidler · 2015
Cited alongside, same era.
rnn : Recurrent Library for Torch
Nicholas Léonard, Sagar Waghmare, Yang Wang, and Jin-Hwa Kim · 2015
Cited alongside, same era.
Bilinear CNN Models for Fine-grained Visual Recognition
Deep Residual Learning for Image Recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Closest in time.
A Focused Dynamic Attention Model for Visual Question Answering
Ilija Ilievski, Shuicheng Yan, and Jiashi Feng · 2016
Closest in time.
Visual genome: Connecting language and vision using crowdsourced dense image annotations
Ranjay Krishna, Yuke Zhu, Oliver Groth, Justin Johnson, Kenji Hata, Joshua Kravitz, Stephanie Chen, Yannis Kalantidis, Li-Jia Li, David A Shamma, Michael Bernstein, and Li Fei-Fei · 2016
Closest in time.
Hierarchical Question-Image Co-Attention for Visual Question Answering
Jiasen Lu, Jianwei Yang, Dhruv Batra, and Devi Parikh · 2016
Closest in time.
Ask Your Neurons: A Deep Learning Approach to Visual Question Answering
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Tsung-Yu Lin, Aruni RoyChowdhury, and Subhransu Maji · 2015
Cited alongside, same era.
Simple Baseline for Visual Question Answering
Bolei Zhou, Yuandong Tian, Sainbayar Sukhbaatar, Arthur Szlam, and Rob Fergus · 2015
Cited alongside, same era.
Learning to Compose Neural Networks for Question Answering
Jacob Andreas, Marcus Rohrbach, Trevor Darrell, and Dan Klein · 2016
Cited alongside, same era.
Multimodal Compact Bilinear Pooling for Visual Question Answering and Visual Grounding
Akira Fukui, Dong Huk Park, Daylen Yang, Anna Rohrbach, Trevor Darrell, and Marcus Rohrbach · 2016
Cited alongside, same era.
Compact Bilinear Pooling
Yang Gao, Oscar Beijbom, Ning Zhang, and Trevor Darrell · 2016
Cited alongside, same era.
Visual Question Answering: Datasets, Algorithms, and Future Challenges
Kushal Kafle and Christopher Kanan
Cited in the paper.
Answer-Type Prediction for Visual Question Answering
Kushal Kafle and Christopher Kanan
Cited in the paper.
Mateusz Malinowski, Marcus Rohrbach, and Mario Fritz · 2016
Closest in time.
Training Recurrent Answering Units with Joint Loss Minimization for VQA
Hyeonwoo Noh and Bohyung Han · 2016
Closest in time.
Image Question Answering using Convolutional Neural Network with Dynamic Parameter Prediction
Hyeonwoo Noh, Paul Hongsuck Seo, and Bohyung Han · 2016
Closest in time.
Dynamic Memory Networks for Visual and Textual Question Answering
Caiming Xiong, Stephen Merity, and Richard Socher · 2016
Closest in time.
Ask, Attend and Answer: Exploring Question-Guided Spatial Attention for Visual Question Answering
Huijuan Xu and Kate Saenko · 2016
Closest in time.
Stacked Attention Networks for Image Question Answering
Zichao Yang, Xiaodong He, Jianfeng Gao, Li Deng, and Alex Smola · 2016
Closest in time.