Fetching the paper…
Reading the bibliography…
We propose a novel memory network model named Read-Write Memory Network (RWMN) to perform question and answering tasks for large-scale, multimodal movie story understanding.
Long Short-term Memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Imagenet: A Large-scale Hierarchical Image Database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Understanding the Difficulty of Training Deep Feedforward Neural Networks
X. Glorot and Y. Bengio · 2010
Earlier work this paper cites.
Recurrent Neural Network Based Language Model
T. Mikolov, M. Karafiát, L. Burget, J. Cernockỳ, and S. Khudanpur · 2010
Earlier work this paper cites.
Rectified Linear Units Improve Restricted Boltzmann Machines
V. Nair and G. E. Hinton · 2010
Earlier work this paper cites.
Adaptive Subgradient Methods for Online Learning and Stochastic Optimization
J. Duchi, E. Hazan, and Y. Singer · 2011
Earlier work this paper cites.
Distributed Representations of Words and Phrases and Their Compositionality
T. Mikolov, I. Sutskever, K. Chen, G. S. Corrado, and J. Dean · 2013
Earlier work this paper cites.
Learning Phrase Representations Using RNN Encoder-Decoder for Statistical Machine Translation
K. Cho, B. Van Merriënboer, C. Gulcehre, D. Bahdanau, F. Bougares, H. Schwenk, and Y. Bengio · 2014
Earlier work this paper cites.
A. Graves, G. Wayne, and I. Danihelka · 2014
Earlier work this paper cites.
Large-scale Video Classification with Convolutional Neural Networks
A. Karpathy, G. Toderici, S. Shetty, T. Leung, R. Sukthankar, and L. Fei-Fei · 2014
Earlier work this paper cites.
End-to-End Memory Networks
S. Sukhbaatar, J. Weston, R. Fergus, et al · 2015
Earlier work this paper cites.
Towards AI-complete Question Answering: A Set of Prerequisite Toy Tasks
J. Weston, A. Bordes, S. Chopra, A. M. Rush, B. van Merriënboer, A. Joulin, and T. Mikolov · 2015
Cited alongside, same era.
Memory Networks
J. Weston, S. Chopra, and A. Bordes · 2015
Cited alongside, same era.
Youtube-8M: A Large-scale Video Classification Benchmark
S. Abu-El-Haija, N. Kothari, J. Lee, P. Natsev, G. Toderici, B. Varadarajan, and S. Vijayanarasimhan · 2016
Cited alongside, same era.
S. Chandar, S. Ahn, H. Larochelle, P. Vincent, G. Tesauro, and Y. Bengio · 2016
Cited alongside, same era.
Multimodal Compact Bilinear Pooling for Visual Question Answering and Visual Grounding
A. Fukui, D. H. Park, D. Yang, A. Rohrbach, T. Darrell, and M. Rohrbach · 2016
Cited alongside, same era.
Squad: 100,000+ Questions for Machine Comprehension of Text
P. Rajpurkar, J. Zhang, K. Lopyrev, and P. Liang · 2016
Later among the works it cites.
MovieQA: Understanding Stories in Movies through Question-answering
M. Tapaswi, Y. Zhu, R. Stiefelhagen, A. Torralba, R. Urtasun, and S. Fidler · 2016
Later among the works it cites.
MSR-VTT: A Large Video Description Dataset for Bridging Video and Language
J. Xu, T. Mei, T. Yao, and Y. Rui · 2016
Later among the works it cites.
Dynamic Neural Turing Machine with Soft and Hard Addressing Schemes
C. Gulcehre, S. Chandar, K. Cho, and Y. Bengio · 2017
Closest in time.
TGIF-QA: Toward Spatio-Temporal Reasoning in Visual Question Answering
Y. Jang, Y. Song, Y. Yu, Y. Kim, and G. Kim · 2017
Closest in time.
Deepstory: video story qa by deep embedded memory networks
K.-M. Kim, M.-O. Heo, S.-H. Choi, and B.-T. Zhang · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Hybrid Computing Using a Neural Network with Dynamic External Memory
A. Graves, G. Wayne, M. Reynolds, T. Harley, I. Danihelka, A. Grabska-Barwińska, S. G. Colmenarejo, E. Grefenstette, T. Ramalho, J. Agapiou, et al · 2016
Cited alongside, same era.
Deep Residual Learning for Image Recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Cited alongside, same era.
Ask Me Anything: Dynamic Memory Networks for Natural Language Processing
A. Kumar, O. Irsoy, J. Su, J. Bradbury, R. English, B. Pierce, P. Ondruska, I. Gulrajani, and R. Socher · 2016
Cited alongside, same era.
Key-value Memory Networks for Directly Reading Documents
A. Miller, A. Fisch, J. Dodge, A.-H. Karimi, A. Bordes, and J. Weston · 2016
Cited alongside, same era.
Scaling Memory-Augmented Neural Networks with Sparse Reads and Writes
J. Rae, J. J. Hunt, I. Danihelka, T. Harley, A. W. Senior, G. Wayne, A. Graves, and T. Lillicrap · 2016
Cited alongside, same era.
Attend to You: Personalized Image Captioning with Context Sequence Memory Networks
C. C. Park, B. Kim, and G. Kim
Cited in the paper.
Closest in time.
Movie Description
A. Rohrbach, A. Torabi, M. Rohrbach, N. Tandon, C. Pal, H. Larochelle, A. Courville, and B. Schiele · 2017
Closest in time.
A Compare-Aggregate Model for Matching Text Sequences
S. Wang and J. Jiang · 2017
Closest in time.
End-to-end Concept Word Detection for Video Captioning, Retrieval, and Question Answering
Y. Yu, H. Ko, Jongwook, and G. Kim · 2017
Closest in time.
Dynamic Key-Value Memory Network for Knowledge Tracing
J. Zhang, X. Shi, I. King, and D.-Y. Yeung · 2017
Closest in time.