Fetching the paper…
Reading the bibliography…
Training recurrent neural networks to model long term dependencies is difficult.
Introduction to wordnet: An on-line lexical database
Miller, George A, Beckwith, Richard, Fellbaum, Christiane, Gross, Derek, and Miller, Katherine J · 1990
Earlier work this paper cites.
Learning long-term dependencies with gradient descent is difficult
Bengio, Yoshua, Simard, Patrice, and Frasconi, Paolo · 1994
Earlier work this paper cites.
Long short-term memory
Hochreiter, Sepp and Schmidhuber, Jürgen · 1997
Earlier work this paper cites.
Freebase: a collaboratively created graph database for structuring human knowledge
Bollacker, Kurt, Evans, Colin, Paritosh, Praveen, Sturge, Tim, and Taylor, Jamie · 2008
Earlier work this paper cites.
The graph neural network model
Scarselli, Franco, Gori, Marco, Tsoi, Ah Chung, Hagenbuchner, Markus, and Monfardini, Gabriele · 2009
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Bahdanau, Dzmitry, Cho, Kyunghyun, and Bengio, Yoshua · 2014
Earlier work this paper cites.
Graves, Alex, Wayne, Greg, and Danihelka, Ivo · 2014
Earlier work this paper cites.
Koutnik, Jan, Greff, Klaus, Gomez, Faustino, and Schmidhuber, Juergen · 2014
Earlier work this paper cites.
Sequence to sequence learning with neural networks
Sutskever, Ilya, Vinyals, Oriol, and Le, Quoc V · 2014
Earlier work this paper cites.
Weston, Jason, Chopra, Sumit, and Bordes, Antoine · 2014
Earlier work this paper cites.
Entity-centric coreference resolution with model stacking
Clark, Kevin and Manning, Christopher D · 2015
Earlier work this paper cites.
Teaching machines to read and comprehend
Hermann, Karl Moritz, Kocisky, Tomas, Grefenstette, Edward, Espeholt, Lasse, Kay, Will, Suleyman, Mustafa, and Blunsom, Phil · 2015
Cited alongside, same era.
Skip-thought vectors
Kiros, Ryan, Zhu, Yukun, Salakhutdinov, Ruslan R, Zemel, Richard, Urtasun, Raquel, Torralba, Antonio, and Fidler, Sanja · 2015
Cited alongside, same era.
End-to-end memory networks
Sukhbaatar, Sainbayar, Weston, Jason, Fergus, Rob, et al · 2015
Cited alongside, same era.
Improved semantic representations from tree-structured long short-term memory networks
Tai, Kai Sheng, Socher, Richard, and Manning, Christopher D · 2015
Cited alongside, same era.
Towards ai-complete question answering: A set of prerequisite toy tasks
Weston, Jason, Bordes, Antoine, Chopra, Sumit, Rush, Alexander M, van Merriënboer, Bart, Joulin, Armand, and Mikolov, Tomas · 2015
Cited alongside, same era.
Gated graph sequence neural networks
Li, Yujia, Tarlow, Daniel, Brockschmidt, Marc, and Zemel, Richard · 2016
Later among the works it cites.
Key-value memory networks for directly reading documents
Miller, Alexander, Fisch, Adam, Dodge, Jesse, Karimi, Amir-Hossein, Bordes, Antoine, and Weston, Jason · 2016
Later among the works it cites.
Reasoning with memory augmented neural networks for language comprehension
Munkhdalai, Tsendsuren and Yu, Hong · 2016
Later among the works it cites.
Who did what: A large-scale person-centered cloze dataset
Onishi, Takeshi, Wang, Hai, Bansal, Mohit, Gimpel, Kevin, and McAllester, David · 2016
Later among the works it cites.
Pixel recurrent neural networks
Oord, Aaron van den, Kalchbrenner, Nal, and Kavukcuoglu, Koray · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ahn, Sungjin, Choi, Heeyoul, Pärnamaa, Tanel, and Bengio, Yoshua · 2016
Cited alongside, same era.
A thorough examination of the cnn/daily mail reading comprehension task
Chen, Danqi, Bolton, Jason, and Manning, Christopher D · 2016
Cited alongside, same era.
Broad context language modeling as reading comprehension
Chu, Zewei, Wang, Hai, Gimpel, Kevin, and McAllester, David · 2016
Cited alongside, same era.
Gated-attention readers for text comprehension
Dhingra, Bhuwan, Liu, Hanxiao, Cohen, William W, and Salakhutdinov, Ruslan · 2016
Cited alongside, same era.
Tracking the world state with recurrent entity networks
Henaff, Mikael, Weston, Jason, Szlam, Arthur, Bordes, Antoine, and LeCun, Yann · 2016
Cited alongside, same era.
Text understanding with the attention sum reader network
Kadlec, Rudolf, Schmid, Martin, Bajgar, Ondrej, and Kleindienst, Jan · 2016
Cited alongside, same era.
On the properties of neural machine translation: Encoder-decoder approaches
Cho, Kyunghyun, Van Merriënboer, Bart, Bahdanau, Dzmitry, and Bengio, Yoshua
Cited in the paper.
Later among the works it cites.
The lambada dataset: Word prediction requiring a broad discourse context
Paperno, Denis, Kruszewski, Germán, Lazaridou, Angeliki, Pham, Quan Ngoc, Bernardi, Raffaella, Pezzelle, Sandro, Baroni, Marco, Boleda, Gemma, and Fernández, Raquel · 2016
Later among the works it cites.
Squad: 100,000+ questions for machine comprehension of text
Rajpurkar, Pranav, Zhang, Jian, Lopyrev, Konstantin, and Liang, Percy · 2016
Later among the works it cites.
Dag-recurrent neural networks for scene labeling
Shuai, Bing, Zuo, Zhen, Wang, Bing, and Wang, Gang · 2016
Later among the works it cites.
Reference-aware language models
Yang, Zichao, Blunsom, Phil, Dyer, Chris, and Ling, Wang · 2016
Later among the works it cites.
Frustratingly short attention spans in neural language modeling
Daniluk, Michał, Rocktäschel, Tim, Welbl, Johannes, and Riedel, Sebastian · 2017
Closest in time.
Emergent logical structure in vector representations of neural readers
Wang, Hai, Onishi, Takeshi, Gimpel, Kevin, and McAllester, David · 2017
Closest in time.