Fetching the paper…
Reading the bibliography…
In this work, we propose a novel method to incorporate corpus-level discourse information into language modelling.
Learning representations by back-propagating errors
Rumelhart, David E, Hinton, Geoffrey E, and Williams, Ronald J · 1988
Earlier work this paper cites.
Learning long-term dependencies with gradient descent is difficult
Bengio, Yoshua, Simard, Patrice, and Frasconi, Paolo · 1994
Earlier work this paper cites.
Improved backing-off for m-gram language modeling
Kneser, Reinhard and Ney, Hermann · 1995
Earlier work this paper cites.
Recursive hetero-associative memories for translation
Forcada, Mikel L and Ñeco, Ramón P · 1997
Earlier work this paper cites.
Long short-term memory
Hochreiter, Sepp and Schmidhuber, Jürgen · 1997
Earlier work this paper cites.
On knowing a word
Miller, George A · 1999
Earlier work this paper cites.
Two decades of statistical language modeling: where do we go from here
Rosenfeld, Ronald · 2000
Earlier work this paper cites.
A neural probabilistic language model
Bengio, Yoshua, Ducharme, Réjean, Vincent, Pascal, and Janvin, Christian · 2003
Earlier work this paper cites.
Latent dirichlet allocation
Blei, David M, Ng, Andrew Y, and Jordan, Michael I · 2003
Earlier work this paper cites.
Feature-rich part-of-speech tagging with a cyclic dependency network
Toutanova, Kristina, Klein, Dan, Manning, Christopher D, and Singer, Yoram · 2003
Earlier work this paper cites.
Latent semantic analysis
Dumais, Susan T · 2004
Earlier work this paper cites.
Practical solutions to the problem of diagonal dominance in kernel document clustering
Greene, Derek and Cunningham, Pádraig · 2006
Earlier work this paper cites.
The psychological functions of function words
Chung, Cindy and Pennebaker, James W · 2007
Cited alongside, same era.
Continuous space language models
Schwenk, Holger · 2007
Cited alongside, same era.
Recurrent neural network based language model
Mikolov, Tomas, Karafiát, Martin, Burget, Lukas, Cernockỳ, Jan, and Khudanpur, Sanjeev · 2010
Cited alongside, same era.
Learning word vectors for sentiment analysis
Maas, Andrew L, Daly, Raymond E, Pham, Peter T, Huang, Dan, Ng, Andrew Y, and Potts, Christopher · 2011
Cited alongside, same era.
Extensions of recurrent neural network language model
Mikolov, Tomáš, Kombrink, Stefan, Burget, Lukáš, Černockỳ, Jan Honza, and Khudanpur, Sanjeev · 2011
Cited alongside, same era.
Context dependent recurrent neural network language model
Mikolov, Tomas and Zweig, Geoffrey · 2012
Cited alongside, same era.
Learning longer memory in recurrent neural networks
Mikolov, Tomas, Joulin, Armand, Chopra, Sumit, Mathieu, Michael, and Ranzato, Marc’Aurelio · 2014
Later among the works it cites.
Sequence to sequence learning with neural networks
Sutskever, Ilya, Vinyals, Oriol, and Le, Quoc VV · 2014
Later among the works it cites.
Gated feedback recurrent neural networks
Chung, Junyoung, Gulcehre, Caglar, Cho, Kyunghyun, and Bengio, Yoshua · 2015
Closest in time.
Greff, Klaus, Srivastava, Rupesh Kumar, Koutník, Jan, Steunebrink, Bas R, and Schmidhuber, Jürgen · 2015
Closest in time.
Teaching machines to read and comprehend
Hermann, Karl Moritz, Kočiskỳ, Tomáš, Grefenstette, Edward, Espeholt, Lasse, Kay, Will, Suleyman, Mustafa, and Blunsom, Phil · 2015
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zeiler, Matthew D · 2012
Cited alongside, same era.
Generating sequences with recurrent neural networks
Graves, Alex · 2013
Cited alongside, same era.
Scalable modified Kneser-Ney language model estimation
Heafield, Kenneth, Pouzyrevsky, Ivan, Clark, Jonathan H., and Koehn, Philipp · 2013
Cited alongside, same era.
Recurrent continuous translation models
Kalchbrenner, Nal and Blunsom, Phil · 2013
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Bahdanau, Dzmitry, Cho, Kyunghyun, and Bengio, Yoshua · 2014
Cited alongside, same era.
Pragmatic neural language modelling in machine translation
Baltescu, Paul and Blunsom, Phil · 2014
Cited alongside, same era.
The goldilocks principle: Reading children’s books with explicit memory representations
Hill, Felix, Bordes, Antoine, Chopra, Sumit, and Weston, Jason · 2015
Closest in time.
An empirical exploration of recurrent network architectures
Jozefowicz, Rafal, Zaremba, Wojciech, and Sutskever, Ilya · 2015
Closest in time.
Kiros, Ryan, Zhu, Yukun, Salakhutdinov, Ruslan, Zemel, Richard S, Torralba, Antonio, Urtasun, Raquel, and Fidler, Sanja · 2015
Closest in time.
Hierarchical neural network generative models for movie dialogues
Serban, Iulian V, Sordoni, Alessandro, Bengio, Yoshua, Courville, Aaron, and Pineau, Joelle · 2015
Closest in time.
Sukhbaatar, Sainbayar, Szlam, Arthur, Weston, Jason, and Fergus, Rob · 2015
Closest in time.
From feedforward to recurrent lstm neural networks for language modeling
Sundermeyer, Martin, Ney, Hermann, and Schluter, Ralf · 2015
Closest in time.