Fetching the paper…
Reading the bibliography…
Deep Neural Networks (DNNs) are powerful models that have achieved excellent performance on difficult learning tasks.
Learning representations by back-propagating errors
D. Rumelhart, G. E. Hinton, and R. J. Williams · 1986
Earlier work this paper cites.
Backpropagation through time: what it does and how to do it
P. Werbos · 1990
Earlier work this paper cites.
Untersuchungen zu dynamischen neuronalen netzen
S. Hochreiter · 1991
Earlier work this paper cites.
On small depth threshold circuits
A. Razborov · 1992
Earlier work this paper cites.
Learning long-term dependencies with gradient descent is difficult
Y. Bengio, P. Simard, and P. Frasconi · 1994
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
LSTM can solve hard long time lag problems
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Gradient flow in recurrent nets: the difficulty of learning long-term dependencies, 2001
S. Hochreiter, Y. Bengio, P. Frasconi, and J. Schmidhuber · 2001
Earlier work this paper cites.
BLEU: a method for automatic evaluation of machine translation
K. Papineni, S. Roukos, T. Ward, and W. J. Zhu · 2002
Earlier work this paper cites.
A neural probabilistic language model
Y. Bengio, R. Ducharme, P. Vincent, and C. Jauvin · 2003
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
A. Graves, S. Fernández, F. Gomez, and J. Schmidhuber · 2006
Cited alongside, same era.
Recurrent neural network based language model
T. Mikolov, M. Karafiát, L. Burget, J. Cernockỳ, and S. Khudanpur · 2010
Cited alongside, same era.
LSTM neural networks for language modeling
M. Sundermeyer, R. Schluter, and H. Ney · 2010
Cited alongside, same era.
Multi-column deep neural networks for image classification
D. Ciresan, U. Meier, and J. Schmidhuber · 2012
Cited alongside, same era.
Context-dependent pre-trained deep neural networks for large vocabulary speech recognition
G. E. Dahl, D. Yu, L. Deng, and A. Acero · 2012
Cited alongside, same era.
Deep neural networks for acoustic modeling in speech recognition
G. Hinton, L. Deng, D. Yu, G. Dahl, A. Mohamed, N. Jaitly, A. Senior, V. Vanhoucke, P. Nguyen, T. Sainath, and B. Kingsbury · 2012
Joint language and translation modeling with recurrent neural networks
M. Auli, M. Galley, C. Quirk, and G. Zweig · 2013
Later among the works it cites.
Generating sequences with recurrent neural networks
A. Graves · 2013
Later among the works it cites.
Recurrent continuous translation models
N. Kalchbrenner and P. Blunsom · 2013
Later among the works it cites.
Neural machine translation by jointly learning to align and translate
D. Bahdanau, K. Cho, and Y. Bengio · 2014
Closest in time.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
K. Cho, B. Merrienboer, C. Gulcehre, F. Bougares, H. Schwenk, and Y. Bengio · 2014
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
ImageNet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Cited alongside, same era.
Building high-level features using large scale unsupervised learning
Q.V. Le, M.A. Ranzato, R. Monga, M. Devin, K. Chen, G.S. Corrado, J. Dean, and A.Y. Ng · 2012
Cited alongside, same era.
Statistical Language Models based on Neural Networks
T. Mikolov · 2012
Cited alongside, same era.
On the difficulty of training recurrent neural networks
R. Pascanu, T. Mikolov, and Y. Bengio · 2012
Cited alongside, same era.
Fast and robust neural network joint models for statistical machine translation
J. Devlin, R. Zbib, Z. Huang, T. Lamar, R. Schwartz, and J. Makhoul · 2014
Closest in time.
Edinburgh’s phrase-based machine translation systems for wmt-14
Nadir Durrani, Barry Haddow, Philipp Koehn, and Kenneth Heafield · 2014
Closest in time.
Multilingual distributed representations without word alignment
K. M. Hermann and P. Blunsom · 2014
Closest in time.
Overcoming the curse of sentence length for neural machine translation using automatic segmentation
J. Pouget-Abadie, D. Bahdanau, B. van Merrienboer, K. Cho, and Y. Bengio · 2014
Closest in time.
University le mans
H. Schwenk · 2014
Closest in time.