Fetching the paper…
Reading the bibliography…
Although Perplexity is a widely used performance metric for language models, the values are highly dependent upon the number of words in the corpus and is useful to compare performance of the same corpus only.
A neural probabilistic language model
Bengio, Y., Ducharme, R., Vincent, P., & Jauvin, C. (2003) · 2003
Earlier work this paper cites.
Extensions of recurrent neural network language model
Mikolov, T., Kombrink, S., Burget, L., Černockỳ, J., & Khudanpur, S. (2011) · 2011
Earlier work this paper cites.
Learning phrase representations using rnn encoder–decoder for statistical machine translation
Cho, K., van Merrienboer, B., Gulcehre, C., Bahdanau, D., Bougares, F., Schwenk, H., & Bengio, Y. (2014) · 2014
Cited alongside, same era.
Deep speech 2: End-to-end speech recognition in english and mandarin
Amodei, D., Ananthanarayanan, S., Anubhai, R., Bai, J., Battenberg, E., Case, C., Casper, J., Catanzaro, B., Cheng, Q., Chen, G. et al. (2016) · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…