Fetching the paper…
Reading the bibliography…
Neural Machine Translation (NMT) is an end-to-end learning approach for automated translation, with the potential to overcome many of the weaknesses of conventional phrase-based translation systems.
A statistical approach to language translation
Brown, P., Cocke, J., Pietra, S. D., Pietra, V. D., Jelinek, F., Mercer, R., and Roossin, P · 1988
Earlier work this paper cites.
A statistical approach to machine translation
Brown, P. F., Cocke, J., Pietra, S. A. D., Pietra, V. J. D., Jelinek, F., Lafferty, J. D., Mercer, R. L., and Roossin, P. S · 1990
Earlier work this paper cites.
The cascade-correlation learning architecture
Fahlman, S. E., and Lebiere, C · 1990
Earlier work this paper cites.
Learning recursive distributed representations for holistic computation
Chrisman, L · 1991
Earlier work this paper cites.
The mathematics of statistical machine translation: Parameter estimation
Brown, P. F., Pietra, V. J. D., Pietra, S. A. D., and Mercer, R. L · 1993
Earlier work this paper cites.
Long short-term memory
Hochreiter, S., and Schmidhuber, J · 1997
Earlier work this paper cites.
Bidirectional recurrent neural networks
Schuster, M., and Paliwal, K · 1997
Earlier work this paper cites.
Learning to forget: Continual prediction with LSTM
Gers, F. A., Schmidhuber, J., and Cummins, F · 2000
Earlier work this paper cites.
Gradient flow in recurrent nets: the difficulty of learning long-term dependencies, 2001
Hochreiter, S., Bengio, Y., Frasconi, P., and Schmidhuber, J · 2001
Earlier work this paper cites.
Statistical phrase-based translation
Koehn, P., Och, F. J., and Marcu, D · 2003
Earlier work this paper cites.
Large scale distributed deep networks
Dean, J., Corrado, G. S., Monga, R., Chen, K., Devin, M., Le, Q. V., Mao, M. Z., Ranzato, M., Senior, A., Tucker, P., Yang, K., and Ng, A. Y · 2012
Earlier work this paper cites.
Understanding the exploding gradient problem
Pascanu, R., Mikolov, T., and Bengio, Y · 2012
Earlier work this paper cites.
Japanese and Korean voice search
Schuster, M., and Nakajima, K · 2012
Earlier work this paper cites.
Recurrent continuous translation models
Kalchbrenner, N., and Blunsom, P · 2013
Earlier work this paper cites.
N-gram counts and language models from the common crawl
Buck, C., Heafield, K., and Van Ooyen, B · 2014
Earlier work this paper cites.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
Cho, K., van Merrienboer, B., Gülçehre, Ç., Bougares, F., Schwenk, H., and Bengio, Y · 2014
Cited alongside, same era.
Fast and robust neural network joint models for statistical machine translation
Devlin, J., Zbib, R., Huang, Z., Lamar, T., Schwartz, R. M., and Makhoul, J · 2014
Cited alongside, same era.
Edinburgh’s phrase-based machine translation systems for WMT-14
Durrani, N., Haddow, B., Koehn, P., and Heafield, K · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Kingma, D. P., and Ba, J · 2014
Cited alongside, same era.
Sequence to sequence learning with neural networks
Sutskever, I., Vinyals, O., and Le, Q. V · 2014
Cited alongside, same era.
Recurrent neural network regularization, 2014
On using very large target vocabulary for neural machine translation
Sébastien, J., Kyunghyun, C., Memisevic, R., and Bengio, Y · 2015
Later among the works it cites.
Srivastava, R. K., Greff, K., and Schmidhuber, J · 2015
Later among the works it cites.
Quantized convolutional neural networks for mobile devices
Wu, J., Leng, C., Wang, Y., Hu, Q., and Cheng, J · 2015
Later among the works it cites.
Tensorflow: A system for large-scale machine learning
Abadi, M., Barham, P., Chen, J., Chen, Z., Davis, A., Dean, J., Devin, M., Ghemawat, S., Irving, G., Isard, M., Kudlur, M., Levenberg, J., Monga, R., Moore, S., Murray, D. G., Steiner, B., Tucker, P., Vasudevan, V., Warden, P., Wicke, M., Yu, Y., and Zheng, X · 2016
Closest in time.
A character-level decoder without explicit segmentation for neural machine translation
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zaremba, W., Sutskever, I., and Vinyals, O · 2014
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Bahdanau, D., Cho, K., and Bengio, Y · 2015
Cited alongside, same era.
Multi-task learning for multiple language translation
Dong, D., Wu, H., He, W., Yu, D., and Wang, H · 2015
Cited alongside, same era.
Deep learning with limited numerical precision
Gupta, S., Agrawal, A., Gopalakrishnan, K., and Narayanan, P · 2015
Cited alongside, same era.
Han, S., Mao, H., and Dally, W. J · 2015
Cited alongside, same era.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S., and Sun, J · 2015
Cited alongside, same era.
Multi-task sequence to sequence learning
Luong, M.-T., Le, Q. V., Sutskever, I., Vinyals, O., and Kaiser, L · 2015
Cited alongside, same era.
Chung, J., Cho, K., and Bengio, Y · 2016
Closest in time.
A character-level decoder without explicit segmentation for neural machine translation
Chung, J., Cho, K., and Bengio, Y · 2016
Closest in time.
Character-based neural machine translation
Costa-Jussà, M. R., and Fonollosa, J. A. R · 2016
Closest in time.
Gülçehre, Ç., Ahn, S., Nallapati, R., Zhou, B., and Bengio, Y · 2016
Closest in time.
Li, F., and Liu, B · 2016
Closest in time.
Achieving open vocabulary neural machine translation with hybrid word-character models
Luong, M., and Manning, C. D · 2016
Closest in time.
Reward augmented maximum likelihood for neural structured prediction
Norouzi, M., Bengio, S., Chen, Z., Jaitly, N., Schuster, M., Wu, Y., and Schuurmans, D · 2016
Closest in time.
Neural machine translation of rare words with subword units
Sennrich, R., Haddow, B., and Birch, A · 2016
Closest in time.
Minimum risk training for neural machine translation
Shen, S., Cheng, Y., He, Z., He, W., Wu, H., Sun, M., and Liu, Y · 2016
Closest in time.
Coverage-based neural machine translation
Tu, Z., Lu, Z., Liu, Y., Liu, X., and Li, H · 2016
Closest in time.
Deep recurrent models with fast-forward connections for neural machine translation
Zhou, J., Cao, Y., Wang, X., Li, P., and Xu, W · 2016
Closest in time.