Fetching the paper…
Reading the bibliography…
Neural Machine Translation (NMT) systems rely on large amounts of parallel data.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
R. Williams · 1992
Earlier work this paper cites.
Catastrophic forgetting in connectionist networks
R. M. French · 1999
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation
R. S. Sutton, D. A. McAllester, S. P. Singh, and Y. Mansour · 2000
Earlier work this paper cites.
Kenlm: Faster and smaller language model queries
K. Heafield · 2011
Earlier work this paper cites.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
K. Cho, B. Van Merriënboer, C. Gulcehre, D. Bahdanau, F. Bougares, H. Schwenk, and Y. Bengio · 2014
Earlier work this paper cites.
On using very large target vocabulary for neural machine translation
S. Jean, K. Cho, R. Memisevic, and Y. Bengio · 2014
Earlier work this paper cites.
Sequence to sequence learning with neural networks
I. Sutskever, O. Vinyals, and Q. V. Le · 2014
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
D. Bahdanau, K. Cho, and Y. Bengio · 2015
Cited alongside, same era.
Tensorflow: A system for large-scale machine learning
M. Abadi, P. Barham, J. Chen, Z. Chen, A. Davis, J. Dean, M. Devin, S. Ghemawat, G. Irving, M. Isard, et al · 2016
Cited alongside, same era.
Dual learning for machine translation
D. He, Y. Xia, T. Qin, L. Wang, N. Yu, T. Liu, and W.-Y. Ma · 2016
Cited alongside, same era.
Google’s multilingual neural machine translation system: enabling zero-shot translation
M. Johnson, M. Schuster, Q. V. Le, M. Krikun, Y. Wu, Z. Chen, N. Thorat, F. Viégas, M. Wattenberg, G. Corrado, et al · 2016
Cited alongside, same era.
Neural machine translation of rare words with subword units
Google’s neural machine translation system: Bridging the gap between human and machine translation
Y. Wu, M. Schuster, Z. Chen, Q. V. Le, M. Norouzi, W. Macherey, M. Krikun, Y. Cao, Q. Gao, K. Macherey, et al · 2016
Later among the works it cites.
The united nations parallel corpus v1. 0
M. Ziemski, M. Junczys-Dowmunt, and B. Pouliquen · 2016
Later among the works it cites.
An actor-critic algorithm for sequence prediction
D. Bahdanau, P. Brakel, K. Xu, A. Goyal, R. Lowe, J. Pineau, A. Courville, and Y. Bengio · 2017
Later among the works it cites.
Dual supervised learning
Y. Xia, T. Qin, W. Chen, J. Bian, N. Yu, and T.-Y. Liu · 2017
Later among the works it cites.
Unsupervised neural machine translation
M. Artetxe, G. Labaka, E. Agirre, and K. Cho · 2018
Closest in time.
Word translation without parallel data
A. Conneau, G. Lample, M. Ranzato, L. Denoyer, and H. Jégou · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
R. Sennrich, B. Haddow, and A. Birch · 2016
Cited alongside, same era.
Unsupervised machine translation using monolingual corpora only
G. Lample, L. Denoyer, and M. Ranzato
Cited in the paper.
Phrase-based & neural unsupervised machine translation
G. Lample, M. Ott, A. Conneau, L. Denoyer, and M. Ranzato
Cited in the paper.
Closest in time.