Fetching the paper…
Reading the bibliography…
Neural Machine Translation (NMT) models are typically trained on heterogeneous data that are concatenated and randomly shuffled.
Maturational constraints on language learning
Elissa L. Newport. 1990 · 1990
Earlier work this paper cites.
Learning and development in neural networks: the importance of starting small
Jeffrey L. Elman. 1993 · 1993
Earlier work this paper cites.
Curriculum learning
Yoshua Bengio, Jérôme Louradour, Ronan Collobert, and Jason Weston. 2009 · 2009
Earlier work this paper cites.
Self-paced learning for latent variable models
M. Kumar, Benjamin Packer, and Daphne Koller. 2010 · 2010
Earlier work this paper cites.
Intelligent selection of language model training data
Robert C. Moore and William Lewis. 2010 · 2010
Earlier work this paper cites.
Domain adaptation via pseudo in-domain data selection
Amittai Axelrod, Xiaodong He, and Jianfeng Gao. 2011 · 2011
Earlier work this paper cites.
Self-paced curriculum learning
Lu Jiang, Deyu Meng, Qian Zhao, Shiguang Shan, and Alexander G. Hauptmann. 2015 · 2015
Earlier work this paper cites.
How to avoid unwanted pregnancies: Domain adaptation using neural network models
Shafiq Joty, Hassan Sajjad, Nadir Durrani, Kamla Al-Mannai, Ahmed Abdelali, and Stephan Vogel. 2015 · 2015
Earlier work this paper cites.
Stanford neural machine translation systems for spoken language domains
Minh-Thang Luong and Christopher Manning. 2015 · 2015
Earlier work this paper cites.
Fast domain adaptation for neural machine translation
Markus Freitag and Yaser Al-Onaizan. 2016 · 2016
Earlier work this paper cites.
Easy questions first? a case study on curriculum learning for question answering
Mrinmaya Sachan and Eric Xing. 2016 · 2016
Earlier work this paper cites.
Learning the curriculum with Bayesian optimization for task-specific word representation learning
Yulia Tsvetkov, Manaal Faruqui, Wang Ling, Brian MacWhinney, and Chris Dyer. 2016 · 2016
Earlier work this paper cites.
Transfer learning for low-resource neural machine translation
Barret Zoph, Deniz Yuret, Jonathan May, and Kevin Knight. 2016 · 2016
Earlier work this paper cites.
Word translation without parallel data
Alexis Conneau, Guillaume Lample, Marc’Aurelio Ranzato, Ludovic Denoyer, and Hervé Jégou. 2017 · 2017
Cited alongside, same era.
Curriculum learning and minibatch bucketing in neural machine translation
Tom Kocmi and Ondřej Bojar. 2017 · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Ł ukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
Achieving human parity on automatic chinese to english news translation
Hany Hassan, Anthony Aue, C. Chen, Vishal Chowdhary, J. Clark, C. Federmann, Xuedong Huang, Marcin Junczys-Dowmunt, W. Lewis, M. Li, Shujie Liu, T. Liu, Renqian Luo, Arul Menezes, Tao Qin, F. Seide, Xu Tan, Fei Tian, Lijun Wu, Shuangzhi Wu, Yingce Xia, Dongdong Zhang, Zhirui Zhang, and M. Zhou. 2018 · 2018
Cited alongside, same era.
Dual conditional cross-entropy filtering of noisy parallel corpora
Marcin Junczys-Dowmunt. 2018 · 2018
Competence-based curriculum learning for neural machine translation
Emmanouil Antonios Platanios, Otilia Stretcu, Graham Neubig, Barnabas Poczos, and Tom Mitchell. 2019 · 2019
Later among the works it cites.
Simple and effective curriculum pointer-generator networks for reading comprehension over long narratives
Yi Tay, Shuohang Wang, Anh Tuan Luu, Jie Fu, Minh C. Phan, Xingdi Yuan, Jinfeng Rao, Siu Cheung Hui, and Aston Zhang. 2019 · 2019
Later among the works it cites.
ParaCrawl: Web-scale acquisition of parallel corpora
Marta Bañón, Pinzhen Chen, Barry Haddow, Kenneth Heafield, Hieu Hoang, Miquel Esplà-Gomis, Mikel L. Forcada, Amir Kamran, Faheem Kirefu, Philipp Koehn, Sergio Ortiz Rojas, Leopoldo Pla Sempere, Gema Ramírez-Sánchez, Elsa Sarrías, Marek Strelec, Brian Thompson, William Waites, Dion Wiggins, and Jaume Zaragoza. 2020 · 2020
Later among the works it cites.
Data Rejuvenation: Exploiting Inactive Training Examples for Neural Machine Translation
Wenxiang Jiao, Xing Wang, Shilin He, Irwin King, Michael Lyu, and Zhaopeng Tu. 2020 · 2020
Later among the works it cites.
Norm-based curriculum learning for neural machine translation
Xuebo Liu, Houtim Lai, Derek F. Wong, and Lidia S. Chao. 2020 · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
On the impact of various types of noise on neural machine translation
Huda Khayrallah and Philipp Koehn. 2018 · 2018
Cited alongside, same era.
A call for clarity in reporting BLEU scores
Matt Post. 2018 · 2018
Cited alongside, same era.
Denoising neural machine translation training with trusted data and online data selection
Wei Wang, Taro Watanabe, Macduff Hughes, Tetsuji Nakagawa, and Ciprian Chelba. 2018 · 2018
Cited alongside, same era.
An empirical exploration of curriculum learning for neural machine translation
Xuan Zhang, Gaurav Kumar, Huda Khayrallah, Kenton Murray, Jeremy Gwinnup, Marianna J Martindale, Paul McNamee, Kevin Duh, and Marine Carpuat. 2018 · 2018
Cited alongside, same era.
Margin-based parallel corpus mining with multilingual sentence embeddings
Mikel Artetxe and Holger Schwenk. 2019 · 2019
Cited alongside, same era.
Low-resource corpus filtering using multilingual sentence embeddings
Vishrav Chaudhary, Yuqing Tang, Francisco Guzmán, Holger Schwenk, and Philipp Koehn. 2019 · 2019
Cited alongside, same era.
On the power of curriculum learning in training deep networks
Guy Hacohen and Daphna Weinshall. 2019 · 2019
Cited alongside, same era.
Later among the works it cites.
Transforming machine translation: a deep learning system reaches news translation quality comparable to human professionals
Martin Popel, Marketa Tomkova, Jakub Tomek, Łukasz Kaiser, Jakob Uszkoreit, Ondřej Bojar, and Zdeněk Žabokrtskỳ. 2020 · 2020
Later among the works it cites.
Self-paced learning for neural machine translation
Yu Wan, Baosong Yang, Derek F. Wong, Yikai Zhou, Lidia S. Chao, Haibo Zhang, and Boxing Chen. 2020 · 2020
Later among the works it cites.
Reinforced curriculum learning on pre-trained neural machine translation models
Mingjun Zhao, Haijiang Wu, Di Niu, and Xiaoli Wang. 2020 · 2020
Later among the works it cites.
Findings of the 2021 conference on machine translation (WMT21)
Farhad Akhbardeh, Arkady Arkhangorodsky, Magdalena Biesialska, Ondřej Bojar, Rajen Chatterjee, Vishrav Chaudhary, Marta R. Costa-jussa, Cristina España-Bonet, Angela Fan, Christian Federmann, Markus Freitag, Yvette Graham, Roman Grundkiewicz, Barry Haddow, Leonie Harter, Kenneth Heafield, Christopher Homan, Matthias Huck, Kwabena Amponsah-Kaakyire, Jungo Kasai, Daniel Khashabi, Kevin Knight, Tom Kocmi, Philipp Koehn, Nicholas Lourie, Christof Monz, Makoto Morishita, Masaaki Nagata, Ajay Nagesh, Toshiaki Nakazawa, Matteo Negri, Santanu Pal, Allahsera Auguste Tapo, Marco Turchi, Valentin Vydrin, and Marcos Zampieri. 2021 · 2021
Later among the works it cites.
Curriculum learning for language modeling
Daniel Campos. 2021 · 2021
Later among the works it cites.
Gradient-guided loss masking for neural machine translation
Xinyi Wang, Ankur Bapna, Melvin Johnson, and Orhan Firat. 2021 · 2021
Later among the works it cites.
Meta-curriculum learning for domain adaptation in neural machine translation
Runzhe Zhan, Xuebo Liu, Derek F. Wong, and Lidia S. Chao. 2021 · 2021
Later among the works it cites.
Reinforcement learning based curriculum optimization for neural machine translation
Gaurav Kumar, George Foster, Colin Cherry, and Maxim Krikun. 2019 · 2061
Closest in time.