Fetching the paper…
Reading the bibliography…
Massively multilingual machine translation (MT) has shown impressive capabilities, including zero and few-shot translation between low-resource language pairs.
A focus on neural machine translation for african languages
Martinus, L. and Abbott, J. Z. (2019) · 1906
Earlier work this paper cites.
Massively multilingual neural machine translation in the wild: Findings and challenges
Arivazhagan, N., Bapna, A., Firat, O., Lepikhin, D., Johnson, M., Krikun, M., Chen, M. X., Cao, Y., Foster, G. F., Cherry, C., Macherey, W., Chen, Z., and Wu, Y. (2019) · 1907
Earlier work this paper cites.
Ccmatrix: Mining billions of high-quality parallel sentences on the web
Schwenk, H., Wenzek, G., Edunov, S., Grave, E., and Joulin, A. (2020) · 1911
Earlier work this paper cites.
The Bible as a Parallel Corpus: Annotating the ‘Book of 2000 Tongues’
Resnik, P., Olsen, M. B., and Diab, M. T. (1999) · 2000
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Papineni, K., Roukos, S., Ward, T., and Zhu, W.-J. (2002) · 2002
Earlier work this paper cites.
Evaluating Amharic Machine Translation
Hadgu, A. T., Beaudoin, A., and Aregawi, A. (2020) · 2003
Earlier work this paper cites.
Improving Yorùbá Diacritic Restoration
Orife, I., Adelani, D., Fasubaa, T. E., Williamson, V., Oyewusi, W. F., Wahab, O., and Túbosún, K. (2020) · 2003
Earlier work this paper cites.
Tailoring and Evaluating the Wikipedia for in-Domain Comparable Corpora Extraction
España-Bonet, C., Barrón-Cedeño, A., and Màrquez, L. (2020) · 2005
Earlier work this paper cites.
Moses: Open source toolkit for statistical machine translation
Koehn, P., Hoang, H., Birch, A., Callison-Burch, C., Federico, M., Bertoldi, N., Cowan, B., Shen, W., Moran, C., Zens, R., Dyer, C., Bojar, O., Constantin, A., and Herbst, E. (2007) · 2007
Earlier work this paper cites.
Multilingual translation with extensible multilingual pretraining and finetuning
Tang, Y., Tran, C., Li, X., Chen, P., Goyal, N., Chaudhary, V., Gu, J., and Fan, A. (2020) · 2008
Earlier work this paper cites.
Introducing the autshumato integrated translation environment
Groenewald, H. J. and Fourie, W. (2009) · 2009
Earlier work this paper cites.
Beyond english–centric multilingual machine translation
Fan, A., Bhosale, S., Schwenk, H., Ma, Z., El-Kishky, A., Goyal, S., Baines, M., Celebi, O., Wenzek, G., Chaudhary, V., Goyal, N., Birch, T., Liptchinsky, V., Edunov, S., Grave, E., Auli, M., and Joulin, A. (2020) · 2010
Earlier work this paper cites.
KenLM: Faster and smaller language model queries
Heafield, K. (2011) · 2011
Earlier work this paper cites.
Polyglot: A massive multilingual natural language processing pipeline
Al-Rfou, R. (2015) · 2015
Earlier work this paper cites.
On using monolingual corpora in neural machine translation
Gulcehre, C., Firat, O., Xu, K., Cho, K., Barrault, L., Lin, H.-C., Bougares, F., Schwenk, H., and Bengio, Y. (2015) · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J. (2015) · 2015
Earlier work this paper cites.
chrF: character n-gram f-score for automatic MT evaluation
Popović, M. (2015) · 2015
Cited alongside, same era.
Neural machine translation of rare words with subword units
Sennrich, R., Haddow, B., and Birch, A. (2016) · 2016
Cited alongside, same era.
LORELEI language packs: Data, tools, and resources for technology development in low resource languages
Strassel, S. and Tracey, J. (2016) · 2016
Cited alongside, same era.
Transfer learning for low-resource neural machine translation
Zoph, B., Yuret, D., May, J., and Knight, K. (2016) · 2016
Cited alongside, same era.
Copied monolingual data improves low-resource neural machine translation
Currey, A., Miceli Barone, A. V., and Heafield, K. (2017) · 2017
Cited alongside, same era.
Google’s Multilingual Neural Machine Translation System: Enabling Zero-Shot Translation
Johnson, M., Schuster, M., Le, Q. V., Krikun, M., Wu, Y., Chen, Z., Thorat, N., Viégas, F. B., Wattenberg, M., Corrado, G., Hughes, M., and Dean, J. (2017) · 2017
Unsupervised pivot translation for distant languages
Leng, Y., Tan, X., Qin, T., Li, X.-Y., and Liu, T.-Y. (2019) · 2019
Later among the works it cites.
Fairseq: A fast, extensible toolkit for sequence modeling
Ott, M., Edunov, S., Baevski, A., Fan, A., Gross, S., Ng, N., Grangier, D., and Auli, M. (2019) · 2019
Later among the works it cites.
Self-supervised neural machine translation
Ruiter, D., España-Bonet, C., and van Genabith, J. (2019) · 2019
Later among the works it cites.
Revisiting low-resource neural machine translation: A case study
Sennrich, R. and Zhang, B. (2019) · 2019
Later among the works it cites.
Massive vs. curated embeddings for low-resourced languages: the case of Yorùbá and Twi
Alabi, J., Amponsah-Kaakyire, K., Adelani, D., and España-Bonet, C. (2020) · 2020
Later among the works it cites.
Proceedings of the Fifth Conference on Machine Translation
Barrault, L., Bojar, O., Bougares, F., Chatterjee, R., Costa-jussà, M. R., Federmann, C., Fishel, M., Fraser, A., Graham, Y., Guzman, P., Haddow, B., Huck, M., Yepes, A. J., Koehn, P., Martins, A., Morishita, M., Monz, C., Nagata, M., Nakazawa, T., and Negri, M., editors (2020) · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Unsupervised pretraining for sequence to sequence learning
Ramachandran, P., Liu, P., and Le, Q. (2017) · 2017
Cited alongside, same era.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, L. u., and Polosukhin, I. (2017) · 2017
Cited alongside, same era.
Sequence-to-Sequence Learning for Automatic Yorùbá Diacritic Restoration
Orife, I. F. d. (2018) · 2018
Cited alongside, same era.
JW300: A wide-coverage parallel corpus for low-resource languages
Agić, Ž. and Vulić, I. (2019) · 2019
Cited alongside, same era.
Margin-based parallel corpus mining with multilingual sentence embeddings
Artetxe, M. and Schwenk, H. (2019) · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Devlin, J., Chang, M.-W., Lee, K., and Toutanova, K. (2019) · 2019
Cited alongside, same era.
Later among the works it cites.
Embedding space correlation as a measure of domain similarity
Beyer, A., Kauermann, G., and Schütze, H. (2020) · 2020
Later among the works it cites.
Igbo-english machine translation: An evaluation benchmark
Ezeani, I., Rayson, P., Onyenwe, I., Chinedu, U., and Hepple, M. (2020) · 2020
Later among the works it cites.
Participatory research for low-resourced machine translation: A case study in african languages
∀ \forall , Nekoto, W., Marivate, V., Matsila, T., Fasubaa, T., Fagbohungbe, T., Akinola, S. O., Muhammad, S., Kabongo Kabenamualu, S., Osei, S., Sackey, F., Niyongabo, R. A., …, Ogueji, K., Siminyu, K., Kreutzer, J., .., and Bashir, A. (2020) · 2020
Later among the works it cites.
OPUS-MT – building open translation services for the world
Tiedemann, J. and Thottingal, S. (2020) · 2020
Later among the works it cites.
Quality at a Glance: An Audit of Web-Crawled Multilingual Datasets
Caswell, I., Kreutzer, J., Wang, L., Wahab, A., van Esch, D., Ulzii-Orshikh, N., Tapo, A., Subramani, N., Sokolov, A., Sikasote, C., Setyawan, M., Sarin, S., Samb, S., Sagot, B., Rivera, C., Rios, A., Papadimitriou, I., Osei, S., Ortiz Suárez, P. J., Orife, I., Ogueji, K., Niyongabo, R. A., Nguyen, T. Q., Müller, M., Müller, A., Hassan Muhammad, S., Muhammad, N., Mnyakeni, A., Mirzakhalov, J., Matangira, T., Leong, C., Lawson, N., Kudugunta, S., Jernite, Y., Jenny, M., Firat, O., Dossou, B. F. P., Dlamini, S., de Silva, N., Çabuk Ballı, S., Biderman, S., Battisti, A., Baruwa, A., Bapna, A., Baljekar, P., Abebe Azime, I., Awokoya, A., Ataman, D., Ahia, O., Ahia, O., Agrawal, S., and Adeyemi, M. (2021) · 2021
Closest in time.
Integrating Unsupervised Data Generation into Self-Supervised Neural Machine Translation for Low-Resource Languages
Ruiter, D., Klakow, D., van Genabith, J., and España-Bonet, C. (2021) · 2021
Closest in time.
WikiMatrix: Mining 135M Parallel Sentences in 1620 Language Pairs from Wikipedia
Schwenk, H., Chaudhary, V., Sun, S., Gong, H., and Guzmán, F. (2021) · 2021
Closest in time.
Ai4d - african language program
Siminyu, K., Kalipe, G., Orlic, D., Abbott, J., Marivate, V., Freshia, S., Sibal, P., Neupane, B., Adelani, D., Taylor, A., Ali, J. T., Degila, K., Balogoun, M., Diop, T. I., David, D., Fourati, C., Haddad, H., and Naski, M. (2021) · 2021
Closest in time.
mT5: A massively multilingual pre-trained text-to-text transformer
Xue, L., Constant, N., Roberts, A., Kale, M., Al-Rfou, R., Siddhant, A., Barua, A., and Raffel, C. (2021) · 2021
Closest in time.