Fetching the paper…
Reading the bibliography…
We demonstrate the potential of few-shot translation systems, trained with unpaired language data, for both high and low-resource language pairs.
The pronouns of power and solidarity
Brown, R. and Gilman, A · 1960
Earlier work this paper cites.
A multilingual view of unsupervised machine translation
Garcia, X., Foret, P., Sellam, T., and Parikh, A. P · 2002
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Papineni, K., Roukos, S., Ward, T., and Zhu, W.-J · 2002
Earlier work this paper cites.
Is MAP decoding all you need? the inadequacy of the mode in neural machine translation
Eikema, B. and Aziz, W · 2005
Earlier work this paper cites.
Harnessing multilinguality in unsupervised machine translation for rare languages
Garcia, X., Siddhant, A., Firat, O., and Parikh, A. P · 2009
Earlier work this paper cites.
Deciphering foreign language
Ravi, S. and Knight, K · 2011
Earlier work this paper cites.
Unsupervised machine translation using monolingual corpora only
Lample, G., Conneau, A., Denoyer, L., and Ranzato, M · 2017
Earlier work this paper cites.
A study of style in machine translation: Controlling the formality of machine translation output
Niu, X., Martindale, M., and Carpuat, M · 2017
Earlier work this paper cites.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., and Polosukhin, I · 2017
Earlier work this paper cites.
Kudo, T. and Richardson, J · 2018
Earlier work this paper cites.
Neural machine translation into language varieties
Lakew, S. M., Erofeeva, A., and Federico, M · 2018
Earlier work this paper cites.
Multi-task neural models for translating between styles within and across languages
Niu, X., Rao, S., and Carpuat, M · 2018
Earlier work this paper cites.
Adafactor: Adaptive learning rates with sublinear memory cost
Shazeer, N. and Stern, M · 2018
Earlier work this paper cites.
Denoising neural machine translation training with trusted data and online data selection
Wang, W., Watanabe, T., Hughes, M., Nakagawa, T., and Chelba, C · 2018
Earlier work this paper cites.
Unsupervised cross-lingual representation learning at scale
Conneau, A., Khandelwal, K., Goyal, N., Chaudhary, V., Wenzek, G., Guzmán, F., Grave, E., Ott, M., Zettlemoyer, L., and Stoyanov, V · 2019
Earlier work this paper cites.
Mass: Masked sequence to sequence pre-training for language generation
Song, K., Tan, X., Qin, T., Lu, J., and Liu, T.-Y · 2019
Earlier work this paper cites.
Paracrawl: Web-scale acquisition of parallel corpora
Bañón, M., Chen, P., Haddow, B., Heafield, K., Hoang, H., Esplà-Gomis, M., Forcada, M. L., Kamran, A., Kirefu, F., Koehn, P., et al · 2020
Earlier work this paper cites.
Language models are few-shot learners
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J. D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al · 2020
Earlier work this paper cites.
Flax: A neural network library and ecosystem for JAX, 2020
Heek, J., Levskaya, A., Oliver, A., Ritter, M., Rondepierre, B., Steiner, A., and van Zee, M · 2020
Earlier work this paper cites.
The state and fate of linguistic diversity and inclusion in the NLP world
Joshi, P., Santy, S., Budhiraja, A., Bali, K., and Choudhury, M · 2020
Cited alongside, same era.
When and why is unsupervised neural machine translation useless?
Kim, Y., Graça, M., and Ney, H · 2020
Cited alongside, same era.
When does unsupervised machine translation work?
Marchisio, K., Duh, K., and Koehn, P · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Raffel, C., Shazeer, N., Roberts, A., Lee, K., Narang, S., Matena, M., Zhou, Y., Li, W., Liu, P. J., et al · 2020
Cited alongside, same era.
Bleurt: Learning robust metrics for text generation
Sellam, T., Das, D., and Parikh, A. P · 2020
Tencent translation system for the WMT21 news translation task
Wang, L., Li, M., Liu, F., Shi, S., Tu, Z., Wang, X., Wu, S., Zeng, J., and Zhang, W · 2021
Later among the works it cites.
A targeted attack on black-box neural machine translation with parallel data poisoning
Xu, C., Wang, J., Tang, Y., Guzmán, F., Rubinstein, B. I., and Cohn, T · 2021
Later among the works it cites.
WeChat neural machine translation systems for WMT21
Zeng, X., Liu, Y., Li, E., Ran, Q., Meng, F., Li, P., Xu, J., and Zhou, J · 2021
Later among the works it cites.
The NiuTrans machine translation systems for WMT21
Zhou, S., Zhou, T., Wei, B., Luo, Y., Mu, Y., Zhou, Z., Wang, C., Zhou, X., Lv, C., Jing, Y., Wang, L., Zhang, J., Huang, C., Yan, Z., Hu, C., Li, B., Xiao, T., and Zhu, J · 2021
Later among the works it cites.
In-context examples selection for machine translation
Agrawal, S., Zhou, C., Lewis, M., Zettlemoyer, L., and Ghazvininejad, M · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
mt5: A massively multilingual pre-trained text-to-text transformer
Xue, L., Constant, N., Roberts, A., Kale, M., Al-Rfou, R., Siddhant, A., Barua, A., and Raffel, C · 2020
Cited alongside, same era.
Jax: Autograd and xla
Bradbury, J., Frostig, R., Hawkins, P., Johnson, M. J., Leary, C., Maclaurin, D., Necula, G., Paszke, A., VanderPlas, J., Wanderman-Milne, S., et al · 2021
Cited alongside, same era.
Experts, Errors, and Context: A Large-Scale Study of Human Evaluation for Machine Translation
Freitag, M., Foster, G., Grangier, D., Ratnakar, V., Tan, Q., and Macherey, W · 2021
Cited alongside, same era.
Towards universality in multilingual text rewriting
Garcia, X., Constant, N., Guo, M., and Firat, O · 2021
Cited alongside, same era.
Unsupervised neural machine translation with generative language models only
Han, J. M., Babuschkin, I., Edwards, H., Neelakantan, A., Xu, T., Polu, S., Ray, A., Shyam, P., Ramesh, A., Radford, A., et al · 2021
Cited alongside, same era.
Hernandez, D., Kaplan, J., Henighan, T., and McCandlish, S · 2021
Cited alongside, same era.
A few more examples may be worth billions of parameters
Kirstain, Y., Lewis, P., Riedel, S., and Levy, O · 2021
Cited alongside, same era.
Findings of the IWSLT 2022 evaluation campaign
Anastasopoulos, A., Barrault, L., Bentivogli, L., Zanon Boito, M., Bojar, O., Cattoni, R., Currey, A., Dinu, G., Duh, K., Elbayad, M., Emmanuel, C., Estève, Y., Federico, M., Federmann, C., Gahbiche, S., Gong, H., Grundkiewicz, R., Haddow, B., Hsu, B., Javorský, D., Kloudová, V., Lakew, S., Ma, X., Mathur, P., McNamee, P., Murray, K., Nǎdejde, M., Nakamura, S., Negri, M., Niehues, J., Niu, X., Ortega, J., Pino, J., Salesky, E., Shi, J., Sperber, M., Stüker, S., Sudoh, K., Turchi, M., Virkar, Y., Waibel, A., Wang, C., and Watanabe, S · 2022
Later among the works it cites.
Data scaling laws in nmt: The effect of noise and architecture
Bansal, Y., Ghorbani, B., Garg, A., Zhang, B., Cherry, C., Neyshabur, B., and Firat, O · 2022
Later among the works it cites.
Palm: Scaling language modeling with pathways
Chowdhery, A., Narang, S., Devlin, J., Bosma, M., Mishra, G., Roberts, A., Barham, P., Chung, H. W., Sutton, C., Gehrmann, S., et al · 2022
Later among the works it cites.
High quality rather than high model probability: Minimum bayes risk decoding with neural metrics
Freitag, M., Grangier, D., Tan, Q., and Liang, B · 2022
Later among the works it cites.
The flores-101 evaluation benchmark for low-resource and multilingual machine translation
Goyal, N., Gao, C., Chaudhary, V., Chen, P.-J., Wenzek, G., Ju, D., Krishnan, S., Ranzato, M., Guzman, F., and Fan, A · 2022
Later among the works it cites.
Training compute-optimal large language models
Hoffmann, J., Borgeaud, S., Mensch, A., Buchatskaya, E., Cai, T., Rutherford, E., Casas, D. d. L., Hendricks, L. A., Welbl, J., Clark, A., et al · 2022
Later among the works it cites.
Finding memo: Extractive memorization in constrained sequence generation tasks
Raunak, V. and Menezes, A · 2022
Later among the works it cites.
Frmt: A benchmark for few-shot region-aware machine translation
Riley, P., Dozat, T., Botha, J. A., Garcia, X., Garrette, D., Riesa, J., Firat, O., and Constant, N · 2022
Later among the works it cites.
Controlling translation formality using pre-trained multilingual language models
Rippeth, E., Agrawal, S., and Carpuat, M · 2022
Later among the works it cites.
Scaling up models and data with t5x and seqio
Roberts, A., Chung, H. W., Levskaya, A., Mishra, G., Bradbury, J., Andor, D., Narang, S., Lester, B., Gaffney, C., Mohiuddin, A., et al · 2022
Later among the works it cites.
Unifying language learning paradigms
Tay, Y., Dehghani, M., Tran, V. Q., Garcia, X., Bahri, D., Schuster, T., Zheng, H. S., Houlsby, N., and Metzler, D · 2022
Later among the works it cites.
Galactica: A large language model for science
Taylor, R., Kardas, M., Cucurull, G., Scialom, T., Hartshorn, A., Saravia, E., Poulton, A., Kerkez, V., and Stojnic, R · 2022
Later among the works it cites.
Prompting palm for translation: Assessing strategies and performance
Vilar, D., Freitag, M., Cherry, C., Luo, J., Ratnakar, V., and Foster, G · 2022
Later among the works it cites.