Language models are few-shot learners
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J. D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al. (2020) · 1901
Earlier work this paper cites.
75 languages, 1 model: Parsing universal dependencies universally
Original
Kondratyuk, D. and Straka, M. (2019) · 1904
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Original
Liu, Y., Ott, M., Goyal, N., Du, J., Joshi, M., Chen, D., Levy, O., Lewis, M., Zettlemoyer, L., and Stoyanov, V. (2019) · 1907
Earlier work this paper cites.
Bart: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension
Original
Lewis, M., Liu, Y., Goyal, N., Ghazvininejad, M., Mohamed, A., Levy, O., Stoyanov, V., and Zettlemoyer, L. (2019) · 1910
Earlier work this paper cites.
Prague arabic dependency treebank: A word on the million words
Smrz, O., Bielickỳ, V., Kourilová, I., Krácmar, J., Hajic, J., and Zemánek, P. (2008) · 2008
Earlier work this paper cites.
Arabic diacritic restoration approach based on maximum entropy models
Zitouni, I. and Sarikaya, R. (2009) · 2009
Earlier work this paper cites.
Using mechanical turk to create a corpus of arabic summaries
El-Haj, M., Kruschwitz, U., and Fox, C. (2010) · 2010
Earlier work this paper cites.
Wikilingua: A new benchmark dataset for cross-lingual abstractive summarization
Original
Ladhak, F., Durmus, E., Cardie, C., and McKeown, K. (2020) · 2010
Earlier work this paper cites.
Transliteration of Arabizi into Arabic orthography: Developing a parallel annotated Arabizi-Arabic script SMS/chat corpus
Bies, A., Song, Z., Maamouri, M., Grimes, S., Lee, H., Wright, J., Strassel, S., Habash, N., Eskander, R., and Rambow, O. (2014) · 2014
Earlier work this paper cites.
The united nations parallel corpus v1. 0
Ziemski, M., Junczys-Dowmunt, M., and Pouliquen, B. (2016) · 2016
Earlier work this paper cites.
Arabic tweets sentimental analysis using machine learning
Alomari, K. M., ElSherif, H. M., and Shaalan, K. (2017) · 2017
Earlier work this paper cites.
Arabic diacritization: Stats, rules, and hacks
Darwish, K., Mubarak, H., and Abdelali, A. (2017) · 2017
Earlier work this paper cites.
Deep contextualized word representations
Peters, M. E., Neumann, M., Iyyer, M., Gardner, M., Clark, C., Lee, K., and Zettlemoyer, L. (2018) · 2018
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Original
Devlin, J., Chang, M.-W., Lee, K., and Toutanova, K. (2019) · 2019
Earlier work this paper cites.
Highly effective Arabic diacritization using sequence to sequence modeling
Mubarak, H., Abdelali, A., Sajjad, H., Samih, Y., and Darwish, K. (2019) · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Radford, A., Wu, J., Child, R., Luan, D., Amodei, D., Sutskever, I., et al. (2019) · 2019
Earlier work this paper cites.