Fetching the paper…
Reading the bibliography…
Transfer learning with a unified Transformer framework (T5) that converts all language problems into a text-to-text format was recently proposed as a simple and effective transfer learning approach.
fairseq: A fast, extensible toolkit for sequence modeling
Myle Ott, Sergey Edunov, Alexei Baevski, Angela Fan, Sam Gross, Nathan Ng, David Grangier, and Michael Auli. 2019 · 1904
Earlier work this paper cites.
Neural arabic question answering
Hussein Mozannar, Karl El Hajal, Elie Maamary, and Hazem Hajj. 2019 · 1906
Earlier work this paper cites.
Multilingual universal sentence encoder for semantic retrieval
Yinfei Yang, Daniel Cer, Amin Ahmad, Mandy Guo, Jax Law, Noah Constant, Gustavo Hernandez Abrego, Steve Yuan, Chris Tar, Yun-Hsuan Sung, et al. 2019 · 1907
Earlier work this paper cites.
Question generation by transformers
Kettip Kriangchaivech and Artit Wangperawong. 2019 · 1909
Earlier work this paper cites.
Mlqa: Evaluating cross-lingual extractive question answering
Patrick Lewis, Barlas Oğuz, Ruty Rinott, Sebastian Riedel, and Holger Schwenk. 2019 · 1910
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu. 2019 · 1910
Earlier work this paper cites.
Multitask learning
Rich Caruana. 1997 · 1997
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Arabic Dialect Identification in the Wild
Ahmed Abdelali, Hamdy Mubarak, Younes Samih, Sabit Hassan, and Kareem Darwish. 2020 · 2005
Earlier work this paper cites.
Romanization, transcription and transliteration
Kenneth Beesley. 1998 · 2006
Earlier work this paper cites.
Using mechanical turk to create a corpus of arabic summaries
Mahmoud El-Haj, Udo Kruschwitz, and Chris Fox. 2010 · 2010
Earlier work this paper cites.
Accuracy and performance of google’s compact language detector
Michael McCandless. 2010 · 2010
Earlier work this paper cites.
Osac: Open source arabic corpora
Motaz K Saad and Wesam M Ashour. 2010 · 2010
Earlier work this paper cites.
mt5: A massively multilingual pre-trained text-to-text transformer
Linting Xue, Noah Constant, Adam Roberts, Mihir Kale, Rami Al-Rfou, Aditya Siddhant, Aditya Barua, and Colin Raffel. 2020 · 2010
Earlier work this paper cites.
Evaluation of topic identification methods on arabic corpora
Mourad Abbas, Kamel Smaïli, and Daoud Berkani. 2011 · 2011
Earlier work this paper cites.
Overview of the iwslt 2012 evaluation campaign
Marcello Federico, Mauro Cettolo, Luisa Bentivogli, Paul Michael, and Stüker Sebastian. 2012 · 2012
Earlier work this paper cites.
Parallel data, tools and interfaces in OPUS
Jörg Tiedemann. 2012 · 2012
Earlier work this paper cites.
Machine translation of arabic dialects
Rabih Zbib, Erika Malchiodi, Jacob Devlin, David Stallard, Spyros Matsoukas, Richard Schwartz, John Makhoul, Omar Zaidan, and Chris Callison-Burch. 2012 · 2012
Earlier work this paper cites.
Report on the 10th iwslt evaluation campaign
Mauro Cettolo, Jan Niehues, Sebastian Stüker, Luisa Bentivogli, and Marcello Federico. 2013 · 2013
Earlier work this paper cites.
The amara corpus: Building parallel language resources for the educational domain
Ahmed Abdelali, Francisco Guzman, Hassan Sajjad, and Stephan Vogel. 2014 · 2014
Earlier work this paper cites.
Report on the 11th iwslt evaluation campaign, iwslt 2014
Mauro Cettolo, Jan Niehues, Sebastian Stüker, Luisa Bentivogli, and Marcello Federico. 2014 · 2014
Earlier work this paper cites.
Development of a TV broadcasts speech recognition system for qatari Arabic
Mohamed Elmahdy, Mark Hasegawa-Johnson, and Eiman Mustafawi. 2014 · 2014
Earlier work this paper cites.
Collecting natural sms and chat conversations in multiple languages: The bolt phase 2 corpus
Zhiyi Song, Stephanie M Strassel, Haejoong Lee, Kevin Walker, Jonathan Wright, Jennifer Garland, Dana Fore, Brian Gainor, Preston Cabe, Thomas Thomas, et al. 2014 · 2014
Cited alongside, same era.
Arabic dialect identification
Omar F Zaidan and Chris Callison-Burch. 2014 · 2014
Cited alongside, same era.
The iwslt 2016 evaluation campaign
Mauro Cettolo, Niehues Jan, Stüker Sebastian, Luisa Bentivogli, Roldano Cattoni, and Marcello Federico. 2016 · 2016
Cited alongside, same era.
1.5 billion words arabic corpus
Ibrahim Abu El-Khair. 2016 · 2016
Cited alongside, same era.
Is neural machine translation ready for deployment? a case study on 30 translation directions
Marcin Junczys-Dowmunt, Tomasz Dwojak, and Hieu Hoang. 2016 · 2016
Cited alongside, same era.
Arabert: Transformer-based model for arabic language understanding
Wissam Antoun, Fady Baly, and Hazem Hajj. 2020 · 2020
Later among the works it cites.
On the cross-lingual transferability of monolingual representations
Mikel Artetxe, Sebastian Ruder, and Dani Yogatama. 2020 · 2020
Later among the works it cites.
Unsupervised cross-lingual representation learning at scale
Alexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary, Guillaume Wenzek, Francisco Guzmán, Edouard Grave, Myle Ott, Luke Zettlemoyer, and Veselin Stoyanov. 2020 · 2020
Later among the works it cites.
Wikilingua: A new benchmark dataset for multilingual abstractive summarization
Claire Cardie Faisal Ladhak, Esin Durmus and Kathleen McKeown. 2020 · 2020
Later among the works it cites.
From Arabic Sentiment Analysis to Sarcasm Detection: The ArSarcasm Dataset
Ibrahim Abu Farha and Walid Magdy. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Pranav Rajpurkar, Jian Zhang, Konstantin Lopyrev, and Percy Liang. 2016 · 2016
Cited alongside, same era.
Neural machine translation of rare words with subword units
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2016 · 2016
Cited alongside, same era.
Semeval-2017 task 1: Semantic textual similarity-multilingual and cross-lingual focused evaluation
Daniel Cer, Mona Diab, Eneko Agirre, Inigo Lopez-Gazpio, and Lucia Specia. 2017 · 2017
Cited alongside, same era.
Ant corpus: an arabic news text collection for textual classification
Amina Chouigui, Oussama Ben Khiroun, and Bilel Elayeb. 2017 · 2017
Cited alongside, same era.
Qcri machine translation systems for iwslt 16
Nadir Durrani, Fahim Dalvi, Hassan Sajjad, and Stephan Vogel. 2017 · 2017
Cited alongside, same era.
An overview of multi-task learning in deep neural networks
Sebastian Ruder. 2017 · 2017
Cited alongside, same era.
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
XTREME: A massively multilingual multi-task benchmark for evaluating cross-lingual generalisation
Junjie Hu, Sebastian Ruder, Aditya Siddhant, Graham Neubig, Orhan Firat, and Melvin Johnson. 2020 · 2020
Later among the works it cites.
Xglue: A new benchmark datasetfor cross-lingual pre-training, understanding and generation
Yaobo Liang, Nan Duan, Yeyun Gong, Ning Wu, Fenfei Guo, Weizhen Qi, Ming Gong, Linjun Shou, Daxin Jiang, Guihong Cao, et al. 2020 · 2020
Later among the works it cites.
Multilingual denoising pre-training for neural machine translation
Yinhan Liu, Jiatao Gu, Naman Goyal, Xian Li, Sergey Edunov, Marjan Ghazvininejad, Mike Lewis, and Luke Zettlemoyer. 2020 · 2020
Later among the works it cites.
Machine generation and detection of arabic manipulated and fake news
El Moatez Billah Nagoudi, AbdelRahim Elmadany, Muhammad Abdul-Mageed, Tariq Alhindi, and Hasan Cavusoglu. 2020 · 2020
Later among the works it cites.
Arabench: Benchmarking dialectal arabic-english machine translation
Hassan Sajjad, Ahmed Abdelali, Nadir Durrani, and Fahim Dalvi. 2020 · 2020
Later among the works it cites.
A unified model for arabizi detection and transliteration using sequence-to-sequence models
Ali Shazal, Aiza Usman, and Nizar Habash. 2020 · 2020
Later among the works it cites.
The united nations parallel corpus v1. 0
Michał Ziemski, Marcin Junczys-Dowmunt, and Bruno Pouliquen. 2016 · 2020
Later among the works it cites.
Pre-training bert on arabic tweets: Practical considerations
Ahmed Abdelali, Sabit Hassan, Hamdy Mubarak, Kareem Darwish, and Younes Samih. 2021 · 2021
Closest in time.
ARBERT & MARBERT: Deep Bidirectional Transformers for Arabic
Muhammad Abdul-Mageed, AbdelRahim Elmadany, and El Moatez Billah Nagoudi. 2021 · 2021
Closest in time.
Unsupervised neural networks for automatic arabic text summarization using document clustering and topic modeling
Nabil Alami, Mohammed Meknassi, Noureddine En-nahnahi, Yassine El Adlouni, and Ouafae Ammor. 2021 · 2021
Closest in time.
Benchmarking transformer-based language models for arabic sentiment and sarcasm detection
Ibrahim Abu Farha and Walid Magdy. 2021 · 2021
Closest in time.
The gem benchmark: Natural language generation, its evaluation and metrics
Sebastian Gehrmann, Tosin Adewumi, Karmanya Aggarwal, Pawan Sasanka Ammanamanchi, Aremu Anuoluwapo, Antoine Bosselut, Khyathi Raghavi Chandu, Miruna Clinciu, Dipanjan Das, Kaustubh D Dhole, et al. 2021 · 2021
Closest in time.
The interplay of variant, size, and task type in Arabic pre-trained language models
Go Inoue, Bashar Alhafni, Nurpeiis Baimukan, Houda Bouamor, and Nizar Habash. 2021 · 2021
Closest in time.
Exploring text-to-text transformers for english to hinglish machine translation with synthetic code-mixing
Ganesh Jawahar, El Moatez Billah Nagoudi, Muhammad Abdul-Mageed, and Laks VS Lakshmanan. 2021 · 2021
Closest in time.
Quality at a glance: An audit of web-crawled multilingual datasets
Julia Kreutzer, Isaac Caswell, Lisa Wang, Ahsan Wahab, Daan van Esch, Nasanbayar Ulzii-Orshikh, Allahsera Tapo, Nishant Subramani, Artem Sokolov, Claytone Sikasote, Monang Setyawan, Supheakmungkol Sarin, Sokhar Samb, Benoît Sagot, Clara Rivera, Annette Rios, Isabel Papadimitriou, Salomey Osei, Pedro Ortiz Suárez, Iroro Orife, Kelechi Ogueji, Andre Niyongabo Rubungo, Toan Q. Nguyen, Mathias Müller, André Müller, Shamsuddeen Hassan Muhammad, Nanda Muhammad, Ayanda Mnyakeni, Jamshidbek Mirzakhalov, Tapiwanashe Matangira, Colin Leong, Nze Lawson, Sneha Kudugunta, Yacine Jernite, Mathias Jenny, Orhan Firat, Bonaventure F. P. Dossou, Sakhile Dlamini, Nisansa de Silva, Sakine Çabuk Ballı, Stella Biderman, Alessia Battisti, Ahmed Baruwa, Ankur Bapna, Pallavi Baljekar, Israel Abebe Azime, Ayodele Awokoya, Duygu Ataman, Orevaoghene Ahia, Oghenefego Ahia, Sweta Agrawal, and Mofetoluwa Adeyemi. 2021 · 2021
Closest in time.
Investigating code-mixed Modern Standard Arabic-Egyptian to English machine translation
El Moatez Billah Nagoudi, AbdelRahim Elmadany, and Muhammad Abdul-Mageed. 2021 · 2021
Closest in time.