Fetching the paper…
Reading the bibliography…
Pretraining Bidirectional Encoder Representations from Transformers (BERT) for downstream NLP tasks is a non-trival task.
Camembert: a tasty french language model
Louis Martin, B. Muller, Pedro Javier Ortiz Suárez, Y. Dupont, L. Romary, ’Eric de la Clergerie, Djamé Seddah, and Benoît Sagot. 2020 · 1911
Earlier work this paper cites.
The emotions
Robert Plutchik. 1991 · 1991
Earlier work this paper cites.
Arabic dialect identification in the wild
Ahmed Abdelali, Hamdy Mubarak, Younes Samih, Sabit Hassan, and Kareem Darwish. 2020 · 2005
Earlier work this paper cites.
Arabic named entity recognition using optimized feature sets
Yassine Benajiba, Mona Diab, and Paolo Rosso. 2008 · 2008
Earlier work this paper cites.
Anchibert: A pre-trained model for ancient chineselanguage understanding and generation
Huishuang Tian, Kexin Yang, Dayiheng Liu, and Jiancheng Lv. 2020 · 2009
Earlier work this paper cites.
Named entity recognition using cross-lingual resources: Arabic as an example
Kareem Darwish. 2013 · 2013
Earlier work this paper cites.
Simple effective microblog named entity recognition: Arabic as an example
Kareem Darwish and Wei Gao. 2014 · 2014
Earlier work this paper cites.
Farasa: A fast and furious segmenter for Arabic
Ahmed Abdelali, Kareem Darwish, Nadir Durrani, and Hamdy Mubarak. 2016a · 2016
Earlier work this paper cites.
Farasa: A fast and furious segmenter for arabic
Ahmed Abdelali, Kareem Darwish, Nadir Durrani, and Hamdy Mubarak. 2016b · 2016
Earlier work this paper cites.
1.5 billion words arabic corpus
Ibrahim Abu El-Khair. 2016 · 2016
Earlier work this paper cites.
Opensubtitles2016: Extracting large parallel corpora from movie and tv subtitles
Pierre Lison and Jörg Tiedemann. 2016 · 2016
Cited alongside, same era.
Neural machine translation of rare words with subword units
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2016 · 2016
Cited alongside, same era.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Yonghui Wu, Mike Schuster, Zhifeng Chen, Quoc V. Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, Jeff Klingner, Apurva Shah, Melvin Johnson, Xiaobing Liu, Łukasz Kaiser, Stephan Gouws, Yoshikiyo Kato, Taku Kudo, Hideto Kazawa, Keith Stevens, George Kurian, Nishant Patil, Wei Wang, Cliff Young, Jason Smith, Jason Riesa, Alex Rudnick, Oriol Vinyals, Greg Corrado, Macduff Hughes, and Jeffrey Dean. 2016 · 2016
Cited alongside, same era.
Arabic tweets sentimental analysis using machine learning
Khaled Mohammad Alomari, Hatem M. ElSherif, and Khaled Shaalan. 2017 · 2017
Cited alongside, same era.
SemEval-2018 task 1: Affect in tweets
Saif Mohammad, Felipe Bravo-Marquez, Mohammad Salameh, and Svetlana Kiritchenko. 2018 · 2018
Arabert: Transformer-based model for arabic language understanding
Wissam Antoun, Fady Baly, and Hazem Hajj. 2020 · 2020
Later among the works it cites.
Spanish pre-trained bert model and evaluation data
José Canete, Gabriel Chaperon, and Rodrigo Fuentes. 2019 · 2020
Later among the works it cites.
Robbert: a dutch roberta-based language model
Pieter Delobelle, Thomas Winters, and Bettina Berendt. 2020 · 2020
Later among the works it cites.
Multi-task learning using AraBert for offensive language detection
Marc Djandji, Fady Baly, Wissam Antoun, and Hazem Hajj. 2020 · 2020
Later among the works it cites.
Arabic offensive language on twitter: Analysis and experiments
Hamdy Mubarak, Ammar Rashed, Kareem Darwish, Younes Samih, and Ahmed Abdelali. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Multi-channel embedding convolutional neural network model for arabic sentiment classification
Abdelghani Dahou, Shengwu Xiong, Junwei Zhou, and Mohamed Abd Elaziz. 2019 · 2019
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Cited alongside, same era.
Multilingual is not enough: Bert for finnish
Antti Virtanen, Jenna Kanerva, Rami Ilo, Jouni Luoma, Juhani Luotolahti, Tapio Salakoski, Filip Ginter, and Sampo Pyysalo. 2019 · 2019
Cited alongside, same era.
Wietse de Vries, Andreas van Cranenburgh, Arianna Bisazza, Tommaso Caselli, Gertjan van Noord, and Malvina Nissim. 2019 · 2019
Cited alongside, same era.
A monolingual approach to contextualized word embeddings for mid-resource languages
Pedro Javier Ortiz Suárez, Laurent Romary, and Benoît Sagot. 2020 · 2020
Later among the works it cites.
Kuisail at semeval-2020 task 12: Bert-cnn for offensive speech identification in social media
Ali Safaya, Moutasem Abdullatif, and Deniz Yuret. 2020 · 2020
Later among the works it cites.
Berturk - bert models for turkish
Stefan Schweter. 2020 · 2020
Later among the works it cites.
Multi-dialect arabic bert for country-level dialect identification
Bashar Talafha, Mohammad Ali, Muhy Eddin Za’ter, Haitham Seelawi, Ibraheem Tuffaha, Mostafa Samir, Wael Farhan, and Hussein T. Al-Natsheh. 2020 · 2020
Later among the works it cites.
Marcos Zampieri, Preslav Nakov, Sara Rosenthal, Pepa Atanasova, Georgi Karadzhov, Hamdy Mubarak, Leon Derczynski, Zeses Pitenis, and Çağrı Çöltekin. 2020 · 2020
Later among the works it cites.