Fetching the paper…
Reading the bibliography…
Since the appearance of BERT, recent works including XLNet and RoBERTa utilize sentence embedding models pre-trained by large corpora and a large number of parameters.
Nltk: the natural language toolkit
Edward Loper and Steven Bird · 2002
Earlier work this paper cites.
The quality of content in open online collaboration platforms: Approaches to NLP-supported information quality management in Wikipedia
Oliver Ferschke · 2014
Earlier work this paper cites.
Neural machine translation of rare words with subword units
Rico Sennrich, Barry Haddow, and Alexandra Birch · 2016
Earlier work this paper cites.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Yonghui Wu, Mike Schuster, Zhifeng Chen, Quoc V Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, et al · 2016
Earlier work this paper cites.
Argumentation mining in user-generated web discourse
Ivan Habernal and Iryna Gurevych · 2017
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Earlier work this paper cites.
Subword regularization: Improving neural network translation models with multiple subword candidates
Taku Kudo · 2018
Cited alongside, same era.
Grapheme-level awareness in word embeddings for morphologically rich languages
Suzi Park and Hyopil Shin · 2018
Cited alongside, same era.
Looking into bert (bert thophapoki)
Sang-kil Park · 2018
Cited alongside, same era.
Pre-training with whole word masking for chinese bert
Yiming Cui, Wanxiang Che, Ting Liu, Bing Qin, Ziqing Yang, Shijin Wang, and Guoping Hu · 2019
Cited alongside, same era.
Bert pretrained model trained on japanese wikipedia articles
Yohei Kikuta · 2019
Cited alongside, same era.
Advanced subword segmentation and interdependent regularization mechanisms for korean language understanding
Mansu Kim, Yunkon Kim, Yeonsoo Lim, and Eui-Nam Huh · 2019
Later among the works it cites.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov · 2019
Later among the works it cites.
Camembert: a tasty french language model
Louis Martin, Benjamin Muller, Pedro Javier Ortiz Suárez, Yoann Dupont, Laurent Romary, Éric Villemonte de la Clergerie, Djamé Seddah, and Benoît Sagot · 2019
Later among the works it cites.
Alberto: Italian bert language understanding model for nlp challenging tasks based on tweets
Marco Polignano, Pierpaolo Basile, Marco de Gemmis, Giovanni Semeraro, and Valerio Basile · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Antti Virtanen, Jenna Kanerva, Rami Ilo, Jouni Luoma, Juhani Luotolahti, Tapio Salakoski, Filip Ginter, and Sampo Pyysalo · 2019
Later among the works it cites.