Fetching the paper…
Reading the bibliography…
There are many Language Models for the English language according to its worldwide relevance.
Introduction to the CoNLL-2002 shared task: Language-independent named entity recognition
E. F. Tjong Kim Sang · 2002
Earlier work this paper cites.
Rcv1: A new benchmark collection for text categorization research
D. D. Lewis, Y. Yang, T. Russell-Rose, and F. Li · 2004
Earlier work this paper cites.
Semeval-2014 task 10: Multilingual semantic textual similarity
E. Agirre, C. Banea, C. Cardie, D. Cer, M. Diab, A. Gonzalez-Agirre, W. Guo, R. Mihalcea, G. Rigau, and J. Wiebe · 2014
Earlier work this paper cites.
Semeval-2015 task 2: Semantic textual similarity, english, spanish and pilot on interpretability
E. Agirre, C. Banea, C. Cardie, D. Cer, M. Diab, A. Gonzalez-Agirre, W. Guo, I. Lopez-Gazpio, M. Maritxalar, R. Mihalcea, et al · 2015
Earlier work this paper cites.
Xnli: Evaluating cross-lingual sentence representations
A. Conneau, R. Rinott, G. Lample, A. Williams, S. R. Bowman, H. Schwenk, and V. Stoyanov · 2018
Earlier work this paper cites.
BERT: pre-training of deep bidirectional transformers for language understanding
J. Devlin, M. Chang, K. Lee, and K. Toutanova · 2018
Cited alongside, same era.
A corpus for multilingual document classification in eight languages
H. Schwenk and X. Li · 2018
Cited alongside, same era.
Roberta: A robustly optimized bert pretraining approach, 2019
Y. Liu, M. Ott, N. Goyal, J. Du, M. Joshi, D. Chen, O. Levy, M. Lewis, L. Zettlemoyer, and V. Stoyanov · 2019
Cited alongside, same era.
fairseq: A fast, extensible toolkit for sequence modeling
M. Ott, S. Edunov, A. Baevski, A. Fan, S. Gross, N. Ng, D. Grangier, and M. Auli · 2019
Cited alongside, same era.
PAWS-X: A Cross-lingual Adversarial Dataset for Paraphrase Identification
Y. Yang, Y. Zhang, C. Tar, and J. Baldridge · 2019
Later among the works it cites.
Legal-ES: A set of large scale resources for Spanish legal text processing
D. Samy, J. Arenas-García, and D. Pérez-Fernández · 2020
Later among the works it cites.
Transformers: State-of-the-art natural language processing
T. Wolf, L. Debut, V. Sanh, J. Chaumond, C. Delangue, A. Moi, P. Cistac, T. Rault, R. Louf, M. Funtowicz, J. Davison, S. Shleifer, P. von Platen, C. Ma, Y. Jernite, J. Plu, C. Xu, T. L. Scao, S. Gugger, M. Drame, Q. Lhoest, and A. M. Rush · 2020
Later among the works it cites.
Spanish language models, 2021
A. Gutiérrez-Fandiño, J. Armengol-Estapé, M. Pàmies, J. Llop-Palao, J. Silveira-Ocampo, C. P. Carrino, A. Gonzalez-Agirre, C. Armentano-Oller, C. Rodriguez-Penagos, and M. Villegas · 2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…