Fetching the paper…
Reading the bibliography…
Large pretrained masked language models have become state-of-the-art solutions for many NLP problems.
Ginter, F., Hajič, J., Luotolahti, J., Straka, M., Zeman, D.: CoNLL 2017 shared task - automatically annotated raw texts and word embeddings (2017), \url
1989
Earlier work this paper cites.
Nivre, J.: Algorithms for Deterministic Incremental Dependency Parsing. Computational Linguistics 34
2008
Earlier work this paper cites.
Paikens, P., Auziņa, I., Garkaje, G., Paegle, M.: Towards named entity annotation of Latvian national library corpus. Frontiers in Artificial Intelligence and Applications 247
2012
Earlier work this paper cites.
Pinnis, M.: Latvian and Lithuanian named entity recognition with TildeNER. In: Proceedings of the 8th international conference on Language Resources and Evaluation LREC 2012. pp. 1258–1265 (01 2012)
2012
Earlier work this paper cites.
Steinberger, R., Eisele, A., Klocek, S., Pilos, S., Schlüter, P.: DGT-TM: A freely available translation memory in 22 languages. In: Proceedings of the 8th international conference on Language Resources and Evaluation LREC (2012)
2012
Earlier work this paper cites.
Jakubíček, M., Kilgarriff, A., Kovář, V., Rychlỳ, P., Suchomel, V.: The TenTen corpus family. In: 7th International Corpus Linguistics Conference CL. pp. 125–127 (2013)
2013
Earlier work this paper cites.
Laur, S.: Nimeüksuste korpus. Center of Estonian Language Resources (2013)
2013
Earlier work this paper cites.
Muischnek, K., Müürisep, K., Puolakainen, T.: Estonian Dependency Treebank: From Constraint Grammar tagset to Universal Dependencies. In: Proceedings of LREC (2016)
2016
Earlier work this paper cites.
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A.N., Kaiser, Ł., Polosukhin, I.: Attention is all you need. In: Advances in neural information processing systems. pp. 5998–6008 (2017)
2017
Earlier work this paper cites.
Darģis, R., Auziņa, I., Bojārs, U., Paikens, P., Znotiņš, A.: Annotation of the corpus of the Saeima with multilingual standards. In: Proceedings of the Eleventh International Conference on Language Resources and Evaluation (LREC) (2018)
2018
Earlier work this paper cites.
Nivre, J., Abrams, M., Agić, Ž.: Universal Dependencies 2.3 (2018), \url
2018
Cited alongside, same era.
Rosa, R.: Plaintext Wikipedia dump 2018 (2018), \url
2018
Cited alongside, same era.
2019
Cited alongside, same era.
Devlin, J., Chang, M.W., Lee, K., Toutanova, K.: BERT: Pre-training of deep bidirectional transformers for language understanding. In: Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers). pp. 4171–4186 (2019). https://doi.org/10.18653/v1/N19-1423
2019
Cited alongside, same era.
Malmsten, M., Börjeson, L., Haffenden, C.: Playing with Words at the National Library of Sweden – Making a Swedish BERT. ArXiv preprint 2007.01658 (2020)
2020
Later among the works it cites.
Raffel, C., Shazeer, N., Roberts, A., Lee, K., Narang, S., Matena, M., Zhou, Y., Li, W., Liu, P.J.: Exploring the limits of transfer learning with a unified text-to-text transformer. Journal of Machine Learning Research 21
2020
Later among the works it cites.
Tanvir, H., Kittask, C., Sirts, K.: EstBERT: A pretrained language-specific BERT for Estonian. arXiv preprint 2011.04784 (2020)
2020
Later among the works it cites.
Ulčar, M., Vaik, K., Lindström, J., Dailidėnaitė, M., Robnik-Šikonja, M.: Multilingual culture-independent word analogy datasets. In: Proceedings of the 12th Language Resources and Evaluation Conference. pp. 4067–4073 (2020)
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
Liu, Y., Ott, M., Goyal, N., Du, J., Joshi, M., Chen, D., Levy, O., Lewis, M., Zettlemoyer, L., Stoyanov, V.: RoBERTa: A robustly optimized BERT pretraining approach. ArXiv preprint 1907.11692 (2019)
2019
Cited alongside, same era.
Ott, M., Edunov, S., Baevski, A., Fan, A., Gross, S., Ng, N., Grangier, D., Auli, M.: Fairseq: A fast, extensible toolkit for sequence modeling. In: Proceedings of NAACL-HLT 2019: Demonstrations (2019)
2019
Cited alongside, same era.
Strzyz, M., Vilares, D., Gómez-Rodríguez, C.: Viable dependency parsing as sequence labeling. In: Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers). pp. 717–723 (2019). https://doi.org/10.18653/v1/N19-1077
2019
Cited alongside, same era.
2019
Cited alongside, same era.
Brown, T., et al.: Language models are few-shot learners. In: Advances in Neural Information Processing Systems. vol. 33, pp. 1877–1901 (2020)
2020
Cited alongside, same era.
Ulčar, M., Robnik-Šikonja, M.: FinEst BERT and CroSloEngual BERT: less is more in multilingual models. In: Proceedings of Text, Speech, and Dialogue, TSD 2020. pp. 104–111 (2020)
2020
Later among the works it cites.
Znotiņš, A., Barzdiņš, G.: LVBERT: Transformer-based model for Latvian language understanding. In: Human Language Technologies–The Baltic Perspective: Proceedings of the Ninth International Conference Baltic HLT 2020. vol. 328, p. 111 (2020)
2020
Later among the works it cites.
Bommasani, R., Hudson, D.A., Adeli, E., Altman, R., Arora, S., von Arx, S., Bernstein, M.S., Bohg, J., Bosselut, A., Brunskill, E., et al.: On the opportunities and risks of foundation models. ArXiv preprint 2108.07258 (2021)
2021
Closest in time.
Marcus, G., Davis, E.: Has AI found a new foundation? The Gradient (2021), 11 September 2021
2021
Closest in time.
Straka, M., Náplava, J., Straková, J., Samuel, D.: RobeCzech: Czech RoBERTa, a monolingual contextualized language representation model (2021)
2021
Closest in time.