Fetching the paper…
Reading the bibliography…
This paper describes the training process of the first Czech monolingual language representation models based on BERT and ALBERT architectures.
Simple bert models for relation extraction and semantic role labeling
Peng Shi and Jimmy Lin. 2019 · 1904
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach. arxiv 2019
Y Liu, M Ott, N Goyal, J Du, M Joshi, D Chen, O Levy, M Lewis, L Zettlemoyer, and V Stoyanov · 1907
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu. 2019 · 1910
Earlier work this paper cites.
Distilbert, a distilled version of bert: smaller, faster, cheaper and lighter
Victor Sanh, Lysandre Debut, Julien Chaumond, and Thomas Wolf. 2019 · 1910
Earlier work this paper cites.
Flaubert: Unsupervised language model pre-training for french
Hang Le, Loïc Vial, Jibril Frej, Vincent Segonne, Maximin Coavoux, Benjamin Lecouteux, Alexandre Allauzen, Benoît Crabbé, Laurent Besacier, and Didier Schwab. 2019 · 1912
Earlier work this paper cites.
Wietse de Vries, Andreas van Cranenburgh, Arianna Bisazza, Tommaso Caselli, Gertjan van Noord, and Malvina Nissim. 2019 · 1912
Earlier work this paper cites.
Named entities in czech: Annotating data and developing NE tagger
Magda Ševčíková, Zdeněk Žabokrtský, and Oldřich Krůza. 2007 · 2007
Earlier work this paper cites.
Czech academic corpus 2.0
Barbora Hladká Vildová, Jan Hajič, Jiří Hana, Jaroslava Hlaváčová, Jiří Mírovský, and Jan Raab. 2008 · 2008
Earlier work this paper cites.
Multilingual dependency learning: A huge feature engineering method to semantic dependency parsing
Hai Zhao, Wenliang Chen, Chunyu Kit, and Guodong Zhou. 2009 · 2009
Earlier work this paper cites.
Sentiment analysis and opinion mining
Bing Liu. 2012 · 2012
Earlier work this paper cites.
Prague dependency treebank 3.0
Eduard Bejček, Eva Hajičová, Jan Hajič, Pavlína Jínová, Václava Kettnerová, Veronika Kolářová, Marie Mikulová, Jiří Mírovský, Anna Nedoluzhko, Jarmila Panevová, Lucie Poláková, Magda Ševčíková, Jan Štěpánek, and Šárka Zikánová. 2013 · 2013
Earlier work this paper cites.
Sentiment analysis in Czech social media using supervised machine learning
Ivan Habernal, Tomáš Ptáček, and Josef Steinberger. 2013 · 2013
Earlier work this paper cites.
Evaluation of the document classification approaches
M. Hrala and P. Král. 2013 · 2013
Earlier work this paper cites.
Crf-based czech named entity recognizer and consolidation of czech ner research
Michal Konkol and Miloslav Konopík. 2013 · 2013
Earlier work this paper cites.
Area under the ROC Curve , pages 38–39. Springer New York, New York, NY
Francisco Melo. 2013 · 2013
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg S Corrado, and Jeff Dean. 2013 · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Cited alongside, same era.
GloVe: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Cited alongside, same era.
TensorFlow: Large-scale machine learning on heterogeneous systems
Martín Abadi, Ashish Agarwal, Paul Barham, Eugene Brevdo, Zhifeng Chen, Craig Citro, Greg S. Corrado, Andy Davis, Jeffrey Dean, Matthieu Devin, Sanjay Ghemawat, Ian Goodfellow, Andrew Harp, Geoffrey Irving, Michael Isard, Yangqing Jia, Rafal Jozefowicz, Lukasz Kaiser, Manjunath Kudlur, Josh Levenberg, Dandelion Mané, Rajat Monga, Sherry Moore, Derek Murray, Chris Olah, Mike Schuster, Jonathon Shlens, Benoit Steiner, Ilya Sutskever, Kunal Talwar, Paul Tucker, Vincent Vanhoucke, Vijay Vasudevan, Fernanda Viégas, Oriol Vinyals, Pete Warden, Martin Wattenberg, Martin Wicke, Yuan Yu, and Xiaoqiang Zheng. 2015 · 2015
Cited alongside, same era.
SYN v4: large corpus of written czech
Michal Křen, Václav Cvrček, Tomáš Čapka, Anna Čermáková, Milena Hnátková, Lucie Chlumská, Tomáš Jelínek, Dominika Kováříková, Vladimír Petkevič, Pavel Procházka, Hana Skoumalová, Michal Škrabal, Petr Truneček, Pavel Vondřička, and Adrian Zasina. 2016 · 2016
Cited alongside, same era.
Tuning multilingual transformers for language-specific named entity recognition
Mikhail Arkhipov, Maria Trofimova, Yuri Kuratov, and Alexey Sorokin. 2019 · 2019
Later among the works it cites.
Cross-lingual language model pretraining
Alexis Conneau and Guillaume Lample. 2019 · 2019
Later among the works it cites.
The second cross-lingual challenge on recognition, normalization, classification, and linking of named entities across Slavic languages
Jakub Piskorski, Laska Laskova, Michał Marcińczuk, Lidia Pivovarova, Pavel Přibáň, Josef Steinberger, and Roman Yangarber. 2019 · 2019
Later among the works it cites.
Curriculum learning in sentiment analysis
Jakub Sido and Miloslav Konopík. 2019 · 2019
Later among the works it cites.
Czech text processing with contextual embeddings: Pos tagging, lemmatization, parsing and ner
Milan Straka, Jana Straková, and Jan Hajič. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yonghui Wu, Mike Schuster, Zhifeng Chen, Quoc V Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, et al. 2016 · 2016
Cited alongside, same era.
Enriching word vectors with subword information
Piotr Bojanowski, Edouard Grave, Armand Joulin, and Tomas Mikolov. 2017 · 2017
Cited alongside, same era.
Learned in translation: Contextualized word vectors
Bryan McCann, James Bradbury, Caiming Xiong, and Richard Socher. 2017 · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Ł ukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
BERT: pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Cited alongside, same era.
Universal language model fine-tuning for text classification
Jeremy Howard and Sebastian Ruder. 2018 · 2018
Cited alongside, same era.
Lda in character-lstm-crf named entity recognition
Miloslav Konopík and Ondřej Pražák. 2018 · 2018
Cited alongside, same era.
Czech legal text treebank 2.0
Vincent Kříž, Barbora Hladká, and Zdeňka Urešová. 2018 · 2018
Cited alongside, same era.
Antti Virtanen, Jenna Kanerva, Rami Ilo, Jouni Luoma, Juhani Luotolahti, Tapio Salakoski, Filip Ginter, and Sampo Pyysalo. 2019 · 2019
Later among the works it cites.
Xlnet: Generalized autoregressive pretraining for language understanding
Zhilin Yang, Zihang Dai, Yiming Yang, Jaime Carbonell, Russ R Salakhutdinov, and Quoc V Le. 2019 · 2019
Later among the works it cites.
Unsupervised cross-lingual representation learning at scale
Alexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary, Guillaume Wenzek, Francisco Guzmán, Edouard Grave, Myle Ott, Luke Zettlemoyer, and Veselin Stoyanov. 2020 · 2020
Later among the works it cites.
The birth of Romanian BERT
Stefan Dumitrescu, Andrei-Marius Avram, and Sampo Pyysalo. 2020 · 2020
Later among the works it cites.
Polbert: Attacking polish nlp tasks with transformers
Dariusz Kłeczek. 2020 · 2020
Later among the works it cites.
ALBERT: A lite BERT for self-supervised learning of language representations
Zhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel, Piyush Sharma, and Radu Soricut. 2020 · 2020
Later among the works it cites.
CamemBERT: a tasty French language model
Louis Martin, Benjamin Muller, Pedro Javier Ortiz Suárez, Yoann Dupont, Laurent Romary, Éric de la Clergerie, Djamé Seddah, and Benoît Sagot. 2020 · 2020
Later among the works it cites.
Kuisail at semeval-2020 task 12: Bert-cnn for offensive speech identification in social media
Ali Safaya, Moutasem Abdullatif, and Deniz Yuret. 2020 · 2020
Later among the works it cites.
Berturk - bert models for turkish
Stefan Schweter. 2020 · 2020
Later among the works it cites.
Czech news dataset for semanic textual similarity
Jakub Sido, Michal Seják, Ondřej Pražák, Miloslav Konopík, and Václav Moravec. 2021 · 2021
Closest in time.