Fetching the paper…
Reading the bibliography…
In recent years, a series of Transformer-based models unlocked major improvements in general natural language understanding (NLU) tasks.
Cross-lingual language model pretraining
Guillaume Lample and Alexis Conneau. 2019 · 1901
Earlier work this paper cites.
Xlnet: Generalized autoregressive pretraining for language understanding
Zhilin Yang, Zihang Dai, Yiming Yang, Jaime Carbonell, Ruslan Salakhutdinov, and Quoc V Le. 2019 · 1906
Earlier work this paper cites.
Spanbert: Improving pre-training by representing and predicting spans
Mandar Joshi, Danqi Chen, Yinhan Liu, Daniel S. Weld, Luke Zettlemoyer, and Omer Levy. 2019 · 1907
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Huggingface’s transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, R’emi Louf, Morgan Funtowicz, and Jamie Brew. 2019 · 1910
Earlier work this paper cites.
CamemBERT: a Tasty French Language Model
Louis Martin, Benjamin Muller, Pedro Javier Ortiz Suárez, Yoann Dupont, Laurent Romary, Éric Villemonte de la Clergerie, Djamé Seddah, and Benoît Sagot. 2019 · 1911
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Narodowy korpus języka polskiego
Adam Przepiórkowski. 2012 · 2012
Earlier work this paper cites.
Open dataset for development of polish question answering systems
Michał Marcinczuk, Marcin Ptak, Adam Radziszewski, and Maciej Piasecki. 2013 · 2013
Earlier work this paper cites.
The sick (sentences involving compositional knowledge) dataset for relatedness and entailment
Marco Marelli, Stefano Menini, Marco Baroni, Luisa Bentivogli, Raffaella Bernardi, and Roberto Zamparelli. 2014 · 2014
Earlier work this paper cites.
The Polish Summaries Corpus
Maciej Ogrodniczuk and Mateusz Kopeć. 2014 · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba. 2014 · 2015
Earlier work this paper cites.
Enriching word vectors with subword information
Piotr Bojanowski, Edouard Grave, Armand Joulin, and Tomas Mikolov. 2016 · 2016
Earlier work this paper cites.
Opensubtitles2016: Extracting large parallel corpora from movie and tv subtitles
Pierre Lison and Jörg Tiedemann. 2016 · 2016
Earlier work this paper cites.
Squad: 100,000+ questions for machine comprehension of text
Pranav Rajpurkar, Jian Zhang, Konstantin Lopyrev, and Percy Liang. 2016 · 2016
Cited alongside, same era.
Neural machine translation of rare words with subword units
Barry Haddow Rico Sennrich and Alexandra Birch. 2016 · 2016
Cited alongside, same era.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Yonghui Wu, Mike Schuster, Zhifeng Chen, Quoc V Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, et al. 2016 · 2016
Cited alongside, same era.
Supervised learning of universal sentence representations from natural language inference data
Alexis Conneau, Douwe Kiela, Holger Schwenk, Loïc Barrault, and Antoine Bordes. 2017 · 2017
Cited alongside, same era.
Results of the poleval 2017 competition: Part-of-speech tagging shared task
Łukasz Kobyliński and Maciej Ogrodniczuk. 2017 · 2017
Cited alongside, same era.
Proceedings of the PolEval 2018 Workshop . Institute of Computer Science, Polish Academy of Sciences, Warsaw, Poland
Maciej Ogrodniczuk and Łukasz Kobyliński, editors. 2018 · 2018
Later among the works it cites.
Deep contextualized word representations
Matthew E. Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018 · 2018
Later among the works it cites.
A broad-coverage challenge corpus for sentence understanding through inference
Adina Williams, Nikita Nangia, and Samuel Bowman. 2018 · 2018
Later among the works it cites.
Tuning multilingual transformers for language-specific named entity recognition
Mikhail Arkhipov, Maria Trofimova, Yuri Kuratov, and Alexey Sorokin. 2019 · 2019
Later among the works it cites.
On the cross-lingual transferability of monolingual representations
Mikel Artetxe, Sebastian Ruder, and Dani Yogatama. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Ł ukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
Results of the poleval 2017 competition: Sentiment analysis shared task
Aleksander Wawer and Maciej Ogrodniczuk. 2017 · 2017
Cited alongside, same era.
Polish evaluation dataset for compositional distributional semantics models
Alina Wróblewska and Katarzyna Krasnowska-Kieraś. 2017 · 2017
Cited alongside, same era.
Senteval: An evaluation toolkit for universal sentence representations
Alexis Conneau and Douwe Kiela. 2018 · 2018
Cited alongside, same era.
Xnli: Evaluating cross-lingual sentence representations
Alexis Conneau, Ruty Rinott, Guillaume Lample, Adina Williams, Samuel R. Bowman, Holger Schwenk, and Veselin Stoyanov. 2018 · 2018
Cited alongside, same era.
Learning word vectors for 157 languages
Edouard Grave, Piotr Bojanowski, Prakhar Gupta, Armand Joulin, and Tomas Mikolov. 2018 · 2018
Cited alongside, same era.
Universal language model fine-tuning for text classification
Jeremy Howard and Sebastian Ruder. 2018 · 2018
Cited alongside, same era.
Evaluation of sentence representations in polish
Sławomir Dadas, Michał Perełkiewicz, and Rafał Poświata. 2019 · 2019
Later among the works it cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
ELMo embeddings for polish
Arkadiusz Janz. 2019 · 2019
Later among the works it cites.
Multi-level sentiment analysis of PolEmo 2.0: Extended corpus of multi-domain consumer reviews
Jan Kocoń, Piotr Miłkowski, and Monika Zaśko-Zielińska. 2019 · 2019
Later among the works it cites.
Empirical linguistic study of sentence embeddings
Katarzyna Krasnowska-Kieraś and Alina Wróblewska. 2019 · 2019
Later among the works it cites.
Proceedings of the PolEval 2019 Workshop . Institute of Computer Science, Polish Academy of Sciences, Warsaw, Poland
Maciej Ogrodniczuk and Łukasz Kobyliński, editors. 2019 · 2019
Later among the works it cites.
Asynchronous Pipeline for Processing Huge Corpora on Medium to Low Resource Infrastructures
Pedro Javier Ortiz Suárez, Benoît Sagot, and Laurent Romary. 2019 · 2019
Later among the works it cites.
Results of the poleval 2019 shared task 6: First dataset and open shared task for automatic cyberbullying detection in polish twitter
Michal Ptaszynski, Agata Pieciukiewicz, and Paweł Dybała. 2019 · 2019
Later among the works it cites.