Fetching the paper…
Reading the bibliography…
Pretrained multilingual models have become a de facto default approach for zero-shot cross-lingual transfer.
Unsupervised cross-lingual representation learning at scale
Alexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary, Guillaume Wenzek, Francisco Guzmán, Edouard Grave, Myle Ott, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1911
Earlier work this paper cites.
How language-neutral is multilingual bert?
Jindřich Libovickỳ, Rudolf Rosa, and Alexander Fraser. 2019 · 1911
Earlier work this paper cites.
Inducing language-agnostic multilingual representations
Wei Zhao, Steffen Eger, Johannes Bjerva, and Isabelle Augenstein. 2020 · 2008
Earlier work this paper cites.
Complex Linguistic Annotation - No Easy Way Out! A Case from Bengali and Hindi POS Labeling Tasks
Priyanka Biswas, Monojit Choudhury, and Kalika Bali. 2009 · 2009
Earlier work this paper cites.
Crowdsourcing research opportunities: Lessons from natural language processing
Marta Sabou, Kalina Bontcheva, and Arno Scharl. 2012 · 2012
Earlier work this paper cites.
Squad: 100, 000+ questions for machine comprehension of text
Pranav Rajpurkar, Jian Zhang, Konstantin Lopyrev, and Percy Liang. 2016 · 2016
Earlier work this paper cites.
XNLI: Evaluating Cross-lingual Sentence Representations
Alexis Conneau, Guillaume Lample, Ruty Rinott, Adina Williams, Samuel R. Bowman, Holger Schwenk, and Veselin Stoyanov. 2018a · 2018
Earlier work this paper cites.
XNLI: Evaluating cross-lingual sentence representations
Alexis Conneau, Ruty Rinott, Guillaume Lample, Adina Williams, Samuel R. Bowman, Holger Schwenk, and Veselin Stoyanov. 2018b · 2018
Earlier work this paper cites.
Annotation artifacts in natural language inference data
Suchin Gururangan, Swabha Swayamdipta, Omer Levy, Roy Schwartz, Samuel Bowman, and Noah A. Smith. 2018 · 2018
Earlier work this paper cites.
A broad-coverage challenge corpus for sentence understanding through inference
Adina Williams, Nikita Nangia, and Samuel Bowman. 2018 · 2018
Earlier work this paper cites.
Cross-lingual language model pretraining
Alexis Conneau and Guillaume Lample. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter. 2019 · 2019
Cited alongside, same era.
How multilingual is Multilingual BERT?
Telmo Pires, Eva Schlinger, and Dan Garrette. 2019 · 2019
Cited alongside, same era.
Faquad: Reading comprehension dataset in the domain of brazilian higher education
H. F. Sayama, A. V. Araujo, and E. R. Fernandes. 2019 · 2019
Cited alongside, same era.
XTREME: A Massively Multilingual Multi-task Benchmark for Evaluating Cross-lingual Generalisation
Junjie Hu, Sebastian Ruder, Aditya Siddhant, Graham Neubig, Orhan Firat, and Melvin Johnson. 2020 · 2020
Later among the works it cites.
PhoBERT: Pre-trained language models for Vietnamese
Dat Quoc Nguyen and Anh Tuan Nguyen. 2020 · 2020
Later among the works it cites.
A Vietnamese dataset for evaluating machine reading comprehension
Kiet Nguyen, Vu Nguyen, Anh Nguyen, and Ngan Nguyen. 2020 · 2020
Later among the works it cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2020 · 2020
Later among the works it cites.
The assin 2 shared task: A quick overview
Livy Real, Erick Fonseca, and Hugo Gonçalo Oliveira. 2020 · 2020
Later among the works it cites.
How Good is Your Tokenizer? On the Monolingual Performance of Multilingual Language Models
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
On the cross-lingual transferability of monolingual representations
Mikel Artetxe, Sebastian Ruder, and Dani Yogatama. 2020 · 2020
Cited alongside, same era.
German’s next language model
Branden Chan, Stefan Schweter, and Timo Möller. 2020 · 2020
Cited alongside, same era.
Can monolingual pretrained models help cross-lingual classification?
Zewen Chi, Li Dong, Furu Wei, Xianling Mao, and Heyan Huang. 2020 · 2020
Cited alongside, same era.
Emerging cross-lingual structure in pretrained language models
Alexis Conneau, Shijie Wu, Haoran Li, Luke Zettlemoyer, and Veselin Stoyanov. 2020 · 2020
Cited alongside, same era.
Phillip Rust, Jonas Pfeiffer, Ivan Vulić, Sebastian Ruder, and Iryna Gurevych. 2020 · 2020
Later among the works it cites.
Pretrained transformers as universal computation engines
Kevin Lu, Aditya Grover, Pieter Abbeel, and Igor Mordatch. 2021 · 2021
Closest in time.
mt5: A massively multilingual pre-trained text-to-text transformer
Linting Xue, Noah Constant, Adam Roberts, Mihir Kale, Rami Al-Rfou, Aditya Siddhant, Aditya Barua, and Colin Raffel. 2021 · 2021
Closest in time.
Language-agnostic representation learning of source code from structure and context
Daniel Zügner, Tobias Kirschstein, Michele Catasta, Jure Leskovec, and Stephan Günnemann. 2021 · 2021
Closest in time.