Fetching the paper…
Reading the bibliography…
An effective method for cross-lingual transfer is to fine-tune a bilingual or multilingual model on a supervised dataset in one language and evaluating it on another language in a zero-shot manner.
PTT5: Pretraining and validating the T5 model on Brazilian Portuguese data
Diedre Carmo, Marcos Piau, Israel Campiotti, Rodrigo Nogueira, and Roberto Lotufo. 2020 · 2008
Earlier work this paper cites.
Complex Linguistic Annotation – No Easy Way Out! A Case from Bangla and Hindi POS Labeling Tasks
Sandipan Dandapat, Priyanka Biswas, Monojit Choudhury, and Kalika Bali. 2009 · 2009
Earlier work this paper cites.
mT5: A massively multilingual pre-trained text-to-text transformer
Linting Xue, Noah Constant, Adam Roberts, Mihir Kale, Rami Al-Rfou, Aditya Siddhant, Aditya Barua, and Colin Raffel. 2021 · 2010
Earlier work this paper cites.
Crowdsourcing research opportunities: lessons from natural language processing
Marta Sabou, Kalina Bontcheva, and Arno Scharl. 2012 · 2012
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Earlier work this paper cites.
SQuAD: 100,000+ Questions for Machine Comprehension of Text
Pranav Rajpurkar, Jian Zhang, and Percy Liang Konstantin Lopyrev. 2016 · 2016
Earlier work this paper cites.
Google’s Neural Machine Translation System: Bridging the Gap between Human and Machine Translation
Yonghui Wu, Mike Schuster, Zhifeng Chen, Quoc V Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, et al · 2016
Earlier work this paper cites.
Attention is all you need. In Advances in neural information processing systems . 5998–6008
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Anserini: Enabling the Use of Lucene for Information Retrieval Research. In Proceedings of the 40th International ACM SIGIR Conference on Research and Development in Information Retrieval (Shinjuku, Tokyo, Japan) (SIGIR ’17) . Association for Computing Machinery, New York, NY, USA, 1253–1256
Peilin Yang, Hui Fang, and Jimmy Lin. 2017 · 2017
Earlier work this paper cites.
MS MARCO: A Human Generated MAchine Reading Comprehension Dataset
Payal Bajaj, Daniel Campos, Nick Craswell, Li Deng, Jianfeng Gao, Xiaodong Liu, Rangan Majumder, Andrew McNamara, Bhaskar Mitra, Tri Nguyen, Mir Rosenber, Xia Song, Alina Stoica, Saurabh Tiwary, and Tong Wang. 2018 · 2018
Earlier work this paper cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018 · 2018
Earlier work this paper cites.
The brWaC Corpus: A New Open Resource for Brazilian Portuguese
Jorge A. Wagner Filho, Rodrigo Wilkens, Marco Idiart, and Aline Villavicencio. 2018 · 2018
Earlier work this paper cites.
Marian: Fast Neural Machine Translation in C++. In Proceedings of ACL 2018, System Demonstrations . Association for Computational Linguistics, Melbourne, Australia, 116–121
Marcin Junczys-Dowmunt, Roman Grundkiewicz, Tomasz Dwojak, Hieu Hoang, Kenneth Heafield, Tom Neckermann, Frank Seide, Ulrich Germann, Alham Fikri Aji, Nikolay Bogoychev, André F. T. Martins, and Alexandra Birch. 2018 · 2018
Earlier work this paper cites.
Adafactor: Adaptive learning rates with sublinear memory cost. In International Conference on Machine Learning . PMLR, 4596–4604
Noam Shazeer and Mitchell Stern. 2018 · 2018
Earlier work this paper cites.
A Broad-Coverage Challenge Corpus for Sentence Understanding through Inference. In Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers) (New Orleans, Louisiana). Association for Computational Linguistics, 1112–1122
Adina Williams, Nikita Nangia, and Samuel Bowman. 2018 · 2018
Earlier work this paper cites.
On the cross-lingual transferability of monolingual representations
Mikel Artetxe, Sebastian Ruder, and Dani Yogatama. 2019 · 2019
Cited alongside, same era.
Cross-lingual Language Model Pretraining
Alexis Conneau and Guillaume Lample. 2019 · 2019
Cited alongside, same era.
XNLI: Evaluating Cross-lingual Sentence Representations
Alexis Conneau, Guillaume Lample, Ruty Rinott, Adina Williams, Samuel R. Bowman, Holger Schwenk, and Veselin Stoyanov. 2019 · 2019
Cited alongside, same era.
What Matters for Neural Cross-Lingual Named Entity Recognition: An Empirical Analysis. In Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP) . 6396–6402
Xiaolei Huang, Jonathan May, and Nanyun Peng. 2019 · 2019
Cited alongside, same era.
Gshard: Scaling giant models with conditional computation and automatic sharding
Dmitry Lepikhin, HyoukJoong Lee, Yuanzhong Xu, Dehao Chen, Orhan Firat, Yanping Huang, Maxim Krikun, Noam Shazeer, and Zhifeng Chen. 2020 · 2020
Later among the works it cites.
A Vietnamese Dataset for Evaluating Machine Reading Comprehension
Kiet Van Nguyen, Duc-Vu Nguyen, Anh Gia-Tuan Nguyen, and Ngan Luu-Thuy Nguyen. 2020 · 2020
Later among the works it cites.
Document Ranking with a Pretrained Sequence-to-Sequence Model. In Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing: Findings . 708–718
Rodrigo Nogueira, Zhiying Jiang, Ronak Pradeep, and Jimmy Lin. 2020 · 2020
Later among the works it cites.
English Intermediate-Task Training Improves Zero-Shot Cross-Lingual Transfer Too
Jason Phang, Iacer Calixto, Phu Mon Htut, Yada Pruksachatkun, Haokun Liu, Clara Vania, Katharina Kann, and Samuel R. Bowman. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Karthikeyan K, Zihan Wang, Stephen Mayhew, and Dan Roth. 2019 · 2019
Cited alongside, same era.
How multilingual is Multilingual BERT?
Telmo Pires, Eva Schlinger, and Dan Garrette. 2019a · 2019
Cited alongside, same era.
How multilingual is multilingual BERT?
Telmo Pires, Eva Schlinger, and Dan Garrette. 2019b · 2019
Cited alongside, same era.
Ipr: The semantic textual similarity and recognizing textual entailment systems
Rui Rodrigues, Paula Couto, and Irene Rodrigues. 2019 · 2019
Cited alongside, same era.
FaQuAD: Reading Comprehension Dataset in the Domain of Brazilian Higher Education
Hélio Fonseca Sayama, Anderson Viçoso Araujo, and Eraldo Rezende Fernandes. 2019 · 2019
Cited alongside, same era.
HuggingFace’s Transformers: State-of-the-art Natural Language Processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander M. Rush. 2019 · 2019
Cited alongside, same era.
Beto, Bentz, Becas: The Surprising Cross-Lingual Effectiveness of BERT
Shijie Wu and Mark Dredze. 2019 · 2019
Cited alongside, same era.
Translation Artifacts in Cross-lingual Transfer Learning
Mikel Artetxe, Gorka Labaka, and Eneko Agirre. 2020 · 2020
Cited alongside, same era.
The ASSIN 2 Shared Task: A Quick Overview
Livy Real, Erick Fonseca, and Hugo Gonçalo Oliveira. 2020 · 2020
Later among the works it cites.
Multilingual transformer ensembles for portuguese natural language tasks
Ruan Rodrigues, Jessica da Silva, Pedro Castro, Nadia Felix, and Anderson Soares. 2020 · 2020
Later among the works it cites.
Cross-Lingual Training of Neural Models for Document Ranking
SPeng Shi, He Bai, and Jimmy Lin. 2020 · 2020
Later among the works it cites.
BERTimbau: Pretrained BERT Models for Brazilian Portuguese
Fábio Souza, Rodrigo Nogueira, and Roberto Lotufo. 2020 · 2020
Later among the works it cites.
OPUS-MT — Building open translation services for the World. In Proceedings of the 22nd Annual Conferenec of the European Association for Machine Translation (EAMT) . Lisbon, Portugal
Jörg Tiedemann and Santhosh Thottingal. 2020 · 2020
Later among the works it cites.
mMARCO: A Multilingual Version of MS MARCO Passage Ranking Dataset
Luiz Henrique Bonifacio, Israel Campiotti, Roberto de Alencar Lotufo, and Rodrigo Nogueira. 2021 · 2021
Closest in time.
On the ability of monolingual models to learn language-agnostic representations
Leandro Rodrigues de Souza, Rodrigo Nogueira, and Roberto Lotufo. 2021 · 2021
Closest in time.
Larger-Scale Transformers for Multilingual Masked Language Modeling
Naman Goyal, Jingfei Du, Myle Ott, Giri Anantharaman, and Alexis Conneau. 2021 · 2021
Closest in time.
Should we Stop Training More Monolingual Models, and Simply Use Machine Translation Instead?
Tim Isbister, Fredrik Carlsson, and Magnus Sahlgren. 2021 · 2021
Closest in time.
GermanQuAD and GermanDPR:Improving Non-English Question Answering and Passage Retrieval
Timo Moller, Julian Risch, and Malte Pietsch. 2021 · 2021
Closest in time.