Fetching the paper…
Reading the bibliography…
This paper presents an approach for adapting the DebertaV3 XSmall model pre-trained in English for Brazilian Portuguese natural language processing (NLP) tasks.
The brwac corpus: A new open resource for brazilian portuguese
Jorge A Wagner Filho, Rodrigo Wilkens, Marco Idiart, and Aline Villavicencio · 2018
Earlier work this paper cites.
LeNER-Br: a dataset for named entity recognition in Brazilian legal text
Pedro H. Luz de Araujo, Teófilo E. de Campos, Renato R. R. de Oliveira, Matheus Stauffer, Samuel Couto, and Paulo Bermejo · 2018
Earlier work this paper cites.
Camembert: a tasty french language model
Louis Martin, Benjamin Muller, Pedro Javier Ortiz Suárez, Yoann Dupont, Laurent Romary, Éric Villemonte de La Clergerie, Djamé Seddah, and Benoît Sagot · 2019
Earlier work this paper cites.
Bertimbau: Pretrained bert models for brazilian portuguese
Fábio Souza, Rodrigo Nogueira, and Roberto Lotufo · 2020
Earlier work this paper cites.
Deberta: Decoding-enhanced bert with disentangled attention
Pengcheng He, Xiaodong Liu, Jianfeng Gao, and Weizhu Chen · 2020
Earlier work this paper cites.
Electra: Pre-training text encoders as discriminators rather than generators
Kevin Clark, Minh-Thang Luong, Quoc V Le, and Christopher D Manning · 2020
Cited alongside, same era.
Ptt5: Pretraining and validating the t5 model on brazilian portuguese data, 2020
Diedre Carmo, Marcos Piau, Israel Campiotti, Rodrigo Nogueira, and Roberto Lotufo · 2020
Cited alongside, same era.
The assin 2 shared task: a quick overview
Livy Real, Erick Fonseca, and Hugo Goncalo Oliveira · 2020
Cited alongside, same era.
Pengcheng He, Jianfeng Gao, and Weizhu Chen · 2021
Cited alongside, same era.
Jordi Armengol-Estapé, Casimiro Pio Carrino, Carlos Rodriguez-Penagos, Ona de Gibert Bonet, Carme Armentano-Oller, Aitor Gonzalez-Agirre, Maite Melero, and Marta Villegas · 2021
Ernie 3.0: Large-scale knowledge enhanced pre-training for language understanding and generation
Yu Sun, Shuohuan Wang, Shikun Feng, Siyu Ding, Chao Pang, Junyuan Shang, Jiaxiang Liu, Xuyi Chen, Yanbin Zhao, Yuxiang Lu, et al · 2021
Later among the works it cites.
Carolina: The open corpus for linguistics and artificial intelligence
Marcelo Finger, Maria Clara Paixão de Sousa, Cristiane Namiuti, Vanessa Martins do Monte, Aline Silva Costa, Felipe Ribas Serras, Mariana Lourenço Sturzeneker, Raquel de Paula Guets, Renata Morais Mesquita, Guilherme Lamartine de Mello, Maria Clara Ramos Morales Crespo, Maria Lina de Souza Jeannine Rocha, Patrícia Brasil, Mariana Marques da Silva, and Mayara Feliciano Palma · 2022
Later among the works it cites.
Hatebr: A large expert annotated corpus of brazilian instagram comments for offensive language and hate speech detection, 2022
Francielle Alves Vargas, Isabelle Carvalho, Fabiana Rodrigues de Góes, Fabrício Benevenuto, and Thiago Alexandre Salgueiro Pardo · 2022
Later among the works it cites.
Sabiá: Portuguese large language models
Ramon Pires, Hugo Abonizio, Thales Rogério, and Rodrigo Nogueira · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Advancing neural encoding of portuguese with transformer albertina pt
João Rodrigues, Luís Gomes, João Silva, António Branco, Rodrigo Santos, Henrique Lopes Cardoso, and Tomás Osório · 2023
Closest in time.