Fetching the paper…
Reading the bibliography…
Pretrained multilingual models are able to perform cross-lingual transfer in a zero-shot setting, even for languages unseen during pretraining.
Roberta: A robustly optimized bert pretraining approach
Y. Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, M. Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Quantifying the carbon emissions of machine learning
Alexandre Lacoste, Alexandra Luccioni, Victor Schmidt, and Thomas Dandres. 2019 · 1910
Earlier work this paper cites.
Universals of language
Joseph Harold Greenberg. 1963 · 1963
Earlier work this paper cites.
Lecciones para el aprendizaje del idioma shipibo-conibo , volume 1 of Documento de Trabajo
Norma Faust. 1973 · 1973
Earlier work this paper cites.
Diccionario raramuri - castellano: Tarahumar
David Brambila. 1976 · 1976
Earlier work this paper cites.
Compendio de la gramática náhuatl , volume 18
Thelma D Sullivan and Miguel León-Portilla. 1976 · 1976
Earlier work this paper cites.
Diccionario de la lengua náhuatl o mexicana , volume 1
Rémi Siméon. 1977 · 1977
Earlier work this paper cites.
Head-marking and dependent-marking grammar
Johanna Nichols. 1986 · 1986
Earlier work this paper cites.
La lengua Guaraní del Paraguay: Historia, sociedad y literatura
Bartomeu Melià. 1992 · 1992
Earlier work this paper cites.
Diccionario Shipibo-Castellano
James Loriot, Erwin Lauriout, and Dwight Day. 1993 · 1993
Earlier work this paper cites.
Introduccion
Ives Goddard. 1996 · 1996
Earlier work this paper cites.
Raíces del Otomí: diccionario
M. Cajero. 1998 · 1998
Earlier work this paper cites.
The argument status of nps in southeast puebla nahuatl: Comments on the polysynthesis parameter
Jeff MacSwan. 1998 · 1998
Earlier work this paper cites.
American Indian languages: the historical linguistics of Native America
Lyle Campbell. 2000 · 2000
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Xtreme: A massively multilingual multi-task benchmark for evaluating cross-lingual generalization
Junjie Hu, Sebastian Ruder, Aditya Siddhant, Graham Neubig, Orhan Firat, and Melvin Johnson. 2020 · 2003
Earlier work this paper cites.
Transitivity in Shipibo-Konibo grammar
Pilar Valenzuela. 2003 · 2003
Earlier work this paper cites.
Curso Básico de Bribri
Adolfo Constenla, Feliciano Elizondo, and Francisco Pereira. 2004 · 2004
Earlier work this paper cites.
Diccionario Fraseológico Bribri-Español Español-Bribri , second edition
Enrique Margery. 2005 · 2005
Earlier work this paper cites.
Choguita rarámuri (tarahumara) phonology and morphology
Gabriela Caballero. 2008 · 2008
Earlier work this paper cites.
Catálogo de las lenguas indígenas nacionales: Variantes lingüísticas de méxico con sus autodenominaciones y referencias geoestadísticas
INEGI. 2008 · 2008
Earlier work this paper cites.
Gramatica wixarika i
José L. Iturrioz and Paula Gómez-López. 2008 · 2008
Earlier work this paper cites.
Farstail: A persian natural language inference dataset
Hossein Amirkhani, Mohammad AzariJafari, Azadeh Amirak, Zohreh Pourjafari, Soroush Faridan Jahromi, and Zeinab Kouhkan. 2020 · 2009
Earlier work this paper cites.
Textual entailment at evalita 2009
Johan Bos, Fabio Massimo Zanzotto, and M. Pennacchiotti. 2009 · 2009
Earlier work this paper cites.
Historia de los Otomíes en Ixtenco , volume 1
Mateo Cajero. 2009 · 2009
Earlier work this paper cites.
Catálogo de las lenguas indígenas nacionales: Variantes lingüísticas de méxico con sus autodenominaciones y referencias geoestadísticas
INALI. 2009 · 2009
Earlier work this paper cites.
Atlas of the World’s Languages in Danger
Christopher Moseley. 2010 · 2010
Earlier work this paper cites.
B. Muller, Antonis Anastasopoulos, Benoît Sagot, and Djamé Seddah. 2020 · 2010
Earlier work this paper cites.
Población total en territorios indígenas por autoidentificación a la etnia indígena y habla de alguna lengua indígena, según pueblo y territorio indígena
INEC. 2011 · 2011
Earlier work this paper cites.
Unknown Mexico: A Record of Five Years’ Exploration Among the Tribes of the Western Sierra Madre , volume 2
Carl Lumholtz. 2011 · 2011
Cited alongside, same era.
Building a formal grammar for a polysynthetic language
Petr Homola. 2012 · 2012
Cited alongside, same era.
Parallel data, tools and interfaces in OPUS
Jörg Tiedemann. 2012 · 2012
Cited alongside, same era.
A dataset for Arabic textual entailment
Maytham Alabbas. 2013 · 2013
Cited alongside, same era.
Ñaantsipeta asháninkaki birakochaki. diccionario asháninka-castellano. versión preliminar
Rubén Cushimariano Romano and Richer C. Sebastián Q. 2008 · 2013
Cited alongside, same era.
Se’ ttö’ bribri ie Hablemos en bribri
Carla Victoria Jara Murillo and Alí García Segura. 2013 · 2013
Cited alongside, same era.
SentencePiece: A simple and language independent subword tokenizer and detokenizer for neural text processing
Taku Kudo and John Richardson. 2018 · 2018
Later among the works it cites.
Word translation without parallel data
Guillaume Lample, Alexis Conneau, Marc’Aurelio Ranzato, Ludovic Denoyer, and Hervé Jégou. 2018b · 2018
Later among the works it cites.
Challenges of language technologies for the indigenous languages of the Americas
Manuel Mager, Ximena Gutierrez-Vasques, Gerardo Sierra, and Ivan Meza-Ruiz. 2018 · 2018
Later among the works it cites.
Toward Universal Dependencies for Shipibo-konibo
Alonso Vasquez, Renzo Ego Aguirre, Candy Angulo, John Miller, Claudia Villanueva, Željko Agić, Roberto Zariquiey, and Arturo Oncevay. 2018 · 2018
Later among the works it cites.
GLUE: A multi-task benchmark and analysis platform for natural language understanding
Alex Wang, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel Bowman. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Distributed representations of words and phrases and their compositionality
Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg S Corrado, and Jeff Dean. 2013 · 2013
Cited alongside, same era.
Lenguas en peligro en Costa Rica: vitalidad, documentación y descripción
Carlos Sánchez Avendaño. 2013 · 2013
Cited alongside, same era.
An analysis of textual inference in German customer emails
Kathrin Eichler, Aleksandra Gabryszak, and Günter Neumann. 2014 · 2014
Cited alongside, same era.
Norma de escritura de la Lengua Hñähñu (Otomí) , 1st edition
INALI. 2014 · 2014
Cited alongside, same era.
GloVe: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Cited alongside, same era.
A large annotated corpus for learning natural language inference
Samuel R. Bowman, Gabor Angeli, Christopher Potts, and Christopher D. Manning. 2015 · 2015
Cited alongside, same era.
A broad-coverage challenge corpus for sentence understanding through inference
Adina Williams, Nikita Nangia, and Samuel Bowman. 2018 · 2018
Later among the works it cites.
JW300: A wide-coverage parallel corpus for low-resource languages
Željko Agić and Ivan Vulić. 2019 · 2019
Later among the works it cites.
Massively multilingual sentence embeddings for zero-shot cross-lingual transfer and beyond
M. Artetxe and Holger Schwenk. 2019 · 2019
Later among the works it cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
A continuous improvement framework of machine translation for Shipibo-konibo
Héctor Erasmo Gómez Montoya, Kervy Dante Rivas Rojas, and Arturo Oncevay. 2019 · 2019
Later among the works it cites.
The flores evaluation datasets for low-resource machine translation: Nepali–english and sinhala–english
Francisco Guzmán, Peng-Jen Chen, Myle Ott, Juan Pino, Guillaume Lample, Philipp Koehn, Vishrav Chaudhary, and Marc’Aurelio Ranzato. 2019 · 2019
Later among the works it cites.
Towards realistic practices in low-resource natural language processing: The development set
Katharina Kann, Kyunghyun Cho, and Samuel R. Bowman. 2019 · 2019
Later among the works it cites.
Cross-lingual language model pretraining
Guillaume Lample and Alexis Conneau. 2019 · 2019
Later among the works it cites.
fairseq: A fast, extensible toolkit for sequence modeling
Myle Ott, Sergey Edunov, Alexei Baevski, Angela Fan, Sam Gross, Nathan Ng, David Grangier, and Michael Auli. 2019 · 2019
Later among the works it cites.
How multilingual is multilingual BERT?
Telmo Pires, Eva Schlinger, and Dan Garrette. 2019 · 2019
Later among the works it cites.
Superglue: A stickier benchmark for general-purpose language understanding systems
Alex Wang, Yada Pruksachatkun, Nikita Nangia, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel Bowman. 2019 · 2019
Later among the works it cites.
Beto, bentz, becas: The surprising cross-lingual effectiveness of BERT
Shijie Wu and Mark Dredze. 2019 · 2019
Later among the works it cites.
With little power comes great responsibility
D. Card, Peter Henderson, Urvashi Khandelwal, Robin Jia, Kyle Mahowald, and Dan Jurafsky. 2020 · 2020
Later among the works it cites.
Parsing with multilingual BERT, a small corpus, and a small treebank
Ethan C. Chau, Lucy H. Lin, and Noah A. Smith. 2020 · 2020
Later among the works it cites.
Development of a Guarani - Spanish parallel corpus
Luis Chiruzzo, Pedro Amarilla, Adolfo Ríos, and Gustavo Giménez Lugo. 2020 · 2020
Later among the works it cites.
Unsupervised cross-lingual representation learning at scale
Alexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary, Guillaume Wenzek, Francisco Guzmán, E. Grave, Myle Ott, Luke Zettlemoyer, and Veselin Stoyanov. 2020 · 2020
Later among the works it cites.
Neural machine translation models with back-translation for the extremely low-resource indigenous language Bribri
Isaac Feldman and Rolando Coto-Solano. 2020 · 2020
Later among the works it cites.
From zero to hero: On the limitations of zero-shot language transfer with multilingual Transformers
Anne Lauscher, Vinit Ravishankar, Ivan Vulić, and Goran Glavaš. 2020 · 2020
Later among the works it cites.
XGLUE: A new benchmark datasetfor cross-lingual pre-training, understanding and generation
Yaobo Liang, Nan Duan, Yeyun Gong, Ning Wu, Fenfei Guo, Weizhen Qi, Ming Gong, Linjun Shou, Daxin Jiang, Guihong Cao, Xiaodong Fan, Ruofei Zhang, Rahul Agrawal, Edward Cui, Sining Wei, Taroon Bharti, Ying Qiao, Jiun-Hung Chen, Winnie Wu, Shuguang Liu, Fan Yang, Daniel Campos, Rangan Majumder, and Ming Zhou. 2020 · 2020
Later among the works it cites.
MAD-X: An Adapter-Based Framework for Multi-Task Cross-Lingual Transfer
Jonas Pfeiffer, Ivan Vulić, Iryna Gurevych, and Sebastian Ruder. 2020a · 2020
Later among the works it cites.
Extending multilingual BERT to low-resource languages
Zihan Wang, Karthikeyan K, Stephen Mayhew, and Dan Roth. 2020 · 2020
Later among the works it cites.
Transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander M. Rush. 2020 · 2020
Later among the works it cites.
Are all languages created equal in multilingual BERT?
Shijie Wu and Mark Dredze. 2020 · 2020
Later among the works it cites.