Fetching the paper…
Reading the bibliography…
Multilingual pre-trained models (mPLMs) have shown impressive performance on cross-lingual transfer tasks.
Cross-lingual name tagging and linking for 282 languages
Xiaoman Pan, Boliang Zhang, Jonathan May, Joel Nothman, Kevin Knight, and Heng Ji. 2017 · 1958
Earlier work this paper cites.
The conversion of scripts, its nature, history, and utilization
Hans H Wellisch, Richard Foreman, Lee Breuer, and Robert Wilson. 1978 · 1978
Earlier work this paper cites.
Linguistic contacts between arabic and other languages
Kees Versteegh. 2001 · 2001
Earlier work this paper cites.
Dict-mlm: Improved multilingual pre-training using bilingual dictionaries
Aditi Chaudhary, Karthik Raman, Krishna Srinivasan, and Jiecao Chen. 2020 · 2010
Earlier work this paper cites.
A simple, fast, and effective reparameterization of IBM model 2
Chris Dyer, Victor Chahuneau, and Noah A. Smith. 2013 · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba. 2015 · 2015
Earlier work this paper cites.
Out-of-the-box universal Romanization tool uroman
Ulf Hermjakob, Jonathan May, and Kevin Knight. 2018 · 2018
Earlier work this paper cites.
Mixed precision training
Paulius Micikevicius, Sharan Narang, Jonah Alben, Gregory F. Diamos, Erich Elsen, David García, Boris Ginsburg, Michael Houston, Oleksii Kuchaiev, Ganesh Venkatesh, and Hao Wu. 2018 · 2018
Earlier work this paper cites.
Pushing the limits of low-resource morphological inflection
Antonios Anastasopoulos and Graham Neubig. 2019 · 2019
Earlier work this paper cites.
Massively multilingual sentence embeddings for zero-shot cross-lingual transfer and beyond
Mikel Artetxe and Holger Schwenk. 2019 · 2019
Earlier work this paper cites.
Cross-lingual language model pretraining
Alexis Conneau and Guillaume Lample. 2019 · 2019
Earlier work this paper cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter. 2019 · 2019
Earlier work this paper cites.
How multilingual is multilingual BERT?
Telmo Pires, Eva Schlinger, and Dan Garrette. 2019 · 2019
Cited alongside, same era.
Cross-lingual alignment of contextual word embeddings, with applications to zero-shot dependency parsing
Tal Schuster, Ori Ram, Regina Barzilay, and Amir Globerson. 2019 · 2019
Cited alongside, same era.
On Romanization for model transfer between scripts in neural machine translation
Chantal Amrhein and Rico Sennrich. 2020 · 2020
Cited alongside, same era.
Multilingual alignment of contextual word representations
Steven Cao, Nikita Kitaev, and Dan Klein. 2020 · 2020
Cited alongside, same era.
Unsupervised cross-lingual representation learning at scale
Alexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary, Guillaume Wenzek, Francisco Guzmán, Edouard Grave, Myle Ott, Luke Zettlemoyer, and Veselin Stoyanov. 2020 · 2020
Cited alongside, same era.
Transliteration for cross-lingual morphological inflection
When being unseen from mBERT is just the beginning: Handling new languages with multilingual language models
Benjamin Muller, Antonios Anastasopoulos, Benoît Sagot, and Djamé Seddah. 2021 · 2021
Later among the works it cites.
Multilingual BERT post-pretraining alignment
Lin Pan, Chung-Wei Hang, Haode Qi, Abhishek Shah, Saloni Potdar, and Mo Yu. 2021 · 2021
Later among the works it cites.
On learning universal representations across languages
Xiangpeng Wei, Rongxiang Weng, Yue Hu, Luxi Xing, Heng Yu, and Weihua Luo. 2021 · 2021
Later among the works it cites.
When is BERT multilingual? isolating crucial ingredients for cross-lingual transfer
Ameet Deshpande, Partha Talukdar, and Karthik Narasimhan. 2022 · 2022
Later among the works it cites.
Align-mlm: Word embedding alignment is crucial for multilingual pre-training
Henry Tang, Ameet Deshpande, and Karthik Narasimhan. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Nikitha Murikinati, Antonios Anastasopoulos, and Graham Neubig. 2020 · 2020
Cited alongside, same era.
Cross-lingual alignment vs joint training: A comparative study and A simple unified framework
Zirui Wang, Jiateng Xie, Ruochen Xu, Yiming Yang, Graham Neubig, and Jaime G. Carbonell. 2020 · 2020
Cited alongside, same era.
Transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Remi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander Rush. 2020 · 2020
Cited alongside, same era.
InfoXLM: An information-theoretic framework for cross-lingual language model pre-training
Zewen Chi, Li Dong, Furu Wei, Nan Yang, Saksham Singhal, Wenhui Wang, Xia Song, Xian-Ling Mao, Heyan Huang, and Ming Zhou. 2021 · 2021
Cited alongside, same era.
Universal Dependencies
Marie-Catherine de Marneffe, Christopher D. Manning, Joakim Nivre, and Daniel Zeman. 2021 · 2021
Cited alongside, same era.
SimCSE: Simple contrastive learning of sentence embeddings
Tianyu Gao, Xingcheng Yao, and Danqi Chen. 2021 · 2021
Cited alongside, same era.
Explicit alignment objectives for multilingual bidirectional encoders
Junjie Hu, Melvin Johnson, Orhan Firat, Aditya Siddhant, and Graham Neubig. 2021 · 2021
Cited alongside, same era.
Glot500: Scaling multilingual corpora and language models to 500 languages
Ayyoob ImaniGooghari, Peiqin Lin, Amir Hossein Kargaran, Silvia Severini, Masoud Jalili Sabet, Nora Kassner, Chunlan Ma, Helmut Schmid, André Martins, François Yvon, and Hinrich Schütze. 2023 · 2023
Later among the works it cites.
Taxi1500: A multilingual dataset for text classification in 1500 languages
Chunlan Ma, Ayyoob ImaniGooghari, Haotian Ye, Ehsaneddin Asgari, and Hinrich Schütze. 2023 · 2023
Later among the works it cites.
Does transliteration help multilingual language modeling?
Ibraheem Muhammad Moosa, Mahmud Elahi Akhter, and Ashfia Binte Habib. 2023 · 2023
Later among the works it cites.
Romanization-based large-scale adaptation of multilingual language models
Sukannya Purkayastha, Sebastian Ruder, Jonas Pfeiffer, Iryna Gurevych, and Ivan Vulić. 2023 · 2023
Later among the works it cites.
Hyperpolyglot LLMs: Cross-lingual interpretability in token embeddings
Andrea W Wen-Yi and David Mimno. 2023 · 2023
Later among the works it cites.
SIB-200: A simple, inclusive, and big evaluation dataset for topic classification in 200+ languages and dialects
David Adelani, Hannah Liu, Xiaoyu Shen, Nikita Vassilyev, Jesujoba Alabi, Yanke Mao, Haonan Gao, and En-Shiun Lee. 2024 · 2024
Closest in time.
Understanding cross-lingual Alignment—A survey
Katharina Hämmerl, Jindřich Libovický, and Alexander Fraser. 2024 · 2024
Closest in time.