Fetching the paper…
Reading the bibliography…
We propose procedures for evaluating and strengthening contextual embedding alignment and show that they are useful in analyzing and improving multilingual BERT.
Cross-lingual language model pretraining
Guillame Lample and Alexis Conneau · 1901
Earlier work this paper cites.
A generalized solution of the orthogonal procrustes problem
Peter H. Schonemann · 1966
Earlier work this paper cites.
Europarl: A parallel corpus for statistical machine translation
Philipp Koehn · 2005
Earlier work this paper cites.
Visualizing data using t-sne
Laurens van der Maaten and Geoffrey Hinton · 2008
Earlier work this paper cites.
MultiUN: A multilingual corpus from united nation documents
Andreas Eisele and Yu Chen · 2010
Earlier work this paper cites.
A universal part-of-speech tagset
Slav Petrov, Dipanjan Das, and Ryan McDonald · 2012
Earlier work this paper cites.
Parallel data, tools and interfaces in OPUS
Jörg Tiedemann · 2012
Earlier work this paper cites.
Polyglot: Distributed word representations for multilingual nlp
Rami Al-Rfou, Bryan Perozzi, and Steven Skiena · 2013
Earlier work this paper cites.
A simple, fast, and effective reparameterization of IBM model 2
Chris Dyer, Victor Chahuneau, and Noah A. Smith · 2013
Earlier work this paper cites.
Learning principled bilingual mappings of word embeddings while preserving monolingual invariance
Mikel Artetxe, Gorka Labaka, and Eneko Agirre · 2016
Earlier work this paper cites.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Yonghui Wu, Mike Schuster, Zhifeng Chen, Quoc V. Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, Jeff Klingner, Apurva Shah, Melvin Johnson, Xiaobing Liu, Lukasz Kaiser, Stephan Gouws, Yoshikiyo Kato, Taku Kudo, Hideto Kazawa, Keith Stevens, George Kurian, Nishant Patil, Wei Wang, Cliff Young, Jason Smith, Jason Riesa, Alex Rudnick, Oriol Vinyals, Greg Corrado, Macduff Hughes, and Jeffrey Dean · 2016
Earlier work this paper cites.
Learning bilingual word embeddings with (almost) no bilingual data
Mikel Artetxe, Gorka Labaka, and Eneko Agirre · 2017
Earlier work this paper cites.
Enriching word vectors with subword information
Piotr Bojanowski, Edouard Grave, Armand Joulin, and Tomas Mikolov · 2017
Earlier work this paper cites.
The mathematics of statistical machine translation: Parameter estimation
Peter F. Brown, Vincent J. Della Pietra, Stephen A. Della Pietra, and Robert L. Mercer · 2017
Cited alongside, same era.
spaCy 2: Natural language understanding with Bloom embeddings, convolutional neural networks and incremental parsing
Matthew Honnibal and Ines Montani · 2017
Cited alongside, same era.
A systematic comparison of various statistical alignment models
Franz Josef Och and Hermann Ney · 2017
Cited alongside, same era.
Offline bilingual word vectors, orthogonal transformations and the inverted softmax
Samuel L. Smith, David H. P. Turban, Steven Hamblin, and Nils Y. Hammerla · 2017
Cited alongside, same era.
A robust self-learning method for fully unsupervised cross-lingual mappings of word embeddings
Mikel Artetxe, Gorka Labaka, and Eneko Agirre · 2018
Cited alongside, same era.
Concatenated p-mean word embeddings as universal cross-lingual sentence representations
Andreas Rücklé, Steffen Eger, Maxime Peyrard, and Iryna Gurevych · 2018
Later among the works it cites.
On the limitations of unsupervised bilingual dictionary induction
Anders Søgaard, Sebastian Ruder, and Ivan Vulić · 2018
Later among the works it cites.
A broad-coverage challenge corpus for sentence understanding through inference
Adina Williams, Nikita Nangia, and Samuel Bowman · 2018
Later among the works it cites.
Unsupervised cross-lingual transfer of word embedding spaces
Ruochen Xu, Yiming Yang, Naoki Otani, and Yuexin Wu · 2018
Later among the works it cites.
Context-aware cross-lingual mapping
Hanan Aldarmaki and Mona Diab · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Xilun Chen and Claire Cardie · 2018
Cited alongside, same era.
Word translation without parallel data
Alexis Conneau, Guillaume Lample, Marc’Aurelio Ranzato, Ludovic Denoyer, and Herve J’egou · 2018
Cited alongside, same era.
XNLI: Evaluating cross-lingual sentence representations
Alexis Conneau, Ruty Rinott, Guillaume Lample, Adina Williams, Samuel Bowman, Holger Schwenk, and Veselin Stoyanov · 2018
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Cited alongside, same era.
Non-adversarial unsupervised word translation
Yedid Hoshen and Lior Wolf · 2018
Cited alongside, same era.
Universal language model fine-tuning for text classification
Jeremy Howard and Sebastian Ruder · 2018
Cited alongside, same era.
Deep contextualized word representations
Matthew Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer · 2018
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Later among the works it cites.
Learning unsupervised multilingual word embeddings with incremental multilingual hubs
Geert Heyman, Bregt Verreet, Ivan Vulić, and Marie-Francine Moens · 2019
Later among the works it cites.
How multilingual is multilingual BERT?
Telmo Pires, Eva Schlinger, and Dan Garrette · 2019
Later among the works it cites.
A survey of cross-lingual word embedding models
Sebastian Ruder, Ivan Vulić, and Anders Søgaard · 2019
Later among the works it cites.
Cross-lingual alignment of contextual word embeddings, with applications to zero-shot dependency parsing
Tal Schuster, Ori Ram, Regina Barzilay, and Amir Globerson · 2019
Later among the works it cites.
Cross-lingual BERT transformation for zero-shot dependency parsing
Yuxuan Wang, Wanxiang Che, Jiang Guo, Yijia Liu, and Ting Liu · 2019
Later among the works it cites.
Simple and effective paraphrastic similarity from parallel translations
John Wieting, Kevin Gimpel, Graham Neubig, and Taylor Berg-Kirkpatrick · 2019
Later among the works it cites.
Moses: Open source toolkit for statistical machine translation
Philipp Koehn, Hieu Hoang, Alexandra Birch, Chris Callison-Burch, Marcello Federico, Nicola Bertoldi, Brooke Cowan, Wade Shen, Christine Moran, Richard Zens, Chris Dyer, Ondřej Bojar, Alexandra Constantin, and Evan Herbst · 2045
Closest in time.