Fetching the paper…
Reading the bibliography…
We introduce BilBOWA (Bilingual Bag-of-Words without Alignments), a simple and computationally-efficient model for learning bilingual distributed representations of words which can scale to large monolingual datasets and does not require word-aligned parallel training data.
An efficient method for generating discrete random variables with general distributions
Walker, Alastair J · 1977
Earlier work this paper cites.
A neural probabilistic language model
Bengio, Y, Ducharme, R, and Vincent, P · 2003
Earlier work this paper cites.
A systematic comparison of various statistical alignment models
Och, Franz Josef and Ney, Hermann · 2003
Earlier work this paper cites.
Rcv1: A new benchmark collection for text categorization research
Lewis, David D, Yang, Yiming, Rose, Tony G, and Li, Fan · 2004
Earlier work this paper cites.
Europarl: A parallel corpus for statistical machine translation
Koehn, Philipp · 2005
Earlier work this paper cites.
Domain adaptation with structural correspondence learning
Blitzer, J., McDonald, R., and Pereira, F · 2006
Earlier work this paper cites.
Adaptive importance sampling to accelerate training of a neural probabilistic language model
Bengio, Yoshua and Senecal, J-S · 2008
Earlier work this paper cites.
Frustratingly easy domain adaptation
Daumé III, Hal · 2009
Cited alongside, same era.
A survey on transfer learning
Pan, Sinno Jialin and Yang, Qiang · 2010
Cited alongside, same era.
Word representations: A simple and general method for semi-supervised learning
Turian, J., Ratinov, L., and Bengio, Y · 2010
Cited alongside, same era.
Natural language processing (almost) from scratch
Collobert, R., Weston, J., Bottou, L., Karlen, M., Kavukcuoglu, K., and Kuksa, P · 2011
Cited alongside, same era.
Inducing crosslingual distributed representations of words
Klementiev, Alexandre, Titov, Ivan, and Bhattarai, Binod · 2012
Cited alongside, same era.
A fast and simple algorithm for training neural probabilistic language models
Mnih, Andriy and Teh, Yee Whye · 2012
Cited alongside, same era.
A simple, fast, and effective reparameterization of ibm model 2
Dyer, Chris, Chahuneau, Victor, and Smith, Noah A · 2013
Later among the works it cites.
Multilingual distributed representations without word alignment
Hermann, Karl Moritz and Blunsom, Phil · 2013
Later among the works it cites.
Bilingual word embeddings for phrase-based machine translation
Zou, Will Y, Socher, Richard, Cer, Daniel, and Manning, Christopher D · 2013
Later among the works it cites.
An autoencoder approach to learning bilingual word representations
Chandar, Sarath, Lauly, Stanislas, Larochelle, Hugo, Khapra, Mitesh M., Ravidran, Balaraman, Raykar, Vikas, and Saha, Amrita · 2014
Closest in time.
Improving vector space word representations using multilingual correlation
Faruqui, Manaal and Dyer, Chris · 2014
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Polyglot: Distributed word representations for multilingual nlp
Al-Rfou’, Rami, Perozzi, Bryan, and Skiena, Steven · 2013
Cited alongside, same era.
Exploiting similarities among languages for machine translation
Mikolov, Tomas, Le, Quoc V, and Sutskever, Ilya
Cited in the paper.
Distributed representations of words and phrases and their compositionality
Mikolov, Tomas, Sutskever, Ilya, Chen, Kai, Corrado, Greg S, and Dean, Jeff
Cited in the paper.
Goldberg, Yoav and Levy, Omer · 2014
Closest in time.
Glove: Global vectors for word representation
Pennington, Jeffrey, Socher, Richard, and Manning, Christopher D · 2014
Closest in time.