Fetching the paper…
Reading the bibliography…
Neural language models learn word representations that capture rich linguistic and conceptual information.
A solution to plato’s problem: The latent semantic analysis theory of acquisition, induction, and representation of knowledge
Thomas K Landauer and Susan T Dumais · 1997
Earlier work this paper cites.
A neural probabilistic language model
Yoshua Bengio, Réjean Ducharme, Pascal Vincent, and Christian Janvin · 2003
Earlier work this paper cites.
A unified architecture for natural language processing: Deep neural networks with multitask learning
Ronan Collobert and Jason Weston · 2008
Earlier work this paper cites.
A study on similarity and relatedness using distributional and wordnet-based approaches
Eneko Agirre, Enrique Alfonseca, Keith Hall, Jana Kravalova, Marius Pasca, and Aitor Soroa · 2009
Earlier work this paper cites.
A scalable hierarchical distributed language model
Andriy Mnih and Geoffrey E Hinton · 2009
Earlier work this paper cites.
From frequency to meaning: Vector space models of semantics
Peter D Turney, Patrick Pantel, et al · 2010
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg S Corrado, and Jeff Dean · 2013
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2014
Cited alongside, same era.
Don’t count, predict! a systematic comparison of context-counting vs. context-predicting semantic vectors
Marco Baroni, Georgiana Dinu, and Germán Kruszewski · 2014
Cited alongside, same era.
Multimodal distributional semantics
Elia Bruni, Nam-Khanh Tran, and Marco Baroni · 2014
Cited alongside, same era.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
Kyunghyun Cho, Bart van Merrienboer, Caglar Gulcehre, Fethi Bougares, Holger Schwenk, and Yoshua Bengio · 2014
Cited alongside, same era.
On the properties of neural machine translation: Encoder–Decoder approaches
Learning abstract concepts from multi-modal data: Since you probably can’t see what i mean
Felix Hill and Anna Korhonen · 2014
Closest in time.
Simlex-999: Evaluating semantic models with (genuine) similarity estimation
Felix Hill, Roi Reichart, and Anna Korhonen · 2014
Closest in time.
Dependency-based word embeddings
Omer Levy and Yoav Goldberg · 2014
Closest in time.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning · 2014
Closest in time.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V Le · 2014
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kyunghyun Cho, Bart van Merriënboer, Dzmitry Bahdanau, and Yoshua Bengio
Cited in the paper.