Fetching the paper…
Reading the bibliography…
The word2vec model and application by Mikolov et al.
Hierarchical probabilistic neural network language model
Morin, F. and Bengio, Y. (2005) · 2005
Earlier work this paper cites.
A scalable hierarchical distributed language model
Mnih, A. and Hinton, G. E. (2009) · 2009
Cited alongside, same era.
Efficient estimation of word representations in vector space
Mikolov, T., Chen, K., Corrado, G., and Dean, J. (2013a)
Cited in the paper.
Distributed representations of words and phrases and their compositionality
Mikolov, T., Sutskever, I., Chen, K., Corrado, G. S., and Dean, J. (2013b)
Cited in the paper.
word2vec explained: deriving mikolov et al.’s negative-sampling word-embedding method
Goldberg, Y. and Levy, O. (2014) · 2014
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…