Fetching the paper…
Reading the bibliography…
Current work in lexical distributed representations maps each word to a point vector in low-dimensional space.
Theoretical foundations of the potential function method in pattern recognition learning
Aizerman, M. A., Braverman, E. A., and Rozonoer, L · 1964
Earlier work this paper cites.
Contextual correlates of synonymy
Rubenstein, Herbert and Goodenough, John B · 1965
Earlier work this paper cites.
Means and variances of stochastic vector products with applications to random linear models
Brown, Gerald G and Rutemiller, Herbert C · 1977
Earlier work this paper cites.
Learning representations by back-propagating errors
Rumelhart, D.E., Hintont, G.E., and Williams, R.J · 1986
Earlier work this paper cites.
Indexing by latent semantic analysis
Deerwester, S., Dumais, S.T., Furnas, G.W., Landauer, T.K., and Harshman, R.A · 1990
Earlier work this paper cites.
Contextual correlates of semantic similarity
Miller, George A and Charles, Walter G · 1991
Earlier work this paper cites.
Mixture density networks, 1994
Bishop, Christopher M · 1994
Earlier work this paper cites.
Placing search in context: The concept revisited
Finkelstein, Lev, Gabrilovich, Evgeniy, Matias, Yossi, Rivlin, Ehud, Solan, Zach, Wolfman, Gadi, and Ruppin, Eytan · 2001
Earlier work this paper cites.
Optimizing search engines using clickthrough data
Joachims, Thorsten · 2002
Earlier work this paper cites.
Distance metric learning with application to clustering with side-information
Xing, Eric P, Jordan, Michael I, Russell, Stuart, and Ng, Andrew Y · 2002
Earlier work this paper cites.
Probability product kernels
Jebara, Tony, Kondor, Risi, and Howard, Andrew · 2004
Earlier work this paper cites.
Learning the kernel matrix with semidefinite programming
Lanckriet, Gert RG, Cristianini, Nello, Bartlett, Peter, Ghaoui, Laurent El, and Jordan, Michael I · 2004
Earlier work this paper cites.
The distributional inclusion hypotheses and lexical entailment
Geffet, Maayan and Dagan, Ido · 2005
Earlier work this paper cites.
Neural probabilistic language models
Bengio, Yoshua, Schwenk, Holger, Senécal, Jean-Sébastien, Morin, Fréderic, and Gauvain, Jean-Luc · 2006
Cited alongside, same era.
A tutorial on energy-based learning
LeCun, Yann, Chopra, Sumit, and Hadsell, Raia · 2006
Cited alongside, same era.
The matrix cookbook
Petersen, Kaare Brandt · 2006
Cited alongside, same era.
Verb similarity on the taxonomy of wordnet
Yang, Dongqiang and Powers, David M. W · 2006
Cited alongside, same era.
Probabilistic matrix factorization
Mnih, Andriy and Salakhutdinov, Ruslan · 2007
Cited alongside, same era.
A scalable hierarchical distributed language model
Mnih, Andriy and Hinton, Geoffrey E · 2008
Cited alongside, same era.
Textual similarity with a bag-of-embedded-words model
Clinchant, Stéphane and Perronnin, Florent · 2013
Later among the works it cites.
Aggregating continuous word embeddings for information retrieval
Clinchant, Stéphane and Perronnin, Florent · 2013
Later among the works it cites.
Distributed representations of words and phrases and their compositionality
Mikolov, Tomas, Sutskever, Ilya, Chen, Kai, Corrado, Greg S, and Dean, Jeff · 2013
Later among the works it cites.
Relation extraction with matrix factorization and universal schemas
Riedel, Sebastian, Yao, Limin, McCallum, Andrew, and Marlin, Benjamin M · 2013
Later among the works it cites.
A new set of norms for semantic relatedness measures
Szumlanski, Sean R, Gomez, Fernando, and Sims, Valerie K · 2013
Later among the works it cites.
Don’t count, predict! a systematic comparison of context-counting vs. context-predicting semantic vectors
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Bayesian probabilistic matrix factorization using markov chain monte carlo
Salakhutdinov, Ruslan and Mnih, Andriy · 2008
Cited alongside, same era.
The wacky wide web: a collection of very large linguistically processed web-crawled corpora
Baroni, Marco, Bernardini, Silvia, Ferraresi, Adriano, and Zanchetta, Eros · 2009
Cited alongside, same era.
Representing words as regions in vector space
Erk, Katrin · 2009
Cited alongside, same era.
Matrix factorization techniques for recommender systems
Koren, Yehuda, Bell, Robert, and Volinsky, Chris · 2009
Cited alongside, same era.
Adaptive subgradient methods for online learning and stochastic optimization
Duchi, John, Hazan, Elad, and Singer, Yoram · 2011
Cited alongside, same era.
Wsabie: Scaling up to large vocabulary image annotation
Weston, Jason, Bengio, Samy, and Usunier, Nicolas · 2011
Cited alongside, same era.
Baroni, Marco, Dinu, Georgiana, and Kruszewski, Germán · 2014
Closest in time.
Multimodal distributional semantics
Bruni, Elia, Tran, Nam-Khanh, and Baroni, Marco · 2014
Closest in time.
Community evaluation and exchange of word vectors at wordvectors.org
Faruqui, Manaal and Dyer, Chris · 2014
Closest in time.
Simlex-999: Evaluating semantic models with (genuine) similarity estimation
Hill, Felix, Reichart, Roi, and Korhonen, Anna · 2014
Closest in time.
A multiplicative model for learning distributed text-based attribute representations
Kiros, Ryan, Zemel, Richard, and Salakhutdinov, Ruslan R · 2014
Closest in time.
Distribution representation of the meaning of words and phrases by a gaussian distribution
Kiyoshiyo, Shimaoka, Masayasu, Muraoka, Futo, Yamamoto, Watanabe, Yotaro, Okazaki, Naoaki, and Inui, Kentaro · 2014
Closest in time.
Neural word embedding as implicit matrix factorization
Levy, Omer and Goldberg, Yoav · 2014
Closest in time.