Fetching the paper…
Reading the bibliography…
Co-occurrence statistics based word embedding techniques have proved to be very useful in extracting the semantic and syntactic representation of words as low dimensional continuous vectors.
Learning effective and interpretable semantic models using non-negative sparse embedding
Brian Murphy, Partha Talukdar, and Tom Mitchell. 2012 · 1950
Earlier work this paper cites.
The sememe
Charles Ernest Bazell. 1966 · 1966
Earlier work this paper cites.
An information-maximization approach to blind separation and blind deconvolution
Anthony J Bell and Terrence J Sejnowski. 1995 · 1995
Earlier work this paper cites.
Emergence of simple-cell receptive field properties by learning a sparse code for natural images
Bruno A Olshausen and David J Field. 1996 · 1996
Earlier work this paper cites.
Sparse coding with an overcomplete basis set: A strategy employed by v1?
Bruno A Olshausen and David J Field. 1997 · 1997
Earlier work this paper cites.
Learning the parts of objects by non-negative matrix factorization
Daniel D Lee and H Sebastian Seung. 1999 · 1999
Earlier work this paper cites.
Non-negative sparse coding
Patrik O Hoyer. 2002 · 2002
Earlier work this paper cites.
On spectral clustering: Analysis and an algorithm
Andrew Y Ng, Michael I Jordan, and Yair Weiss. 2002 · 2002
Earlier work this paper cites.
Stochastic neighbor embedding
Geoffrey E Hinton and Sam T Roweis. 2003 · 2003
Earlier work this paper cites.
A tutorial on spectral clustering
Ulrike Von Luxburg. 2007 · 2007
Earlier work this paper cites.
Visualizing data using t-sne
Laurens van der Maaten and Geoffrey Hinton. 2008 · 2008
Earlier work this paper cites.
A fast iterative shrinkage-thresholding algorithm for linear inverse problems
Amir Beck and Marc Teboulle. 2009 · 2009
Earlier work this paper cites.
Adaptive subgradient methods for online learning and stochastic optimization
John Duchi, Elad Hazan, and Yoram Singer. 2011 · 2011
Cited alongside, same era.
Efficient estimation of word representations in vector space
Tomas Mikolov, Kai Chen, Greg Corrado, and Jeffrey Dean. 2013a · 2013
Cited alongside, same era.
Linguistic regularities in continuous space word representations
Tomas Mikolov, Wen-tau Yih, and Geoffrey Zweig. 2013c · 2013
Cited alongside, same era.
Interpretable semantic vectors from a joint model of brain-and text-based meaning
Alona Fyshe, Partha P Talukdar, Brian Murphy, and Tom M Mitchell. 2014 · 2014
Cited alongside, same era.
Linguistic regularities in sparse and explicit word representations
Omer Levy and Yoav Goldberg. 2014 · 2014
Cited alongside, same era.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Later among the works it cites.
Lexical sememe prediction via word embeddings and matrix factorization
Ruobing Xie, Xingchi Yuan, Zhiyuan Liu, and Maosong Sun. 2017 · 2017
Later among the works it cites.
Linear algebraic structure of word senses, with applications to polysemy
Sanjeev Arora, Yuanzhi Li, Yingyu Liang, Tengyu Ma, and Andrej Risteski. 2018 · 2018
Later among the works it cites.
The sparse manifold transform
Yubei Chen, Dylan Paiton, and Bruno Olshausen. 2018 · 2018
Later among the works it cites.
Deep contextualized word representations
Matthew Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018 · 2018
Later among the works it cites.
Spine: Sparse interpretable neural embeddings
Anant Subramanian, Danish Pruthi, Harsh Jhamtani, Taylor Berg-Kirkpatrick, and Eduard Hovy. 2018 · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Sparse overcomplete word vector representations
Manaal Faruqui, Yulia Tsvetkov, Dani Yogatama, Chris Dyer, and Noah A. Smith. 2015 · 2015
Cited alongside, same era.
Bag of tricks for efficient text classification
Armand Joulin, Edouard Grave, Piotr Bojanowski, and Tomas Mikolov. 2016 · 2016
Cited alongside, same era.
Towards universal paraphrastic sentence embeddings
John Wieting, Mohit Bansal, Kevin Gimpel, and Karen Livescu · 2016
Cited alongside, same era.
Enriching word vectors with subword information
Piotr Bojanowski, Edouard Grave, Armand Joulin, and Tomas Mikolov. 2017 · 2017
Cited alongside, same era.
Visual exploration of semantic relationships in neural word embeddings
Shusen Liu, Peer-Timo Bremer, Jayaraman J Thiagarajan, Vivek Srikumar, Bei Wang, Yarden Livnat, and Valerio Pascucci. 2017 · 2017
Cited alongside, same era.
Improved word representation learning with sememes
Yilin Niu, Ruobing Xie, Zhiyuan Liu, and Maosong Sun. 2017 · 2017
Cited alongside, same era.
Later among the works it cites.
Pretrained fasttext english word vectors
2019
Closest in time.
Analogies explained: Towards understanding word embeddings
Carl Allen and Timothy M. Hospedales. 2019 · 2019
Closest in time.
BERT: pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Closest in time.
Towards understanding linear word analogies
Kawin Ethayarajh, David Duvenaud, and Graeme Hirst. 2019 · 2019
Closest in time.
Wikimedia downloads
Wikimedia. 2019 · 2019
Closest in time.