Fetching the paper…
Reading the bibliography…
We present hash embeddings, an efficient method for representing words in a continuous vector form.
Quantization
Gray, R. M. and Neuhoff, D. L. (1998) · 1998
Earlier work this paper cites.
Foundations of statistical natural language processing
Manning, C. D., Schütze, H., et al. (1999) · 1999
Earlier work this paper cites.
Entropy-based pruning of backoff language models
Stolcke, A. (2000) · 2000
Earlier work this paper cites.
Supervised semantic indexing
Bai, B., Weston, J., Grangier, D., Collobert, R., Sadamasa, K., Qi, Y., Chapelle, O., and Weinberger, K. (2009) · 2009
Earlier work this paper cites.
Learning to hash with binary reconstructive embeddings
Kulis, B. and Darrell, T. (2009) · 2009
Earlier work this paper cites.
Feature hashing for large scale multitask learning
Weinberger, K. Q., Dasgupta, A., Attenberg, J., Langford, J., and Smola, A. J. (2009) · 2009
Earlier work this paper cites.
Multi-prototype vector-space models of word meaning
Reisinger, J. and Mooney, R. J. (2010) · 2010
Earlier work this paper cites.
Product quantization for nearest neighbor search
Jegou, H., Douze, M., and Schmid, C. (2011) · 2011
Cited alongside, same era.
Improving word representations via global context and multiple word prototypes
Huang, E. H., Socher, R., Manning, C. D., and Ng, A. Y. (2012) · 2012
Cited alongside, same era.
Learning deep structured semantic models for web search using clickthrough data
Huang, P.-S., He, X., Gao, J., Deng, L., Acero, A., and Heck, L. (2013) · 2013
Cited alongside, same era.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J. (2014) · 2014
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Ioffe, S. and Szegedy, C. (2015) · 2015
Cited alongside, same era.
Very deep convolutional networks for natural language processing
Conneau, A., Schwenk, H., Barrault, L., and LeCun, Y. (2016) · 2016
Later among the works it cites.
Neural machine translation with characters and hierarchical encoding
Johansen, A. R., Hansen, J. M., Obeid, E. K., Sønderby, C. K., and Winther, O. (2016) · 2016
Later among the works it cites.
Convolutional neural networks for text categorization: Shallow word-level vs. deep character-level
Johnson, R. and Zhang, T. (2016) · 2016
Later among the works it cites.
Virtual adversarial training for semi-supervised text classification
Miyato, T., Dai, A. M., and Goodfellow, I. (2016) · 2016
Later among the works it cites.
Efficient character-level document classification by combining convolution and recurrent layers
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zhang, X., Zhao, J. J., and LeCun, Y. (2015) · 2015
Cited alongside, same era.
Hash2vec, feature hashing for word embeddings
Argerich, L., Zaffaroni, J. T., and Cano, M. J. (2016) · 2016
Cited alongside, same era.
Fasttext.zip: Compressing text classification models
Joulin, A., Grave, E., Bojanowski, P., Douze, M., Jégou, H., and Mikolov, T. (2016a)
Cited in the paper.
Bag of tricks for efficient text classification
Joulin, A., Grave, E., Bojanowski, P., and Mikolov, T. (2016b)
Cited in the paper.
Xiao, Y. and Cho, K. (2016) · 2016
Later among the works it cites.
Google’s trained word2vec model in python
Miháltz, M. (2016) · 2017
Closest in time.
Generative and discriminative text classification with recurrent neural networks
Yogatama, D., Dyer, C., Ling, W., and Blunsom, P. (2017) · 2017
Closest in time.