Fetching the paper…
Reading the bibliography…
Continuous word representation (aka word embedding) is a basic building block in many neural network-based models used in natural language processing tasks.
Bleu: a method for automatic evaluation of machine translation
K. Papineni, S. Roukos, T. Ward, and W.-J. Zhu · 2002
Earlier work this paper cites.
Introduction to information retrieval
R. R. Larson · 2010
Earlier work this paper cites.
Recurrent neural network based language model
T. Mikolov, M. Karafiát, L. Burget, J. Cernocký, and S. Khudanpur · 2010
Earlier work this paper cites.
Learning word vectors for sentiment analysis
A. L. Maas, R. E. Daly, P. T. Pham, D. Huang, A. Y. Ng, and C. Potts · 2011
Earlier work this paper cites.
Polyglot: Distributed word representations for multilingual nlp
R. Al-Rfou, B. Perozzi, and S. Skiena · 2013
Earlier work this paper cites.
Document classification by topic labeling
S. Hingmire, S. Chougule, G. K. Palshikar, and S. Chakraborti · 2013
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
T. Mikolov, I. Sutskever, K. Chen, G. S. Corrado, and J. Dean · 2013
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
D. Bahdanau, K. Cho, and Y. Bengio · 2014
Earlier work this paper cites.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
K. Cho, B. Van Merriënboer, C. Gulcehre, D. Bahdanau, F. Bougares, H. Schwenk, and Y. Bengio · 2014
Earlier work this paper cites.
Generative adversarial nets
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio · 2014
Earlier work this paper cites.
Glove: Global vectors for word representation
J. Pennington, R. Socher, and C. D. Manning · 2014
Earlier work this paper cites.
Semi-supervised sequence learning
A. M. Dai and Q. V. Le · 2015
Earlier work this paper cites.
Unsupervised domain adaptation by backpropagation
Y. Ganin and V. Lempitsky · 2015
Earlier work this paper cites.
Recurrent convolutional neural networks for text classification
S. Lai, L. Xu, K. Liu, and J. Zhao · 2015
Earlier work this paper cites.
Unsupervised representation learning with deep convolutional generative adversarial networks
A. Radford, L. Metz, and S. Chintala · 2015
Earlier work this paper cites.
Sequence level training with recurrent neural networks
M. Ranzato, S. Chopra, M. Auli, and W. Zaremba · 2015
Cited alongside, same era.
Neural machine translation of rare words with subword units
R. Sennrich, B. Haddow, and A. Birch · 2015
Cited alongside, same era.
A simple but tough-to-beat baseline for sentence embeddings
S. Arora, Y. Liang, and T. Ma · 2016
Cited alongside, same era.
An actor-critic algorithm for sequence prediction
D. Bahdanau, P. Brakel, K. Xu, A. Goyal, R. Lowe, J. Pineau, A. Courville, and Y. Bengio · 2016
Cited alongside, same era.
A convolutional encoder model for neural machine translation
J. Gehring, M. Auli, D. Grangier, and Y. N. Dauphin · 2016
Classical structured prediction losses for sequence to sequence learning
S. Edunov, M. Ott, M. Auli, D. Grangier, and M. Ranzato · 2017
Later among the works it cites.
Convolutional sequence to sequence learning
J. Gehring, M. Auli, D. Grangier, D. Yarats, and Y. N. Dauphin · 2017
Later among the works it cites.
Neural phrase-based machine translation
P.-S. Huang, C. Wang, D. Zhou, and L. Deng · 2017
Later among the works it cites.
Dynamic evaluation of neural sequence models
B. Krause, E. Kahembwe, I. Murray, and S. Renals · 2017
Later among the works it cites.
Unsupervised machine translation using monolingual corpora only
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Improving neural language models with a continuous cache
E. Grave, A. Joulin, and N. Usunier · 2016
Cited alongside, same era.
Exploring the limits of language modeling
R. Jozefowicz, O. Vinyals, M. Schuster, N. Shazeer, and Y. Wu · 2016
Cited alongside, same era.
Neural machine translation in linear time
N. Kalchbrenner, L. Espeholt, K. Simonyan, A. v. d. Oord, A. Graves, and K. Kavukcuoglu · 2016
Cited alongside, same era.
Character-aware neural language models
Y. Kim, Y. Jernite, D. Sontag, and A. M. Rush · 2016
Cited alongside, same era.
Professor forcing: A new algorithm for training recurrent networks
A. M. Lamb, A. G. A. P. GOYAL, Y. Zhang, S. Zhang, A. C. Courville, and Y. Bengio · 2016
Cited alongside, same era.
Pointer sentinel mixture models
S. Merity, C. Xiong, J. Bradbury, and R. Socher · 2016
Cited alongside, same era.
Improved techniques for training gans
T. Salimans, I. Goodfellow, W. Zaremba, V. Cheung, A. Radford, and X. Chen · 2016
Cited alongside, same era.
G. Lample, L. Denoyer, and M. Ranzato · 2017
Later among the works it cites.
Regularizing and optimizing LSTM language models
S. Merity, N. S. Keskar, and R. Socher · 2017
Later among the works it cites.
All-but-the-top: simple and effective postprocessing for word representations
J. Mu, S. Bhat, and P. Viswanath · 2017
Later among the works it cites.
All-but-the-top: Simple and effective postprocessing for word representations
J. Mu, S. Bhat, and P. Viswanath · 2017
Later among the works it cites.
Adversarial discriminative domain adaptation
E. Tzeng, J. Hoffman, K. Saenko, and T. Darrell · 2017
Later among the works it cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin · 2017
Later among the works it cites.
Breaking the softmax bottleneck: A high-rank RNN language model
Z. Yang, Z. Dai, R. Salakhutdinov, and W. W. Cohen · 2017
Later among the works it cites.
Unpaired image-to-image translation using cycle-consistent adversarial networks
J.-Y. Zhu, T. Park, P. Isola, and A. A. Efros · 2017
Later among the works it cites.
Fix your classifier: the marginal value of training the last weight layer
E. Hoffer, I. Hubara, and D. Soudry · 2018
Closest in time.
Analyzing uncertainty in neural machine translation
M. Ott, M. Auli, D. Granger, and M. Ranzato · 2018
Closest in time.
Tensor2tensor for neural machine translation
A. Vaswani, S. Bengio, E. Brevdo, F. Chollet, A. N. Gomez, S. Gouws, L. Jones, L. Kaiser, N. Kalchbrenner, N. Parmar, R. Sepassi, N. Shazeer, and J. Uszkoreit · 2018
Closest in time.