Fetching the paper…
Reading the bibliography…
In this paper, we propose a new method for calculating the output layer in neural machine translation systems.
A mathematical theory of communication
Claude E. Shannon. 1948 · 1948
Earlier work this paper cites.
Notes on digital coding
Marcel J. E. Golay. 1949 · 1949
Earlier work this paper cites.
Human behavior and the principle of least effort
George. K. Zipf. 1949 · 1949
Earlier work this paper cites.
A method for the construction of minimum-redundancy codes
David A. Huffman. 1952 · 1952
Earlier work this paper cites.
Dropout: a simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey E Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov. 2014 · 1958
Earlier work this paper cites.
Error bounds for convolutional codes and an asymptotically optimum decoding algorithm
Andrew Viterbi. 1967 · 1967
Earlier work this paper cites.
Strategies for training large vocabulary neural language models
Wenlin Chen, David Grangier, and Michael Auli. 2016 · 1985
Earlier work this paper cites.
Class-based n-gram models of natural language
Peter F Brown, Peter V Desouza, Robert L Mercer, Vincent J Della Pietra, and Jenifer C Lai. 1992 · 1992
Earlier work this paper cites.
Solving multiclass learning problems via error-correcting output codes
Thomas G. Dietterich and Ghulum Bakiri. 1995 · 1995
Earlier work this paper cites.
Building a bilingual travel conversation database for speech translation research
Toshiyuki Takezawa. 1999 · 1999
Earlier work this paper cites.
Learning to forget: Continual prediction with LSTM
Felix A Gers, Jürgen Schmidhuber, and Fred Cummins. 2000 · 2000
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Cited alongside, same era.
On nearest-neighbor error-correcting output codes with application to all-pairs multiclass support vector machines
Aldebaro Klautau, Nikola Jevtić, and Alon Orlitsky. 2003 · 2003
Cited alongside, same era.
Hierarchical probabilistic neural network language model
Frederic Morin and Yoshua Bengio. 2005 · 2005
Cited alongside, same era.
Using svm and error-correcting codes for multiclass dialog act classification in meeting corpus
Yang Liu. 2006 · 2006
Cited alongside, same era.
Moses: Open source toolkit for statistical machine translation
Philipp Koehn, Hieu Hoang, Alexandra Birch, Chris Callison-Burch, Marcello Federico, Nicola Bertoldi, Brooke Cowan, Wade Shen, Christine Moran, Richard Zens, Chris Dyer, Ondrej Bojar, Alexandra Constantin, and Evan Herbst. 2007 · 2007
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio. 2014 · 2014
Later among the works it cites.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba. 2014 · 2014
Later among the works it cites.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V Le. 2014 · 2014
Later among the works it cites.
Character-based neural machine translation
Wang Ling, Isabel Trancoso, Chris Dyer, and Alan W Black. 2015 · 2015
Later among the works it cites.
Effective approaches to attention-based neural machine translation
Thang Luong, Hieu Pham, and Christopher D. Manning. 2015 · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Multilabel classification by bch code and random forests
Abbas Z Kouzani and Gulisong Nasireding. 2009 · 2009
Cited alongside, same era.
Multilabel classification using error correction codes
Abbas Z Kouzani. 2010 · 2010
Cited alongside, same era.
Multi-label classification with error-correcting codes
Chun-Sung Ferng and Hsuan-Tien Lin. 2011 · 2011
Cited alongside, same era.
Pointwise prediction for robust, adaptable japanese morphological analysis
Graham Neubig, Yosuke Nakata, and Shinsuke Mori. 2011 · 2011
Cited alongside, same era.
A fast and simple algorithm for training neural probabilistic language models
Andriy Mnih and Yee Whye Teh. 2012 · 2012
Cited alongside, same era.
Distributed representations of words and phrases and their compositionality
Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg S Corrado, and Jeff Dean. 2013 · 2013
Cited alongside, same era.
Multilabel classification using error-correcting codes of hard or soft bits
Chun-Sung Ferng and Hsuan-Tien Lin. 2013
Cited in the paper.
Vocabulary selection strategies for neural machine translation
Gurvan L’Hostis, David Grangier, and Michael Auli. 2016 · 2016
Later among the works it cites.
Vocabulary manipulation for neural machine translation
Haitao Mi, Zhiguo Wang, and Abe Ittycheriah. 2016 · 2016
Later among the works it cites.
Aspec: Asian scientific paper excerpt corpus
Toshiaki Nakazawa, Manabu Yaguchi, Kiyotaka Uchimoto, Masao Utiyama, Eiichiro Sumita, Sadao Kurohashi, and Hitoshi Isahara. 2016 · 2016
Later among the works it cites.
Neural machine translation of rare words with subword units
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2016 · 2016
Later among the works it cites.
Dynet: The dynamic neural network toolkit
Graham Neubig, Chris Dyer, Yoav Goldberg, Austin Matthews, Waleed Ammar, Antonios Anastasopoulos, Miguel Ballesteros, David Chiang, Daniel Clothiaux, Trevor Cohn, Kevin Duh, Manaal Faruqui, Cynthia Gan, Dan Garrette, Yangfeng Ji, Lingpeng Kong, Adhiguna Kuncoro, Gaurav Kumar, Chaitanya Malaviya, Paul Michel, Yusuke Oda, Matthew Richardson, Naomi Saphra, Swabha Swayamdipta, and Pengcheng Yin. 2017 · 2017
Closest in time.
Variable-length word encodings for neural translation models
Rohan Chitnis and John DeNero. 2015 · 2093
Closest in time.