Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Building large monolingual dictionaries at the leipzig corpora collection: From 100 to 200 languages
Dirk Goldhahn, Thomas Eckart, and Uwe Quasthoff · 2012
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
Original
Tomas Mikolov, Ilya Sutskever, Kai Chen, Gregory S. Corrado, and Jeffrey Dean · 2013
Earlier work this paper cites.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher D. Manning · 2014
Earlier work this paper cites.
Sequence to sequence learning with neural networks
Original
Ilya Sutskever, Oriol Vinyals, and Quoc V. Le · 2014
Earlier work this paper cites.
Semi-supervised sequence learning
Original
Andrew M. Dai and Quoc V. Le · 2015
Earlier work this paper cites.
A hybrid method for persian named entity recognition
Farid Ahmadi and Hamed Moradi · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
Original
Diederik P. Kingma and Jimmy Ba · 2015
Earlier work this paper cites.
Unsupervised pretraining for sequence to sequence learning
Original
Prajit Ramachandran, Peter J. Liu, and Quoc V. Le · 2016
Earlier work this paper cites.
Neural machine translation of rare words with subword units
Original
Rico Sennrich, Barry Haddow, and Alexandra Birch · 2016
Earlier work this paper cites.
Personer: Persian named-entity recognition
Original
Hanieh Poostchi, Ehsan Zare Borzeshi, Mohammad Abdous, and Massimo Piccardi · 2016
Earlier work this paper cites.
Attention is all you need
Original
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Persian named entity recognition
Kia Dashtipour, Mandar Gogate, Ahsan Adeel, Abdulrahman Algarafi, Newton Howard, and Amir Hussain · 2017
Earlier work this paper cites.
Deep contextualized word representations
Original
Matthew E. Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer · 2018
Earlier work this paper cites.
Improving language understanding by generative pre-training
Alec Radford · 2018
Earlier work this paper cites.
c-rnn: A fine-grained language model for image captioning
Gengshi Huang and Haifeng Hu · 2018
Earlier work this paper cites.
Multi-task character-level attentional networks for medical concept normalization
Jinghao Niu, Yehui Yang, Siheng Zhang, Zhengya Sun, and Wensheng Zhang · 2018
Earlier work this paper cites.