Fetching the paper…
Reading the bibliography…
We propose a large margin criterion for training neural language models.
An empirical study of smoothing techniques for language modeling
Stanley F. Chen and Joshua Goodman. 1996 · 1996
Earlier work this paper cites.
A bit of progress in language modeling
Joshua T. Goodman. 2001 · 2001
Earlier work this paper cites.
Discriminative training of language models for speech recognition in acoustics
Hong-Kwang Jeff Kuo, Eric Fosler-Lussier, Hui Jiang, and Chin-Hui Lee. 2002 · 2002
Earlier work this paper cites.
A neural probabilistic language model
Yoshua Bengio, Réjean Ducharme, Pascal Vincent, and Christian Jauvin. 2003 · 2003
Earlier work this paper cites.
Discriminative language modeling with conditional random fields and the perceptron algorithm
Brian Roark, Murat Saraclar, and Michael Collins Mark Johnson. 2004 · 2004
Earlier work this paper cites.
Discriminative reranking for natural language parsing
Michael Collins and Terry Koo. 2005 · 2005
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
Alex Graves, Santiago Fernández, Faustino Gomez, and Jürgen Schmidhuber. 2006 · 2006
Earlier work this paper cites.
Discriminative n-gram language modeling
Brian Roark, Murat Saraclar, and Michael Collins. 2007 · 2007
Earlier work this paper cites.
Statistical analysis of bayes optimal subset ranking
David Cossock and Tong Zhang. 2008 · 2008
Earlier work this paper cites.
Large-scale discriminative n-gram language models for statistical machine translation
Zhifei Li and Sanjeev Khudanpur. 2008 · 2008
Earlier work this paper cites.
Statistical machine translation
Philipp Koehn. 2009 · 2009
Earlier work this paper cites.
Learning to rank for information retrieval , volume 3
Tie-Yan Liu. 2009 · 2009
Cited alongside, same era.
Recurrent neural network based language model
Tomas Mikolov, Martin Karafiát, Lukas Burget, Jan Černocký, and Sanjeev Khudanpur. 2010 · 2010
Cited alongside, same era.
Improved neural network based language modelling and adaptation
Junho Park, Xunying Liu, Mark J. F. Gales, and Phil C. Woodland. 2010 · 2010
Cited alongside, same era.
Tuning as ranking
Mark Hopkins and Jonathan May. 2011 · 2011
Cited alongside, same era.
Lstm neural networks for language modeling
Martin Sundermeyer, Ralf Schlüter, and Hermann Ney. 2012 · 2012
Cited alongside, same era.
Classification and ranking approaches to discriminative language modeling for asr
Erinç Dikic, Murat Semerci, Murat Saraçlar, and Ethem Alpaydin. 2013 · 2013
Cited alongside, same era.
Eesen: End-to-end speech recognition using deep rnn models and wfst-based decoding, pages 167–174. ieee, 2015
Yajie Miao, Mohammad Gowayyed, and Florian Metze. 2015 · 2015
Later among the works it cites.
Discriminative method for recurrent neural network language models
Yuuki Tachioka and Shinji Watanabe. 2015 · 2015
Later among the works it cites.
Deep speech 2: End-to-end speech recognition in english and mandarin
Dario Amodei, Sundaram Ananthanarayanan, Rishita Anubhai, Jingliang Bai, Eric Battenberg, Carl Case, and Jared Casper et al. 2016 · 2016
Later among the works it cites.
End-to-end attention-based large vocabulary speech recognition
Dzmitry Bahdanau, Jan Chorowski, Dmitriy Serdyuk, Philemon Brakel, and Yoshua Bengio. 2016 · 2016
Later among the works it cites.
Exploring the limits of language modeling
Rafal Jozefowicz, Oriol Vinyals, Mike Schuster, Noam Shazeer, and Yonghui Wu. 2016 · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Decoder integration and expected bleu training for recurrent neural network language models
Michael Auli and Jianfeng Gao. 2014 · 2014
Cited alongside, same era.
Automatic Speech Recognition: A Deep Learning Approach
Dong Yu and Li Deng. 2014 · 2014
Cited alongside, same era.
Recurrent neural network regularization
Wojciech Zaremba, Ilya Sutskever, and Oriol Vinyals. 2014 · 2014
Cited alongside, same era.
Recurrent neural network language model adaptation for multi-genre broadcast speech recognition
Xie Chen, Tian Tan, Xunying Liu, Pierre Lanchantin, Moquan Wan, Mark J.F. Gales, and Philip C. Woodland. 2015 · 2015
Cited alongside, same era.
On using monolingual corpora in neural machine translation
Caglar Gulcehre, Orhan Firat, Kelvin Xu, Kyunghyun Cho, Loic Barrault, Huei-Chi Lin, Fethi Bougares, Holger Schwenk, and Yoshua Bengio. 2015 · 2015
Cited alongside, same era.
Minimum risk training for neural machine translation
Shiqi Shen, Yong Cheng, Zhongjun He, Wei He, Hua Wu, Maosong Sun, and Yang Liu. 2016 · 2016
Later among the works it cites.
Sequence-to-sequence learning as beam-search optimization
Sam Wiseman and Alexander M. Rush. 2016 · 2016
Later among the works it cites.
Gram-ctc: Automatic unit selection and target decomposition for sequence labelling
Hairong Liu, Zhenyao Zhu, Xiangang Li, and Sanjeev Satheesh. 2017 · 2017
Later among the works it cites.
Approaches for neural-network language model adaptation
Min Ma, Michael Nirschl, Fadi Biadsy, and Shankar Kumar. 2017 · 2017
Later among the works it cites.
Cold fusion: Training seq2seq models together with language models
Anuroop Sriram, Heewoo Jun, Sanjeev Satheesh, and Adam Coates. 2017 · 2017
Later among the works it cites.