Fetching the paper…
Reading the bibliography…
We explore neural language modeling for speech recognition where the context spans multiple sentences.
“LSTM Language Models for LVCSR in First-Pass Decoding and Lattice-Rescoring”, 2019
Eugen Beck, Wei Zhou, Ralf Schlüter and Hermann Ney · 1907
Earlier work this paper cites.
“A Cache-Based Natural Language Model for Speech Recognition”
R. Kuhn and R. De · 1990
Earlier work this paper cites.
“Trigger-based Language Models: A Maximum Entropy Approach”
Raymond Lau, Ronald Rosenfeld and Salim Roukos · 1993
Earlier work this paper cites.
“Exploiting Latent Semantic Information in Statistical Language Modeling”
Jerome Bellegarda · 2000
Earlier work this paper cites.
“Language modeling for dialog system.”, 2000, pp. 118–121
Wei Xu and Alexander Rudnicky · 2000
Earlier work this paper cites.
“Multi-speaker Language Modeling”
Gang Ji and Jeff Bilmes · 2004
Earlier work this paper cites.
“Recurrent neural network based language model”
Tomas Mikolov et al · 2010
Earlier work this paper cites.
“LSTM Neural Networks for Language Modeling”
Martin Sundermeyer, Ralf Schlüter and Hermann Ney · 2012
Earlier work this paper cites.
“Noise-contrastive Estimation of Unnormalized Statistical Models, with Applications to Natural Image Statistics”
Michael. Gutmann and Aapo Hyvärinen · 2012
Earlier work this paper cites.
“Context dependent recurrent neural network language model”
Tomas Mikolov and Geoffrey Zweig · 2012
Cited alongside, same era.
“Efficient lattice rescoring using recurrent neural network language models”
X. Liu et al · 2014
Cited alongside, same era.
“Cache Based Recurrent Neural Network Language Model Inference for First Pass Speech Recognition”
Zhiheng Huang, Geoffrey Zweig and Benoit Dumoulin · 2014
Cited alongside, same era.
“Notes on Noise Contrastive Estimation and Negative Sampling”, 2014
Chris Dyer · 2014
Cited alongside, same era.
“Hierarchical Recurrent Neural Network for Document Modeling”
Rui Lin et al · 2015
Cited alongside, same era.
“Librispeech: An ASR corpus based on public domain audio books”
“Session-level Language Modeling for Conversational Speech”
Wayne Xiong, Lingfeng Wu, Jun Zhang and Andreas Stolcke · 2018
Later among the works it cites.
“Bert: Pre-training of deep bidirectional transformers for language understanding”
Jacob Devlin, Ming-Wei Chang, Kenton Lee and Kristina Toutanova · 2018
Later among the works it cites.
“Improving language understanding by generative pre-training”, 2018
Alec Radford, Karthik Narasimhan, Tim Salimans and Ilya Sutskever · 2018
Later among the works it cites.
“Training tips for the transformer model”
Martin Popel and Ondřej Bojar · 2018
Later among the works it cites.
“Sharp Nearby, Fuzzy Far Away: How Neural Language Models Use Context”
Urvashi Khandelwal, He He, Peng Qi and Dan Jurafsky · 2018
Later among the works it cites.
“Language modeling: Attention mechanisms for extending context-awareness of LSTM”, 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
V. Panayotov, G. Chen, D. Povey and S. Khudanpur · 2015
Cited alongside, same era.
“Exploring the limits of language modeling”, 2016
Rafal Jozefowicz et al · 2016
Cited alongside, same era.
“Effectively Building Tera Scale MaxEnt Language Models Incorporating Non-Linguistic Signals”
Fadi Biadsy, Mohammadreza Ghodsi and Diamantino Caseiro · 2017
Cited alongside, same era.
“Attention is all you need”
Ashish Vaswani et al · 2017
Cited alongside, same era.
URL: https://github.com/kaldi-asr/kaldi/tree/master/egs/librispeech
Cited in the paper.
Jurik Juraska, Sarangarajan Parthasarathy and William Gale · 2018
Later among the works it cites.
“Language Modeling with Deep Transformers”
Kazuki Irie, Albert Zeyer, Ralf Schlüter and Hermann Ney · 2019
Closest in time.
“Language modeling with deep Transformers”
Kazuki Irie, Albert Zeyer, Ralf Schluter and Hermann Ney · 2019
Closest in time.
“Transformer-XL: Attentive Language Models beyond a Fixed-Length Context”
Zihang Dai et al · 2019
Closest in time.