Fetching the paper…
Reading the bibliography…
Automatic speech recognition (ASR) systems often make unrecoverable errors due to subsystem pruning (acoustic, language and pronunciation models); for example pruning words due to acoustics using short-term context, prior to rescoring with long-term context based on linguistics.
“A tutorial on hidden Markov models and selected applications in speech recognition,”
Lawrence R Rabiner, · 1989
Earlier work this paper cites.
“A tree-based statistical language model for natural language speech recognition,”
Lalit R Bahl, Peter F Brown, Peter V de Souza, and Robert L Mercer, · 1989
Earlier work this paper cites.
“Some statistical issues in the comparison of speech recognition algorithms,”
Laurence Gillick and Stephen J Cox, · 1989
Earlier work this paper cites.
“A statistical approach to machine translation,”
Peter F Brown, John Cocke, Stephen A Della Pietra, Vincent J Della Pietra, Fredrick Jelinek, John D Lafferty, Robert L Mercer, and Paul S Roossin, · 1990
Earlier work this paper cites.
“Feedback strategies for error correction in speech recognition systems,”
William A Ainsworth and SR Pratt, · 1992
Earlier work this paper cites.
Fundamentals of Speech Recognition
Lawrence Rabiner and Biing-Hwang Juang, · 1993
Earlier work this paper cites.
“Errors and error correction in automatic speech recognition systems,”
JM Noyes and CR Frankish, · 1994
Earlier work this paper cites.
“Combining linguistic and statistical knowledge sources in natural-language processing for ATIS,”
Robert Moore, Douglas Appelt, John Dowding, J Mark Gawron, and Douglas Moran, · 1995
Earlier work this paper cites.
“Error correction via a post-processor for continuous speech recognition,”
Eric K Ringger and James F Allen, · 1996
Earlier work this paper cites.
“The CMU pronouncing dictionary,”
Robert L Weide, · 1998
Earlier work this paper cites.
“Modeling pronunciation variation for ASR: A survey of the literature,”
Helmer Strik and Catia Cucchiarini, · 1999
Earlier work this paper cites.
“Phrase-based language models for speech recognition,”
Hong-Kwang Jeff Kuo and Wolfgang Reichl, · 1999
Earlier work this paper cites.
“Two decades of statistical language modeling: Where do we go from here?,”
Ronald Rosenfeld, · 2000
Earlier work this paper cites.
“Multimodal error correction for speech user interfaces,”
Bernhard Suhm, Brad Myers, and Alex Waibel, · 2001
Earlier work this paper cites.
“Large scale discriminative training of hidden Markov models for speech recognition,”
P.C. Woodland and D. Povey, · 2002
Earlier work this paper cites.
“SRILM-an extensible language modeling toolkit,”
Andreas Stolcke, · 2002
Earlier work this paper cites.
“BLEU: a method for automatic evaluation of machine translation,”
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu, · 2002
Earlier work this paper cites.
“Pronunciation modeling for ASR–knowledge-based and data-derived methods,”
Mirjam Wester, · 2003
Earlier work this paper cites.
“Statistical phrase-based translation,”
Philipp Koehn, Franz Josef Och, and Daniel Marcu, · 2003
Earlier work this paper cites.
“Minimum error rate training in statistical machine translation,”
Franz Josef Och, · 2003
Earlier work this paper cites.
“A systematic comparison of various statistical alignment models,”
Franz Josef Och and Hermann Ney, · 2003
Cited alongside, same era.
“Context-based speech recognition error detection and correction,”
Arup Sarma and David D Palmer, · 2004
Cited alongside, same era.
“Speech recognition error correction using maximum entropy language model,”
Minwoo Jeong, Sangkeun Jung, and Gary Geunbae Lee, · 2004
Cited alongside, same era.
“The Fisher Corpus: a resource for the next generations of speech-to-text.,”
Christopher Cieri, David Miller, and Kevin Walker, · 2004
Cited alongside, same era.
“Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,”
Alex Graves, Santiago Fernández, Faustino Gomez, and Jürgen Schmidhuber, · 2006
Cited alongside, same era.
“Discriminative n-gram language modeling,”
“Performance comparison of training algorithms for semi-supervised discriminative language modeling,”
Erinç Dikici, Arda Celebi, and Murat Saraçlar, · 2012
Later among the works it cites.
“Speech recognition with deep recurrent neural networks,”
Alex Graves, Abdel-rahman Mohamed, and Geoffrey Hinton, · 2013
Later among the works it cites.
“Statistical error correction methods for domain-specific ASR systems,”
Horia Cucu, Andi Buzo, Laurent Besacier, and Corneliu Burileanu, · 2013
Later among the works it cites.
“Two-step correction of speech recognition errors based on n-gram and long contextual information.,”
Ryohei Nakatani, Tetsuya Takiguchi, and Yasuo Ariki, · 2013
Later among the works it cites.
“Decoding with large-scale neural language models improves translation.,”
Ashish Vaswani, Yinggong Zhao, Victoria Fossum, and David Chiang, · 2013
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Brian Roark, Murat Saraclar, and Michael Collins, · 2007
Cited alongside, same era.
“Moses: Open source toolkit for statistical machine translation,”
Philipp Koehn, Hieu Hoang, Alexandra Birch, Chris Callison-Burch, Marcello Federico, Nicola Bertoldi, Brooke Cowan, Wade Shen, Christine Moran, Richard Zens, et al., · 2007
Cited alongside, same era.
“Improved minimum error rate training in Moses,”
Nicola Bertoldi, Barry Haddow, and Jean-Baptiste Fouet, · 2009
Cited alongside, same era.
“Recurrent neural network based language model.,”
Tomas Mikolov, Martin Karafiát, Lukas Burget, Jan Cernockỳ, and Sanjeev Khudanpur, · 2010
Cited alongside, same era.
“Automatic speech recognition system channel modeling,”
Qun Feng Tan, Kartik Audhkhasi, Panayiotis G Georgiou, Emil Ettelaie, and Shrikanth S Narayanan, · 2010
Cited alongside, same era.
“The Kaldi speech recognition toolkit,”
Daniel Povey, Arnab Ghoshal, Gilles Boulianne, Lukas Burget, Ondrej Glembek, Nagendra Goel, Mirko Hannemann, Petr Motlicek, Yanmin Qian, Petr Schwarz, et al., · 2011
Cited alongside, same era.
“Training of error-corrective model for ASR without using audio data,”
Gakuto Kurata, Nobuyasu Itoh, and Masafumi Nishimura, · 2011
Cited alongside, same era.
Daniel M Bikel and Keith B Hall, · 2013
Later among the works it cites.
“Improving speech recognition for children using acoustic adaptation and pronunciation modeling.,”
Prashanth Gurunath Shivakumar, Alexandros Potamianos, Sungbok Lee, and Shrikanth Narayanan, · 2014
Later among the works it cites.
“Sequence to sequence learning with neural networks,”
Ilya Sutskever, Oriol Vinyals, and Quoc V Le, · 2014
Later among the works it cites.
“Asr error detection using recurrent neural network language model and complementary ASR,”
Yik-Cheung Tam, Yun Lei, Jing Zheng, and Wen Wang, · 2014
Later among the works it cites.
“Error correction of automatic speech recognition based on normalized web distance,”
E Byambakhishig, Katsuyuki Tanaka, Ryo Aihara, Toru Nakashika, Tetsuya Takiguchi, and Yasuo Ariki, · 2014
Later among the works it cites.
“Fast and robust neural network joint models for statistical machine translation.,”
Jacob Devlin, Rabih Zbib, Zhongqiang Huang, Thomas Lamar, Richard M Schwartz, and John Makhoul, · 2014
Later among the works it cites.
“Enhancing the TED-LIUM corpus with selected data for language modeling and more TED talks.,”
Anthony Rousseau, Paul Deléglise, and Yannick Estève, · 2014
Later among the works it cites.
“On the properties of neural machine translation: Encoder-decoder approaches,”
Kyunghyun Cho, Bart Van Merriënboer, Dzmitry Bahdanau, and Yoshua Bengio, · 2014
Later among the works it cites.
“Word-error correction of continuous speech recognition based on normalized relevance distance.,”
Yohei Fusayasu, Katsuyuki Tanaka, Tetsuya Takiguchi, and Yasuo Ariki, · 2015
Later among the works it cites.
“Investigations on phrase-based decoding with recurrent neural network language and translation models,”
Tamer Alkhouli, Felix Rietig, and Hermann Ney, · 2015
Later among the works it cites.
“End-to-end attention-based large vocabulary speech recognition,”
Dzmitry Bahdanau, Jan Chorowski, Dmitriy Serdyuk, Philemon Brakel, and Yoshua Bengio, · 2016
Later among the works it cites.
“Listen, attend and spell: A neural network for large vocabulary conversational speech recognition,”
William Chan, Navdeep Jaitly, Quoc Le, and Oriol Vinyals, · 2016
Later among the works it cites.
“Automatic correction of ASR outputs by using machine translation,”
Luis Fernando D’Haro and Rafael E Banchs, · 2016
Later among the works it cites.
“Directed automatic speech transcription error correction using bidirectional LSTM,”
D. Zheng, Z. Chen, Y. Wu, and K. Yu, · 2016
Later among the works it cites.
“Optimization for statistical machine translation: A survey,”
Graham Neubig and Taro Watanabe, · 2016
Later among the works it cites.