Fetching the paper…
Reading the bibliography…
Recently, data-driven based Automatic Speech Recognition (ASR) systems have achieved state-of-the-art results.
C. Stein, “Inadmissibility of the usual estimator for the mean of a multivariate normal distribution,” STANFORD UNIVERSITY STANFORD United States, Tech. Rep., 1956
1956
Earlier work this paper cites.
L. O. Chua and L. Yang, “Cellular neural networks: Theory,”
1988
Earlier work this paper cites.
J. J. Godfrey, E. C. Holliman, and J. McDaniel, “SWITCHBOARD: Telephone speech corpus for research and development,” in
1992
Earlier work this paper cites.
R. Hecht-Nielsen, “Theory of the backpropagation neural network,” in
1992
Earlier work this paper cites.
H. J. Sussmann, “Uniqueness of the weights for minimal feedforward nets with a given input-output map,”
1992
Earlier work this paper cites.
V. Valtchev, J. Odell, P. C. Woodland, and S. J. Young, “Lattice-based discriminative training for large vocabulary speech recognition,” in
1996
Earlier work this paper cites.
R. Caruana, “Multitask learning,”
1997
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,”
1997
Earlier work this paper cites.
R. M. French, “Catastrophic forgetting in connectionist networks,”
1999
Earlier work this paper cites.
M. Paulik, C. Fügen, S. Stüker, T. Schultz, T. Schaaf, and A. Waibel, “Document driven machine translation enhanced ASR,” in
2005
Earlier work this paper cites.
A. Graves, S. Fernández, F. J. Gomez, and J. Schmidhuber, “Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,” in
2006
Cited alongside, same era.
D. Yu, L. Deng, and G. Dahl, “Roles of pre-training and fine-tuning in context-dependent dbn-hmms for real-world speech recognition,” in
2010
Cited alongside, same era.
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlicek, Y. Qian, P. Schwarz
2011
Cited alongside, same era.
K. Veselỳ, A. Ghoshal, L. Burget, and D. Povey, “Sequence-discriminative training of deep neural networks.” in
2013
Cited alongside, same era.
T. N. Sainath, O. Vinyals, A. W. Senior, and H. Sak, “Convolutional, Long Short-Term Memory, fully connected Deep Neural Networks,” in
2015
Cited alongside, same era.
D. Povey, V. Peddinti, D. Galvez, P. Ghahremani, V. Manohar, X. Na, Y. Wang, and S. Khudanpur, “Purely Sequence-Trained Neural Networks for ASR Based on Lattice-Free MMI.” in
2016
Later among the works it cites.
G. Pundak and T. N. Sainath, “Lower Frame Rate Neural Network Acoustic Models,” in
2016
Later among the works it cites.
M. Abadi, P. Barham, J. Chen
2016
Later among the works it cites.
J. Kirkpatrick, R. Pascanu, N. Rabinowitz, J. Veness, G. Desjardins, A. A. Rusu, K. Milan, J. Quan, T. Ramalho, A. Grabska-Barwinska
2017
Later among the works it cites.
G. Saon and M. Picheny, “Recent advances in conversational speech recognition using convolutional and recurrent neural networks,”
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: An ASR corpus based on public domain audio books,” in
2015
Cited alongside, same era.
D. P. Kingma and J. Ba, “Adam: A Method for Stochastic Optimization,” in
2015
Cited alongside, same era.
D. Amodei, S. Ananthanarayanan, R. Anubhai, J. Bai, E. Battenberg, C. Case, J. Casper, B. Catanzaro, Q. Cheng, G. Chen
2016
Cited alongside, same era.
2016
Cited alongside, same era.
W. Xiong, J. Droppo, X. Huang, F. Seide, M. Seltzer, A. Stolcke, D. Yu, and G. Zweig, “The Microsoft 2016 conversational speech recognition system,” in
2017
Later among the works it cites.
W. Hartmann, R. Hsiao, T. Ng, J. Ma, F. Keith, and M.-H. Siu, “Improved Single System Conversational Telephone Speech Recognition with VGG Bottleneck Features,” in
2017
Later among the works it cites.
J. Yoon, E. Yang, J. Lee, and S. J. Hwang, “Lifelong Learning with Dynamically Expandable Networks,” in
2018
Later among the works it cites.
K. Audhkhasi, B. Kingsbury, B. Ramabhadran, G. Saon, and M. Picheny, “Building Competitive Direct Acoustics-to-Word Models for English Conversational Speech Recognition,” in
2018
Later among the works it cites.