“Subphonetic modeling with markov states-senone,”
M.-Y. Hwang and X. Huang, · 1992
Earlier work this paper cites.
“Learning long-term dependencies with gradient descent is difficult,”
Y. Bengio, P. Simard, et al., · 1994
Earlier work this paper cites.
“Long short-term memory,”
S. Hochreiter and J. Schmidhuber, · 1997
Earlier work this paper cites.
“Conversational speech transcription using context-dependent deep neural networks,”
F. Seide, G. Li, and D. Yu, · 2011
Earlier work this paper cites.
“The kaldi speech recognition toolkit,”
D. Povey, A. Ghoshal, G. Boulianne, et al., · 2011
Earlier work this paper cites.
“Deep neural networks for acoustic modeling in speech recognition,”
G. Hinton, L. Deng, D. Yu, et al., · 2012
Earlier work this paper cites.
“Sequence-discriminative training of deep neural networks.,”
K. Veselỳ, A. Ghoshal, L. Burget, and D. Povey, · 2013
Earlier work this paper cites.
“Long short-term memory recurrent neural network architectures for large scale acoustic modeling,”
H. Sak, A. Senior, and F. Beaufays, · 2014
Earlier work this paper cites.
“Convolutional neural networks for speech recognition,”
O. Abdel-Hamid, A. Mohamed, H. Jiang, et al., · 2014
Earlier work this paper cites.
“Adam: A method for stochastic optimization,”
Original
D. P. Kingma and J. Ba, · 2014
Earlier work this paper cites.
“Very deep convolutional networks for large-scale image recognition,”
Original
K. Simonyan and A. Zisserman, · 2014
Earlier work this paper cites.
“A time delay neural network architecture for efficient modeling of long temporal contexts,”
V. Peddinti, D. Povey, and S. Khudanpur, · 2015
Earlier work this paper cites.
“Feedforward sequential memory neural networks without recurrent feedback,”
Original
S. Zhang, H. Jiang, S. Wei, and L. Dai, · 2015
Earlier work this paper cites.
“Librispeech: an asr corpus based on public domain audio books,”
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, · 2015
Earlier work this paper cites.