“Low latency acoustic modeling using temporal convolution and lstms,”
V. Peddinti, Y. Wang, D. Povey, and S. Khudanpur, · 2018
Later among the works it cites.
“Semi-orthogonal low-rank matrix factorization for deep neural networks,”
D. Povey, G. Cheng, Y. Wang, K. Li, H. Xu, M. Yarmohamadi, and S. Khudanpur, · 2018
Later among the works it cites.
“End-to-end speech recognition using lattice-free mmi,”
H. Hadian, H. Sameti, D. Povey, and S. Khudanpur, · 2018
Later among the works it cites.
“Phonetic and graphemic systems for multi-genre broadcast transcription,”
Y. Wang, X. Chen, M. Gales, A. Ragni, and J. Wong, · 2018
Later among the works it cites.
“State-of-the-art speech recognition with sequence-to-sequence models,”
C. Chiu, T. N. Sainath, Y. Wu, R. Prabhavalkar, P. Nguyen, Z. Chen, A. Kannan, R. J. Weiss, K. Rao, E. Gonina, N. Jaitly, B. Li, J. Chorowski, and M. Bacchiani, · 2018
Later among the works it cites.
“Improved training of end-to-end attention models for speech recognition,”
A. Zeyer, K. Irie, R. Schlüter, and H. Ney, · 2018
Later among the works it cites.
“No need for a lexicon? evaluating the value of the pronunciation lexica in end-to-end models,”
T. Sainath, R. Prabhavalkar, S. Kumar, S. Lee, A. Kannan, D. Rybach, V. Schogol, P. Nguyen, B. Li, Y. Wu, et al., · 2018
Later among the works it cites.
“Building competitive direct acoustics-to-word models for english conversational speech recognition,”
K. Audhkhasi, B. Kingsbury, B. Ramabhadran, G. Saon, and M. Picheny, · 2018
Later among the works it cites.
“Advancing acoustic-to-word ctc model,”
J. Li, G. Ye, A. Das, R. Zhao, and Y. Gong, · 2018
Later among the works it cites.
“Acoustic-to-word attention-based model complemented with character-level ctc-based model,”
S. Ueno, H. Inaguma, M. Mimura, and T. Kawahara, · 2018
Later among the works it cites.
“The CAPIO 2017 conversational speech recognition system,”
Original
Kyu J. Han, Akshay Chandrashekaran, Jungsuk Kim, and Ian R. Lane, · 2018
Later among the works it cites.
“Streaming end-to-end speech recognition for mobile devices,”
Y. He, T. N. Sainath, R. Prabhavalkar, I. McGraw, R. Alvarez, D. Zhao, D. Rybach, A. Kannan, Y. Wu, R. Pang, Q. Liang, D. Bhatia, Y. Shangguan, B. Li, G. Pundak, K. C. Sim, T. Bagby, S. Chang, K. Rao, and A. Gruenstein, · 2019
Closest in time.
“Specaugment: A simple data augmentation method for automatic speech recognition,”
D. Park, W. Chan, Y. Zhang, C. Chiu, B. Zoph, E. Cubuk, and Q. Le, · 2019
Closest in time.
“RWTH ASR Systems for LibriSpeech: Hybrid vs Attention,”
C. Lüscher, E. Beck, K. Irie, M. Kitza, W. Michel, A. Zeyer, R. Schlüter, and H. Ney, · 2019
Closest in time.
“Sequence-to-sequence speech recognition with time-depth separable convolutions,”
A. Hannun, A. Lee, Q. Xu, and R. Collobert, · 2019
Closest in time.
“Jasper: An end-to-end convolutional neural acoustic model,”
J. Li, V. Lavrukhin, B. Ginsburg, R. Leary, O. Kuchaiev, J. M. Cohen, H. Nguyen, and R. T. Gadde, · 2019
Closest in time.