M. J. F. Gales, “Maximum likelihood linear transformations for HMM-based speech recognition,”
1997
Earlier work this paper cites.
J. F. William, W. M. Fisher, A. F. Martin, M. A. Przybocki, and D. S. Pallett, “NIST evaluation of conversational speech recognition over the telephone: English and mandarin performance results,” in
2000
Earlier work this paper cites.
P. Thompson and H. Nesi, “Research in progress, the British Academic Spoken English (BASE) corpus project.”
2001
Earlier work this paper cites.
A. Stolcke, “SRILM – An extensible language modeling toolkit,” in
2002
Earlier work this paper cites.
R. C. Simpson, S. L. Briggs, J. Ovens, and J. M. Swales, “The Michigan Corpus of Academic Spoken English.”
2002
Earlier work this paper cites.
A. Krizhevsky and G. Hinton, “Learning multiple layers of features from tiny images,” in
2009
Earlier work this paper cites.
S. Zagoruyko and N. Komodakis, “Wide residual networks,” in
Original
2009
Earlier work this paper cites.
B. Li and K. C. Sim, “Comparison of discriminative input and output transformations for speaker adaptation in the hybrid NN/HMM systems,” in
2010
Earlier work this paper cites.
N. Dehak, P. Kenny, R. Dehak, P. Dumouchel, and P. Ouellet, “Front-end factor analysis for speaker verification,”
2011
Earlier work this paper cites.
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlicek, Y. Qian, P. Schwarz, J. Silovsky, G. Stemmer, and K. Vesely, “The Kaldi speech recognition toolkit,” in
2011
Earlier work this paper cites.
H. Xu, D. Povey, L. Mangu, and J. Zhu, “Minimum bayes risk decoding and system combination based on a recursion for edit distance,”
2011
Earlier work this paper cites.
J. Bergstra, R. Bardenet, Y. Bengio, and B. Kégl, “Algorithms for hyper-parameter optimization,” in
2011
Earlier work this paper cites.
A. Rousseau, P. Deléglise, and Y. Esteve, “TED-LIUM: an automatic speech recognition dedicated corpus.” in
2012
Earlier work this paper cites.
A. Rousseau, P. Deléglise, and Y. Estève, “TED-LIUM: an automatic speech recognition dedicated corpus,” in
2012
Earlier work this paper cites.