Fetching the paper…
Reading the bibliography…
In this paper, we present our overall efforts to improve the performance of a code-switching speech recognition system using semi-supervised training methods from lexicon learning to acoustic modeling, on the South East Asian Mandarin-English (SEAME) data.
A. Waibel, “Modular construction of time-delay neural networks for speech recognition,”
1989
Earlier work this paper cites.
A. Stolcke, “Srilm-an extensible language modeling toolkit,” in
2002
Earlier work this paper cites.
B. Mak and E. Barnard, “Phone clustering using the bhattacharyya distance,” in
2008
Earlier work this paper cites.
R. Hsiao, M. Fuhs, Y.-C. Tam, Q. Jin, and T. Schultz, “The cmu-interact 2008 mandarin transcription system,” in
2008
Earlier work this paper cites.
M. Bisani and H. Ney, “Joint-sequence models for grapheme-to-phoneme conversion,”
2008
Earlier work this paper cites.
H. Lin, L. Deng, D. Yu, Y.-f. Gong, A. Acero, and C.-H. Lee, “A study on multilingual acoustic modeling for large vocabulary asr,” in
2009
Earlier work this paper cites.
T. Mikolov, M. Karafiát, L. Burget, J. Černockỳ, and S. Khudanpur, “Recurrent neural network based language model,” in
2010
Earlier work this paper cites.
D.-C. Lyu, T.-P. Tan, E.-S. Chng, and H. Li, “An analysis of a mandarin-english code-switching speech corpus: Seame,”
2010
Earlier work this paper cites.
A. Laurent, S. Meignier, T. Merlin, and P. Deléglise, “Acoustics-based phonetic transcription method for proper nouns,” in
2010
Earlier work this paper cites.
Y. Li, P. Fung, P. Xu, and Y. Liu, “Asymmetric acoustic modeling of mixed language speech,” in
2011
Cited alongside, same era.
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlicek, Y. Qian, P. Schwarz
2011
Cited alongside, same era.
N. T. Vu, D.-C. Lyu, J. Weiner, D. Telaar, T. Schlippe, F. Blaicher, E.-S. Chng, T. Schultz, and H. Li, “A first speech recognition system for mandarin-english code-switch conversational speech,” in
2012
Cited alongside, same era.
D. Povey, M. Hannemann, G. Boulianne, L. Burget, A. Ghoshal, M. Janda, M. Karafiát, S. Kombrink, P. Motlíček, Y. Qian
2012
Cited alongside, same era.
T. Mikolov and G. Zweig, “Context dependent recurrent neural network language model.”
2012
Cited alongside, same era.
G. Saon, H. Soltau, D. Nahamoo, and M. Picheny, “Speaker adaptation of neural network acoustic models using i-vectors.” in
2013
Later among the works it cites.
K. Vesely, M. Hannemann, and L. Burget, “Semi-supervised training of deep neural networks,” in
2013
Later among the works it cites.
S. Thomas, M. L. Seltzer, K. Church, and H. Hermansky, “Deep neural network features and semi-supervised training for low resource speech recognition,” in
2013
Later among the works it cites.
P. Bell, M. J. Gales, T. Hain, J. Kilgour, P. Lanchantin, X. Liu, A. McParland, S. Renals, O. Saz, M. Wester
2015
Later among the works it cites.
V. Peddinti, D. Povey, and S. Khudanpur, “A time delay neural network architecture for efficient modeling of long temporal contexts,” in
2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
P. Auer,
2013
Cited alongside, same era.
J.-T. Huang, J. Li, D. Yu, L. Deng, and Y. Gong, “Cross-language knowledge transfer using multilingual deep neural network with shared hidden layers,” in
2013
Cited alongside, same era.
H. Adel, N. T. Vu, F. Kraus, T. Schlippe, H. Li, and T. Schultz, “Recurrent neural network language modeling for code switching conversational speech,” in
2013
Cited alongside, same era.
H. Adel, N. T. Vu, and T. Schultz, “Combination of recurrent neural networks and factored language models for code-switching language modeling,” in
2013
Cited alongside, same era.
“Cmu pronunciation dictionary for american english,” http://www.speech.cs.cmu.edu/cgi-bin/cmudict
Cited in the paper.
V. Manohar, H. Hadian, D. Povey, and S. Khudanpur, “Semi-supervised training of acoustic models using lattice-free mmi.”
Cited in the paper.
E. Yılmaz, H. van den Heuvel, and D. van Leeuwen, “Investigating bilingual deep neural networks for automatic recognition of code-switching frisian speech,”
2016
Later among the works it cites.
D. Povey, V. Peddinti, D. Galvez, P. Ghahremani, V. Manohar, X. Na, Y. Wang, and S. Khudanpur, “Purely sequence-trained neural networks for asr based on lattice-free mmi.” in
2016
Later among the works it cites.
2017
Later among the works it cites.
H. Xu, K. Li, Y. Wang, J. Wang, S. Kang, X. Chen, D. Povey, and S. Khudanpur, “Neural network language modeling with letter-based features and importance sampling,” in
2018
Closest in time.