Fetching the paper…
Reading the bibliography…
Multilingual models for Automatic Speech Recognition (ASR) are attractive as they have been shown to benefit from more training data, and better lend themselves to adaptation to under-resourced languages.
L. F. Lamel, J.-L. Gauvain, M. Eskénazi
1991
Earlier work this paper cites.
D. B. Paul and J. M. Baker, “The design for the Wall Street Journal-based CSR corpus,” in
1992
Earlier work this paper cites.
T. Schultz and A. Waibel, “Polyphone decision tree specialization for language adaptation,” in
2000
Earlier work this paper cites.
A. Graves, S. Fernández, F. Gomez, and J. Schmidhuber, “Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,” in
2006
Earlier work this paper cites.
K. Veselỳ, M. Karafiát, F. Grézl, M. Janda, and E. Egorova, “The language-independent bottleneck features,” in
2012
Earlier work this paper cites.
S. Thomas, S. Ganapathy, and H. Hermansky, “Multilingual MLP features for low-resource LVCSR systems,” in
2012
Earlier work this paper cites.
H. A. Bourlard and N. Morgan,
2012
Earlier work this paper cites.
G. Hinton, L. Deng, D. Yu, G. E. Dahl, A.-r. Mohamed, N. Jaitly, A. Senior, V. Vanhoucke, P. Nguyen, T. N. Sainath
2012
Earlier work this paper cites.
J.-T. Huang, J. Li, D. Yu, L. Deng, and Y. Gong, “Cross-language knowledge transfer using multilingual deep neural network with shared hidden layers,” in
2013
Earlier work this paper cites.
N. T. Vu and T. Schultz, “Multilingual multilayer perceptron for rapid language adaptation between and across language families,” in
2013
Earlier work this paper cites.
Z. Tüske, J. Pinto, D. Willett, and R. Schlüter, “Investigation on cross-and multilingual MLP features under matched and mismatched acoustical conditions,” in
2013
Cited alongside, same era.
A. Ghoshal, P. Swietojanski, and S. Renals, “Multilingual training of deep neural networks,” in
2013
Cited alongside, same era.
T. Schultz, N. T. Vu, and T. Schlippe, “GlobalPhone: A multilingual text & speech database in 20 languages,” in
2013
Cited alongside, same era.
M. J. F. Gales, K. M. Knill, A. Ragni, and S. P. Rath, “Speech recognition and keyword spotting for low resource languages: Babel project research at CUED,”
2014
Cited alongside, same era.
F. Grézl, M. Karafiát, and K. Veselỳ, “Adaptation of multilingual stacked bottle-neck neural network structure for new language,” in
2014
Cited alongside, same era.
——, “SAT-LHUC: Speaker adaptive training for learning hidden unit contributions,” in
2016
Later among the works it cites.
S. Semeniuta, A. Severyn, and E. Barth, “Recurrent dropout without memory loss,” 2016
2016
Later among the works it cites.
W. Xiong, J. Droppo, X. Huang, F. Seide, M. Seltzer, A. Stolcke, D. Yu, and G. Zweig, “The Microsoft 2016 conversational speech recognition system,” in
2017
Closest in time.
S. Tong, P. N. Garner, and H. Bourlard, “An investigation of deep neural networks for multilingual speech recognition training and adaptation,” in
2017
Closest in time.
S. Kim and M. L. Seltzer, “Towards language-universal end-to-end speech recognition,”
2017
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
P. Swietojanski and S. Renals, “Learning hidden unit contributions for unsupervised speaker adaptation of neural network acoustic models,” in
2014
Cited alongside, same era.
N. Srivastava, G. E. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov, “Dropout: a simple way to prevent neural networks from overfitting.”
2014
Cited alongside, same era.
2014
Cited alongside, same era.
H. Sak, A. Senior, K. Rao, O. Irsoy, A. Graves, F. Beaufays, and J. Schalkwyk, “Learning acoustic frame labeling for speech recognition with recurrent neural networks,” in
2015
Cited alongside, same era.
Y. Miao, M. Gowayyed, X. Na, T. Ko, F. Metze, and A. Waibel, “An empirical exploration of CTC acoustic models,” in
2016
Cited alongside, same era.
2017
Closest in time.
2017
Closest in time.