Fetching the paper…
Reading the bibliography…
Multilingual end-to-end (E2E) models have shown great promise in expansion of automatic speech recognition (ASR) coverage of the world's languages.
C. Fugen, S. Stuker, H. Soltau, F. Metze, and T. Schultz, “Efficient handling of multilingual language models,” in
2003
Earlier work this paper cites.
A. Graves, S. Fernández, F. Gomez, and J. Schmidhuber, “Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,” in
2006
Earlier work this paper cites.
B. Kingsbury, “Lattice-based optimization of sequence classification criteria for neural-network acoustic modeling,” in
2009
Earlier work this paper cites.
I. Garcia-Moral, R. Solera-Urena, C. Palaez-Moreno, and F. Diaz-de Maria, “Data balancing for efficient training of hybrid ann/hmm automatic speech recognition systems,” in
2011
Earlier work this paper cites.
S. Thomas, S. Ganapathy, and H. Hermansky, “Multilingual mlp features for low resource lvcsr systems,” in
2012
Earlier work this paper cites.
A. Graves, “Sequence transduction with recurrent neural networks,” in
2012
Earlier work this paper cites.
Z. Tuske, J. Pinto, D. Willett, and R. Schluter, “Investigation on cross and multilingual mlp features under matched and mismatched acoustical conditions,” in
2013
Earlier work this paper cites.
A. Ghoshal, P. Swietojanski, and S. Renals, “Multilingual training of deep neural networks,” in
2013
Earlier work this paper cites.
G. Heigold, V. Vanhoucke, A. Senior, P. Nguyen, M. Ranzato, M. Devin, and J. Dean, “Multilingual acoustic models using distributed deep neural networks,” in
2013
Earlier work this paper cites.
P. Swietojanski and S. Renals, “Learning hidden unit contributions for unsupervised speaker adaptation of neural network acoustic models,”
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
J. Cui, B. Kingsbury, B. Ramabhadran, and et al, “Multilingual representations for low resource speech recognition and keyword search,” in
2015
Earlier work this paper cites.
D. Chen and B. Mak, “Multitask learning of deep neural networks for low-resource speech recognition,” in
2015
Cited alongside, same era.
W. Chan, N. Jaitly, Q. V. Le, and O. Vinyals, “Listen, attend and spell,”
2015
Cited alongside, same era.
T. Alumae, S. Tsakalidis, and R. Schwartz, “Improved multilingual training of stacked neural network acoustic models for low resource languages,” in
2016
Cited alongside, same era.
T. Sercu, C. Puhrsch, B. Kingsbury, and Y. LeCun, “Very deep convolutional neural networks for multilingual lvcsr,” in
2016
Cited alongside, same era.
T. Tan, Y. Qian, and K. Yu, “Cluster adaptive training for deep neural network based acoustic model,” in
2016
Cited alongside, same era.
2017
Later among the works it cites.
S. Toshniwal, T. Sainath, R. Weiss, B. Li, P. Moreno, E. Weinstein, and K. Rao, “Multilingual speech recognition with a single end-to-end model,” in
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. Sercu, G. Saon, J. Cui, B. Ramabhadran, B. Kingsbury, and A. Sethy, “Network architectures for multilingual speech representation learning,” in
2017
Cited alongside, same era.
J. Cui, B. Kingsbury, B. Ramabhadran, and et al, “Knowledge distillation across ensembles of multilingual models for low-resource languages,” in
2017
Cited alongside, same era.
S. Watanabe, T. Hori, and J. Hershey, “Language independent end-to-end architecture for joint language identification and speech recognition,” in
2017
Cited alongside, same era.
S. Rebuffi, H. Bilen, and A. Vedaldi, “Learning multiple visual domains with residual adapters,” in
2017
Cited alongside, same era.
K. Rao, H. Sak, and R. Prabhavalkar, “Exploring architectures, data and units for streaming end-to-end speech recognition with rnn-transducer,” in
2017
Cited alongside, same era.
2017
Cited alongside, same era.
C. Kim, A. Misra, K. Chin, T. Hughes, A. Narayanan, T. N. Sainath, and M. Bacchiani, “Generated of large-scale simulated utterances in virtual rooms to train deep-neural networks for far-field speech recognition in Google Home,” in
2017
Cited alongside, same era.
2018
Later among the works it cites.
M. Grace, M. Bastani, and E. Weinstein, “Occam’s adaptation: A comparison of interpolation of bases adaptation methods for multi-dialect acoustic modeling with lstms.” in
2018
Later among the works it cites.
J. Emond, B. Ramabhadran, B. Roark, P. Moreno, and M. Ma, “Transliteration based approaches to improve code-switched speech recognition performance,” in
2018
Later among the works it cites.
Y. He, T. N. Sainath, R. Prabhavalkar, I. McGraw, R. Alvarez, D. Zhao, D. Rybach, A. Kannan, Y. Wu, R. Pang, Q. Liang, D. Bhatia, Y. Shangguan, B. Li, G. Pundak, K. Sim, T. Bagby, S. Chang, K. Rao, and A. Gruenstein, “Streaming end-to-end speech recognition for mobile devices,” in
2019
Closest in time.
N. Houlsby, A. Giurgiu, S. Jastrze¸bski, B. Morrone, Q. de Laroussilhe, A. Gesmundo, M. Attariyan, and S. Gelly, “Parameter-efficient transfer learning for nlp.” in
2019
Closest in time.
Anonymous authors., “Simple, scalable adaptation for neural machine translation,” in
2019
Closest in time.
J. Shen, P. Nguyen, Y. Wu, Z. Chen
2019
Closest in time.