Fetching the paper…
Reading the bibliography…
In this paper, we demonstrate the efficacy of transfer learning and continuous learning for various automatic speech recognition (ASR) tasks.
D. B. Paul and J. M. Baker, “The design for the Wall Street Journal based CSR corpus,” in
1992
Earlier work this paper cites.
J. J. Godfrey and E. Holliman, “Switchboard-1 release 2 LDC97S62,” 1993
1993
Earlier work this paper cites.
O. Anderson, P. Dalsgaard, and W. Barry, “On the use of data-driven clustering technique for identification of poly- and mono-phonemes for four european languages,” in
1994
Earlier work this paper cites.
T. Schultz and A. H. Waibel, “Language-independent and language-adaptive acoustic modeling for speech recognition,”
2001
Earlier work this paper cites.
C. Cieri, D. Miller, and K. Walker, “The Fisher corpus: a resource for the next generations of speech-to-text,” in
2004
Earlier work this paper cites.
A. Graves, S. Fernández, F. Gomez, and J. Schmidhuber, “Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,” in
2006
Earlier work this paper cites.
P. Swietojanski, A. Ghoshal, and S. Renals, “Unsupervised cross-lingual knowledge transfer in DNN-based LVCSR,” in
2012
Earlier work this paper cites.
J.-T. Huang, J. Li, D. Yu, L. Deng, and Y. Gong, “Cross-language knowledge transfer using multilingual deep neural network with shared hidden layers,” in
2013
Earlier work this paper cites.
A. Ghoshal, P. Swietojanski, and S. Renals, “Multilingual training of deep neural networks,” in
2013
Cited alongside, same era.
D. Wang and T. F. Zheng, “Transfer learning for speech and language processing,” in
2015
Cited alongside, same era.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: an ASR corpus based on public domain audio books,” in
2015
Cited alongside, same era.
W. Chan, N. Jaitly, Q. V. Le, and O. Vinyals, “Listen, attend and spell,” in
2016
Cited alongside, same era.
J. Kunze, L. Kirsch, I. Kurenkov, A. Krug, J. Johannsmeier, and S. Stober, “Transfer learning for speech recognition on a budget,” in
2017
Cited alongside, same era.
D. Bukhari, Y. Wang, and H. Wang, “Multilingual convolutional, long short-term memory, deep neural networks for low resource speech recognition,”
S. Ueno, T. Moriya, M. Mimura, S. Sakai, Y. Shinohara, Y. Yamaguchi, Y. Aono, and T. Kawahara, “Encoder transfer for attention-based acoustic-to-word speech recognition,” in
2018
Later among the works it cites.
T. Moriya, R. Masumura, T. Asami, Y. Shinohara, M. Delcroix, Y. Yamaguchi, and Y. Aono, “Progressive neural network-based knowledge transfer in acoustic models,” in
2018
Later among the works it cites.
T. Kendall and C. Farrington, “The Corpus of Regional African American Language,” Oct 2018. [Online]. Available:
2018
Later among the works it cites.
2019
Later among the works it cites.
J. X. Koh, A. Mislan, K. Khoo, B. Ang, W. Ang, C. Ng, and Y.-Y. Tan, “Building the Singapore English National Speech Corpus,” in
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
S. Tong, P. Garner, and H. Bourlard, “Multilingual training and cross-lingual adaptation on CTC-based acoustic model,”
2017
Cited alongside, same era.
F. Chollet, “Xception: Deep learning with depthwise separable convolutions,” in
2017
Cited alongside, same era.
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
S. Kriman, S. Beliaev, B. Ginsburg, J. Huang, O. Kuchaiev, V. Lavrukhin, R. Leary, J. Li, and Y. Zhang, “QuartzNet: Deep automatic speech recognition with 1D time-channel separable convolutions,”
2020
Closest in time.