Fetching the paper…
Reading the bibliography…
Unsupervised domain adaptation of speech signal aims at adapting a well-trained source-domain acoustic model to the unlabeled data from target domain.
“Learning small-size DNN with output-distribution-based criteria.,”
Jinyu Li, Rui Zhao, Jui-Ting Huang, and Yifan Gong, · 1914
Earlier work this paper cites.
“Linear hidden transformations for adaptation of hybrid ann/hmm models,”
Roberto Gemello, Franco Mana, Stefano Scanzio, Pietro Laface, and Renato De Mori, · 2007
Earlier work this paper cites.
“Conversational speech transcription using context-dependent deep neural networks,”
Frank Seide, Gang Li, and Dong Yu, · 2011
Earlier work this paper cites.
“Making deep belief networks effective for large vocabulary continuous speech recognition,”
Tara N Sainath, Brian Kingsbury, Bhuvana Ramabhadran, Petr Fousek, Petr Novak, and Abdel-rahman Mohamed, · 2011
Earlier work this paper cites.
“Feature engineering in context-dependent deep neural networks for conversational speech transcription,”
F. Seide, G. Li, X. Chen, and D. Yu, · 2011
Earlier work this paper cites.
“Application of pretrained deep neural networks to large vocabulary speech recognition,”
Navdeep Jaitly, Patrick Nguyen, Andrew Senior, and Vincent Vanhoucke, · 2012
Earlier work this paper cites.
“Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups,”
Geoffrey Hinton, Li Deng, Dong Yu, George E Dahl, Abdel-rahman Mohamed, Navdeep Jaitly, Andrew Senior, Vincent Vanhoucke, Patrick Nguyen, Tara N Sainath, et al., · 2012
Earlier work this paper cites.
“Improving wideband speech recognition using mixed-bandwidth training data in cd-dnn-hmm,”
Jinyu Li, Dong Yu, Jui-Ting Huang, and Yifan Gong, · 2012
Earlier work this paper cites.
“Recent advances in deep learning for speech research at microsoft,”
Li Deng, Jinyu Li, Jui-Ting Huang, Kaisheng Yao, Dong Yu, Frank Seide, Michael Seltzer, Geoff Zweig, Xiaodong He, Jason Williams, et al., · 2013
Earlier work this paper cites.
“Kl-divergence regularized deep neural network adaptation for improved large vocabulary speech recognition,”
D. Yu, K. Yao, H. Su, G. Li, and F. Seide, · 2013
Earlier work this paper cites.
“Speaker adaptation of context dependent deep neural networks,”
H. Liao, · 2013
Earlier work this paper cites.
“Restructuring of deep neural network acoustic models with singular value decomposition.,”
Jian Xue, Jinyu Li, and Yifan Gong, · 2013
Earlier work this paper cites.
“An overview of noise-robust automatic speech recognition,”
J. Li, L. Deng, Y. Gong, and R. Haeb-Umbach, · 2014
Cited alongside, same era.
“Singular value decomposition based low-footprint speaker adaptation and personalization for deep neural network,”
J. Xue, J. Li, D. Yu, M. Seltzer, and Y. Gong, · 2014
Cited alongside, same era.
“Generative adversarial nets,”
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio, · 2014
Cited alongside, same era.
“An introduction to computational networks and the computational network toolkit,”
Dong Yu, Adam Eversole, Mike Seltzer, Kaisheng Yao, Zhiheng Huang, Brian Guenter, Oleksii Kuchaiev, Yu Zhang, Frank Seide, Huaming Wang, et al., · 2014
Cited alongside, same era.
“Long short-term memory recurrent neural network architectures for large scale acoustic modeling,”
H. Sak, A. Senior, and F. Beaufays, · 2014
Cited alongside, same era.
“Low-rank plus diagonal adaptation for deep neural networks,”
Y. Zhao, J. Li, and Y. Gong, · 2016
Later among the works it cites.
“Learning hidden unit contributions for unsupervised acoustic model adaptation,”
P. Swietojanski, J. Li, and S. Renals, · 2016
Later among the works it cites.
“Adversarial multi-task learning of deep neural networks for robust speech recognition.,”
Yusuke Shinohara, · 2016
Later among the works it cites.
“Invariant representations for noisy speech recognition,”
Dmitriy Serdyuk, Kartik Audhkhasi, Philémon Brakel, Bhuvana Ramabhadran, Samuel Thomas, and Yoshua Bengio, · 2016
Later among the works it cites.
“Domain separation networks,”
Konstantinos Bousmalis, George Trigeorgis, Nathan Silberman, Dilip Krishnan, and Dumitru Erhan, · 2016
Later among the works it cites.
“Multi-channel speech recognition: Lstms all the way through,”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Robust Automatic Speech Recognition: A Bridge to Practical Applications
J. Li, L. Deng, R. Haeb-Umbach, and Y. Gong, · 2015
Cited alongside, same era.
“Maximum a posteriori adaptation of network parameters in deep models,”
Zhen Huang, Sabato Marco Siniscalchi, I-Fan Chen, Jinyu Li, Jiadong Wu, and Chin-Hui Lee, · 2015
Cited alongside, same era.
“Rapid adaptation for deep neural networks through multi-task learning.,”
Zhen Huang, Jinyu Li, Sabato Marco Siniscalchi, I-Fan Chen, Ji Wu, and Chin-Hui Lee, · 2015
Cited alongside, same era.
“Differentiable pooling for unsupervised speaker adaptation,”
P. Swietojanski and S. Renals, · 2015
Cited alongside, same era.
“Unsupervised domain adaptation by backpropagation,”
Yaroslav Ganin and Victor Lempitsky, · 2015
Cited alongside, same era.
“The third chime speech separation and recognition challenge: Dataset, task and baselines,”
J. Barker, R. Marxer, E. Vincent, and S. Watanabe, · 2015
Cited alongside, same era.
“The merl/sri system for the 3rd chime challenge using beamforming, robust feature extraction, and advanced speech recognition,”
T. Hori, Z. Chen, H. Erdogan, J. R. Hershey, J. Le Roux, V. Mitra, and S. Watanabe, · 2015
Cited alongside, same era.
H. Erdogan, T. Hayashi, J. R. Hershey, et al., · 2016
Later among the works it cites.
“Recent progresses in deep learning based acoustic models,”
Dong Yu and Jinyu Li, · 2017
Closest in time.
“Unsupervised speaker adaptation of batch normalized acoustic models for robust asr,”
Z. Q. Wang and D. Wang, · 2017
Closest in time.
“Large-scale domain adaptation via teacher-student learning,”
Jinyu Li, Michael L Seltzer, Xi Wang, Rui Zhao, and Yifan Gong, · 2017
Closest in time.
“An unsupervised deep domain adaptation approach for robust speech recognition,”
Sining Sun, Binbin Zhang, Lei Xie, and Yanning Zhang, · 2017
Closest in time.
“Deep long short-term memory adaptive beamforming networks for multichannel robust speech recognition,”
Z. Meng, S. Watanabe, J. R. Hershey, and H. Erdogan, · 2017
Closest in time.