Fetching the paper…
Reading the bibliography…
We propose a novel adversarial multi-task learning scheme, aiming at actively curtailing the inter-talker feature variability while maximizing its senone discriminability so as to enhance the performance of a deep neural network (DNN) based ASR system.
“A compact model for speaker-adaptive training,”
Tasos Anastasakos, John McDonough, Richard Schwartz, and John Makhoul, · 1996
Earlier work this paper cites.
“Maximum likelihood linear transformations for hmm-based speech recognition,”
M.J.F. Gales, · 1998
Earlier work this paper cites.
“Cluster adaptive training of hidden markov models,”
M. J. F. Gales, · 2000
Earlier work this paper cites.
“Visualizing data using t-sne,”
Laurens van der Maaten and Geoffrey Hinton, · 2008
Earlier work this paper cites.
“Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups,”
G. Hinton, L. Deng, D. Yu, et al., · 2012
Earlier work this paper cites.
“Improving wideband speech recognition using mixed-bandwidth training data in CD-DNN-HMM,”
J. Li, D. Yu, J.-T. Huang, and Y. Gong, · 2012
Earlier work this paper cites.
“Speaker adaptation of neural network acoustic models using i-vectors,”
G. Saon, H. Soltau, D. Nahamoo, and M. Picheny, · 2013
Earlier work this paper cites.
“Singular value decomposition based low-footprint speaker adaptation and personalization for deep neural network,”
J. Xue, J. Li, D. Yu, M. Seltzer, and Y. Gong, · 2014
Earlier work this paper cites.
“Fast adaptation of deep neural network based on discriminant codes for speech recognition,”
S. Xue, O. Abdel-Hamid, H. Jiang, L. Dai, and Q. Liu, · 2014
Earlier work this paper cites.
“Generative adversarial nets,”
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio, · 2014
Earlier work this paper cites.
“Long short-term memory recurrent neural network architectures for large scale acoustic modeling,”
F. Beaufays H. Sak, A. Senior, · 2014
Earlier work this paper cites.
“Speaker adaptive training of deep neural network acoustic models using i-vectors,”
Y. Miao, H. Zhang, and F. Metze, · 2015
Cited alongside, same era.
“Multi-basis adaptive neural network for rapid adaptation in speech recognition,”
C. Wu and M. J. F. Gales, · 2015
Cited alongside, same era.
“Maximum a posteriori adaptation of network parameters in deep models,”
Z. Huang, S. Siniscalchi, I. Chen, et al., · 2015
Cited alongside, same era.
“Rapid adaptation for deep neural networks through multi-task learning,”
Z. Huang, J. Li, S. Siniscalchi, et al., · 2015
Cited alongside, same era.
“Unsupervised domain adaptation by backpropagation,”
Yaroslav Ganin and Victor Lempitsky, · 2015
Cited alongside, same era.
“The third chime speech separation and recognition challenge: Dataset, task and baselines,”
J. Barker, R. Marxer, E. Vincent, and S. Watanabe, · 2015
“Adversarial multi-task learning of deep neural networks for robust speech recognition.,”
Yusuke Shinohara, · 2016
Later among the works it cites.
“Invariant representations for noisy speech recognition,”
Dmitriy Serdyuk, Kartik Audhkhasi, Philémon Brakel, Bhuvana Ramabhadran, Samuel Thomas, and Yoshua Bengio, · 2016
Later among the works it cites.
“Domain separation networks,”
Konstantinos Bousmalis, George Trigeorgis, Nathan Silberman, Dilip Krishnan, and Dumitru Erhan, · 2016
Later among the works it cites.
“Multi-channel speech recognition: Lstms all the way through,”
Hakan Erdogan, Tomoki Hayashi, John R Hershey, Takaaki Hori, Chiori Hori, Wei-Ning Hsu, Suyoun Kim, Jonathan Le Roux, Zhong Meng, and Shinji Watanabe, · 2016
Later among the works it cites.
“Recent progresses in deep learning based acoustic models,”
Dong Yu and Jinyu Li, · 2017
Later among the works it cites.
“An unsupervised deep domain adaptation approach for robust speech recognition,”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
“The merl/sri system for the 3rd chime challenge using beamforming, robust feature extraction, and advanced speech recognition,”
T. Hori, Z. Chen, H. Erdogan, et al., · 2015
Cited alongside, same era.
“Low-rank plus diagonal adaptation for deep neural networks,”
Y. Zhao, J. Li, and Y. Gong, · 2016
Cited alongside, same era.
“Learning hidden unit contributions for unsupervised acoustic model adaptation,”
P. Swietojanski, J. Li, and S. Renals, · 2016
Cited alongside, same era.
“Cluster adaptive training for deep neural network based acoustic model,”
T. Tan, Y. Qian, and K. Yu, · 2016
Cited alongside, same era.
“Factorized hidden layer adaptation for deep neural network based acoustic modeling,”
L. Samarakoon and K. C. Sim, · 2016
Cited alongside, same era.
Sining Sun, Binbin Zhang, Lei Xie, and Yanning Zhang, · 2017
Later among the works it cites.
“Unsupervised adaptation with domain separation networks for robust speech recognition,”
Z. Meng, Z. Chen, V. Mazalov, J. Li, and Y. Gong, · 2017
Later among the works it cites.
“English conversational telephone speech recognition by humans and machines,”
George Saon, Gakuto Kurata, Tom Sercu, et al., · 2017
Later among the works it cites.
“Deep long short-term memory adaptive beamforming networks for multichannel robust speech recognition,”
Z. Meng, S. Watanabe, J. R. Hershey, and H. Erdogan, · 2017
Later among the works it cites.
“Adversarial teacher-student learning for unsupervised domain adaptation,”
Zhong Meng, Jinyu Li, Yifan Gong, and Biing-Hwang (Fred) Juang, · 2018
Closest in time.