Fetching the paper…
Reading the bibliography…
We apply multilayer bootstrap network (MBN), a recent proposed unsupervised learning method, to unsupervised speaker recognition.
“Speaker verification using adapted gaussian mixture models,”
Douglas A Reynolds, Thomas F Quatieri, and Robert B Dunn, · 2000
Earlier work this paper cites.
“Cluster ensembles—a knowledge reuse framework for combining multiple partitions,”
Alexander Strehl and Joydeep Ghosh, · 2003
Earlier work this paper cites.
“SVM based speaker verification using a GMM supervector kernel and NAP variability compensation,”
William M Campbell, Douglas E Sturim, Douglas A Reynolds, and Alex Solomonoff, · 2006
Earlier work this paper cites.
“Speech separation challenge,” http://staffwww.dcs.shef.ac.uk/people/M.Cooke/SpeechSeparationChallenge.htm
Martin Cooke and Te-Won Lee, · 2006
Earlier work this paper cites.
“Joint factor analysis versus eigenchannels in speaker recognition,”
Patrick Kenny, Gilles Boulianne, Pierre Ouellet, and Pierre Dumouchel, · 2007
Earlier work this paper cites.
“The ICSI RT07s speaker diarization system,”
Chuck Wooters and Marijn Huijbregts, · 2008
Earlier work this paper cites.
“Speaker clustering using vector quantization and spectral clustering,”
Ken-ichi Iso, · 2010
Earlier work this paper cites.
“Front-end factor analysis for speaker verification,”
Najim Dehak, Patrick Kenny, Réda Dehak, Pierre Dumouchel, and Pierre Ouellet, · 2011
Earlier work this paper cites.
“Learning speaker-specific characteristics with a deep neural architecture,”
Ke Chen and Ahmad Salman, · 2011
Cited alongside, same era.
“Speaker clustering and cluster purification methods for rt07 and rt09 evaluation meeting data,”
Tin Lay Nwe, Hanwu Sun, Bin Ma, and Haizhou Li, · 2012
Cited alongside, same era.
“Intra-conversation intra-speaker variability compensation for speaker clustering,”
Kui Wu, Yan Song, Wu Guo, and Lirong Dai, · 2012
Cited alongside, same era.
“Context-dependent pre-trained deep neural networks for large-vocabulary speech recognition,”
George E Dahl, Dong Yu, Li Deng, and Alex Acero, · 2012
Cited alongside, same era.
“Unsupervised methods for speaker diarization: An integrated and iterative approach,”
Stephen H Shum, Najim Dehak, Réda Dehak, and James R Glass, · 2013
Cited alongside, same era.
“Towards scaling up classification-based speech separation,”
“Nonlinear dimensionality reduction of data by multilayer bootstrap networks,”
Xiao-Lei Zhang, · 2014
Later among the works it cites.
“Cochannel speaker identification in anechoic and reverberant conditions,”
Xiaojia Zhao, Yuxuan Wang, and DeLiang Wang, · 2015
Closest in time.
“A regression approach to speech enhancement based on deep neural networks,”
Yong Xu, Jun Du, Li-Rong Dai, and Chin-Hui Lee, · 2015
Closest in time.
“Joint optimization of masks and deep recurrent neural networks for monaural source separation,”
Po-Sen Huang, Minje Kim, Mark Hasegawa-Johnson, and Paris Smaragdis, · 2015
Closest in time.
“Deep ensemble learning for monaural speech separation,”
Xiao-Lei Zhang and DeLiang Wang, · 2015
Closest in time.
“Boosting contextual information for deep neural network based voice activity detection,”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yuxuan Wang and DeLiang Wang, · 2013
Cited alongside, same era.
“Modeling spectral envelopes using restricted boltzmann machines and deep belief networks for statistical parametric speech synthesis,”
Zhen-Hua Ling, Li Deng, and Dong Yu, · 2013
Cited alongside, same era.
“Deep belief networks based voice activity detection,”
Xiao-Lei Zhang and Ji Wu, · 2013
Cited alongside, same era.
Xiao-Lei Zhang and DeLiang Wang, · 2015
Closest in time.
“A comparative study of spectral clustering for i-vector-based speaker clustering under noisy conditions,”
Naohiro Tawara, Tetsuji Ogawa, and Tetsunori Kobayashi, · 2045
Closest in time.