Fetching the paper…
Reading the bibliography…
Deep clustering is a recently introduced deep learning architecture that uses discriminatively trained embeddings as the basis for clustering.
A. S. Bregman,
1990
Earlier work this paper cites.
M. P. Cooke, “Modelling auditory processing and organisation,” Ph.D. dissertation, Univ. of Sheffield, 1991
1991
Earlier work this paper cites.
D. P. W. Ellis, “Prediction-driven computational auditory scene analysis,” Ph.D. dissertation, MIT, 1996
1996
Earlier work this paper cites.
D. Wang, “On ideal binary mask as the computational goal of auditory scene analysis,” in
2005
Earlier work this paper cites.
F. R. Bach and M. I. Jordan, “Learning spectral clustering, with application to speech separation,”
2006
Earlier work this paper cites.
T. Virtanen, “Speech recognition using factorial hidden Markov models for separation in the feature space,” in
2006
Earlier work this paper cites.
P. Smaragdis, “Convolutive speech bases and their application to supervised speech separation,”
2007
Earlier work this paper cites.
R. J. Weiss, “Underdetermined source separation using speaker subspace models,” Ph.D. dissertation, Columbia University, 2009
2009
Earlier work this paper cites.
Y. Bengio, J. Louradour, R. Collobert, and J. Weston, “Curriculum learning,” in
2009
Earlier work this paper cites.
J. R. Hershey, S. J. Rennie, P. A. Olsen, and T. T. Kristjansson, “Super-human multi-talker speech recognition: A graphical modeling approach,”
2010
Earlier work this paper cites.
S. J. Rennie, J. R. Hershey, and P. A. Olsen, “Single-channel multitalker speech recognition,”
2010
Cited alongside, same era.
M. Cooke, J. R. Hershey, and S. J. Rennie, “Monaural speech separation and recognition challenge,”
2010
Cited alongside, same era.
D. M. Blei, P. R. Cook, and M. Hoffman, “Bayesian nonparametric matrix factorization for recorded music,” in
2010
Cited alongside, same era.
M. Nakano, J. Le Roux, H. Kameoka, T. Nakamura, N. Ono, and S. Sagayama, “Bayesian nonparametric spectrogram modeling based on infinite factorial infinite hidden markov model,” in
2011
Cited alongside, same era.
2011
Cited alongside, same era.
Y. Xu, J. Du, L.-R. Dai, and C.-H. Lee, “An experimental study on speech enhancement based on deep neural networks,”
2014
Later among the works it cites.
W. Zaremba, I. Sutskever, and O. Vinyals, “Recurrent neural network regularization,”
2014
Later among the works it cites.
2014
Later among the works it cites.
F. Weninger, H. Erdogan, S. Watanabe, E. Vincent, J. Le Roux, J. R. Hershey, and B. Schuller, “Speech enhancement with LSTM recurrent neural networks and its application to noise-robust ASR,” in
2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. Tieleman and G. Hinton, “Lecture 6.5—RmsProp: Divide the gradient by a running average of its recent magnitude,” COURSERA: Neural Networks for Machine Learning, 2012
2012
Cited alongside, same era.
K. Hu and D. Wang, “An unsupervised approach to cochannel speech separation,”
2013
Cited alongside, same era.
U. Simsekli, J. Le Roux, and J. Hershey, “Non-negative source-filter dynamical system for speech enhancement,” in
2014
Cited alongside, same era.
Y. Wang, A. Narayanan, and D. Wang, “On training targets for supervised speech separation,”
2014
Cited alongside, same era.
2015
Later among the works it cites.
T. Moon, H. Choi, H. Lee, and I. Song, “RnnDrop: A novel dropout for RNNs in ASR,”
2015
Later among the works it cites.
H. Erdogan, J. R. Hershey, S. Watanabe, and J. Le Roux, “Phase-sensitive and recognition-boosted speech separation using deep recurrent neural networks,” in
2015
Later among the works it cites.
J. R. Hershey, Z. Chen, J. Le Roux, and S. Watanabe, “Deep clustering: Discriminative embeddings for segmentation and separation,” in
2016
Closest in time.
Y. Isik, J. Le Roux, Z. Chen, S. Watanabe, and J. R. Hershey, “Deep clustering: Audio examples,”
2016
Closest in time.