Fetching the paper…
Reading the bibliography…
Singing voice separation based on deep learning relies on the usage of time-frequency masking.
“Signal estimation from modified short-time Fourier transform,”
D. Griffin and J. Lim, · 1984
Earlier work this paper cites.
“Understanding the difficulty of training deep feedforward neural networks,”
X. Glorot and Y. Bengio, · 2010
Earlier work this paper cites.
“On the use of masking filters in sound source separation,”
D. FitzGerald and R. Jaiswal, · 2012
Earlier work this paper cites.
“Generalized denoising auto-encoders as generative models,”
Y. Bengio, L. Yao, G. Alain, and P. Vincent, · 2013
Earlier work this paper cites.
“Exact solutions to the nonlinear dynamics of learning in deep linear neural networks,”
A.-M. Saxe, J.-L. McClelland, and S. Ganguli, · 2013
Earlier work this paper cites.
“Medleydb: A multitrack dataset for annotation-intensive MIR research,”
R. M. Bittner, J. Salamon, M. Tierney, M. Mauch, C. Cannam, and J. P. Bello, · 2014
Earlier work this paper cites.
“Adam: A method for stochastic optimization,”
D.-P. Kingma and J. Ba, · 2014
Earlier work this paper cites.
“Generalized Wiener filtering with fractional power spectrograms,”
A. Liutkus and R. Badeau, · 2015
Earlier work this paper cites.
“Deep neural network based instrument extraction from music,”
S. Uhlich, F. Giron, and Y. Mitsufuji, · 2015
Cited alongside, same era.
“Joint optimization of masks and deep recurrent neural networks for monaural source separation,”
P.-S. Huang, M. Kim, M. Hasegawa-Johnson, and P. Smaragdis, · 2015
Cited alongside, same era.
R.-K. Srivastava, K. Greff, and J. Schmidhuber, · 2015
Cited alongside, same era.
“Cauchy nonnegative matrix factorization,”
A. Liutkus, D. Fitzgerald, and R. Badeau, · 2015
Cited alongside, same era.
“Phase-sensitive and recognition-boosted speech separation using deep recurrent neural networks,”
H. Erdogan, J. R. Hershey, S. Watanabe, and J. Le Roux, · 2015
Cited alongside, same era.
“Single-channel audio source separation using deep neural network ensembles,”
“Deep networks with stochastic depth,”
G. Huang, Y. Sun, Z. Liu, D. Sedra, and K. Q. Weinberger, · 2016
Later among the works it cites.
Y. Wu et al, · 2016
Later among the works it cites.
“The 2016 signal separation evaluation campaign,”
A. Liutkus, F.-R. Stöter, Z. Rafii, D. Kitamura, B. Rivet, N. Ito, N. Ono, and J. Fontecave, · 2017
Closest in time.
“Monoaural audio source separation using deep convolutional neural networks,”
P. Chandna, M. Miron, J. Janer, and E. Gómez, · 2017
Closest in time.
“Multi-scale multi-band densenets for audio source separation,”
N. Takahashi and Y. Mitsufuji, · 2017
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
E.-M. Grais, G. Roma, A.J.R. Simpson, and M.-D. Plumbley, · 2016
Cited alongside, same era.
“Multichannel music separation with deep neural networks,”
A.-A. Nugraha, A. Liutkus, and E. Vincent, · 2016
Cited alongside, same era.
“New sonorities for jazz recordings: Separation and mixing using deep neural networks,”
S.-I. Mimilakis, E. Cano, J. Abeßer, and G. Schuller, · 2016
Cited alongside, same era.
“A recurrent encoder-decoder approach with skip-filtering connections for monaural singing voice separation,”
S.-I. Mimilakis, K. Drossos, G. Schuller, and T. Virtanen, · 2017
Closest in time.
“Convolutional neural networks analyzed via convolutional sparse coding,”
V. Papyan, Y. Romano, and M. Elad, · 2017
Closest in time.