Fetching the paper…
Reading the bibliography…
Most of the currently successful source separation techniques use the magnitude spectrogram as input, and are therefore by default omitting part of the signal: the phase.
B. A. Olshausen and D. J. Field, “Sparse coding with an overcomplete basis set: A strategy employed by v1?”
1997
Earlier work this paper cites.
A. Hyvärinen and E. Oja, “Independent component analysis: algorithms and applications,”
2000
Earlier work this paper cites.
T. Virtanen and A. Klapuri, “Separation of harmonic sound sources using sinusoidal modeling,” in
2000
Earlier work this paper cites.
D. D. Lee and H. S. Seung, “Algorithms for non-negative matrix factorization,” in
2001
Earlier work this paper cites.
S. Dubnov, “Extracting sound objects by independent subspace analysis,” in
2002
Earlier work this paper cites.
G.-J. Jang and T.-W. Lee, “A maximum likelihood approach to single-channel source separation,”
2003
Earlier work this paper cites.
T. Blumensath and M. Davies, “Unsupervised learning of sparse and shift-invariant decompositions of polyphonic music,” in
2004
Earlier work this paper cites.
C. Févotte, R. Gribonval, and E. Vincent, “Bss_eval toolbox user guide–revision 2.0,” 2005
2005
Earlier work this paper cites.
T. Virtanen, “Unsupervised learning methods for source separation in monaural music signals,” in
2006
Earlier work this paper cites.
H. Kameoka, N. Ono, K. Kashino, and S. Sagayama, “Complex nmf: A new sparse representation for acoustic signals,” in
2009
Earlier work this paper cites.
P.-S. Huang, M. Kim, M. Hasegawa-Johnson, and P. Smaragdis, “Singing-voice separation from monaural recordings using deep recurrent neural networks.” in
2014
Earlier work this paper cites.
A. Roebel, J. Pons, M. Liuni, and M. Lagrangey, “On automatic drum transcription using non-negative matrix deconvolution and itakura saito divergence,” in
2015
Cited alongside, same era.
O. Ronneberger, P. Fischer, and T. Brox, “U-net: Convolutional networks for biomedical image segmentation,” in
2015
Cited alongside, same era.
2016
Cited alongside, same era.
P. Chandna, M. Miron, J. Janer, and E. Gómez, “Monoaural audio source separation using deep convolutional neural networks,” in
2017
Cited alongside, same era.
A. Jansson, E. Humphrey, N. Montecchio, R. Bittner, A. Kumar, and T. Weyde, “Singing voice separation with deep u-net convolutional networks,”
2017
J. Pons and X. Serra, “Designing efficient architectures for modeling temporal features with convolutional neural networks,” in
2017
Later among the works it cites.
2017
Later among the works it cites.
2017
Later among the works it cites.
F.-R. Stöter, A. Liutkus, and N. Ito, “The 2018 signal separation evaluation campaign,”
2018
Closest in time.
J. Pons, O. Nieto, M. Prockup, E. M. Schmidt, A. F. Ehmann, and X. Serra, “End-to-end learning for music audio tagging at scale,”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
A. Liutkus, F.-R. Stöter, Z. Rafii, D. Kitamura, B. Rivet, N. Ito, N. Ono, and J. Fontecave, “The 2016 signal separation evaluation campaign,” in
2017
Cited alongside, same era.
D. Rethage, J. Pons, and X. Serra, “A wavenet for speech denoising,”
2017
Cited alongside, same era.
J. Pons, O. Slizovskaia, R. Gong, E. Gómez, and X. Serra, “Timbre analysis of music audio signals with convolutional neural networks,” in
2017
Cited alongside, same era.
2018
Closest in time.
D. Stoller, S. Ewert, and S. Dixon, “Wave-u-net: A multi-scale neural network for end-to-end audio source separation,”
2018
Closest in time.
2018
Closest in time.
2018
Closest in time.
J. Le Roux, G. Wichern, S. Watanabe, A. Sarroff, and J. R. Hershey, “Phasebook and friends: Leveraging discrete representations for source separation,”
2019
Closest in time.
J. Pons and X. Serra, “Randomly weighted cnns for (music) audio classification,”
2019
Closest in time.