Fetching the paper…
Reading the bibliography…
Separating two sources from an audio mixture is an important task with many applications.
“Performance measurement in blind audio source separation,”
E. Vincent, R. Gribonval, and C. Fevotte, · 2006
Earlier work this paper cites.
Audio Cover Song Identification and Similarity: Background, Approaches, Evaluation, and Beyond
J. Serrà, E. Gómez, and P. Herrera, · 2010
Earlier work this paper cites.
“On the improvement of singing voice separation for monaural recordings using the mir-1k dataset,”
C.-L. Hsu and J.-S. R. Jang, · 2010
Earlier work this paper cites.
“A musically motivated mid-level representation for pitch estimation and musical audio source separation,”
J. L. Durrieu, B. David, and G. Richard, · 2011
Earlier work this paper cites.
“A tandem algorithm for singing pitch extraction and voice separation from music accompaniment,”
C. L. Hsu, D. Wang, J.-S. R. Jang, and K. Hu, · 2012
Earlier work this paper cites.
“Singing-voice separation from monaural recordings using robust principal component analysis,”
P.-S. Huang, S. D. Chen, P. Smaragdis, and M. Hasegawa-Johnson, · 2012
Earlier work this paper cites.
“Real-time online singing voice separation from monaural recordings using robust low-rank modeling,”
P. Sprechmann, A. M. Bronstein, and G. Sapiro, · 2012
Earlier work this paper cites.
“Recurrent neural networks for noise reduction in robust ASR,”
A. L. Maas, Q. V. Le, T. M. O’Neil, O. Vinyals, P. Nguyen, and A. Y. Ng, · 2012
Earlier work this paper cites.
“REpeating Pattern Extraction Technique (REPET): A simple method for music/voice separation,”
Z. Rafii and B. Pardo, · 2013
Earlier work this paper cites.
“Low-rank representation of both singing voice and music accompaniment via learned dictionaries,”
Y.-H. Yang, · 2013
Cited alongside, same era.
“Multi-stage non-negative matrix factorization for monaural singing voice separation,”
B. Zhu, W. Li, R. Li, and X. Xue, · 2013
Cited alongside, same era.
“Singing-voice separation from monaural recordings using deep recurrent neural networks,”
P.-S. Huang, M. Kim, M. Hasegawa-Johnson, and P. Smaragdis, · 2014
Cited alongside, same era.
“Generative adversarial nets,”
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, · 2014
Cited alongside, same era.
“MedleyDB: A multitrack dataset for annotation-intensive mir research,”
R. M. Bittner, J. Salamon, M. Tierney, M. Mauch, C. Cannam, and J. P. Bello, · 2014
Cited alongside, same era.
“Deep neural network based instrument extraction from music,”
“Singing voice separation and vocal f0 estimation based on mutual combination of robust principal component analysis and subharmonic summation,”
Y. Ikemiya, K. Itoyama, and K. Yoshii, · 2016
Later among the works it cites.
“Complex NMF under phase constraints based on signal modeling: Application to audio source separation,”
P. Magron, R. Badeau, and B. David, · 2016
Later among the works it cites.
“A deep ensemble learning method for monaural speech separation,”
X.-L. Zhang and D. Wang, · 2016
Later among the works it cites.
“Image-to-image translation with conditional adversarial networks,”
P. Isola, J.-Y. Zhu, T. Zhou, and A. A. Efros, · 2016
Later among the works it cites.
[Online] https://sisec.inria.fr/sisec-2016/2016-professionally-produced-music-recordings/
“SiSEC MUS Homepage,” 2016, · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. Uhlich, F. Giron, and Y. Mitsufuji, · 2015
Cited alongside, same era.
“Vocal activity informed singing voice separation with the iKala dataset,”
T.-S. Chan, T.-C. Yeh, Z.-C. Fan, H.-W. Chen, L. Su, Y.-H. Yang, and R. Jang, · 2015
Cited alongside, same era.
“Multi-resolution stacking for speech separation based on boosted dnn,”
X.-L. Zhang and D. Wang, · 2015
Cited alongside, same era.
“Singing voice separation and pitch extraction from monaural polyphonic audio music via DNN and adaptive pitch tracking,”
Z.-C. Fan, J.-S. R. Jang, and C.-L. Lu, · 2016
Cited alongside, same era.
“Deep clustering and conventional networks for music separation: Stronger together,”
Y. Luo, Z. Chen, J. R. Hershey, J. L. Roux, and N. Mesgarani, · 2017
Closest in time.
“Imporving music source separation based on deep neural networs through data augmentation and network blenging,”
S. Uhlich, M. Porcu, F. Giron, M. Enenkl, T. Kemp, N. Takahashi, and Y. Mitsufuji, · 2017
Closest in time.
“Segan: Speech enhancement generative adversarial network,”
S. Pascual, A. Bonafonte, and J. Serrà, · 2017
Closest in time.
M. Arjovsky, S. Chintala, and L. Bottou, · 2017
Closest in time.