Fetching the paper…
Reading the bibliography…
Supervised multi-channel audio source separation requires extracting useful spectral, temporal, and spatial features from the mixed signals.
D. Griffin and J. Lim, “Signal estimation from modified short-time Fourier transform,” IEEE Trans. on Acoustics, Speech, and Signal Processing , vol. 32, no. 1, pp. 236–243, 1984
1984
Earlier work this paper cites.
E. Vincent, “Musical source separation using time-frequency source priors,” IEEE Trans. on Audio, Speech, and Language Processing , vol. 14, no. 1, pp. 91–98, 2006
2006
Earlier work this paper cites.
E. Vincent, R. Gribonval, and C. Fevotte, “Performance measurement in blind audio source separation,” IEEE Trans. on Audio, Speech, and Language Processing , vol. 14, no. 4, pp. 1462–69, Jul. 2006
2006
Earlier work this paper cites.
J.-L. Durrieu, B. David, and G. Richard, “A musically motivated mid-level representation for pitch estimation and musical audio source separation,” IEEE Trans. on on Selected Topics on Signal Processing , vol. 5, no. 6, pp. 1118–1133, Oct. 2011
2011
Earlier work this paper cites.
A. Ozerov, E. Vincent, and F. Bimbot, “A general flexible framework for the handling of prior information in audio source separation,” IEEE Trans. on Audio, Speech, and Language Processing , vol. 20, no. 4, pp. 1118–1133, Oct. 2012
2012
Earlier work this paper cites.
P. Huang, S. Chen, P. Smaragdis, and M. Hasegawa-Johnson, “Singing-voice separation from monaural recordings using robust principal component analysis,” in Proc. ICASSP , 2012, pp. 57–60
2012
Earlier work this paper cites.
B. Gao, W. L. Woo, and S. S. Dlay, “Unsupervised single-channel separation of nonstationary signals using Gammatone filterbank and itakura–saito nonnegative matrix two-dimensional factorizations,” IEEE Trans. on Circuits and Systems I , vol. 60, no. 3, pp. 662–675, 2013
2013
Earlier work this paper cites.
Z. Rafii and B. Pardo, “REpeating pattern extraction technique (REPET): A simple method for music/voice separation,” IEEE Trans. on Audio, Speech, and Language Processing , vol. 21, no. 1, pp. 71–82, Jan. 2013
2013
Earlier work this paper cites.
M. Krawczyk and T. Gerkmann, “STFT phase reconstruction in voiced speech for an improved single-channel speech enhancement,” IEEE Trans. on Audio, Speech, and Language Processing , vol. 22, no. 12, pp. 1–10, 2014
2014
Earlier work this paper cites.
S. Dieleman and B. Schrauwen, “End-to-end learning for music audio,” in Proc. ICASSP , 2014, pp. 6964–6968
2014
Earlier work this paper cites.
T. N. Sainath, R. J. Weiss, A. W. Senior, K. W. Wilson, and O. Vinyals, “Learning the speech front-end with raw waveform CLDNNs,” in Proc. InterSpeech , 2015
2015
Earlier work this paper cites.
Y. Hoshen, R. Weiss, and K. W. Wilson, “Speech acoustic modeling from raw multichannel waveforms,” in Proc. ICASSP , 2015
2015
Cited alongside, same era.
2015
Cited alongside, same era.
2015
Cited alongside, same era.
F. Chollet et al. , “Keras,” https://github.com/keras-team/keras
2015
Cited alongside, same era.
A. Liutkus, D. FitzGerald, Z. Rafii, and L. Daudet, “Scalable audio separation with light kernel additive modelling,” in Proc. ICASSP , 2015, pp. 76–80
2015
Cited alongside, same era.
A. Zermini, Q. Liu, X. Yong, M. Plumbley, D. Betts, and W. Wang, “Binaural and log-power spectra features with deep neural networks for speech-noise separation,” in Proc. International Workshop on Multimedia Signal Processing , 2017
2017
Later among the works it cites.
2017
Later among the works it cites.
M. Dubey, G. Kenyon, N. Carlson, and A. Thresher, “Does phase matter for monaural source separation?” in Proc. NIPS , 2017
2017
Later among the works it cites.
T. N. Sainath, R. J. Weiss, K. W. Wilson, B. Li, A. Narayanan, E. Variani, M. Bacchiani, I. Shafran, A. Senior, K. Chin, A. Misra, and C. Kim, “Multichannel signal processing with deep neural networks for automatic speech recognition,” IEEE/ACM Trans. on Audio, Speech, and Language Processing. , vol. 25, no. 5, pp. 965–979, May 2017
2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A. A. Nugraha, A. Liutkus, and E. Vincent, “Multichannel audio source separation with deep neural networks,” IEEE/ACM Trans. on Audio, Speech, and Language Processing , vol. 24, no. 9, pp. 1652–1664, 2016
2016
Cited alongside, same era.
Y. Yu, W. Wang, and P. Han, “Localization based stereo speech source separation using probabilistic time-frequency masking and deep neural networks,” EURASIP Journal on Audio, Speech, and Music Processing , pp. 1–18, 2016
2016
Cited alongside, same era.
E. M. Grais, G. Roma, A. J. R. Simpson, and M. D. Plumbley, “Single channel audio source separation using deep neural network ensembles,” in Proc. 140th Audio Engineering Society Convention , 2016
2016
Cited alongside, same era.
2016
Cited alongside, same era.
F.-R. Stoter, A. Liutkus, R. Badeau, B. Edler, and P. Magron, “Common fate model for unison source separation,” in Proc. ICASSP , 2016
2016
Cited alongside, same era.
S. Uhlich, M. Porcu, F. Giron, M. Enenkl, T. Kemp, N. Takahashi, and Y. Mitsufuji, “Improving music source separation based on deep neural networks through data augmentation and network blending,” in Proc. ICASSP , 2017
2017
Cited alongside, same era.
Later among the works it cites.
S. Venkataramani, J. Casebeer, and P. Smaragdis, “Adaptive front-ends for end-to-end source separation,” in Proc. NIPS , 2017
2017
Later among the works it cites.
2017
Later among the works it cites.
E. M. Grais and M. D. Plumbley, “Single channel audio source separation using convolutional denoising autoencoders,” in Proc. GlobalSIP , 2017
2017
Later among the works it cites.
A. Liutkus, F. Stoter, Z. Rafii, D. Kitamura, B. Rivet, N. Ito, N. Ono, and J. Fontecave, “The 2016 signal separation evaluation campaign,” in Proc. LVA/ICA , 2017, pp. 323–332
2017
Later among the works it cites.
P. Chandna, M. Miron, J. Janer, and E. Gomez, “Monoaural audio source separation using deep convolutional neural networks,” in Proc. LVA/ICA , 2017, pp. 258–266
2017
Later among the works it cites.
I.-Y. Jeong and K. Lee, “Singing voice separation using RPCA with weighted l1-norm,” in Proc. LVA/ICA , 2017, pp. 553–562
2017
Later among the works it cites.