Fetching the paper…
Reading the bibliography…
Speaker-aware source separation methods are promising workarounds for major difficulties such as arbitrary source permutation and unknown number of sources.
P. K. Kuhl, “Human adults and human infants show a perceptual magnet effect,”
1991
Earlier work this paper cites.
E. Vincent, R. Gribonval, and C. Fevotte, “Performance measurement in blind audio source separation,”
2006
Earlier work this paper cites.
A. Narayanan and D. Wang, “Improving robustness of deep neural network acoustic models via speech separation and joint adaptive training,”
2015
Earlier work this paper cites.
P. S. Huang, M. Kim, M. H. Johnson, and P. Smaragdis, “Joint optimization of masks and deep recurrent neural networks for monaural source separation,”
2015
Earlier work this paper cites.
C. Weng, D. Yu, M. L. Seltzer, and J. Droppo, “Deep neural networks for single-channel multi-talker speech recognition,”
2015
Earlier work this paper cites.
A. W. Bronkhorst, “The cocktail-party problem revisited: early processing and selection of multi-talker speech,”
2015
Earlier work this paper cites.
J. R. Hershey, Z. Chen, J. L. Roux, and S. Watanabe, “Deep clustering: Discriminative embeddings for segmentation and separation,”
2016
Earlier work this paper cites.
Y. Isik, J. L. Roux, Z. Chen, S. Watanabe, and J. R. Hershey, “Single-channel multi-speaker separation using deep clustering,”
2016
Earlier work this paper cites.
M. Delcroix, K. Kinoshita, C. Yu, A. Ogawa, T. Yoshioka, and T. Nakatani, “Context adaptive deep neural networks for fast acoustic model adaptation in noisy conditions,”
2016
Cited alongside, same era.
X. L. Zhang and D. Wang, “A deep ensemble learning method for monaural speech separation,”
2016
Cited alongside, same era.
K. Vesely, S. Watanabe, K. Zmolikova, M. Karafiat, L. Burget, and J. H. Cernocky, “Sequence summarizing neural network for speaker adaptation,”
2016
Cited alongside, same era.
D. Yu and J. D. Li, “Recent progresses in deep learning based acoustic models,”
2017
Cited alongside, same era.
Y. Dong, K. Morten, T. Zheng-Hua, and J. Jesper, “Permutation invariant training of deep models for speaker-independent multi-talker speech separation,”
2017
Cited alongside, same era.
2017
Later among the works it cites.
Z. Chen, Y. Luo, and N. Mesgarani, “Deep attractor network for single-microphone speaker separation,”
2017
Later among the works it cites.
D. Wang and J. Chen, “Supervised speech separation based deep learning: An overview,”
2017
Later among the works it cites.
K. Zmolikova, M. Delcroix, K. Kinoshita, T. Higuchi, A. Ogawa, and T. Nakatani, “Speaker-aware neural network based beamformer for speaker extraction in speech mixtures,”
2017
Later among the works it cites.
K. Zmolikova, M. Delcroix, K. Kinoshita, T. Higuchi, A. Ogawa, and T. Nakatani, “Learning speaker representation for neural network based multichannel speaker extractions,”
2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. Yu, X. Chang, and Y. M. Qian, “Recognizing multi-talker speech with permutation invariant training,”
2017
Cited alongside, same era.
2017
Cited alongside, same era.
M. Kolbaek, D. Yu, Z. H. Tan, and J. Jensen, “Joint separation and denoising of noisy multi-talker speech using recurrent neural networks and permutation invariant training,”
2017
Cited alongside, same era.
J. Heymann, L. Drude, C. Boeddeker, P. Hanebrink, and R. Haeb-Umbach, “Beamnet: End-to-end training of a beamformer-supported multi-channel asr system,”
Cited in the paper.
R. Maas, S. H. K. Parthasarathi, B. King, R. Huang, and B. Hoffmeister, “Anchored speech detection,”
Cited in the paper.
Cited in the paper.
Later among the works it cites.
B. King, I. Chen, Y. Vaizman, Y. Liu, R. Maas, S. H. K. Parthasarathi, and B. Hoffmeister, “Robust speech recognition via anchor word representations,”
2017
Later among the works it cites.
Y. Qian, C. Weng, X. Chang, S. Wang, and D. Yu, “Past review, current progress, and challenges ahead on the cocktail party problem,”
2018
Closest in time.