Fetching the paper…
Reading the bibliography…
Speech separation is an essential task for multi-talker speech recognition.
“The numpy array: a structure for efficient numerical computation,”
Stefan Van Der Walt, S Chris Colbert, and Gael Varoquaux, · 2011
Earlier work this paper cites.
“librosa: Audio and music signal analysis in python,”
Brian McFee, Colin Raffel, Dawen Liang, Daniel PW Ellis, Matt McVicar, Eric Battenberg, and Oriol Nieto, · 2015
Earlier work this paper cites.
“Deep clustering: Discriminative embeddings for segmentation and separation,”
John R Hershey, Zhuo Chen, Jonathan Le Roux, and Shinji Watanabe, · 2016
Earlier work this paper cites.
“Permutation invariant training of deep models for speaker-independent multi-talker speech separation,”
Dong Yu, Morten Kolbæk, Zheng-Hua Tan, and Jesper Jensen, · 2017
Earlier work this paper cites.
“Deep clustering and conventional networks for music separation: Stronger together,”
Yi Luo, Zhuo Chen, John R Hershey, Jonathan Le Roux, and Nima Mesgarani, · 2017
Earlier work this paper cites.
“Noisy speech database for training speech enhancement algorithms and tts models,”
Cassia Valentini-Botinhao et al., · 2017
Earlier work this paper cites.
“Improving mask learning based speech enhancement system with restoration layers and residual connection.,”
Zhuo Chen, Yan Huang, Jinyu Li, and Yifan Gong, · 2017
Cited alongside, same era.
“Alternative objective functions for deep clustering,”
Zhong-Qiu Wang, Jonathan Le Roux, and John R Hershey, · 2018
Cited alongside, same era.
“End-to-end speech separation with unfolded iterative phase reconstruction,”
Zhong-Qiu Wang, Jonathan Le Roux, DeLiang Wang, and John R Hershey, · 2018
Cited alongside, same era.
“Deep learning based phase reconstruction for speaker separation: A trigonometric perspective,”
Zhong-Qiu Wang, Ke Tan, and DeLiang Wang, · 2018
Cited alongside, same era.
“Tasnet: time-domain audio separation network for real-time, single-channel speech separation,”
Yi Luo and Nima Mesgarani, · 2018
Cited alongside, same era.
“The northwestern university source separation library,”
Ethan Manilow, Prem Seetharaman, and Bryan Pardo, · 2018
Later among the works it cites.
“Phasebook and friends: Leveraging discrete representations for source separation,”
Jonathan Le Roux, Gordon Wichern, Shinji Watanabe, Andy Sarroff, and John R Hershey, · 2019
Closest in time.
Ziqiang Shi, Huibin Lin, Liu Liu, Rujie Liu, and Jiqing Han, · 2019
Closest in time.
“Dual-path rnn: efficient long sequence modeling for time-domain single-channel speech separation,”
Yi Luo, Zhuo Chen, and Takuya Yoshioka, · 2019
Closest in time.
“Divide and conquer: A deep casa approach to talker-independent monaural speaker separation,”
Yuzhou Liu and DeLiang Wang, · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“Tasnet: Surpassing ideal time-frequency masking for speech separation,”
Yi Luo and Nima Mesgarani, · 2018
Cited alongside, same era.
Closest in time.