T. Ochiai, S. Watanabe, T. Hori, J. R. Hershey, and X. Xiao, “Unified architecture for multichannel end-to-end speech recognition with neural beamforming,” IEEE Journal of Selected Topics in Signal Processing , vol. 11, no. 8, pp. 1274–1288, Dec. 2017
2017
Cited alongside, same era.
Z.-Q. Wang, J. Le Roux, and J. R. Hershey, “Alternative objective functions for deep clustering,” in Proc. IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , Apr. 2018
2018
Cited alongside, same era.
E. Vincent, T. Virtanen, and S. Gannot, Audio Source Separation and Speech Enhancement , 1st ed. Wiley Publishing, 2018
2018
Cited alongside, same era.
Z.-Q. Wang, J. Le Roux, and J. R. Hershey, “Multi-channel deep clustering: Discriminative spectral and spatial embeddings for speaker-independent speech separation,” in Proc. IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , Apr. 2018
2018
Cited alongside, same era.
R. Scheibler, E. Bezzam, and I. Dokmanić, “Pyroomacoustics: A python package for audio room simulation and array processing algorithms,” in Proc. IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , Apr. 2018, pp. 351–355
2018
Cited alongside, same era.
Y. Luo and N. Mesgarani, “TasNet: Time-domain audio separation network for real-time, single-channel speech separation,” in Proc. IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , Apr. 2018
2018
Cited alongside, same era.
Y. Luo and N. Mesgarani, “Real-time single-channel dereverberation and separation with time-domain audio separation network.” in Interspeech , 2018, pp. 342–346
2018
Cited alongside, same era.
Y. Zhao, Z.-Q. Wang, and D. Wang, “Two-stage deep learning for noisy-reverberant speech enhancement,” IEEE/ACM Transactions on Audio, Speech and Language Processing , vol. 27, no. 1, pp. 53–62, 2018
2018
Cited alongside, same era.
S. Settle, J. Le Roux, T. Hori, S. Watanabe, and J. R. Hershey, “End-to-end multi-speaker speech recognition,” in Proc. IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , Apr. 2018
2018
Cited alongside, same era.