Fetching the paper…
Reading the bibliography…
Speech enhancement algorithms based on deep learning have been improved in terms of speech intelligibility and perceptual quality greatly.
N. Ahmed, T. Natarajan, and K. R. Rao, “Discrete cosine transform,” IEEE transactions on Computers , vol. 100, no. 1, pp. 90–93, 1974
1974
Earlier work this paper cites.
J. B. Allen and D. A. Berkley, “Image method for efficiently simulating small-room acoustics,” The Journal of the Acoustical Society of America , vol. 65, no. 4, pp. 943–950, 1979
1979
Earlier work this paper cites.
J. S. Garofolo, L. F. Lamel, W. M. Fisher, J. G. Fiscus, and D. S. Pallett, “Darpa timit acoustic-phonetic continous speech corpus cd-rom. nist speech disc 1-1.1,” NASA STI/Recon technical report n , vol. 93, p. 27403, 1993
1993
Earlier work this paper cites.
A. Varga and H. J. Steeneken, “Assessment for automatic speech recognition: Ii. noisex-92: A database and an experiment to study the effect of additive noise on speech recognition systems,” Speech communication , vol. 12, no. 3, pp. 247–251, 1993
1993
Earlier work this paper cites.
G. Hu and D. Wang, “Speech segregation based on pitch tracking and amplitude modulation,” in Proceedings of the 2001 IEEE Workshop on the Applications of Signal Processing to Audio and Acoustics (Cat. No. 01TH8575) . IEEE, 2001, pp. 79–82
2001
Earlier work this paper cites.
S. Srinivasan, N. Roman, and D. Wang, “Binary and ratio time-frequency masks for robust speech recognition,” Speech Communication , vol. 48, no. 11, pp. 1486–1501, 2006
2006
Earlier work this paper cites.
K. Paliwal, K. Wójcicki, and B. Shannon, “The importance of phase in speech enhancement,” speech communication , vol. 53, no. 4, pp. 465–494, 2011
2011
Earlier work this paper cites.
Y. Xu, J. Du, L.-R. Dai, and C.-H. Lee, “An experimental study on speech enhancement based on deep neural networks,” IEEE Signal processing letters , vol. 21, no. 1, pp. 65–68, 2013
2013
Earlier work this paper cites.
Y. Wang, A. Narayanan, and D. Wang, “On training targets for supervised speech separation,” IEEE/ACM transactions on audio, speech, and language processing , vol. 22, no. 12, pp. 1849–1858, 2014
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
H. Erdogan, J. R. Hershey, S. Watanabe, and J. Le Roux, “Phase-sensitive and recognition-boosted speech separation using deep recurrent neural networks,” in 2015 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2015, pp. 708–712
2015
Cited alongside, same era.
D. S. Williamson, Y. Wang, and D. Wang, “Complex ratio masking for monaural speech separation,” IEEE/ACM transactions on audio, speech, and language processing , vol. 24, no. 3, pp. 483–492, 2015
2015
Cited alongside, same era.
O. Ronneberger, P. Fischer, and T. Brox, “U-net: Convolutional networks for biomedical image segmentation,” in International Conference on Medical image computing and computer-assisted intervention . Springer, 2015, pp. 234–241
2015
Cited alongside, same era.
P. G. Shivakumar and P. G. Georgiou, “Perception optimized deep denoising autoencoders for speech enhancement.” in Interspeech , 2016, pp. 3743–3747
2016
Cited alongside, same era.
J. M. Martin-Donas, A. M. Gomez, J. A. Gonzalez, and A. M. Peinado, “A deep learning loss function based on the perceptual evaluation of the speech quality,” IEEE Signal processing letters , vol. 25, no. 11, pp. 1680–1684, 2018
2018
Later among the works it cites.
M. Kolbæk, Z.-H. Tan, and J. Jensen, “Monaural speech enhancement using deep neural networks by maximizing a short-time objective intelligibility measure,” in 2018 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2018, pp. 5059–5063
2018
Later among the works it cites.
K. Tan and D. Wang, “Complex spectral mapping with a convolutional recurrent network for monaural speech enhancement,” in ICASSP 2019-2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2019, pp. 6865–6869
2019
Later among the works it cites.
Y. Luo and N. Mesgarani, “Conv-tasnet: Surpassing ideal time–frequency magnitude masking for speech separation,” IEEE/ACM transactions on audio, speech, and language processing , vol. 27, no. 8, pp. 1256–1266, 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
M. Delfarah and D. Wang, “Features for masking-based monaural speech separation in reverberant conditions,” IEEE/ACM Transactions on Audio, Speech, and Language Processing , vol. 25, no. 5, pp. 1085–1094, 2017
2017
Cited alongside, same era.
Q. Liu, W. Wang, P. J. Jackson, and Y. Tang, “A perceptually-weighted deep neural network for monaural speech enhancement in various background noise conditions,” in 2017 25th European Signal Processing Conference (EUSIPCO) . IEEE, 2017, pp. 1270–1274
2017
Cited alongside, same era.
H.-S. Choi, J.-H. Kim, J. Huh, A. Kim, J.-W. Ha, and K. Lee, “Phase-aware speech enhancement with deep complex u-net,” in International Conference on Learning Representations , 2018
2018
Cited alongside, same era.
K. Tan and D. Wang, “A convolutional recurrent neural network for real-time speech enhancement.” in Interspeech , 2018, pp. 3229–3233
2018
Cited alongside, same era.
2019
Later among the works it cites.
J. Le Roux, S. Wisdom, H. Erdogan, and J. R. Hershey, “Sdr–half-baked or well done?” in ICASSP 2019-2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2019, pp. 626–630
2019
Later among the works it cites.
S. B. L. M. T. X. S. M. Z. Y. H. F. J. W. B. H. Z. Y. X. Hu, Y. Liu and L. Xie, “Dccrn: Deep complex convolution recurrent network for phase-aware speech enhancement,” in Interspeech , 2020, pp. 2472–2476
2020
Later among the works it cites.
C. Geng and L. Wang, “End-to-end speech enhancement based on discrete cosine transform,” in 2020 IEEE International Conference on Artificial Intelligence and Computer Applications (ICAICA) . IEEE, 2020, pp. 379–383
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.