Fetching the paper…
Reading the bibliography…
Speech restoration aims to remove distortions in speech signals.
A. Erell and M. Weintraub, “Estimation using log-spectral-distance criterion for noise-robust speech recognition,” in Proceedings of the IEEE Conference on Acoustics, Speech, and Signal Processing , 1990, pp. 853–856
1990
Earlier work this paper cites.
A. W. Rix, J. G. Beerends, M. P. Hollier, and A. P. Hekstra, “Perceptual evaluation of speech quality (PESQ) - a new method for speech quality assessment of telephone networks and codecs,” in Proceedings of the IEEE Conference on Acoustics, Speech, and Signal Processing , 2001, pp. 749–752
2001
Earlier work this paper cites.
Z. Wang, A. C. Bovik, H. R. Sheikh, and E. P. Simoncelli, “Image quality assessment: from error visibility to structural similarity,” IEEE transactions on Image Processing , vol. 13, no. 4, pp. 600–612, 2004
2004
Earlier work this paper cites.
P. C. Loizou, Speech enhancement: theory and practice . CRC press, 2007
2007
Earlier work this paper cites.
T. Van den Bogaert, S. Doclo, J. Wouters, and M. Moonen, “Speech enhancement with multichannel wiener filter techniques in multimicrophone binaural hearing aids,” The Journal of the Acoustical Society of America , vol. 125, no. 1, pp. 360–371, 2009
2009
Earlier work this paper cites.
C. H. Taal, R. C. Hendriks, R. Heusdens, and J. Jensen, “An algorithm for intelligibility prediction of time-frequency weighted noisy speech,” IEEE/ACM Transactions on Audio, Speech, and Language Processing , vol. 19, no. 7, pp. 2125–2136, 2011
2011
Earlier work this paper cites.
2017
Earlier work this paper cites.
C. Valentini-Botinhao et al. , “Noisy speech database for training speech enhancement algorithms and TTS models,” 2017
2017
Earlier work this paper cites.
S. Pascual, A. Bonafonte, and J. Serra, “SEGAN: Speech enhancement generative adversarial network,” in INTERSPEECH , 2017
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
2019
Cited alongside, same era.
J. Yamagishi, C. Veaux, K. MacDonald et al. , “CSTR VCTK corpus: English multi-speaker corpus for cstr voice cloning toolkit,” 2019
2019
Cited alongside, same era.
P. Záviška, P. Rajmic, O. Mokrỳ, and Z. Prŭša, “A proper version of synthesis-based sparse audio declipper,” in Proceedings of the IEEE Conference on Acoustics, Speech, and Signal Processing , 2019, pp. 591–595
2019
Cited alongside, same era.
2020
Cited alongside, same era.
2020
Later among the works it cites.
2020
Later among the works it cites.
J. Zhang, M. D. Plumbley, and W. Wang, “Weighted magnitude-phase loss for speech dereverberation,” in Proceedings of the IEEE Conference on Acoustics, Speech, and Signal Processing , 2021, pp. 5794–5798
2021
Later among the works it cites.
Y. Ai, H. Li, X. Wang, J. Yamagishi, and Z. Ling, “Denoising-and-dereverberation hierarchical neural vocoder for robust waveform generation,” in 2021 IEEE Spoken Language Technology Workshop (SLT) . IEEE, 2021, pp. 477–484
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
P. Záviška, P. Rajmic, A. Ozerov, and L. Rencker, “A survey and an extensive evaluation of popular audio declipping methods,” IEEE Journal of Selected Topics in Signal Processing , vol. 15, no. 1, pp. 5–24, 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2020
Cited alongside, same era.
K. Tan, Y. Xu, S.-X. Zhang, M. Yu, and D. Yu, “Audio-visual speech separation and dereverberation with a two-stage multimodal network,” IEEE Journal of Selected Topics in Signal Processing , vol. 14, no. 3, pp. 542–553, 2020
2020
Cited alongside, same era.
S. Sulun and M. E. Davies, “On filter generalization for music bandwidth extension using deep neural networks,” IEEE Journal of Selected Topics in Signal Processing , vol. 15, no. 1, pp. 132–142, 2020
2020
Cited alongside, same era.
S. Maiti and M. I. Mandel, “Speaker independence of neural vocoders and their effect on parametric resynthesis speech enhancement,” in Proceedings of the IEEE Conference on Acoustics, Speech, and Signal Processing , 2020, pp. 206–210
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2021
Later among the works it cites.
H. Wang and D. Wang, “Towards robust speech super-resolution,” IEEE/ACM Transactions on Audio, Speech, and Language Processing , vol. 29, pp. 2058–2066, 2021
2021
Later among the works it cites.
X. Liu, T. Iqbal, J. Zhao, Q. Huang, M. D. Plumbley, and W. Wang, “Conditional sound generation using neural discrete time-frequency representation learning,” in IEEE International Workshop on Machine Learning for Signal Processing , 2021, pp. 1–6
2021
Later among the works it cites.
Q. Kong, Y. Cao, H. Liu, K. Choi, and Y. Wang, “Decoupling magnitude and phase estimation with deep resunet for music source separation.” in The International Society for Music Information Retrieval , 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
Q. Kong, H. Liu, X. Du, L. Chen, R. Xia, and Y. Wang, “Speech enhancement with weakly labelled data from AudioSet,” in INTERSPEECH , 2021
2021
Later among the works it cites.