Fetching the paper…
Reading the bibliography…
There are many deterministic mathematical operations (e.g.
T. E. Tremain, “The government standard linear predictive coding algorithm: Lpc-10,” Speech Technology , pp. 40–49, 1982
1982
Earlier work this paper cites.
J. S. Garofolo, L. F. Lamel, W. M. Fisher, J. G. Fiscus, and D. S. Pallett, “Darpa timit acoustic-phonetic continous speech corpus cd-rom. nist speech disc 1-1.1,” NASA STI/Recon technical report n , vol. 93, p. 27403, 1993
1993
Earlier work this paper cites.
A. W. Rix, J. G. Beerends, M. P. Hollier, and A. P. Hekstra, “Perceptual evaluation of speech quality (pesq)-a new method for speech quality assessment of telephone networks and codecs,” in 2001 IEEE International Conference on Acoustics, Speech, and Signal Processing. Proceedings (Cat. No. 01CH37221) , vol. 2. IEEE, 2001, pp. 749–752
2001
Earlier work this paper cites.
M. K. Hasan, S. Salahuddin, and M. R. Khan, “A modified a priori snr for speech enhancement using spectral subtraction rules,” IEEE Signal Processing Letters , vol. 11, no. 4, pp. 450–453, 2004
2004
Earlier work this paper cites.
Y. Hu and P. C. Loizou, “Evaluation of objective quality measures for speech enhancement,” IEEE Transactions on Audio, Speech, and Language Processing , vol. 16, no. 1, pp. 229–238, 2007
2007
Earlier work this paper cites.
M. Abd El-Fattah, M. I. Dessouky, S. M. Diab, and F. E.-S. Abd El-Samie, “Speech enhancement using an adaptive wiener filtering approach,” Progress in Electromagnetics Research , vol. 4, pp. 167–184, 2008
2008
Earlier work this paper cites.
3GPP, “3gpp ts 26.090 - mandatory speech codec speech processing functions; adaptive multi-rate (amr) speech codec; transcoding functions,” 3GPP , Retrieved 2010-07-21
2010
Earlier work this paper cites.
P. C. Loizou, Speech enhancement: theory and practice . CRC Press, 2013
2013
Cited alongside, same era.
2014
Cited alongside, same era.
2016
Cited alongside, same era.
2017
Cited alongside, same era.
T. Lan, Y. Lyu, W. Ye, G. Hui, Z. Xu, and Q. Liu, “Combining multi-perspective attention mechanism with convolutional networks for monaural speech enhancement,” IEEE Access , vol. 8, pp. 78 979–78 991, 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
W. Ping, K. Peng, K. Zhao, and Z. Song, “Waveflow: A compact flow-based model for raw audio,” in International Conference on Machine Learning . PMLR, 2020, pp. 7706–7716
2020
Later among the works it cites.
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
A. S. Subramanian, X. Wang, M. K. Baskar, S. Watanabe, T. Taniguchi, D. Tran, and Y. Fujita, “Speech enhancement using end-to-end speech recognition objectives,” in 2019 IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA) . IEEE, 2019, pp. 234–238
2019
Cited alongside, same era.
J.-M. Valin and J. Skoglund, “Lpcnet: Improving neural speech synthesis through linear prediction,” in ICASSP 2019-2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2019, pp. 5891–5895
2019
Cited alongside, same era.
S. Nanavati, “Diffwave,” https://github.com/lmnt-com/diffwave/, Sep 2020
2020
Later among the works it cites.
R. Ardila, M. Branson, K. Davis, M. Henretty, M. Kohler, J. Meyer, R. Morais, L. Saunders, F. M. Tyers, and G. Weber, “Common voice: A massively-multilingual speech corpus,” in Proceedings of the 12th Conference on Language Resources and Evaluation (LREC 2020) , 2020, pp. 4211–4215
2020
Later among the works it cites.