Perceptual loss based speech denoising with an ensemble of audio pattern recognition and self-supervised models
S. Kataria, J. Villalba, and N. Dehak · 2021
Later among the works it cites.
SE-Conformer: time-domain speech enhancement using conformer
E. Kim and H. Seo · 2021
Later among the works it cites.
DiffWave: a versatile diffusion model for audio synthesis
Z. Kong, W. Ping, J. Huang, K. Zhao, and B. Catanzaro · 2021
Later among the works it cites.
NU-Wave: a diffusion probabilistic model for neural audio upsampling
Original
J. Lee and S. Han · 2021
Later among the works it cites.
VoiceFixer: toward general speech restoration with neural vocoder
H. Liu, Q. Kong, Q. Tian, Y. Zhao, D. Wang, C. Huang, and Y. Wang · 2021
Later among the works it cites.
A study on speech enhancement based on diffusion probabilistic model
Y.-J. Lu, Y. Tsao, and S. Watanabe · 2021
Later among the works it cites.
DCCRN+: channel-wise subband DCCRN with SNR estimation for speech enhancement
S. Lv, Y. Hu, S. Zhang, and L. Xie · 2021
Later among the works it cites.
Cascaded time + time-frequency UNet for speech enhancement: jointly addressing clipping, codec distortions, and gaps
A. A. Nair and K. Koishida · 2021
Later among the works it cites.
High fidelity speech regeneration with application to speech enhancement
A. Polyak, L. Wolf, Y. Adi, O. Kabeli, and Y. Taigman · 2021
Later among the works it cites.
Upsampling artifacts in neural audio synthesis
J. Pons, S. Pascual, G. Cengarle, and J. Serrà · 2021
Later among the works it cites.
Grad-TTS: a diffusion probabilistic model for text-to-speech
Original
V. Popov, I. Vovk, V. Gogoryan, T. Sadekova, and M. Kudinov · 2021
Later among the works it cites.
CRASH: raw audio score-based generative modeling for controllable high-resolution drum sound synthesis
S. Rouard and G. Hadjeres · 2021
Later among the works it cites.
Score-based generative modeling through stochastic differential equations
Y. Song, J. Sohl-Dickstein, D. P. Kingma, A. Kumar, S. Ermon, and B. Poole · 2021
Later among the works it cites.
A flow-based neural network for time domain speech enhancement
M. Strauss and B. Edler · 2021
Later among the works it cites.
HiFi-GAN-2: studio-quality speech enhancement via generative adversarial networks conditioned on acoustic features
J. Su, Z. Jin, and A. Finkelstein · 2021
Later among the works it cites.
A two-stage complex network using cycle-consistent generative adversarial networks for speech enhancement
G. Yu, Y. Wang, H. Wang, Q. Zhang, and C. Zheng · 2021
Later among the works it cites.
Interactive speech and noise modeling for speech enhancement
C. Zheng, X. Peng, Y. Zhang, S. Srinivasan, and Y. Lu · 2021
Later among the works it cites.
Conditional diffusion probabilistic model for speech enhancement
Y.-J. Lu, Z.-Q. Wang, S. Watanabe, A. Richard, C. Yu, and Y. Tsao · 2022
Closest in time.
Phase sensitive masking-based single channel speech enhancement using conditional generative adversarial network
S. Routray and Q. Mao · 2022
Closest in time.
Speech enhancement with score-based generative models in the complex STFT domain
S. Welker, J. Richter, and T. Gerkmann · 2022
Closest in time.