Fetching the paper…
Reading the bibliography…
With the development of audio playback devices and fast data transmission, the demand for high sound quality is rising for both entertainment and communications.
Springer, 1998
S. J. Godsill and P. J. W. Rayner, Digital Audio Restoration—A Statistical Model Based Approach · 1998
Earlier work this paper cites.
Cambridge University Press, 3rd ed., 2007
W. H. Press, S. A. Teukolsky, W. T. Vetterling, and B. P. Flannery, Numerical Recipes: The Art of Scientific Computing · 2007
Earlier work this paper cites.
T. Gerkmann and R. Martin, “Empirical distributions of DFT-domain speech coefficients based on estimated speech variances,” in Proc. Int. Workshop Acoustic Signal Enhancement
2010
Earlier work this paper cites.
P. Vincent, “A connection between score matching and denoising autoencoders,” Neural Computation
2011
Earlier work this paper cites.
Springer, 2013
B. Øksendal, Stochastic Differential Equations: An Introduction with Applications · 2013
Earlier work this paper cites.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial networks,” in Proc. Neural Inf. Process. Syst
2014
Earlier work this paper cites.
D. P. Kingma and M. Welling, “Auto-encoding variational Bayes,” Proc. Int. Conf. Learning Repr
2014
Earlier work this paper cites.
T. Gerkmann and E. Vincent, “Spectral masking and filtering,” in Audio Source Separation and Speech Enhancement
2018
Earlier work this paper cites.
D. Wang and J. Chen, “Supervised speech separation based on deep learning: An overview,” IEEE Trans. Audio Speech Lang. Process
2018
Earlier work this paper cites.
Y. Song and S. Ermon, “Generative modeling by estimating gradients of the data distribution,” in Proc. Neural Inf. Process. Syst
2019
Earlier work this paper cites.
J. Ho, A. Jain, and P. Abbeel, “Denoising diffusion probabilistic models,” in Proc. Neural Inf. Process. Syst
2020
Earlier work this paper cites.
Y. Song, J. Sohl-Dickstein, D. P. Kingma, A. Kumar, S. Ermon, and B. Poole, “Score-based generative modeling through stochastic differential equations,” in Proc. Int. Conf. Learning Repr
2021
Earlier work this paper cites.
N. Chen, Y. Zhang, H. Zen, R. J. Weiss, M. Norouzi, and W. Chan, “WaveGrad: Estimating gradients for waveform generation,” Proc. Int. Conf. Learning Repr
2021
Earlier work this paper cites.
Z. Kong, W. Ping, J. Huang, K. Zhao, and B. Catanzaro, “DiffWave: A versatile diffusion model for audio synthesis,” Proc. Int. Conf. Learning Repr
2021
Earlier work this paper cites.
Y.-J. Lu, Y. Tsao, and S. Watanabe, “A study on speech enhancement based on diffusion probabilistic model,” in Proc. Asia-Pacific Signal and Information Processing Association (APSIPA)
2021
Cited alongside, same era.
Y.-J. Lu, Z.-Q. Wang, S. Watanabe, A. Richard, C. Yu, and Y. Tsao, “Conditional diffusion probabilistic model for speech enhancement,” in Proc. IEEE Int. Conf. Acoust. Speech Signal Process
2022
Cited alongside, same era.
S. Welker, J. Richter, and T. Gerkmann, “Speech enhancement with score-based generative models in the complex STFT domain,” in Proc. Interspeech
2022
Cited alongside, same era.
E. Nachmani, R. S. Roman, and L. Wolf, “Denoising diffusion gamma models,” in Proc. Int. Conf. Learning Repr
2022
Cited alongside, same era.
J. Song, C. Meng, and S. Ermon, “Denoising diffusion implicit models,” in Proc. Int. Conf. Learning Repr
2022
B. Lay, S. Welker, J. Richter, and T. Gerkmann, “Reducing the prior mismatch of stochastic differential equations for diffusion-based speech enhancement,” in Proc. Interspeech
2023
Later among the works it cites.
H. Chung, J. Kim, M. T. Mccann, M. L. Klasky, and J. C. Ye, “Diffusion posterior sampling for general noisy inverse problems,” in Proc. Int. Conf. Learning Repr
2023
Later among the works it cites.
J.-M. Lemercier, S. Welker, and T. Gerkmann, “Diffusion posterior sampling for informed single-channel dereverberation,” in Proc. IEEE Workshop Appl. Signal Process. Audio Acoust
2023
Later among the works it cites.
R. Scheibler, Y. Ji, S.-W. Chung, J. Byun, S. Choe, and M.-S. Choi, “Diffusion-based generative speech source separation,” in Proc. IEEE Int. Conf. Acoust. Speech Signal Process
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
R. Huang, Z. Zhao, H. Liu, J. Liu, C. Cui, and Y. Ren, “ProDiff: Progressive fast diffusion model for high-quality text-to-speech,” in ACM Multimedia
2022
Cited alongside, same era.
M. W. Y. Lam, J. Wang, D. Su, and D. Yu, “BDDM: Bilateral denoising diffusion models for fast and high-quality speech synthesis,” in Proc. Int. Conf. Learning Repr
2022
Cited alongside, same era.
MIT Press, 2023
K. P. Murphy, Probabilistic Machine Learning: Advanced Topics · 2023
Cited alongside, same era.
E. Moliner, J. Lehtinen, and V. Välimäki, “Solving audio inverse problems with a diffusion model,” in Proc. IEEE Int. Conf. Acoust. Speech Signal Process
2023
Cited alongside, same era.
J. Richter, S. Welker, J.-M. Lemercier, B. Lay, and T. Gerkmann, “Speech enhancement and dereverberation with diffusion-based generative models,” IEEE/ACM Trans. Audio Speech Lang. Process
2023
Cited alongside, same era.
J.-M. Lemercier, J. Richter, S. Welker, and T. Gerkmann, “StoRM: A diffusion-based stochastic regeneration model for speech enhancement and dereverberation,” IEEE Trans. Audio Speech Lang. Process
2023
Cited alongside, same era.
H. Liu, Z. Chen, Y. Yuan, X. Mei, X. Liu, D. Mandic, W. Wang, and M. D. Plumbley, “AudioLDM: Text-to-audio generation with latent diffusion models,” in Proc. Int. Conf. Machine Learning
2023
Cited alongside, same era.
2023
Later among the works it cites.
Y. Wang, Z. Ju, X. Tan, L. He, Z. Wu, J. Bian, and S. Zhao, “Audit: Audio editing by following instructions with latent diffusion models,” in Proc. Neural Inf. Process. Syst
2023
Later among the works it cites.
J. Richter, S. Frintrop, and T. Gerkmann, “Audio-visual speech enhancement with score-based generative models,” in Proc. ITG Conf. Speech Communication
2023
Later among the works it cites.
E. Moliner and V. Välimäki, “Diffusion-based audio inpainting,” J. Audio Eng. Soc
2024
Closest in time.
2024
Closest in time.
A. H. Liu, M. Le, A. Vyas, B. Shi, A. Tjandra, and W.-N. Hsu, “Generative pre-training for speech with flow matching,” in Proc. Int. Conf. Learning Repr
2024
Closest in time.
E. Moliner, J.-M. Lemercier, S. Welker, T. Gerkmann, and V. Välimäki, “BUDDy: Single-channel blind unsupervised dereverberation with diffusion models,” in Proc. Int. Workshop Acoustic Signal Enhancement
2024
Closest in time.
B. Lay, J.-M. Lemercier, J. Richter, and T. Gerkmann, “Single and few-step diffusion for generative speech enhancement,” in Proc. IEEE Int. Conf. Acoust. Speech Signal Process
2024
Closest in time.
J. Richter, S. Welker, J.-M. Lemercier, B. Lay, T. Peer, and T. Gerkmann, “Causal diffusion models for generalized speech enhancement,” IEEE Open Journal of Signal Processing
2024
Closest in time.