Fetching the paper…
Reading the bibliography…
Real-world audio recordings are often degraded by factors such as noise, reverberation, and equalization distortion.
Y. Ephraim and D. Malah, “Speech enhancement using a minimum-mean square error short-time spectral amplitude estimator,”
1984
Earlier work this paper cites.
P. Scalart
1996
Earlier work this paper cites.
P. A. Naylor and N. D. Gaubitch,
2010
Earlier work this paper cites.
T. Nakatani, T. Yoshioka, K. Kinoshita, M. Miyoshi, and B.-H. Juang, “Speech dereverberation based on variance-normalized delayed linear prediction,”
2010
Earlier work this paper cites.
Z. Duan, G. J. Mysore, and P. Smaragdis, “Speech enhancement by online non-negative spectrogram decomposition in nonstationary noise environments,” in
2012
Earlier work this paper cites.
K. Kinoshita, M. Delcroix, T. Yoshioka, T. Nakatani, A. Sehr, W. Kellermann, and R. Maas, “The reverb challenge: A common evaluation framework for dereverberation and recognition of reverberant speech,” in
2013
Earlier work this paper cites.
K. Han, Y. Wang, D. Wang, W. S. Woods, I. Merks, and T. Zhang, “Learning spectral mapping for speech dereverberation and denoising,”
2015
Earlier work this paper cites.
Y. Xu, J. Du, L.-R. Dai, and C.-H. Lee, “A regression approach to speech enhancement based on deep neural networks,”
2015
Earlier work this paper cites.
L. A. Gatys, A. S. Ecker, and M. Bethge, “A neural algorithm of artistic style,”
2015
Earlier work this paper cites.
G. J. Mysore, “Can we automatically transform speech recorded on common consumer devices in real-world environments into professional production quality speech? A dataset, insights, and challenges,”
2015
Earlier work this paper cites.
F. G. Germain, G. J. Mysore, and T. Fujioka, “Equalization matching of speech recordings in real-world environments,” in
2016
Earlier work this paper cites.
A. Van Den Oord, S. Dieleman, H. Zen, K. Simonyan, O. Vinyals, A. Graves, N. Kalchbrenner, A. W. Senior, and K. Kavukcuoglu, “Wavenet: A generative model for raw audio.” in
2016
Earlier work this paper cites.
I. Goodfellow, “Nips 2016 tutorial: Generative adversarial networks,”
2016
Earlier work this paper cites.
J. Traer and J. H. McDermott, “Statistics of natural reverberation enable perceptual separation of sound and space,”
2016
Earlier work this paper cites.
J. Eaton, N. D. Gaubitch, A. H. Moore, and P. A. Naylor, “Estimation of room acoustic parameters: The ace challenge,”
2016
Cited alongside, same era.
K. Kinoshita, M. Delcroix, S. Gannot, E. A. Habets, R. Haeb-Umbach, W. Kellermann, V. Leutnant, R. Maas, T. Nakatani, B. Raj
2016
Cited alongside, same era.
C. Valentini-Botinhao, X. Wang, S. Takaki, and J. Yamagishi, “Investigating rnn-based speech enhancement methods for noise-robust text-to-speech.” in
2016
Cited alongside, same era.
D. S. Williamson and D. Wang, “Speech dereverberation and denoising using complex ratio masks,” in
2017
Cited alongside, same era.
S. Pascual, A. Bonafonte, and J. Serrà, “Segan: Speech enhancement generative adversarial network,”
2017
Cited alongside, same era.
2018
Later among the works it cites.
J. Su, A. Finkelstein, and Z. Jin, “Perceptually-motivated environment-specific speech enhancement,” in
2019
Later among the works it cites.
R. Giri, U. Isik, and A. Krishnaswamy, “Attention wave-u-net for speech enhancement,” in
2019
Later among the works it cites.
2019
Later among the works it cites.
K. Kumar, R. Kumar, T. de Boissiere, L. Gestin, W. Z. Teoh, J. Sotelo, A. de Brébisson, Y. Bengio, and A. C. Courville, “Melgan: Generative adversarial networks for conditional waveform synthesis,” in
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
K. Qian, Y. Zhang, S. Chang, X. Yang, D. Florêncio, and M. Hasegawa-Johnson, “Speech enhancement using bayesian wavenet.” in
2017
Cited alongside, same era.
J. H. Lim and J. C. Ye, “Geometric gan,”
2017
Cited alongside, same era.
C. Donahue, B. Li, and R. Prabhavalkar, “Exploring speech enhancement with generative adversarial networks for robust speech recognition,” in
2018
Cited alongside, same era.
H. Kagami, H. Kameoka, and M. Yukawa, “Joint separation and dereverberation of reverberant mixtures with determined multichannel non-negative matrix factorization,” in
2018
Cited alongside, same era.
W. Mack, S. Chakrabarty, F.-R. Stöter, S. Braun, B. Edler, and E. Habets, “Single-channel dereverberation using direct mmse optimization and bidirectional lstm networks,”
2018
Cited alongside, same era.
D. Rethage, J. Pons, and X. Serra, “A wavenet for speech denoising,” in
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2019
Later among the works it cites.
S. Pascual, J. Serrà, and A. Bonafonte, “Towards generalized speech enhancement with generative adversarial networks,”
2019
Later among the works it cites.
S.-W. Fu, C.-F. Liao, and Y. Tsao, “Learning with learned loss function: Speech enhancement with quality-net to improve perceptual evaluation of speech quality,”
2019
Later among the works it cites.
F. G. Germain, Q. Chen, and V. Koltun, “Speech denoising with deep feature losses,”
2019
Later among the works it cites.
2019
Later among the works it cites.
2019
Later among the works it cites.
2020
Closest in time.
J. Su, Z. Jin, and A. Finkelstein, “Acoustic matching by embedding impulse responses,” in
2020
Closest in time.
S.-W. Fu, C.-F. Liao, Y. Tsao, and S.-D. Lin, “Metricgan: Generative adversarial networks based black-box metric scores optimization for speech enhancement,” in
2041
Closest in time.