Fetching the paper…
Reading the bibliography…
The speech enhancement task usually consists of removing additive noise or reverberation that partially mask spoken utterances, affecting their intelligibility.
1904
Earlier work this paper cites.
R. Kubichek, “Mel-cepstral distance measure for objective speech quality assessment,” in
1993
Earlier work this paper cites.
L.-P. Yang and Q.-J. Fu, “Spectral subtraction-based speech enhancement for cochlear implant patients in background noise,”
2005
Earlier work this paper cites.
D. Erro, I. Sainz, E. Navas, and I. Hernáez, “Improved hnm-based vocoder for statistical synthesizers,” in
2011
Earlier work this paper cites.
P. C. Loizou,
2013
Earlier work this paper cites.
X. Lu, Y. Tsao, S. Matsuda, and C. Hori, “Speech enhancement based on deep denoising autoencoder.” in
2013
Earlier work this paper cites.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” in
2014
Earlier work this paper cites.
Y. Xu, J. Du, L.-R. Dai, and C.-H. Lee, “A regression approach to speech enhancement based on deep neural networks,”
2015
Earlier work this paper cites.
F. Weninger, H. Erdogan, S. Watanabe, E. Vincent, J. Le Roux, J. R. Hershey, and B. Schuller, “Speech enhancement with LSTM recurrent neural networks and its application to noise-robust ASR,” in
2015
Earlier work this paper cites.
H. Erdogan, J. R. Hershey, S. Watanabe, and J. Le Roux, “Phase-sensitive and recognition-boosted speech separation using deep recurrent neural networks,” in
2015
Earlier work this paper cites.
A. Radford, L. Metz, and S. Chintala, “Unsupervised representation learning with deep convolutional generative adversarial networks,”
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Delving deep into rectifiers: Surpassing human-level performance on imagenet classification,” in
2015
Earlier work this paper cites.
C. Veaux, J. Yamagishi, K. MacDonald
2016
Cited alongside, same era.
S. Pascual and A. Bonafonte, “Multi-output rnn-lstm for multiple speaker speech synthesis and adaptation,” in
2016
Cited alongside, same era.
D. S. Williamson and D. Wang, “Time-frequency masking in the complex domain for speech dereverberation and denoising,”
2017
Cited alongside, same era.
S. R. Park and J. Lee, “A fully convolutional neural network for speech enhancement,” in
2017
Cited alongside, same era.
S. Pascual, A. Bonafonte, and J. Serrà, “Segan: Speech enhancement generative adversarial network,” in
2017
Cited alongside, same era.
T. Higuchi, K. Kinoshita, M. Delcroix, and T. Nakatani, “Adversarial training for data-driven speech enhancement without parallel corpus,” in
D. Rethage, J. Pons, and X. Serra, “A wavenet for speech denoising,” in
2018
Later among the works it cites.
S. Qin and T. Jiang, “Improved wasserstein conditional generative adversarial network speech enhancement,”
2018
Later among the works it cites.
C. Donahue, B. Li, and R. Prabhavalkar, “Exploring speech enhancement with generative adversarial networks for robust speech recognition,” in
2018
Later among the works it cites.
S. Pascual, A. Bonafonte, J. Serrà, and J. A.-G. López, “Whispered-to-voiced alaryngeal speech conversion with generative adversarial networks,”
2018
Later among the works it cites.
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
A. Odena, C. Olah, and J. Shlens, “Conditional image synthesis with auxiliary classifier gans,” in
2017
Cited alongside, same era.
2017
Cited alongside, same era.
X. Mao, Q. Li, H. Xie, R. Y. Lau, Z. Wang, and S. P. Smolley, “Least Squares Generative Adversarial Networks,” in
2017
Cited alongside, same era.
P. Isola, J.-Y. Zhu, T. Zhou, and A. A. Efros, “Image-to-image translation with conditional adversarial networks,” in
2017
Cited alongside, same era.
M. Heusel, H. Ramsauer, T. Unterthiner, B. Nessler, and S. Hochreiter, “Gans trained by a two time-scale update rule converge to a local nash equilibrium,” in
2017
Cited alongside, same era.
S. Pascual, M. Park, J. Serrà, A. Bonafonte, and K.-H. Ahn, “Language and noise transfer in speech enhancement generative adversarial network,” in
2018
Later among the works it cites.
W. Ping, K. Peng, and J. Chen, “Clarinet: Parallel wave generation in end-to-end text-to-speech,”
2018
Later among the works it cites.
H. Zhang, I. Goodfellow, D. Metaxas, and A. Odena, “Self-attention generative adversarial networks,”
2018
Later among the works it cites.
2018
Later among the works it cites.
C. Donahue, J. McAuley, and M. Puckette, “Synthesizing audio with generative adversarial networks,”
2018
Later among the works it cites.