Fetching the paper…
Reading the bibliography…
Improving speech system performance in noisy environments remains a challenging task, and speech enhancement (SE) is one of the effective techniques to solve the problem.
Y. Ephraim and D. Malah, “Speech enhancement using a minimum-mean square error short-time spectral amplitude estimator,”
1984
Earlier work this paper cites.
J. S. Garofolo, L. F. Lamel, W. M. Fisher, J. G. Fiscus, and D. S. Pallett, “Darpa timit acoustic-phonetic continous speech corpus cd-rom. nist speech disc 1-1.1,”
1993
Earlier work this paper cites.
A. W. Rix, J. G. Beerends, M. P. Hollier, and A. P. Hekstra, “Perceptual evaluation of speech quality (pesq)-a new method for speech quality assessment of telephone networks and codecs,” in
2001
Earlier work this paper cites.
D. Wang and G. J. Brown,
2006
Earlier work this paper cites.
R. C. Hendriks, R. Heusdens, and J. Jensen, “Mmse based noise psd tracking with low complexity,” in
2010
Earlier work this paper cites.
C. H. Taal, R. C. Hendriks, R. Heusdens, and J. Jensen, “An algorithm for intelligibility prediction of time–frequency weighted noisy speech,”
2011
Earlier work this paper cites.
M. L. Seltzer, D. Yu, and Y. Wang, “An investigation of deep neural networks for noise robust speech recognition,” in
2013
Earlier work this paper cites.
X. Lu, Y. Tsao, S. Matsuda, and C. Hori, “Speech enhancement based on deep denoising autoencoder,” in
2013
Earlier work this paper cites.
P. C. Loizou,
2013
Earlier work this paper cites.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” in
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
D. Kingma and J. Ba, “Adam: A method for stochastic optimization,”
2014
Earlier work this paper cites.
M. Mirza and S. Osindero, “Conditional generative adversarial nets,”
2014
Earlier work this paper cites.
Y. Wang, A. Narayanan, and D. Wang, “On training targets for supervised speech separation,”
2014
Cited alongside, same era.
A. Larcher, K. A. Lee, B. Ma, and H. Li, “Text-dependent speaker verification: Classifiers, databases and rsr2015,”
2014
Cited alongside, same era.
Y. Xu, J. Du, L.-R. Dai, and C.-H. Lee, “A regression approach to speech enhancement based on deep neural networks,”
2015
Cited alongside, same era.
2015
Cited alongside, same era.
2015
Cited alongside, same era.
2016
Later among the works it cites.
S. Mobin and J. Bruna, “Voice conversion using convolutional neural networks,”
2016
Later among the works it cites.
O. Mogren, “C-rnn-gan: Continuous recurrent neural networks with adversarial training,”
2016
Later among the works it cites.
X. Wang and A. Gupta, “Generative image modeling using style and structure adversarial networks,” in
2016
Later among the works it cites.
D. Pathak, P. Krahenbuhl, J. Donahue, T. Darrell, and A. A. Efros, “Context encoders: Feature learning by inpainting,” in
2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. E. Reed, Y. Zhang, Y. Zhang, and H. Lee, “Deep visual analogy-making,” in
2015
Cited alongside, same era.
O. Ronneberger, P. Fischer, and T. Brox, “U-net: Convolutional networks for biomedical image segmentation,” in
2015
Cited alongside, same era.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: an asr corpus based on public domain audio books,” in
2015
Cited alongside, same era.
2016
Cited alongside, same era.
M. Kolbœk, Z.-H. Tan, and J. Jensen, “Speech enhancement using long short-term memory based recurrent neural networks for noise robust speaker verification,” in
2016
Cited alongside, same era.
J. Chen, Y. Wang, S. E. Yoho, D. Wang, and E. W. Healy, “Large-scale training to increase speech intelligibility for hearing-impaired listeners in novel noises,”
2016
Cited alongside, same era.
S. R. Park and J. Lee, “A fully convolutional neural network for speech enhancement,”
2016
Cited alongside, same era.
A. K. Sarkar and Z.-H. Tan, “Text dependent speaker verification using un-supervised hmm-ubm and temporal gmm-ubm,”
2016
Later among the works it cites.
M. Falcone, B. Fauve, M. Cornacchia
2016
Later among the works it cites.
M. Kolbæk, Z.-H. Tan, and J. Jensen, “Speech intelligibility potential of general and specialized deep neural network based speech enhancement systems,”
2017
Closest in time.
2017
Closest in time.
Y.-C. Lin, “pix2pix-tensorflow,” Github repository: https://github.com/yenchenlin/pix2pix-tensorflow, 2016, accessed: March 2017
2017
Closest in time.
ITU, “Wideband extension to recommendation p.862 for the assessment of wideband telephone networks and speech codecs,” Available: https://www.itu.int/rec/T-REC-P.862.2-200511-S/en, 2005, accessed: March 2017
2017
Closest in time.