Fetching the paper…
Reading the bibliography…
In this paper we consider the problem of speech enhancement in real-world like conditions where multiple noises can simultaneously corrupt speech.
H. Fletcher, “Auditory patterns,”
1940
Earlier work this paper cites.
D. D. Greenwood, “Critical bandwidth and the frequency coordinates of the basilar membrane,”
1961
Earlier work this paper cites.
J. Zwislocki, “Analysis of some auditory characteristics.” DTIC Document, Tech. Rep., 1963
1963
Earlier work this paper cites.
B. Scharf, “Critical bands,”
1970
Earlier work this paper cites.
R. P. Hellman, “Asymmetry of masking between noise and tone,”
1972
Earlier work this paper cites.
S. F. Boll, “Suppression of acoustic noise in speech using spectral subtraction,”
1979
Earlier work this paper cites.
J. S. Lim and A. V. Oppenheim, “Enhancement and bandwidth compression of noisy speech,”
1979
Earlier work this paper cites.
E. Terhardt, “Calculating virtual pitch,”
1979
Earlier work this paper cites.
Y. Ephraim and D. Malah, “Speech enhancement using a minimum-mean square error short-time spectral amplitude estimator,”
1984
Earlier work this paper cites.
D. W. Griffin and J. S. Lim, “Signal estimation from modified short-time fourier transform,”
1984
Earlier work this paper cites.
——, “Speech enhancement using a minimum mean-square error log-spectral amplitude estimator,”
1985
Earlier work this paper cites.
J. S. Garofolo, L. F. Lamel, W. M. Fisher, J. G. Fiscus, and D. S. Pallett, “Darpa timit acoustic-phonetic continous speech corpus cd-rom. nist speech disc 1-1.1,”
1993
Earlier work this paper cites.
Y. Ephraim and H. L. Van Trees, “A signal subspace approach for speech enhancement,”
1995
Cited alongside, same era.
T. Painter and A. Spanias, “A review of algorithms for perceptual coding of digital audio signals,” IEEE, pp. 179–208, 1997
1997
Cited alongside, same era.
A. W. Rix, J. G. Beerends, M. P. Hollier, and A. P. Hekstra, “Perceptual evaluation of speech quality (pesq)-a new method for speech quality assessment of telephone networks and codecs,” IEEE, pp. 749–752, 2001
2001
Cited alongside, same era.
Y. Hu and P. C. Loizou, “A generalized subspace approach for enhancing speech corrupted by colored noise,”
2003
Cited alongside, same era.
I. Cohen and S. Gannot, “Spectral enhancement methods,” in
2008
Cited alongside, same era.
X. Lu, Y. Tsao, S. Matsuda, and C. Hori, “Speech enhancement based on deep denoising autoencoder.” pp. 436–440, 2013
2013
Later among the works it cites.
M. Seltzer, D. Yu, and Y. Wang, “An investigation of deep neural networks for noise robust speech recognition,” IEEE, pp. 7398–7402, 2013
2013
Later among the works it cites.
A. Narayanan and D. Wang, “Ideal ratio mask estimation using deep neural networks for robust speech recognition,” IEEE, pp. 7092–7096, 2013
2013
Later among the works it cites.
T. Gerkmann and M. Krawczyk, “Mmse-optimal spectral amplitude estimation given the stft-phase,”
2013
Later among the works it cites.
A. Graves and N. Jaitly, “Towards end-to-end speech recognition with recurrent neural networks,” pp. 1764–1772, 2014
2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
C. H. Taal, R. C. Hendriks, R. Heusdens, and J. Jensen, “An algorithm for intelligibility prediction of time–frequency weighted noisy speech,”
2011
Cited alongside, same era.
G. Hinton, L. Deng, D. Yu, G. Dahl, A. rahman Mohamed, N. Jaitly, A. Senior, V. Vanhoucke, P. Nguyen, T. Sainath, and B. Kingsbury, “Deep neural networks for acoustic modeling in speech recognition,”
2012
Cited alongside, same era.
G. E. Dahl, D. Yu, L. Deng, and A. Acero, “Context-dependent pre-trained deep neural networks for large-vocabulary speech recognition,”
2012
Cited alongside, same era.
A. L. Maas, Q. V. Le, T. M. O’Neil, O. Vinyals, P. Nguyen, and A. Y. Ng, “Recurrent neural networks for noise reduction in robust asr.” Citeseer, 2012
2012
Cited alongside, same era.
2012
Cited alongside, same era.
P. C. Loizou,
2013
Cited alongside, same era.
A. Graves, A.-r. Mohamed, and G. Hinton, “Speech recognition with deep recurrent neural networks,” IEEE, pp. 6645–6649, 2013
2013
Cited alongside, same era.
X. Lu, Y. Tsao, S. Matsuda, and C. Hori, “Ensemble modeling of denoising autoencoder for speech spectrum restoration,” pp. 885–889, 2014
2014
Later among the works it cites.
B. Xia and C. Bao, “Wiener filtering based speech enhancement with weighted denoising auto-encoder and noise classification,”
2014
Later among the works it cites.
Y. Wang, A. Narayanan, and D. Wang, “On training targets for supervised speech separation,”
2014
Later among the works it cites.
Y. Xu, J. Du, L.-R. Dai, and C.-H. Lee, “A regression approach to speech enhancement based on deep neural networks,”
2015
Later among the works it cites.
Y. Xu, J. Du, L.-R. Dai, and C.-H. Lee, “A regression approach to speech enhancement based on deep neural networks,”
2015
Later among the works it cites.
H.-W. Tseng, M. Hong, and Z.-Q. Luo, “Combining sparse nmf with deep neural network: A new classification-based approach for speech enhancement,” IEEE, pp. 2145–2149, 2015
2015
Later among the works it cites.
FreeSound, “
2015
Later among the works it cites.