Fetching the paper…
Reading the bibliography…
In this paper, we propose the coarse-to-fine optimization for the task of speech enhancement.
J. B. Allen, “Short term spectral analysis, synthesis, and modification by discrete fourier transform,”
1977
Earlier work this paper cites.
J. Lim and A. Oppenheim, “All-pole modeling of degraded speech,”
1978
Earlier work this paper cites.
M. Berouti, R. Schwartz, and J. Makhoul, “Enhancement of speech corrupted by acoustic noise,”
1979
Earlier work this paper cites.
D. J. Fleet and Y. Weiss, “Optical flow estimation,” in
2005
Earlier work this paper cites.
“P.862.2: Wideband extension to recommendation p. 862 for the assessment of wideband telephone networks and speech codecs,” in
2007
Earlier work this paper cites.
Y. Hu and P. C. Loizou, “Evaluation of objective quality measures for speech enhancement,”
2008
Earlier work this paper cites.
X. Lu, Y. Tsao, S. Matsuda, and C. Hori, “Speech enhancement based on deep denoising autoencoder,” in
2013
Earlier work this paper cites.
A. L. Maas, A. Y. Hannun, and A. Y. Ng, “Rectifier nonlinearities improve neural network acoustic models,” in
2013
Earlier work this paper cites.
J. Thiemann, N. Ito, and E. Vincent, “The diverse environments multi-channel acoustic noise database: A database of multichannel environmental noise recordings,”
2013
Earlier work this paper cites.
C. Veaux, J. Yamagishi, and S. King, “The voice bank corpus: Design, collection and data analysis of a large regional accent speech database,” in
2013
Earlier work this paper cites.
P. C. Loizou,
2013
Earlier work this paper cites.
I. J. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” in
2014
Earlier work this paper cites.
M. Mirza and S. Osindero, “Conditional generative adversarial nets,” in
2014
Earlier work this paper cites.
Y. Taigman, M. Yang, M. A. Ranzato, and L. Wolf, “Deepface: Closing the gap to human-level performance in face verification,” in
2014
Cited alongside, same era.
Y. Xu, J. Du, L.-R. Dai, and C.-H. Lee, “A regression approach to speech enhancement based on deep neural networks,” in
2015
Cited alongside, same era.
F. Weninger, H. Erdogan, S. Watanabe, E. Vincent, J. L. Roux, J. R. Hershey, and B. Schuller, “Speech enhancement with lstm recurrent neural networks and its application to noise-robust asr,” in
2015
Cited alongside, same era.
J. Yao, M. Boben, S. Fidler, and R. Urtasun, “Real-time coarse-to-fine topologically preserving segmentation,” in
2015
Cited alongside, same era.
V. Dumoulin and F. Visin, “A guide to convolution arithmetic for deep learning,” in
2015
D. Michelsanti and Z.-H. Tan, “Conditional generative adversarial networks for speech enhancement and noise-robust speaker verification,” in
2017
Later among the works it cites.
T. Kaneko, S. Takaki, H. Kameoka, and J. Yamagishi, “Generative adversarial network based postfilter for stft spectrograms,” in
2017
Later among the works it cites.
P. Isola, J. Zhu, T. Zhou, and A. A. Efros, “Image-to-image translation with conditional adversarial networks,” in
2017
Later among the works it cites.
S. R. Park and J. W. Lee, “A fully convolutional neural network for speech enhancement,” in
2017
Later among the works it cites.
D. S. Williamson and D. Wang, “Time-frequency masking in the complex domain for speech dereverberation and denoising,”
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
S. Ioffe and C. Szegedy, “Batch normalization: Accelerating deep network training by reducing internal covariate shift,” in
2015
Cited alongside, same era.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,” in
2015
Cited alongside, same era.
A. Kuman and D. Florencio, “Speech enhancement in multiple noise conditions using deep neural networks,” in
2016
Cited alongside, same era.
M. Mathieu, C. Couprie, and Y. LeCun, “Deep multi-scale video prediction beyond mean square error,” in
2016
Cited alongside, same era.
I. J. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative image modeling using style and structure adversarial networks,” in
2016
Cited alongside, same era.
C. Valentini-Botinhao, “Noisy speech database for training speech enhancement algorithms and tts models,” in
2016
Cited alongside, same era.
S. Pascual, A. Bonafonte, and J. Serra, “Segan: Speech enhancement generative adversarial network,” in
2017
Cited alongside, same era.
S. Novoselov, V. Shchemelinin, A. Shulipa, A. Kozlov, and I. Kremnev, “Triplet loss based cosine similarity metric learning for text-independent speaker recognition,” in
2018
Later among the works it cites.
F. G. Germain, Q. Chen, and V. Koltun, “Speech denoising with deep feature losses,” in
2018
Later among the works it cites.
M. H. Soni, N. Shah, and H. A. Patil, “Time-frequency masking-based speech enhancement using generative adversarial network,” in
2018
Later among the works it cites.
K. Tan and D. Wang, “A convolutional recurrent neural network for real-time speech enhancement,” in
2018
Later among the works it cites.
S. Fu, T. Wang, Y. Tsao, X. Lu, and H. Kawai, “End-to-end waveform utterance enhancement for direct evaluation metrics optimization by fully convolutional neural networks,”
2018
Later among the works it cites.
D. Rethage, J. Pons, and X. Serra, “A wavenet for speech denoising,” in
2018
Later among the works it cites.
H.-S. Choi, J.-H. Kim, J. Huh, A. Kim, J.-W. Ha, and K. Lee, “Phase-aware speech enhancement with deep complex u-net,” in
2019
Closest in time.