Fetching the paper…
Reading the bibliography…
In this work, we tackle a denoising and dereverberation problem with a single-stage framework.
A. W. Rix, J. G. Beerends, M. P. Hollier, and A. P. Hekstra, “Perceptual evaluation of speech quality (pesq)-a new method for speech quality assessment of telephone networks and codecs,” in
2001
Earlier work this paper cites.
J. S. Bradley, H. Sato, and M. Picard, “On the importance of early reflections for speech in rooms,”
2003
Earlier work this paper cites.
I. McCowan, D. Dean, M. McLaren, R. Vogt, and S. Sridharan, “The delta-phase spectrum with application to voice activity detection and speaker recognition,”
2011
Earlier work this paper cites.
P. Mowlaee, R. Saeidi, and R. Martin, “Phase estimation for signal reconstruction in single-channel source separation,” in
2012
Earlier work this paper cites.
Y. Hu and K. Kokkinakis, “Effects of early and late reflections on intelligibility of reverberated speech by cochlear implant listeners,”
2014
Earlier work this paper cites.
P. Mowlaee and R. Saeidi, “Time-frequency constraints for phase estimation in single-channel speech enhancement,” in
2014
Earlier work this paper cites.
H. Erdogan, J. R. Hershey, S. Watanabe, and J. Le Roux, “Phase-sensitive and recognition-boosted speech separation using deep recurrent neural networks,” in
2015
Earlier work this paper cites.
D. S. Williamson, Y. Wang, and D. Wang, “Complex ratio masking for monaural speech separation,”
2016
Earlier work this paper cites.
E. Jang, S. Gu, and B. Poole, “Categorical reparameterization with gumbel-softmax,”
2016
Earlier work this paper cites.
J. F. Gemmeke, D. P. Ellis, D. Freedman, A. Jansen, W. Lawrence, R. C. Moore, M. Plakal, and M. Ritter, “Audio set: An ontology and human-labeled dataset for audio events,” in
2017
Earlier work this paper cites.
E. Fonseca, J. Pons Puig, X. Favory, F. Font Corbera, D. Bogdanov, A. Ferraro, S. Oramas, A. Porter, and X. Serra, “Freesound datasets: a platform for the creation of open audio datasets,” in
2017
Earlier work this paper cites.
N. Takahashi and Y. Mitsufuji, “Multi-scale multi-band densenets for audio source separation,” in
2017
Cited alongside, same era.
Y. Sun, W. Wang, J. A. Chambers, and S. M. Naqvi, “Enhanced time-frequency masking by using neural networks for monaural source separation in reverberant room environments,” in
2018
Cited alongside, same era.
2018
Cited alongside, same era.
N. Takahashi, P. Agrawal, N. Goswami, and Y. Mitsufuji, “Phasenet: Discretized phase modeling with deep neural networks for audio source separation,”
2018
Cited alongside, same era.
T. Afouras, J. S. Chung, and A. Zisserman, “The conversation: Deep audio-visual speech enhancement,”
2018
Cited alongside, same era.
J. Yao and A. Al-Dahle, “Coarse-to-fine optimization for speech enhancement,”
2019
Later among the works it cites.
J. Le Roux, S. Wisdom, H. Erdogan, and J. R. Hershey, “Sdr–half-baked or well done?” in
2019
Later among the works it cites.
K. Tan and D. Wang, “Complex spectral mapping with a convolutional recurrent network for monaural speech enhancement,” in
2019
Later among the works it cites.
2019
Later among the works it cites.
Z.-Q. Wang, K. Tan, and D. Wang, “Deep learning based phase reconstruction for speaker separation: A trigonometric perspective,” in
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Y. Zhao, Z.-Q. Wang, and D. Wang, “Two-stage deep learning for noisy-reverberant speech enhancement,”
2018
Cited alongside, same era.
R. Scheibler, E. Bezzam, and I. Dokmanić, “Pyroomacoustics: A python package for audio room simulation and array processing algorithms,” in
2018
Cited alongside, same era.
“Itu-t recommendation p.808, subjective evaluation of speech quality with a crowdsourcing approach.”
2018
Cited alongside, same era.
X. Wang and C. Bao, “Masking estimation with phase restoration of clean speech for monaural speech enhancement,”
2019
Cited alongside, same era.
2019
Cited alongside, same era.
H. McGuire, “Librivox: Free public domain audiobooks,”
Cited in the paper.
2019
Later among the works it cites.
S. J. Reddi, S. Kale, and S. Kumar, “On the convergence of adam and beyond,”
2019
Later among the works it cites.
Y. Koizumi, K. Yaiabe, M. Delcroix, Y. Maxuxama, and D. Takeuchi, “Speech enhancement using self-adaptation and multi-head self-attention,” in
2020
Closest in time.
2020
Closest in time.