Fetching the paper…
Reading the bibliography…
Speech restoration (SR) is a task of converting degraded speech signals into high-quality ones.
J. B. Allen and D. A. Berkley, “Image method for efficiently simulating small-room acoustics,”
1979
Earlier work this paper cites.
T. Nakatani, T. Yoshioka,
2010
Earlier work this paper cites.
J. Thiemann, N. Ito, and E. Vincent, “The diverse environments multi-channel acoustic noise database (DEMAND): A database of multichannel environmental noise recordings,”
2013
Earlier work this paper cites.
K. Han, Y. Wang, and D. Wang, “Learning spectral mapping for speech dereverberation,” in
2014
Earlier work this paper cites.
K. Li and C.-H. Lee, “A deep neural network approach to speech bandwidth expansion,” in
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
K. Kinoshita, M. Delcroix,
2017
Earlier work this paper cites.
V. Kuleshov, S. Z. Enam, and S. Ermon, “Audio super resolution using neural networks,” in
2017
Earlier work this paper cites.
K. Ito and L. Johnson, “The LJ speech dataset,”
2017
Earlier work this paper cites.
D. Wang and J. Chen, “Supervised speech separation based on deep learning: An overview,”
2018
Earlier work this paper cites.
J. Shen, R. Pang,
2018
Earlier work this paper cites.
E. Perez, F. Strub,
2018
Earlier work this paper cites.
Y. Jia, Y. Zhang,
2018
Earlier work this paper cites.
N. Kalchbrenner, W. Elsen,
2018
Earlier work this paper cites.
S. Maiti and M. I. Mandel, “Parametric resynthesis with neural vocoders,” in
2019
Earlier work this paper cites.
F. Biadsy, R. J. Weiss,
2019
Earlier work this paper cites.
W.-N. Hsu, Y. Zhang,
2019
Cited alongside, same era.
——, “Disentangling correlated speaker and noise for speech synthesis via data augmentation and adversarial factorization,” in
2019
Cited alongside, same era.
Y. Chen, Y. Assael,
2019
Cited alongside, same era.
H. Zen, R. Clark,
2019
Cited alongside, same era.
A. Baevski, H. Zhou,
2020
Cited alongside, same era.
R. Ardila, M. Branson,
2020
Cited alongside, same era.
R. Yamamoto, E. Song, and J.-M. Kim, “Parallel WaveGAN: A fast waveform generation model based on generative adversarial networks with multi-resolution spectrogram,” in
C. Wang, M. Riviere,
2021
Later among the works it cites.
Y. Jia, H. Zen,
2021
Later among the works it cites.
A. Polyak, Y. Adi,
2021
Later among the works it cites.
Y. Koizumi, S. Karita,
2021
Later among the works it cites.
N. Chen, Y. Zhang,
2021
Later among the works it cites.
J. Pelecanos, Q. Wang, and I. L. Moreno, “Dr-Vectors: Decision residual networks and an improved loss for speaker recognition,” in
2021
Later among the works it cites.
S. Wang, A. Mesaros,
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2020
Cited alongside, same era.
J. Kong, J. Kim, and J. Bae, “HiFi-GAN: Generative adversarial networks for efficient and high fidelity speech synthesis,” in
2020
Cited alongside, same era.
A. Gulati, C.-C. Chiu,
2020
Cited alongside, same era.
Y. Zhang, J. Qin,
2020
Cited alongside, same era.
2020
Cited alongside, same era.
J. Su, Z. Jin, and A. Finkelstein, “HiFi-GAN-2: Studio-quality speech enhancement via generative adversarial networks conditioned on acoustic features,” in
2021
Cited alongside, same era.
T. Saeki, S. Takamichi,
2022
Later among the works it cites.
2022
Later among the works it cites.
C.-C. Chiu, J. Qin,
2022
Later among the works it cites.
Z. Borsos, M. Sharifi, and M. Tagliasacchi, “SpeechPainter: Text-conditioned speech inpainting,” in
2022
Later among the works it cites.
Y. Koizumi, H. Zen,
2022
Later among the works it cites.
Q. Wang, Y. Yu,
2022
Later among the works it cites.
J. Pelecanos, Q. Wang,
2022
Later among the works it cites.
S.-g. Lee, W. Ping,
2023
Closest in time.
Y. Koizumi, K. Yatabe,
2023
Closest in time.