Fetching the paper…
Reading the bibliography…
This paper presents a novel approach to audio restoration, focusing on the enhancement of low-quality music recordings, and in particular historical ones.
“John McCormack: The story of a singer,”
L. A. G. Strong, · 1941
Earlier work this paper cites.
Enrico Caruso: His Recorded Legacy
J. Freestone, · 1961
Earlier work this paper cites.
“Blind deconvolution through digital signal processing,”
T. G. Stockham, T. M. Cannon, and R. B. Ingebretsen, · 1975
Earlier work this paper cites.
Digital Audio Restoration—A Statistical Model Based Approach
S. J. Godsill and P. J. W. Rayner, · 1998
Earlier work this paper cites.
“Gigli, Beniamino,”
D. Shawe-Taylor and A. Blyth, · 2001
Earlier work this paper cites.
“Patti, Adelina,”
E. Forbes, · 2001
Earlier work this paper cites.
“Digital audio antiquing—Signal processing methods for imitating the sound quality of historical recordings,”
V. Välimäki, S. González, O. Kimmelma, and J. Parviainen, · 2008
Earlier work this paper cites.
“Melba, Dame Nellie,”
D. Shawe-Taylor, · 2009
Earlier work this paper cites.
“Caruso, Enrico,”
R. Celletti and A. Blyth, · 2013
Earlier work this paper cites.
“The NUS sung and spoken lyrics corpus: A quantitative comparison of singing and speech,”
Z. Duan, H. Fang, B. Li, K. C. Sim, and Y. Wang, · 2013
Earlier work this paper cites.
“Adam: A method for stochastic optimization,”
D. P. Kingma and J. Ba, · 2015
Earlier work this paper cites.
“Vocalset: A singing voice dataset,”
J. Wilkins, P. Seetharaman, A. Wahl, and B. Pardo, · 2018
Earlier work this paper cites.
“How to interpret early recordings? Artefacts and resonances in recording and reproduction of singing voices,”
M. Kob and T. A. Weege, · 2019
Earlier work this paper cites.
“Enabling factorized piano music modeling and generation with the MAESTRO dataset,”
C. Hawthorne, A. Stasyuk, A. Roberts, et al., · 2019
Earlier work this paper cites.
“Fréchet audio distance: A reference-free metric for evaluating music enhancement algorithms,”
K. Kilgour, M. Zuluaga, D. Roblek, and M. Sharifi, · 2019
Cited alongside, same era.
“Analyzing and improving the image quality of StyleGAN,”
T. Karras, S. Laine, M. Aittala, et al., · 2020
Cited alongside, same era.
“Perceptual loss function for neural modeling of audio systems,”
A. Wright and V. Välimäki, · 2020
Cited alongside, same era.
“PJS: Phoneme-balanced japanese singing-voice corpus,”
J. Koguchi, S. Takamichi, and M. Morise, · 2020
Cited alongside, same era.
“Children’s song dataset for singing voice research,”
S. Choi, W. Kim, S. Park, S. Yong, and J. Nam, · 2020
Cited alongside, same era.
“Score-based generative modeling through stochastic differential equations,”
Y. Song, J. Sohl-Dickstein, D. P Kingma, et al., · 2021
“Solving audio inverse problems with a diffusion model,”
E. Moliner, J. Lehtinen, and V. Välimäki, · 2023
Later among the works it cites.
“Diffusion posterior sampling for general noisy inverse problems,”
H. Chung, J. Kim, M. T. Mccann, et al., · 2023
Later among the works it cites.
“Parallel diffusion models of operator and image for blind inverse problems,”
H. Chung, J. Kim, S. Kim, and J. C. Ye, · 2023
Later among the works it cites.
“Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation,”
Y. Wu, K. Chen, T. Zhang, Y. Hui, T. Berg-Kirkpatrick, and S. Dubnov, · 2023
Later among the works it cites.
“High fidelity neural audio compression,”
A. Défossez, J. Copet, G. Synnaeve, and Y. Adi, · 2023
Later among the works it cites.
“BEHM-GAN: Bandwidth extension of historical music using generative adversarial networks,”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
“The birth of ‘modern’ vocalism: The paradigmatic case of Enrico Caruso,”
B. Gentili, · 2021
Cited alongside, same era.
“Multi-singer: Fast multi-singer singing voice vocoder with a large-scale corpus,”
R. Huang, F. Chen, Y. Ren, J. Liu, C. Cui, and Z. Zhao, · 2021
Cited alongside, same era.
“A two-stage U-net for high-fidelity denoising of historical recordings,”
E. Moliner and V. Välimäki, · 2022
Cited alongside, same era.
“Elucidating the design space of diffusion-based generative models,”
T. Karras, M. Aittala, T. Aila, and S. Laine, · 2022
Cited alongside, same era.
“Opencpop: A high-quality open source chinese popular song corpus for singing voice synthesis,”
Y. Wang, X. Wang, P. Zhu, J. Wu, H. Li, H. Xue, Y. Zhang, L. Xie, and M. Bi, · 2022
Cited alongside, same era.
“M4singer: A multi-style, multi-singer and musical score provided mandarin singing corpus,”
L. Zhang, R. Li, S. Wang, L. Deng, J. Liu, Y. Ren, J. He, R. Huang, J. Zhu, X. Chen, et al., · 2022
Cited alongside, same era.
E. Moliner and V. Välimäki, · 2023
Later among the works it cites.
“Hybrid transformers for music source separation,”
S. Rouard, F. Massa, and A. Défossez, · 2023
Later among the works it cites.
“Diffusion models for audio restoration,”
J.-M. Lemercier, J. Richter, S. Welker, E. Moliner, V. Välimäki, and T. Gerkmann, · 2024
Closest in time.
“Blind audio bandwidth extension: A diffusion-based zero-shot approach,”
E. Moliner, F. Elvander, and V. Välimäki, · 2024
Closest in time.
“CADS: Unleashing the diversity of diffusion models through condition-annealed sampling,”
S. Sadat, J. Buhmann, D. Bradely, O. Hilliges, and R.M Weber, · 2024
Closest in time.
“Diffusion-based audio inpainting,”
E. Moliner and V. Välimäki, · 2024
Closest in time.
“Adapting Frechet audio distance for generative music evaluation,”
A. Gui, H. Gamper, S. Braun, and D. Emmanouilidou, · 2024
Closest in time.
“Guidance with spherical gaussian constraint for conditional diffusion,”
Lingxiao Yang, Shutong Ding, Yifan Cai, Jingyi Yu, Jingya Wang, and Ye Shi, · 2024
Closest in time.