Fetching the paper…
Reading the bibliography…
We propose the Fr\'echet Audio Distance (FAD), a novel, reference-free evaluation metric for music enhancement algorithms.
“Phase vocoder”
James Flanagan and RM Golden · 1966
Earlier work this paper cites.
“The Analysis of Permutations”
R.. Plackett · 1975
Earlier work this paper cites.
“The Fréchet distance between multivariate normal distributions”
DC Dowson and BV Landau · 1982
Earlier work this paper cites.
“Signal estimation from modified short-time Fourier transform”
D. Griffin and Jae Lim · 1984
Earlier work this paper cites.
“Perceptual evaluation of speech quality (PESQ)-a new method for speech quality assessment of telephone networks and codecs”
Antony Rix, John Beerends, Michael Hollier and Andries Hekstra · 2001
Earlier work this paper cites.
“Performance measurement in blind audio source separation”
Emmanuel Vincent, R“’emi Gribonval and C“’edric F“’evotte · 2006
Earlier work this paper cites.
“Evaluation of objective quality measures for speech enhancement”
Yi Hu and Philipos Loizou · 2008
Earlier work this paper cites.
“Evaluation of algorithms using games : The case of music tagging”
Edith Law et al · 2009
Earlier work this paper cites.
“A short-time objective intelligibility measure for time-frequency weighted noisy speech”
Cees Taal, Richard Hendriks, Richard Heusdens and Jesper Jensen · 2010
Cited alongside, same era.
“Speech Enhancement: Theory and Practice”
Philipos. Loizou · 2013
Cited alongside, same era.
“mir_eval: A transparent implementation of common MIR metrics”
Colin Raffel et al · 2014
Cited alongside, same era.
“Very Deep Convolutional Networks for Large-Scale Image Recognition”
Karen Simonyan and Andrew Zisserman · 2014
Cited alongside, same era.
“Going deeper with convolutions”
Christian Szegedy et al · 2015
Cited alongside, same era.
“Gans trained by a two time-scale update rule converge to a local nash equilibrium”
Martin Heusel et al · 2017
Later among the works it cites.
“Singing voice separation with deep U-Net convolutional networks”
Andreas Jansson et al · 2017
Later among the works it cites.
“Supervised speech separation based on deep learning: an overview”
DeLiang Wang and Jitong Chen · 2017
Later among the works it cites.
Li Chai, Jun Du and Chin-Hui Lee · 2018
Closest in time.
“SDR-half-baked or well done?”
Jonathan Le, Scott Wisdom, Hakan Erdogan and John Hershey · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sami Abu-El-Haija et al · 2016
Cited alongside, same era.
“CNN architectures for large-scale audio classification”
Shawn Hershey et al · 2017
Cited alongside, same era.
“Vimeo”, http://www.vimeo.com
Cited in the paper.
“YouTube”, http://www.youtube.com
Cited in the paper.
“Music Source Separation Using Stacked Hourglass Networks”
Sungheon Park, Taehoon Kim, Kyogu Lee and Nojun Kwak · 2018
Closest in time.
“Towards Accurate Generative Models of Video: A New Metric & Challenges”
Thomas Unterthiner et al · 2018
Closest in time.