Fetching the paper…
Reading the bibliography…
In this paper we introduce the Frechet Music Distance (FMD), a novel evaluation metric for generative symbolic music models, inspired by the Frechet Inception Distance (FID) in computer vision and Frechet Audio Distance (FAD) in generative audio.
The Fréchet distance between multivariate normal distributions
Dowson, D.; and Landau, B. 1982 · 1982
Earlier work this paper cites.
Least median of squares regression
Rousseeuw, P. J. 1984 · 1984
Earlier work this paper cites.
A fast algorithm for the minimum covariance determinant estimator
Rousseeuw, P. J.; and Driessen, K. V. 1999 · 1999
Earlier work this paper cites.
A well-conditioned estimator for large-dimensional covariance matrices
Ledoit, O.; and Wolf, M. 2004 · 2004
Earlier work this paper cites.
Sparse inverse covariance estimation with the graphical lasso
Friedman, J.; Hastie, T.; and Tibshirani, R. 2008 · 2008
Earlier work this paper cites.
Pop909: A pop-song dataset for music arrangement generation
Wang, Z.; Chen, K.; Jiang, J.; Zhang, Y.; Xu, M.; Dai, S.; Gu, X.; and Xia, G. 2020 · 2008
Earlier work this paper cites.
Wu, S.-L.; and Yang, Y.-H. 2020 · 2008
Earlier work this paper cites.
Shrinkage Algorithms for MMSE Covariance Estimation
Chen, Y.; Wiesel, A.; Eldar, Y. C.; and Hero, A. O. 2010 · 2010
Earlier work this paper cites.
Scikit-learn: Machine Learning in Python
Pedregosa, F.; Varoquaux, G.; Gramfort, A.; Michel, V.; Thirion, B.; Grisel, O.; Blondel, M.; Prettenhofer, P.; Weiss, R.; Dubourg, V.; Vanderplas, J.; Passos, A.; Cournapeau, D.; Brucher, M.; Perrot, M.; and Duchesnay, E. 2011 · 2011
Earlier work this paper cites.
Intuitive analysis, creation and manipulation of MIDI data with pretty_midi
Raffel, C.; and Ellis, D. P. 2014 · 2014
Earlier work this paper cites.
C-RNN-GAN: Continuous recurrent neural networks with adversarial training
Mogren, O. 2016 · 2016
Earlier work this paper cites.
Music transcription modelling and composition using deep learning
Sturm, B. L.; Santos, J. F.; Ben-Tal, O.; and Korshunova, I. 2016 · 2016
Cited alongside, same era.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
Heusel, M.; Ramsauer, H.; Unterthiner, T.; Nessler, B.; and Hochreiter, S. 2017 · 2017
Cited alongside, same era.
Crestel, L.; Esling, P.; Heng, L.; and McAdams, S. 2018 · 2018
Cited alongside, same era.
Musegan: Multi-track sequential generative adversarial networks for symbolic music generation and accompaniment
Dong, H.-W.; Hsiao, W.-Y.; Yang, L.-C.; and Yang, Y.-H. 2018 · 2018
Cited alongside, same era.
Masked autoencoders are scalable vision learners
He, K.; Chen, X.; Xie, S.; Li, Y.; Dollár, P.; and Girshick, R. 2022 · 2022
Later among the works it cites.
Theme transformer: Symbolic music generation with theme-conditioned transformer
Shih, Y.-J.; Wu, S.-L.; Zalkow, F.; Müller, M.; and Yang, Y.-H. 2022 · 2022
Later among the works it cites.
Multitrack music transformer
Dong, H.-W.; Chen, K.; Dubnov, S.; McAuley, J.; and Berg-Kirkpatrick, T. 2023 · 2023
Later among the works it cites.
A survey on deep learning for symbolic music generation: Representations, algorithms, evaluations, and challenges
Ji, S.; Yang, X.; and Luo, J. 2023 · 2023
Later among the works it cites.
Clamp: Contrastive language-music pre-training for cross-modal symbolic music information retrieval
Wu, S.; Yu, D.; Tan, X.; and Sun, M. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Huang, C.-Z. A.; Vaswani, A.; Uszkoreit, J.; Shazeer, N.; Simon, I.; Hawthorne, C.; Dai, A. M.; Hoffman, M. D.; Dinculescu, M.; and Eck, D. 2018 · 2018
Cited alongside, same era.
Fr \ \backslash ’echet audio distance: A metric for evaluating music enhancement algorithms
Kilgour, K.; Zuluaga, M.; Roblek, D.; and Sharifi, M. 2018 · 2018
Cited alongside, same era.
Enabling Factorized Piano Music Modeling and Generation with the MAESTRO Dataset
Hawthorne, C.; Stasyuk, A.; Roberts, A.; Simon, I.; Huang, C.-Z. A.; Dieleman, S.; Elsen, E.; Engel, J.; and Eck, D. 2019 · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
Radford, A.; Wu, J.; Child, R.; Luan, D.; Amodei, D.; Sutskever, I.; et al. 2019 · 2019
Cited alongside, same era.
Muspy: A toolkit for symbolic music generation
Dong, H.-W.; Chen, K.; McAuley, J.; and Berg-Kirkpatrick, T. 2020 · 2020
Cited alongside, same era.
Interacting with GPT-2 to generate controlled and believable musical sequences in ABC notation
Geerlings, C.; and Merono-Penuela, A. 2020 · 2020
Cited alongside, same era.
Array programming with NumPy
Harris, C. R.; Millman, K. J.; van der Walt, S. J.; Gommers, R.; Virtanen, P.; Cournapeau, D.; Wieser, E.; Taylor, J.; Berg, S.; Smith, N. J.; Kern, R.; Picus, M.; Hoyer, S.; van Kerkwijk, M. H.; Brett, M.; Haldane, A.; del Río, J. F.; Wiebe, M.; Peterson, P.; Gérard-Marchant, P.; Sheppard, K.; Reddy, T.; Weckesser, W.; Abbasi, H.; Gohlke, C.; and Oliphant, T. E. 2020 · 2020
Cited alongside, same era.
Adapting frechet audio distance for generative music evaluation
Gui, A.; Gamper, H.; Braun, S.; and Emmanouilidou, D. 2024 · 2024
Closest in time.
MidiCaps–A large-scale MIDI dataset with text captions
Melechovsky, J.; Roy, A.; and Herremans, D. 2024 · 2024
Closest in time.
Symbotunes: unified hub for symbolic music generative models
Skierś, P.; Łazarski, M.; Kopeć, M.; and Modrzejewski, M. 2024 · 2024
Closest in time.
Tailleur, M.; Lee, J.; Lagrange, M.; Choi, K.; Heller, L. M.; Imoto, K.; and Okamoto, Y. 2024 · 2024
Closest in time.
CLaMP 2: Multimodal Music Information Retrieval Across 101 Languages Using Large Language Models
Wu, S.; Wang, Y.; Yuan, R.; Guo, Z.; Tan, X.; Zhang, G.; Zhou, M.; Chen, J.; Mu, X.; Gao, Y.; et al. 2024 · 2024
Closest in time.