Fetching the paper…
Reading the bibliography…
We present SingSong, a system that generates instrumental music to accompany input vocals, potentially offering musicians and non-musicians alike an intuitive new way to create music featuring their own voice.
Sur la distance de deux lois de probabilité
Fréchet, M · 1957
Earlier work this paper cites.
Probabilistic melodic harmonization
Paiement, J.-F., Eck, D., and Bengio, S · 2006
Earlier work this paper cites.
MySong: automatic accompaniment generation for vocal melodies
Simon, I., Morris, D., and Basu, S · 2008
Earlier work this paper cites.
Accurate tempo estimation based on recurrent neural networks and resonating comb filters
Böck, S., Krebs, F., and Widmer, G · 2015
Earlier work this paper cites.
WaveNet: A generative model for raw audio
van den Oord, A., Dieleman, S., Zen, H., Simonyan, K., Vinyals, O., Graves, A., Kalchbrenner, N., Senior, A., and Kavukcuoglu, K · 2016
Earlier work this paper cites.
CNN architectures for large-scale audio classification
Hershey, S., Chaudhuri, S., Ellis, D. P., Gemmeke, J. F., Jansen, A., Moore, R. C., Plakal, M., Platt, D., Saurous, R. A., Seybold, B., et al · 2017
Earlier work this paper cites.
GANs trained by a two time-scale update rule converge to a local Nash equilibrium
Heusel, M., Ramsauer, H., Unterthiner, T., Nessler, B., and Hochreiter, S · 2017
Earlier work this paper cites.
SampleRNN: An unconditional end-to-end neural audio generation model
Mehri, S., Kumar, K., Gulrajani, I., Kumar, R., Jain, S., Sotelo, J., Courville, A., and Bengio, Y · 2017
Earlier work this paper cites.
MUSDB18 - a corpus for music separation
Rafii, Z., Liutkus, A., Stöter, F.-R., Mimilakis, S. I., and Bittner, R · 2017
Earlier work this paper cites.
Neural discrete representation learning
van den Oord, A., Vinyals, O., and Kavukcuoglu, K · 2017
Earlier work this paper cites.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., and Polosukhin, I · 2017
Earlier work this paper cites.
The challenge of realistic music generation: modelling raw audio at scale
Dieleman, S., van den Oord, A., and Simonyan, K · 2018
Earlier work this paper cites.
Enabling factorized piano music modeling and generation with the MAESTRO dataset
Hawthorne, C., Stasyuk, A., Roberts, A., Simon, I., Huang, C.-Z. A., Dieleman, S., Elsen, E., Engel, J., and Eck, D · 2018
Cited alongside, same era.
Genre-agnostic key classification with convolutional neural networks
Korzeniowski, F. and Widmer, G · 2018
Cited alongside, same era.
Adversarial audio synthesis
Donahue, C., McAuley, J., and Puckette, M · 2019
Cited alongside, same era.
GANSynth: Adversarial neural audio synthesis
Engel, J., Agrawal, K. K., Chen, S., Gulrajani, I., Donahue, C., and Roberts, A · 2019
Cited alongside, same era.
Fréchet Audio Distance: A reference-free metric for evaluating music enhancement algorithms
Kilgour, K., Zuluaga, M., Roblek, D., and Sharifi, M · 2019
Cited alongside, same era.
High-level control of drum track generation using learned patterns of rhythmic interaction
RAVE: A variational autoencoder for fast and high-quality neural audio synthesis
Caillon, A. and Esling, P · 2021
Later among the works it cites.
w2v-BERT: Combining contrastive learning and masked language modeling for self-supervised speech pre-training
Chung, Y.-A., Zhang, Y., Han, W., Chiu, C.-C., Qin, J., Pang, R., and Wu, Y · 2021
Later among the works it cites.
KUIELab-MDX-Net: A two-stream neural network for music demixing
Kim, M., Choi, W., Chung, J., Lee, D., and Jung, S · 2021
Later among the works it cites.
SoundStream: An end-to-end neural audio codec
Zeghidour, N., Luebs, A., Omran, A., Skoglund, J., and Tagliasacchi, M · 2021
Later among the works it cites.
AudioLM: a language modeling approach to audio generation
Borsos, Z., Marinier, R., Vincent, D., Kharitonov, E., Pietquin, O., Sharifi, M., Teboul, O., Grangier, D., Tagliasacchi, M., and Zeghidour, N · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Lattner, S. and Grachten, M · 2019
Cited alongside, same era.
Jukebox: A generative model for music
Dhariwal, P., Jun, H., Payne, C., Kim, J. W., Radford, A., and Sutskever, I · 2020
Cited alongside, same era.
DDSP: Differentiable digital signal processing
Engel, J., Hantrakul, L., Gu, C., and Roberts, A · 2020
Cited alongside, same era.
BassNet: A variational gated autoencoder for conditional generation of bass guitar tracks with learned interactive control
Grachten, M., Lattner, S., and Deruty, E · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Raffel, C., Shazeer, N., Roberts, A., Lee, K., Narang, S., Matena, M., Zhou, Y., Li, W., Liu, P. J., et al · 2020
Cited alongside, same era.
Automatic melody harmonization with triad chords: A comparative study
Yeh, Y.-C., Hsiao, W.-Y., Fukayama, S., Kitahara, T., Genchel, B., Liu, H.-M., Dong, H.-W., Chen, Y., Leong, T., and Yang, Y.-H · 2020
Cited alongside, same era.
vocadito: A dataset of solo vocals with f 0 f_{0} , note, and lyric annotations
Bittner, R. M., Pasalo, K., Bosch, J. J., Meseguer-Brocal, G., and Rubinstein, D · 2021
Cited alongside, same era.
Later among the works it cites.
Quantifying memorization across neural language models
Carlini, N., Ippolito, D., Jagielski, M., Lee, K., Tramer, F., and Zhang, C · 2022
Later among the works it cites.
Redefining relationships in music
Detweiler, C., Coleman, B., Diaz, F., Dom, L., Donahue, C., Engel, J., Huang, C.-Z. A., James, L., Manilow, E., McCroskery, A., Pedersen, K., et al · 2022
Later among the works it cites.
It’s raw! audio generation with state-space models
Goel, K., Gu, A., Donahue, C., and Ré, C · 2022
Later among the works it cites.
Scaling up models and data with t5x
Roberts, A., Chung, H. W., Levskaya, A., Mishra, G., Bradbury, J., Andor, D., Narang, S., Lester, B., Gaffney, C., Mohiuddin, A., Hawthorne, C., et al · 2022
Later among the works it cites.
JukeDrummer: Conditional beat-aware audio-domain drum accompaniment generation via transformer VQ-VAE
Wu, Y.-K., Chiu, C.-Y., and Yang, Y.-H · 2022
Later among the works it cites.
MusicLM: Generating music from text
Agostinelli, A., Denk, T. I., Borsos, Z., Engel, J., Verzetti, M., Caillon, A., Huang, Q., Jansen, A., Roberts, A., Tagliasacchi, M., Sharifi, M., Zeghidour, N., and Frank, C · 2023
Closest in time.