Fetching the paper…
Reading the bibliography…
Controllable music generation plays a vital role in human-AI music co-creation.
Rwc music database: Popular, classical and jazz music databases
Masataka Goto, Hiroki Hashiguchi, Takuichi Nishimura, and Ryuichi Oka · 2002
Earlier work this paper cites.
Madmom: A new python audio and music signal processing library
Sebastian Böck, Filip Korzeniowski, Jan Schlüter, Florian Krebs, and Gerhard Widmer · 2016
Earlier work this paper cites.
Neural discrete representation learning
Aaron Van Den Oord, Oriol Vinyals, et al · 2017
Earlier work this paper cites.
Cheng-Zhi Anna Huang, Ashish Vaswani, Jakob Uszkoreit, Noam Shazeer, Ian Simon, Curtis Hawthorne, Andrew M Dai, Matthew D Hoffman, Monica Dinculescu, and Douglas Eck · 2018
Earlier work this paper cites.
Infilling piano performances
Daphne Ippolito, Anna Huang, Curtis Hawthorne, and Douglas Eck · 2018
Earlier work this paper cites.
Inpainting of long audio segments with similarity graphs
Nathanael Perraudin, Nicki Holighaus, Piotr Majdak, and Peter Balazs · 2018
Earlier work this paper cites.
Large-vocabulary chord transcription via chord structure decomposition
Junyan Jiang, Ke Chen, Wei Li, and Gus Xia · 2019
Earlier work this paper cites.
Cutting music source separation some slakh: A dataset to study the impact of training data quality and quantity
Ethan Manilow, Gordon Wichern, Prem Seetharaman, and Jonathan Le Roux · 2019
Earlier work this paper cites.
A context encoder for audio inpainting
Andrés Marafioti, Nathanaël Perraudin, Nicki Holighaus, and Piotr Majdak · 2019
Earlier work this paper cites.
Vision-infused deep audio inpainting
Hang Zhou, Ziwei Liu, Xudong Xu, Ping Luo, and Xiaogang Wang · 2019
Earlier work this paper cites.
Music sketchnet: Controllable music generation via factorized representations of pitch and rhythm
Ke Chen, Cheng-i Wang, Taylor Berg-Kirkpatrick, and Shlomo Dubnov · 2020
Earlier work this paper cites.
Jukebox: A generative model for music
Prafulla Dhariwal, Heewoo Jun, Christine Payne, Jong Wook Kim, Alec Radford, and Ilya Sutskever · 2020
Earlier work this paper cites.
Pop music transformer: Beat-based modeling and generation of expressive pop piano compositions
Yu-Siang Huang and Yi-Hsuan Yang · 2020
Cited alongside, same era.
Gacela: A generative adversarial context encoder for long audio inpainting of music
Andres Marafioti, Piotr Majdak, Nicki Holighaus, and Nathanaël Perraudin · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu · 2020
Cited alongside, same era.
Variable-length music score infilling via xlnet and musically specialized positional encoding
Chin-Jui Chang, Chun-Yi Lee, and Yi-Hsuan Yang · 2021
Cited alongside, same era.
Lora: Low-rank adaptation of large language models
Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen · 2021
Cited alongside, same era.
Simple and controllable music generation
Jade Copet, Felix Kreuk, Itai Gat, Tal Remez, David Kant, Gabriel Synnaeve, Yossi Adi, and Alexandre Défossez · 2023
Later among the works it cites.
High fidelity neural audio compression
Alexandre Défossez, Jade Copet, Gabriel Synnaeve, and Yossi Adi · 2023
Later among the works it cites.
Llama-adapter v2: Parameter-efficient visual instruction model
Peng Gao, Jiaming Han, Renrui Zhang, Ziyi Lin, Shijie Geng, Aojun Zhou, Wei Zhang, Pan Lu, Conghui He, Xiangyu Yue, et al · 2023
Later among the works it cites.
Vampnet: Music generation via masked acoustic token modeling
Hugo F Flores Garcia, Prem Seetharaman, Rithesh Kumar, and Bryan Pardo · 2023
Later among the works it cites.
Polyffusion: A diffusion model for polyphonic score generation with internal and external controls
Lejun Min, Junyan Jiang, Gus Xia, and Jingwei Zhao · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Prefix-tuning: Optimizing continuous prompts for generation
Xiang Lisa Li and Percy Liang · 2021
Cited alongside, same era.
Musebert: Pre-training music representation for music understanding and controllable generation
Ziyu Wang and Gus Xia · 2021
Cited alongside, same era.
Accomontage: Accompaniment arrangement via phrase selection and style transfer
Jingwei Zhao and Gus Xia · 2021
Cited alongside, same era.
High fidelity neural audio compression
Alexandre Défossez, Jade Copet, Gabriel Synnaeve, and Yossi Adi · 2022
Cited alongside, same era.
P-tuning: Prompt tuning can be comparable to fine-tuning across scales and tasks
Xiao Liu, Kaixuan Ji, Yicheng Fu, Weng Tam, Zhengxiao Du, Zhilin Yang, and Jie Tang · 2022
Cited alongside, same era.
Music demixing challenge 2021
Yuki Mitsufuji, Giorgio Fabbro, Stefan Uhlich, Fabian-Robert Stöter, Alexandre Défossez, Minseok Kim, Woosung Choi, Chin-Yun Yu, and Kin-Wai Cheuk · 2022
Cited alongside, same era.
MusicLM: Generating music from text
Andrea Agostinelli, Timo I Denk, Zalán Borsos, Jesse Engel, Mauro Verzetti, Antoine Caillon, Qingqing Huang, Aren Jansen, Adam Roberts, Marco Tagliasacchi, et al · 2023
Cited alongside, same era.
Mo \ \backslash ˆ usai: Text-to-music generation with long-context latent diffusion
Flavio Schneider, Ojasv Kamal, Zhijing Jin, and Bernhard Schölkopf · 2023
Later among the works it cites.
Controllable music inpainting with mixed-level and disentangled representation
Shiqi Wei, Ziyu Wang, Weiguo Gao, and Gus Xia · 2023
Later among the works it cites.
Music controlnet: Multiple time-varying controls for music generation, 2023
Shih-Lun Wu, Chris Donahue, Shinji Watanabe, and Nicholas J. Bryan · 2023
Later among the works it cites.
Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation
Yusong Wu, Ke Chen, Tianyu Zhang, Yuchen Hui, Taylor Berg-Kirkpatrick, and Shlomo Dubnov · 2023
Later among the works it cites.
Llama-adapter: Efficient fine-tuning of language models with zero-init attention
Renrui Zhang, Jiaming Han, Aojun Zhou, Xiangfei Hu, Shilin Yan, Pan Lu, Hongsheng Li, Peng Gao, and Yu Qiao · 2023
Later among the works it cites.
Content-based controls for music large language modeling, 2024
Liwei Lin, Gus Xia, Junyan Jiang, and Yixiao Zhang · 2024
Closest in time.