Fetching the paper…
Reading the bibliography…
In this work, we provide a comprehensive survey of AI music generation tools, including both research projects and commercialized applications.
A circumplex model of affect
James Russell · 1980
Earlier work this paper cites.
The markov process as a compositional model: A survey and tutorial
Charles Ames · 1989
Earlier work this paper cites.
Style and music: Theory, history, and ideology
Leonard B Meyer · 1989
Earlier work this paper cites.
Genjam: A genetic algorithm for generating jazz solos
John A Biles · 1994
Earlier work this paper cites.
5 - exploration of timbre by analysis and synthesis
Jean-Claude Risset and David L. Wessel · 1999
Earlier work this paper cites.
Rule-based analysis and generation of music
Randall Richard Spangler · 1999
Earlier work this paper cites.
Music composition using genetic evolutionary algorithms
Manuel Marques, V Oliveira, S Vieira, and AC Rosa · 2000
Earlier work this paper cites.
The Science of Sound
Thomas D. Rossing, F. Richard Moore, and Paul A. Wheeler · 2002
Earlier work this paper cites.
Melisma stochastic melody generator
Daniel Sleator and David Temperley · 2003
Earlier work this paper cites.
Digital audio workstation
Colby Leider · 2004
Earlier work this paper cites.
Hearing in time: Psychological aspects of musical meter
Justin London · 2004
Earlier work this paper cites.
A diffusion model account of response time and accuracy in a brightness discrimination task: Fitting real data and failing to fit fake but plausible data
Roger Ratcliff, Pablo Gómez, and Gail McKoon · 2004
Earlier work this paper cites.
What’s that sound? An introduction to rock and its history
John Covach and Andrew Flory · 2005
Earlier work this paper cites.
Tuning, Timbre, Spectrum, Scale
W.A. Sethares · 2005
Earlier work this paper cites.
Exploring the musical mind: Cognition, emotion, ability, function
John A Sloboda · 2005
Earlier work this paper cites.
Visualizing data using t-sne
Laurens Van der Maaten and Geoffrey Hinton · 2008
Earlier work this paper cites.
Engineering acoustics
Michael Möser · 2009
Earlier work this paper cites.
Zero-shot learning with semantic output codes
Mark Palatucci, Dean Pomerleau, Geoffrey E Hinton, and Tom M Mitchell · 2009
Earlier work this paper cites.
Apopcaleaps: Automatic music generation with chrism
Jon Sneyers and Danny De Schreye · 2010
Earlier work this paper cites.
Musescore
Werner Schweer and Others · 2011
Earlier work this paper cites.
Automatic music generation using evolutionary algorithms and neural networks
Ali Çağatay Yiiksel, Mehmet Melih Karci, and A. Şima Uyar · 2011
Earlier work this paper cites.
The complete musician: An integrated approach to tonal theory, analysis, and listening
Steven G Laitz · 2012
Earlier work this paper cites.
The physics and psychophysics of music: An introduction
Juan G Roederer · 2013
Earlier work this paper cites.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Chapter 9 - big data driven natural language processing research and applications
Venkat N. Gudivada, Dhana Rao, and Vijay V. Raghavan · 2015
Earlier work this paper cites.
Deep learning
Yann LeCun, Yoshua Bengio, and Geoffrey Hinton · 2015
Earlier work this paper cites.
Understanding lstm networks, 2015
Christopher Olah · 2015
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics
Jascha Sohl-Dickstein, Eric Weiss, Niru Maheswaranathan, and Surya Ganguli · 2015
Earlier work this paper cites.
Learning-based methods for comparing sequences, with applications to audio-to-MIDI alignment and matching, June 2016
Colin Raffel · 2016
Earlier work this paper cites.
Learning-Based Methods for Comparing Sequences, with Applications to Audio-to-MIDI Alignment and Matching
Colin Raffel · 2016
Earlier work this paper cites.
Wavenet: A generative model for raw audio, 2016
Aaron van den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew Senior, and Koray Kavukcuoglu · 2016
Earlier work this paper cites.
The Evolution of Music: Culture, Technology, and Society
Elena Denisova-Schmidt · 2017
Earlier work this paper cites.
Neural audio synthesis of musical notes with wavenet autoencoders, 2017
Jesse Engel, Cinjon Resnick, Adam Roberts, Sander Dieleman, Douglas Eck, Karen Simonyan, and Mohammad Norouzi · 2017
Earlier work this paper cites.
Multiplicative LSTM for sequence modelling
Ben Krause, Iain Murray, Steve Renals, and Liang Lu · 2017
Earlier work this paper cites.
Learning to generate reviews and discovering sentiment
Alec Radford, Rafal Jozefowicz, and Ilya Sutskever · 2017
Earlier work this paper cites.
Markov chain based procedural music generator with user chosen mood compatibility
Adhika Sigit Ramanto and Nur Ulfa Maulidevi · 2017
Cited alongside, same era.
Attention is all you need, 2017
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
mixup: Beyond empirical risk minimization
Hongyi Zhang, Moustapha Cisse, Yann N Dauphin, and David Lopez-Paz · 2017
Cited alongside, same era.
Musegan: Multi-track sequential generative adversarial networks for symbolic music generation and accompaniment
Hao-Wen Dong, Wen-Yi Hsiao, Li-Chia Yang, and Yi-Hsuan Yang · 2018
Cited alongside, same era.
Music composition with artificial intelligence system based on markov chain and genetic algorithm
Sudhanshu Gautam and Sarita Soni · 2018
Cited alongside, same era.
Markov chains for computer music generation
Ilana Shapiro and Mark Huber · 2021
Later among the works it cites.
Score-based generative modeling through stochastic differential equations, 2021
Yang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar, Stefano Ermon, and Ben Poole · 2021
Later among the works it cites.
Soundstream: An end-to-end neural audio codec, 2021
Neil Zeghidour, Alejandro Luebs, Ahmed Omran, Jan Skoglund, and Marco Tagliasacchi · 2021
Later among the works it cites.
Scaling instruction-finetuned language models, 2022
Hyung Won Chung, Le Hou, Shayne Longpre, Barret Zoph, Yi Tay, William Fedus, Yunxuan Li, Xuezhi Wang, Mostafa Dehghani, Siddhartha Brahma, Albert Webson, Shixiang Shane Gu, Zhuyun Dai, Mirac Suzgun, Xinyun Chen, Aakanksha Chowdhery, Alex Castro-Ros, Marie Pellat, Kevin Robinson, Dasha Valter, Sharan Narang, Gaurav Mishra, Adams Yu, Vincent Zhao, Yanping Huang, Andrew Dai, Hongkun Yu, Slav Petrov, Ed H. Chi, Jeff Dean, Jacob Devlin, Adam Roberts, Denny Zhou, Quoc V. Le, and Jason Wei · 2022
Later among the works it cites.
Riffusion - Stable diffusion for real-time music generation, 2022
Seth* Forsgren and Hayk* Martiros · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Curtis Hawthorne, Andriy Stasyuk, Adam Roberts, Ian Simon, Cheng-Zhi Anna Huang, Sander Dieleman, Erich Elsen, Jesse Engel, and Douglas Eck · 2018
Cited alongside, same era.
Music transformer, 2018
Cheng-Zhi Anna Huang, Ashish Vaswani, Jakob Uszkoreit, Noam Shazeer, Ian Simon, Curtis Hawthorne, Andrew M. Dai, Matthew D. Hoffman, Monica Dinculescu, and Douglas Eck · 2018
Cited alongside, same era.
Neural discrete representation learning, 2018
Aaron van den Oord, Oriol Vinyals, and Koray Kavukcuoglu · 2018
Cited alongside, same era.
Generating long sequences with sparse transformers, 2019
Rewon Child, Scott Gray, Alec Radford, and Ilya Sutskever · 2019
Cited alongside, same era.
GANSynth: Adversarial neural audio synthesis
Jesse Engel, Kumar Krishna Agrawal, Shuo Chen, Ishaan Gulrajani, Chris Donahue, and Adam Roberts · 2019
Cited alongside, same era.
Music generation using an interactive evolutionary algorithm
Majid Farzaneh and Rahil Mahdian Toroghi · 2019
Cited alongside, same era.
Learning to groove with inverse sequence transformations
Jon Gillick, Adam Roberts, Jesse Engel, Douglas Eck, and David Bamman · 2019
Cited alongside, same era.
Mulan: A joint embedding of music audio and natural language
Qingqing Huang, Aren Jansen, Joonseok Lee, Ravi Ganti, Judith Yue Li, and Daniel P. W. Ellis · 2022
Later among the works it cites.
Mulan: A joint embedding of music audio and natural language
Qingqing Huang, Aren Jansen, Joonseok Lee, Ravi Ganti, Judith Yue Li, and Daniel PW Ellis · 2022
Later among the works it cites.
Auto-encoding variational bayes, 2022
Diederik P Kingma and Max Welling · 2022
Later among the works it cites.
Diffusion autoencoders: Toward a meaningful and decodable representation
Konpat Preechakul, Nattanat Chatthee, Suttisak Wizadwongsa, and Supasorn Suwajanakorn · 2022
Later among the works it cites.
High-resolution image synthesis with latent diffusion models
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer · 2022
Later among the works it cites.
High-resolution image synthesis with latent diffusion models, 2022
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer · 2022
Later among the works it cites.
It’s time for artistic correspondence in music and video, 2022
Didac Suris, Carl Vondrick, Bryan Russell, and Justin Salamon · 2022
Later among the works it cites.
Sdmuse: Stochastic differential music editing and generation via hybrid representation, 2022
Chen Zhang, Yi Ren, Kejun Zhang, and Shuicheng Yan · 2022
Later among the works it cites.
Stochastic differential music editing and generation via hybrid representation, Sep 2022
Chen Zhang, Yi Ren, Kejun Zhang, and Shuicheng Yan · 2022
Later among the works it cites.
A review of intelligent music generation systems, 2022
Ziyi Zhao, Hanwei Liu, Song Li, Junwei Pang, Maoqing Zhang, Yi Qin, Lei Wang, and Qidi Wu · 2022
Later among the works it cites.
Quantized gan for complex music generation from dance videos
Ye Zhu, Kyle Olszewski, Yu Wu, Panos Achlioptas, Menglei Chai, Yan Yan, and Sergey Tulyakov · 2022
Later among the works it cites.
Video background music generation: Dataset, method and evaluation, 2022
Le Zhuo, Zhaokai Wang, Baisen Wang, Yue Liao, Stanley Peng, Chenxi Bao, Miao Lu, Xiaobo Li, and Si Liu · 2022
Later among the works it cites.
ABC Notation
ABC Notation · 2023
Closest in time.
Bitmidi: Free midi files, 2023
Feross Aboukhadijeh · 2023
Closest in time.
Musiclm: Generating music from text, 2023
Andrea Agostinelli, Timo I. Denk, Zalán Borsos, Jesse Engel, Mauro Verzetti, Antoine Caillon, Qingqing Huang, Aren Jansen, Adam Roberts, Marco Tagliasacchi, Matt Sharifi, Neil Zeghidour, and Christian Frank · 2023
Closest in time.
Musiclm - pytorch
Louis Bouchard · 2023
Closest in time.
Simple and controllable music generation, 2023
Jade Copet, Felix Kreuk, Itai Gat, Tal Remez, David Kant, Gabriel Synnaeve, Yossi Adi, and Alexandre Défossez · 2023
Closest in time.
Simple and controllable music generation
Jade Copet, Felix Kreuk, Itai Gat, Tal Remez, David Kant, Gabriel Synnaeve, Yossi Adi, and Alexandre Défossez · 2023
Closest in time.
Magenta improv rnn
Magenta Developers · 2023
Closest in time.
Magenta melody rnn
Magenta Developers · 2023
Closest in time.
Magenta polyphony rnn
Magenta Developers · 2023
Closest in time.
Magenta model directory
Magenta Developers · 2023
Closest in time.
Markov melody generator
Simone Hill · 2023
Closest in time.
Noise2music: Text-conditioned music generation with diffusion models, 2023
Qingqing Huang, Daniel S. Park, Tao Wang, Timo I. Denk, Andy Ly, Nanxin Chen, Zhengdong Zhang, Zhishuai Zhang, Jiahui Yu, Christian Frank, Jesse Engel, Quoc V. Le, William Chan, Zhifeng Chen, and Wei Han · 2023
Closest in time.
Make-an-audio: Text-to-audio generation with prompt-enhanced diffusion models, 2023
Rongjie Huang, Jiawei Huang, Dongchao Yang, Yi Ren, Luping Liu, Mingze Li, Zhenhui Ye, Jinglin Liu, Xiang Yin, and Zhou Zhao · 2023
Closest in time.
Classical archives: The largest classical music site in the world, 2023
Classical Archives LLC · 2023
Closest in time.
MusicXML
Recordare LLC · 2023
Closest in time.
Moûsai: Text-to-music generation with long-context latent diffusion, 2023
Flavio Schneider, Zhijing Jin, and Bernhard Schölkopf · 2023
Closest in time.
Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation, 2023
Yusong Wu, Ke Chen, Tianyu Zhang, Yuchen Hui, Taylor Berg-Kirkpatrick, and Shlomo Dubnov · 2023
Closest in time.