Fetching the paper…
Reading the bibliography…
Musical expression requires control of both what notes are played, and how they are performed.
Tentative standards for sound level meters
RG McCurdy · 1936
Earlier work this paper cites.
The synthesis of complex audio spectra by means of frequency modulation
John M Chowning · 1973
Earlier work this paper cites.
Introduction to granular synthesis
Curtis Roads · 1988
Earlier work this paper cites.
Spectral modeling synthesis: A sound analysis/synthesis system based on a deterministic plus stochastic decomposition
Xavier Serra and Julius Smith · 1990
Earlier work this paper cites.
The complete midi 1.0 detailed specification
MIDI Manufacturers Association et al · 1996
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. Lecun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Modeling and control of expressiveness in music performance
Sergio Canazza, Giovanni De Poli, Carlo Drioli, Antonio Roda, and Alvise Vidolin · 2004
Earlier work this paper cites.
Corpus-based concatenative synthesis
Diemo Schwarz · 2007
Earlier work this paper cites.
Rectifier nonlinearities improve neural network acoustic models
Andrew L Maas, Awni Y Hannun, and Andrew Y Ng · 2013
Earlier work this paper cites.
Generative adversarial networks. arxiv e-prints
Ian J Goodfellow, J Pouget-Abadie, M Mirza, B Xu, D Warde-Farley, S Ozair, A Courville, and Y Bengio · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
The sense of ensemble: a machine learning approach to expressive performance modelling in string quartets
Marco Marchini, Rafael Ramirez, Panos Papiotis, and Esteban Maestre · 2014
Earlier work this paper cites.
Analysis of expressive musical terms in violin using score-informed and expression-based audio features
Pei-Ching Li, Li Su, Yi-Hsuan Yang, Alvin WY Su, et al · 2015
Earlier work this paper cites.
Empirical evaluation of rectified activations in convolutional network
Bing Xu, Naiyan Wang, Tianqi Chen, and Mu Li · 2015
Earlier work this paper cites.
Jimmy Lei Ba, Jamie Ryan Kiros, and Geoffrey E Hinton · 2016
Earlier work this paper cites.
Wavenet: A generative model for raw audio
Aäron van den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew Senior, and Koray Kavukcuoglu · 2016
Earlier work this paper cites.
Automatic violin synthesis using expressive musical term features
Chih-Hong Yang, Pei-Ching Li, AW Su, Li Su, Yi-Hsuan Yang, et al · 2016
Cited alongside, same era.
Deep salience representations for f0 estimation in polyphonic music
Rachel M Bittner, Brian McFee, Justin Salamon, Peter Li, and Juan Pablo Bello · 2017
Cited alongside, same era.
Counterpoint by convolution
Cheng-Zhi Anna Huang, Tim Cooijmans, Adam Roberts, Aaron Courville, and Douglas Eck · 2017
Cited alongside, same era.
Least squares generative adversarial networks
Xudong Mao, Qing Li, Haoran Xie, Raymond YK Lau, Zhen Wang, and Stephen Paul Smolley · 2017
Cited alongside, same era.
Analysis and synthesis of the violin playing style of heifetz and oistrakh
Chi-Ching Shih, Pei-Ching Li, Yi-Ju Lin, Yu-Lin Wang, Alvin WY Su, Li Su, and Yi-Hsuan Yang · 2017
Cited alongside, same era.
Neural source-filter waveform models for statistical parametric speech synthesis
Xin Wang, Shinji Takaki, and Junichi Yamagishi · 2019
Later among the works it cites.
Faceswapnet: Landmark guided many-to-many face reenactment
Jiangning Zhang, Xianfang Zeng, Yusu Pan, Yong Liu, Yu Ding, and Changjie Fan · 2019
Later among the works it cites.
The sound of motions
Hang Zhao, Chuang Gan, Wei-Chiu Ma, and Antonio Torralba · 2019
Later among the works it cites.
Vector-quantized timbre representation
Adrien Bitton, Philippe Esling, and Tatsuya Harada · 2020
Later among the works it cites.
Language models are few-shot learners
Tom B Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Alexandre Défossez, Neil Zeghidour, Nicolas Usunier, Léon Bottou, and Francis Bach · 2018
Cited alongside, same era.
The challenge of realistic music generation: modelling raw audio at scale
Sander Dieleman, Aaron van den Oord, and Karen Simonyan · 2018
Cited alongside, same era.
Crepe: A convolutional representation for pitch estimation
Jong Wook Kim, Justin Salamon, Peter Li, and Juan Pablo Bello · 2018
Cited alongside, same era.
Creating a multitrack classical music performance dataset for multimodal music analysis: Challenges, insights, and applications
Bochen Li, Xinzhao Liu, Karthik Dinesh, Zhiyao Duan, and Gaurav Sharma · 2018
Cited alongside, same era.
Conditioning deep generative raw audio models for structured automatic music
Rachel Manzelli, Vijay Thakkar, Ali Siahkamari, and Brian Kulis · 2018
Cited alongside, same era.
Everybody dance now
Caroline Chan, Shiry Ginosar, Tinghui Zhou, and Alexei A Efros · 2019
Cited alongside, same era.
Gansynth: Adversarial neural audio synthesis
Jesse Engel, Kumar Krishna Agrawal, Shuo Chen, Ishaan Gulrajani, Chris Donahue, and Adam Roberts · 2019
Cited alongside, same era.
Later among the works it cites.
Towards realistic midi instrument synthesizers
Rodrigo Castellon, Chris Donahue, and Percy Liang · 2020
Later among the works it cites.
Jukebox: A generative model for music
Prafulla Dhariwal, Heewoo Jun, Christine Payne, Jong Wook Kim, Alec Radford, and Ilya Sutskever · 2020
Later among the works it cites.
The curious case of neural text degeneration
Ari Holtzman, Jan Buys, Li Du, Maxwell Forbes, and Yejin Choi · 2020
Later among the works it cites.
Ai song contest: Human-ai co-creation in songwriting
Cheng-Zhi Anna Huang, Hendrik Vincent Koops, Ed Newton-Rex, Monica Dinculescu, and Carrie J. Cai · 2020
Later among the works it cites.
The control-synthesis approach for making expressive and controllable neural music synthesizers
Nicolas Jonason, Bob Sturm, and Carl Thomé · 2020
Later among the works it cites.
Panns: Large-scale pretrained audio neural networks for audio pattern recognition
Qiuqiang Kong, Yin Cao, Turab Iqbal, Yuxuan Wang, Wenwu Wang, and Mark D Plumbley · 2020
Later among the works it cites.
Controllable neural prosody synthesis
Max Morrison, Zeyu Jin, Justin Salamon, Nicholas J Bryan, and Gautham J Mysore · 2020
Later among the works it cites.
Fastspeech 2: Fast and high-quality end-to-end text to speech
Yi Ren, Chenxu Hu, Xu Tan, Tao Qin, Sheng Zhao, Zhou Zhao, and Tie-Yan Liu · 2020
Later among the works it cites.
Sequence-to-sequence piano transcription with transformers
Curtis Hawthorne, Ian Simon, Rigel Swavely, Ethan Manilow, and Jesse Engel · 2021
Closest in time.
Ben Hayes, Charalampos Saitis, and György Fazekas · 2021
Closest in time.
Zero-shot text-to-image generation
Aditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray, Chelsea Voss, Alec Radford, Mark Chen, and Ilya Sutskever · 2021
Closest in time.
A Neural Parametric Singing Synthesizer Modeling Timbre and Expression from Natural Songs
Merlijn Blaauw and Jordi Bonada · 2076
Closest in time.