Fetching the paper…
Reading the bibliography…
Generating musical audio directly with neural networks is notoriously difficult because it requires coherently modeling structure at many different timescales.
Dynamic programming algorithm optimization for spoken word recognition
Hiroaki Sakoe and Seibi Chiba · 1978
Earlier work this paper cites.
Calculation of a constant q spectral transform
Judith C. Brown · 1991
Earlier work this paper cites.
S Young, G Evermann, M Gales, T Hain, D Kershaw, X Liu, G Moore, J Odell, D Ollason, D Povey, et al · 2006
Earlier work this paper cites.
Evaluation of multiple-f0 estimation and tracking systems
Mert Bay, Andreas F Ehmann, and J Stephen Downie · 2009
Earlier work this paper cites.
A modal-based real-time piano synthesizer
B. Bank, S. Zambon, and F. Fontana · 2010
Earlier work this paper cites.
MAPS - A piano database for multipitch estimation and automatic transcription of music
Valentin Emiya, Nancy Bertin, Bertrand David, and Roland Badeau · 2010
Earlier work this paper cites.
Constant-q transform toolbox for music processing
Christian Schörkhuber and Anssi Klapuri · 2010
Earlier work this paper cites.
Saarland music data (SMD)
Meinard Müller, Verena Konz, Wolfgang Bogler, and Vlora Arifi-Müller · 2011
Earlier work this paper cites.
Fifty years of artificial reverberation
Vesa Valimaki, Julian D Parker, Lauri Savioja, Julius O Smith, and Jonathan S Abel · 2012
Cited alongside, same era.
Intuitive analysis, creation and manipulation of midi data with pretty_midi
Colin Raffel and Daniel P W Ellis · 2014
Cited alongside, same era.
Fundamentals of music processing: Audio, analysis, algorithms, applications
Meinard Müller · 2015
Cited alongside, same era.
Pysox: Leveraging the audio signal processing power of sox in python
Rachel Bittner, Eric Humphrey, and Juan Bello · 2016
Cited alongside, same era.
WaveNet: A generative model for raw audio
Aäron van den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew W Senior, and Koray Kavukcuoglu · 2016
Cited alongside, same era.
librosa 0.5.1, May 2017
Brian McFee, Matt McVicar, Oriol Nieto, Stefan Balke, Carl Thome, Dawen Liang, Eric Battenberg, Josh Moore, Rachel Bittner, Ryuichi Yamamoto, Dan Ellis, Fabian-Robert Stoter, Douglas Repetto, Simon Waloschek, CJ Carr, Seth Kranzler, Keunwoo Choi, Petr Viktorin, Joao Felipe Santos, Adrian Holovaty, Waldir Pimenta, Hojin Lee, and Paul Brossier · 2017
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Later among the works it cites.
The challenge of realistic music generation: modelling raw audio at scale
Sander Dieleman, Aäron van den Oord, and Karen Simonyan · 2018
Closest in time.
Onsets and frames: Dual-objective piano transcription
Curtis Hawthorne, Erich Elsen, Jialin Song, Adam Roberts, Ian Simon, Colin Raffel, Jesse Engel, Sageev Oore, and Douglas Eck · 2018
Closest in time.
An improved relative self-attention mechanism for transformer with application to music generation
Cheng-Zhi Anna Huang, Ashish Vaswani, Jakob Uszkoreit, Noam Shazeer, Curtis Hawthorne, Andrew M Dai, Matthew D Hoffman, and Douglas Eck · 2018
Closest in time.
Deep polyphonic adsr piano note transcription
Rainer Kelz, Sebastian Bock, and Gerhard Widmer · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Learning features of music from scratch
John Thickstun, Zaid Harchaoui, and Sham Kakade · 2017
Cited alongside, same era.
Combining deep generative raw audio models for structured automatic music
Thakkar Vijay Manzelli, Rachel and, Ali Siahkamari, and Brian Kulis · 2018
Closest in time.
Parallel WaveNet: Fast high-fidelity speech synthesis
Aäron van den Oord, Yazhe Li, Igor Babuschkin, Karen Simonyan, Oriol Vinyals, Koray Kavukcuoglu, George van den Driessche, Edward Lockhart, Luis Cobo, Florian Stimberg, Norman Casagrande, Dominik Grewe, Seb Noury, Sander Dieleman, Erich Elsen, Nal Kalchbrenner, Heiga Zen, Alex Graves, Helen King, Tom Walters, Dan Belov, and Demis Hassabis · 2018
Closest in time.