Fetching the paper…
Reading the bibliography…
Following their success in Computer Vision and other areas, deep learning techniques have recently become widely adopted in Music Information Retrieval (MIR) research.
Learning internal representations by error propagation
D. E. Rumelhart, G. E. Hinton, and R. J. Williams · 1985
Earlier work this paper cites.
Speech communication: human and machine
D. O’shaughnessy · 1987
Earlier work this paper cites.
Modèles connexionnistes de l’apprentissage
L. Yann · 1987
Earlier work this paper cites.
Backpropagation applied to handwritten zip code recognition
Y. LeCun, B. Boser, J. S. Denker, D. Henderson, R. E. Howard, W. Hubbard, and L. D. Jackel · 1989
Earlier work this paper cites.
Learning long-term dependencies with gradient descent is difficult
Y. Bengio, P. Simard, and P. Frasconi · 1994
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Bidirectional recurrent neural networks
M. Schuster and K. K. Paliwal · 1997
Earlier work this paper cites.
Realtime chord recognition of musical sound: a system using common lisp music
T. Fujishima · 1999
Earlier work this paper cites.
Mathematical representation of joint time-chroma distributions
G. H. Wakefield · 1999
Earlier work this paper cites.
Gtzan genre collection
G. Tzanetakis and P. Cook · 2001
Earlier work this paper cites.
Music information retrieval
J. S. Downie · 2003
Earlier work this paper cites.
A comparative study of filter bank spacing for speech recognition
B. J. Shannon and K. K. Paliwal · 2003
Earlier work this paper cites.
Extreme learning machine: a new learning scheme of feedforward neural networks
G.-B. Huang, Q.-Y. Zhu, and C.-K. Siew · 2004
Earlier work this paper cites.
Social tagging and music information retrieval
P. Lamere · 2008
Earlier work this paper cites.
Unsupervised feature learning for audio classification using convolutional deep belief networks
H. Lee, P. Pham, Y. Largman, and A. Y. Ng · 2009
Earlier work this paper cites.
Universal onset detection with bidirectional long short-term memory neural networks
F. Eyben, S. Böck, B. W. Schuller, and A. Graves · 2010
Earlier work this paper cites.
Learning features from music audio with deep belief networks
P. Hamel and D. Eck · 2010
Earlier work this paper cites.
Audio musical genre classification using convolutional neural networks and pitch and tempo transformations
L. Li · 2010
Earlier work this paper cites.
Automatic musical pattern feature extraction using convolutional neural network
T. L. Li, A. B. Chan, and A. Chun · 2010
Earlier work this paper cites.
Melody line estimation in homophonic music audio signals based on temporal-variability of melodic source
H. Tachibana, T. Ono, N. Ono, and S. Sagayama · 2010
Earlier work this paper cites.
Simple methods for improving speaker-similarity of hmm-based speech synthesis
J. Yamagishi and S. King · 2010
Earlier work this paper cites.
The million song dataset
T. Bertin-Mahieux, D. P. Ellis, B. Whitman, and P. Lamere · 2011
Earlier work this paper cites.
Deep sparse rectifier neural networks
X. Glorot, A. Bordes, and Y. Bengio · 2011
Earlier work this paper cites.
Rethinking automatic chord recognition with convolutional neural networks
E. J. Humphrey and J. P. Bello · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
Efficient backprop
Y. A. LeCun, L. Bottou, G. B. Orr, and K.-R. Müller · 2012
Earlier work this paper cites.
An introduction to the psychology of hearing
B. C. Moore · 2012
Earlier work this paper cites.
Unsupervised learning of local features for music classification
J. Wülfing and M. Riedmiller · 2012
Earlier work this paper cites.
Improving deep neural networks for lvcsr using rectified linear units and dropout
G. E. Dahl, T. N. Sainath, and G. E. Hinton · 2013
Earlier work this paper cites.
Multiscale approaches to music audio feature learning
S. Dieleman and B. Schrauwen · 2013
Earlier work this paper cites.
Transfer learning in mir: Sharing learned latent representations for music audio classification and similarity
P. Hamel, M. E. Davies, K. Yoshii, and M. Goto · 2013
Earlier work this paper cites.
M. Lin, Q. Chen, and S. Yan · 2013
Earlier work this paper cites.
Learning binary codes for efficient large-scale music similarity search
J. Schlüter · 2013
Earlier work this paper cites.
Musical onset detection with convolutional neural networks
J. Schlüter and S. Böck · 2013
Earlier work this paper cites.
Deep content-based music recommendation
A. Van den Oord, S. Dieleman, and B. Schrauwen · 2013
Earlier work this paper cites.
On rectified linear units for speech processing
M. D. Zeiler, M. Ranzato, R. Monga, M. Mao, K. Yang, Q. V. Le, P. Nguyen, A. Senior, V. Vanhoucke, J. Dean, et al · 2013
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
D. Bahdanau, K. Cho, and Y. Bengio · 2014
Earlier work this paper cites.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
K. Cho, B. van Merrienboer, C. Gulcehre, F. Bougares, H. Schwenk, and Y. Bengio · 2014
Earlier work this paper cites.
End-to-end learning for music audio
S. Dieleman and B. Schrauwen · 2014
Earlier work this paper cites.
Generative adversarial nets
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio · 2014
Earlier work this paper cites.
Singing-voice separation from monaural recordings using deep recurrent neural networks
P.-S. Huang, M. Kim, M. Hasegawa-Johnson, and P. Smaragdis · 2014
Earlier work this paper cites.
From music audio to chord tablature: Teaching deep convolutional networks toplay guitar
E. J. Humphrey and J. P. Bello · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Cited alongside, same era.
Conditional generative adversarial nets
M. Mirza and S. Osindero · 2014
Cited alongside, same era.
Long short-term memory recurrent neural network architectures for large scale acoustic modeling
H. Sak, A. Senior, and F. Beaufays · 2014
Cited alongside, same era.
Improved musical onset detection with convolutional neural networks
J. Schluter and S. Bock · 2014
Cited alongside, same era.
Y.-A. Chung, C.-C. Wu, C.-H. Shen, H.-Y. Lee, and L.-S. Lee · 2016
Later among the works it cites.
Automatic chord estimation on seventhsbass chord vocabulary using deep neural network
J. Deng and Y.-K. Kwok · 2016
Later among the works it cites.
A primer on neural network models for natural language processing
Y. Goldberg · 2016
Later among the works it cites.
Deep learning
I. Goodfellow, Y. Bengio, and A. Courville · 2016
Later among the works it cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Later among the works it cites.
Learning temporal features using a deep neural network and its application to music genre classification
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. Sigtia and S. Dixon · 2014
Cited alongside, same era.
Chord recognition with stacked denoising autoencoders
N. Steenbergen, T. Gevers, and J. Burgoyne · 2014
Cited alongside, same era.
Sequence to sequence learning with neural networks
I. Sutskever, O. Vinyals, and Q. V. Le · 2014
Cited alongside, same era.
Boundary detection in music structure analysis using convolutional neural networks
K. Ullrich, J. Schlüter, and T. Grill · 2014
Cited alongside, same era.
Transfer learning by supervised pre-training for audio-based music classification
A. Van Den Oord, S. Dieleman, and B. Schrauwen · 2014
Cited alongside, same era.
Towards time-varying music auto-tagging based on cal500 expansion
S.-Y. Wang, J.-C. Wang, Y.-H. Yang, and H.-M. Wang · 2014
Cited alongside, same era.
Visualizing and understanding convolutional networks
M. D. Zeiler and R. Fergus · 2014
Cited alongside, same era.
I.-Y. Jeong and K. Lee · 2016
Later among the works it cites.
On the potential of simple framewise approaches to piano transcription
R. Kelz, M. Dorfer, F. Korzeniowski, S. Böck, A. Arzt, and G. Widmer · 2016
Later among the works it cites.
Feature learning for chord recognition: The deep chroma extractor
F. Korzeniowski and G. Widmer · 2016
Later among the works it cites.
A fully convolutional deep auditory model for musical chord recognition
F. Korzeniowski and G. Widmer · 2016
Later among the works it cites.
A deep bidirectional long short-term memory based multi-scale approach for music dynamic emotion prediction
X. Li, H. Xianyu, J. Tian, W. Chen, F. Meng, M. Xu, and L. Cai · 2016
Later among the works it cites.
Deep convolutional networks on the pitch spiral for musical instrument recognition
V. Lostanlen and C.-E. Cella · 2016
Later among the works it cites.
Samplernn: An unconditional end-to-end neural audio generation model
S. Mehri, K. Kumar, I. Gulrajani, R. Kumar, S. Jain, J. Sotelo, A. Courville, and Y. Bengio · 2016
Later among the works it cites.
Learning high-level features for chord recognition using autoencoder
V. Phongthongloa, S. Kamonsantiroj, and L. Pipanmaekaporn · 2016
Later among the works it cites.
Singing voice melody transcription using deep neural networks
F. Rigaud and M. Radenen · 2016
Later among the works it cites.
An overview of gradient descent optimization algorithms
S. Ruder · 2016
Later among the works it cites.
Learning to pinpoint singing voice from weakly labeled examples
J. Schlüter · 2016
Later among the works it cites.
Mastering the game of go with deep neural networks and tree search
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, et al · 2016
Later among the works it cites.
Wavenet: A generative model for raw audio
A. van den Oord, S. Dieleman, H. Zen, K. Simonyan, O. Vinyals, A. Graves, N. Kalchbrenner, A. Senior, and K. Kavukcuoglu · 2016
Later among the works it cites.
M. Arjovsky, S. Chintala, and L. Bottou · 2017
Closest in time.
Began: Boundary equilibrium generative adversarial networks
D. Berthelot, T. Schumm, and L. Metz · 2017
Closest in time.
Deep salience representations for f 0 f_{0} estimation in polyphonic music
R. Bittner, B. McFee, J. Salamon, P. Li, and J. Bello · 2017
Closest in time.
A neural parametric singing synthesizer
M. Blaauw and J. Bonada · 2017
Closest in time.
Deep cross-modal audio-visual generation
L. Chen, S. Srivastava, Z. Duan, and C. Xu · 2017
Closest in time.
A comparison on audio signal preprocessing methods for deep neural networks on music tagging
K. Choi, G. Fazekas, K. Cho, and M. Sandler · 2017
Closest in time.
The effects of noisy labels on deep convolutional neural networks for music classification
K. Choi, G. Fazekas, K. Cho, and M. Sandler · 2017
Closest in time.
Transfer learning for music classification and regression tasks
K. Choi, G. Fazekas, M. Sandler, and K. Cho · 2017
Closest in time.
Kapre: On-gpu audio preprocessing layers for a quick implementation of deep neural network models with keras
K. Choi, D. Joo, and J. Kim · 2017
Closest in time.
Neural audio synthesis of musical notes with wavenet autoencoders
J. Engel, C. Resnick, A. Roberts, S. Dieleman, D. Eck, K. Simonyan, and M. Norouzi · 2017
Closest in time.
Deep convolutional neural networks for predominant instrument recognition in polyphonic music
Y. Han, J. Kim, and K. Lee · 2017
Closest in time.
J. Lee and J. Nam · 2017
Closest in time.
Sample-level deep convolutional neural networks for music auto-tagging using raw waveforms
J. Lee, J. Park, K. L. Kim, and J. Nam · 2017
Closest in time.
Deep ranking: triplet matchnet for music metric learning
R. Lu, K. Wu, Z. Duan, and C. Zhang · 2017
Closest in time.
Stacked convolutional and recurrent neural networks for music emotion recognition
M. Malik, S. Adavanne, K. Drossos, T. Virtanen, D. Ticha, and R. Jarina · 2017
Closest in time.
Structured training for large-vocabulary chord recognition
B. McFee and J. Bello · 2017
Closest in time.
Segan: Speech enhancement generative adversarial network
S. Pascual, A. Bonafonte, and J. Serrà · 2017
Closest in time.
Designing efficient architectures for modeling temporal features with convolutional neural networks
J. Pons and X. Serra · 2017
Closest in time.
Timbre analysis of music audio signals with convolutional neural networks, 2017
J. Pons, O. Slizovskaia, R. Gong, E. Gómez, and X. Serra · 2017
Closest in time.
Music transcription with convolutional sequence-to-sequence models
K. Ullrich and E. van der Wel · 2017
Closest in time.
Revisiting the problem of audio-based hit song prediction using convolutional neural networks
L.-C. Yang, S.-Y. Chou, J.-Y. Liu, Y.-H. Yang, and Y.-A. Chen · 2017
Closest in time.
L.-C. Yang, S.-Y. Chou, and Y.-H. Yang · 2017
Closest in time.