Fetching the paper…
Reading the bibliography…
Self-attention is an attention mechanism that learns a representation by relating different positions in the sequence.
The relationship between precision-recall and roc curves
J. Davis and M. Goadrich · 2006
Earlier work this paper cites.
Evaluation of algorithms using games: The case of music tagging
E. Law, K. West, M. I. Mandel, M. Bay, and J. S. Downie · 2009
Earlier work this paper cites.
The million song dataset
T. Bertin-Mahieux, D. P. Ellis, B. Whitman, and P. Lamere · 2011
Earlier work this paper cites.
Essentia: An audio analysis library for music information retrieval
D. Bogdanov, N. Wack, E. Gómez Gutiérrez, S. Gulati, P. Herrera Boyer, O. Mayor, G. Roma Trepat, J. Salamon, J. R. Zapata González, and X. Serra · 2013
Earlier work this paper cites.
On the difficulty of training recurrent neural networks
R. Pascanu, T. Mikolov, and Y. Bengio · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2015
Earlier work this paper cites.
librosa: Audio and music signal analysis in python
B. McFee, C. Raffel, D. Liang, D. P. Ellis, M. McVicar, E. Battenberg, and O. Nieto · 2015
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2015
Earlier work this paper cites.
Document modeling with gated recurrent neural network for sentiment classification
D. Tang, B. Qin, and T. Liu · 2015
Earlier work this paper cites.
Convolutional recurrent neural networks: Learning spatial dependencies for image representation
Z. Zuo, B. Shuai, G. Wang, X. Liu, X. Wang, B. Wang, and Y. Chen · 2015
Earlier work this paper cites.
Automatic tagging using deep convolutional neural networks
K. Choi, G. Fazekas, and M. Sandler · 2016
Earlier work this paper cites.
Explaining deep convolutional neural networks on music classification
K. Choi, G. Fazekas, and M. Sandler · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
Downbeat tracking using beat synchronous features with recurrent neural networks
F. Krebs, S. Böck, M. Dorfer, and G. Widmer · 2016
Earlier work this paper cites.
An end-to-end neural network for polyphonic piano music transcription
S. Sigtia, E. Benetos, and S. Dixon · 2016
Earlier work this paper cites.
Wavenet: A generative model for raw audio
A. Van Den Oord, S. Dieleman, H. Zen, K. Simonyan, O. Vinyals, A. Graves, N. Kalchbrenner, A. W. Senior, and K. Kavukcuoglu · 2016
Cited alongside, same era.
Convolutional recurrent neural networks for music classification
K. Choi, G. Fazekas, M. Sandler, and K. Cho · 2017
Cited alongside, same era.
Transfer learning for music classification and regression tasks
K. Choi, G. Fazekas, M. Sandler, and K. Cho · 2017
Cited alongside, same era.
Onsets and frames: Dual-objective piano transcription
C. Hawthorne, E. Elsen, J. Song, A. Roberts, I. Simon, C. Raffel, J. Engel, S. Oore, and D. Eck · 2017
Cited alongside, same era.
Train longer, generalize better: closing the generalization gap in large batch training of neural networks
E. Hoffer, I. Hubara, and D. Soudry · 2017
Cited alongside, same era.
Drum transcription via joint beat and drum modeling using convolutional recurrent neural networks
R. Vogl, M. Dorfer, G. Widmer, and P. Knees · 2017
Later among the works it cites.
The marginal value of adaptive gradient methods in machine learning
A. C. Wilson, R. Roelofs, M. Stern, N. Srebro, and B. Recht · 2017
Later among the works it cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova · 2018
Later among the works it cites.
NSML: Meet the mlaas platform with a real-world case study
H. Kim, M. Kim, D. Seo, J. Kim, H. Park, S. Park, H. Jo, K. Kim, Y. Yang, Y. Kim, et al · 2018
Later among the works it cites.
Sample-level cnn architectures for music auto-tagging using raw waveforms
T. Kim, J. Lee, and J. Nam · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
N. S. Keskar and R. Socher · 2017
Cited alongside, same era.
Sample-level deep convolutional neural networks for music auto-tagging using raw waveforms
J. Lee, J. Park, K. L. Kim, and J. Nam · 2017
Cited alongside, same era.
Interactive visualization and manipulation of attention-based neural machine translation
J. Lee, J.-H. Shin, and J.-S. Kim · 2017
Cited alongside, same era.
A structured self-attentive sentence embedding
Z. Lin, M. Feng, C. N. d. Santos, M. Yu, B. Xiang, B. Zhou, and Y. Bengio · 2017
Cited alongside, same era.
Local interpretable model-agnostic explanations for music content analysis
S. Mishra, B. L. Sturm, and S. Dixon · 2017
Cited alongside, same era.
Designing efficient architectures for modeling temporal features with convolutional neural networks
J. Pons and X. Serra · 2017
Cited alongside, same era.
Timbre analysis of music audio signals with convolutional neural networks
J. Pons, O. Slizovskaia, R. Gong, E. Gómez, and X. Serra · 2017
Cited alongside, same era.
Understanding a deep machine listening model through feature inversion
S. Mishra, B. L. Sturm, and S. Dixon · 2018
Later among the works it cites.
Image transformer
N. Parmar, A. Vaswani, J. Uszkoreit, Ł. Kaiser, N. Shazeer, A. Ku, and D. Tran · 2018
Later among the works it cites.
End-to-end learning for music audio tagging at scale
J. Pons, O. Nieto, M. Prockup, E. Schmidt, A. Ehmann, and X. Serra · 2018
Later among the works it cites.
Language models are unsupervised multitask learners
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, and I. Sutskever · 2018
Later among the works it cites.
Towards flatter loss surface via nonmonotonic learning rate scheduling
S. Seong, Y. Lee, Y. Kee, D. Han, and J. Kim · 2018
Later among the works it cites.
Non-local neural networks
X. Wang, R. Girshick, A. Gupta, and K. He · 2018
Later among the works it cites.
Compact generalized non-local network
K. Yue, M. Sun, Y. Yuan, F. Zhou, E. Ding, and F. Xu · 2018
Later among the works it cites.
Enabling factorized piano music modeling and generation with the MAESTRO dataset
C. Hawthorne, A. Stasyuk, A. Roberts, I. Simon, C.-Z. A. Huang, S. Dieleman, E. Elsen, J. Engel, and D. Eck · 2019
Closest in time.
Music transformer
C.-Z. A. Huang, A. Vaswani, J. Uszkoreit, I. Simon, C. Hawthorne, N. Shazeer, A. M. Dai, M. D. Hoffman, M. Dinculescu, and D. Eck · 2019
Closest in time.
Self-attention generative adversarial networks
H. Zhang, I. Goodfellow, D. Metaxas, and A. Odena · 2019
Closest in time.