Fetching the paper…
Reading the bibliography…
Feature learning and deep learning have drawn great attention in recent years as a way of transforming input data into more effective representations using learning algorithms.
D. J. F. Bruno. A. Olshausen, “Emergence of simple-cellreceptive field properties by learning a sparse code for natural images,” Nature , pp. 607–609, 1996
1996
Earlier work this paper cites.
G. Tzanetakis and P. Cook, “Musical genre classification of audio signals,” IEEE Transaction on Speech and Audio Processing , 2002
2002
Earlier work this paper cites.
D.-N. Jiang, L. Lu, H.-J. Zhang, and J.-H. Tao, “Music type classification by spectral contrast feature,” in Proceedings of International Conference on Multimedia Expo (ICME) , 2002
2002
Earlier work this paper cites.
M. S. Lewicki, “Efficient coding of natural sounds,” Nature Neuroscience , 2002
2002
Earlier work this paper cites.
M. F. McKinney and J. Breebaart, “Features for audio and music classification,” in Proceedings of the 4th International Conference on Music Information Retrieval (ISMIR) , 2003
2003
Earlier work this paper cites.
T. Li, M. Ogihara, and Q. Li, “A comparative study of content-based music genre classification,” in Proceedings of the 26th international ACM SIGIR conference on Research and development in informaion retrieval , 2003
2003
Earlier work this paper cites.
J. Bergstra, N. Casagrande, D. Erhan, D. Eck, and B. Kegl, “Aggregate features and adaboost for music classification,” Machine Learning , 2006
2006
Earlier work this paper cites.
G. E. Hinton, S. Osindero, and Y.-W. Teh, “A fast learning algorithm for deep belief nets,” Neural computation , vol. 18, pp. 1527–1554, 2006
2006
Earlier work this paper cites.
R. Grosse, R. Raina, H. Kwong, and A. Y. Ng, “Shift-invariant sparse coding for audio classification,” in Proceedings of the Conference on Uncertainty in AI , 2007
2007
Earlier work this paper cites.
Y. Bengio, P. Lamblin, D. Popovici, and H. Larochelle, “Greedy layer-wise training of deep networks,” Advances in Neural Information Processing Systems 19 , 2007
2007
Earlier work this paper cites.
C. N. Silla, A. Koerich, and C. Kaestner, “A feature selection approach for automatic music genre classification,” International Journal of Semantic Computing , 2008
2008
Earlier work this paper cites.
P.-A. Manzagol, T. Bertin-Mahieux, and D. Eck, “On the use of sparse time-relative auditory codes for music,” in Proceedings of the 9th International Conference on Music Information Retrieval (ISMIR) , 2008
2008
Earlier work this paper cites.
H. Lee, C. Ekanadham, and A. Y. Ng, “Sparse deep belief net model for visual area V2,” in Advances in Neural Information Processing Systems 20 , 2008, pp. 873–880
2008
Earlier work this paper cites.
Y. Bengio, “Learning deep architectures for ai,” Foundations and trends in Machine Learning , 2009
2009
Earlier work this paper cites.
H. Lee, Y. Largman, P. Pham, and A. Y. Ng, “Unsupervised feature learning for audio classification using convolutional deep belief networks,” in Advances in Neural Information Processing Systems 22 , 2009, pp. 1096–1104
2009
Earlier work this paper cites.
A. Hyvärinen, J. Hurri, and P. O. Hoyer, Natural Image Statistics . Springer-Verlag, 2009
2009
Earlier work this paper cites.
E. Law and L. V. Ahn, “Input-agreement: a new mechanism for collecting data using human computation games,” in Proc. Intl. Conf. on Human factors in computing systems, CHI. ACM , 2009
2009
Earlier work this paper cites.
Y. Panagakis, C. Kotropoulos, and G. R. Arce, “Non-negative multilinear principal component analysis of auditory temporal modulations for music genre classification,” IEEE Transaction on Audio, Speech and Language Processing , 2010
2010
Cited alongside, same era.
T. Bertin-Mahieux, D. Eck, F. Maillet, and P. Lamere, “Autotagger: a model for predicting social tags from acoustic features on large music databases,” in Journal of New Music Research , 2010
2010
Cited alongside, same era.
K. K. C. J.-S. R. Jang and C. S. Iliopoulos, “Music genre classification via compressive sampling,” in Proceedings of the 11th International Conference on Music Information Retrieval (ISMIR) , 2010
2010
Cited alongside, same era.
P. Hamel and D. Eck, “Learning features from music audio with deep belief networks,” in In Proceedings of the 11th International Conference on Music Information Retrieval (ISMIR) , 2010
2010
Cited alongside, same era.
J. Wülfing and M. Riedmiller, “Unsupervised learning of local features for music classification,” in Proceedings of the 13th International Conference on Music Information Retrieval (ISMIR) , 2012
2012
Later among the works it cites.
E. M. Schmidt, J. Scott, and Y. E. Kim, “Feature learning in dynamic environments:modeling the acoustic structure of musical emotion,” in Proceedings of the 13th International Conference on Music Information Retrieval (ISMIR) , 2012
2012
Later among the works it cites.
P. Hamel, Y. Bengio, and D. Eck, “Building musically-relevant audio features through multiple timescale representations,” in Proceedings of the 13th International Conference on Music Information Retrieval (ISMIR) , 2012
2012
Later among the works it cites.
M. D. Zeiler, “Adadelta: An adaptive learning rate method,” in arXiv:1212.5701v1 , 2012
2012
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. Ellis, “Time-frequency automatic gain control,” web resource, available, http://labrosa.ee.columbia.edu/matlab/tf_agc/
2010
Cited alongside, same era.
G. E. Hinton, “A practical guide to training restricted boltzmann machines,” UTML Technical Report , vol. 2010-003, 2010
2010
Cited alongside, same era.
V. Nair and G. Hinton, “Rectified linear units improve restricted boltzmann machines,” in Proceedings of the 27th International Conference on Machine Learning (ICML) , 2010
2010
Cited alongside, same era.
M. Henaff, K. Jarrett, K. Kavukcuoglu, and Y. LeCun, “Unsupervised learning of sparse features for scalable audio classification,” in Proceedings of the 12th International Conference on Music Information Retrieval (ISMIR) , 2011
2011
Cited alongside, same era.
J. Schlüter and C. Osendorfer, “Music Similarity Estimation with the Mean-Covariance Restricted Boltzmann Machine,” in Proceedings of the 10th International Conference on Machine Learning and Applications , 2011
2011
Cited alongside, same era.
A. Coates, H. Lee, and A. Ng, “An analysis of single-layer networks in unsupervised feature learning,” Journal of Machine Learning Research , 2011
2011
Cited alongside, same era.
E. M. Schmidt and Y. E. Kim, “Learning emotion-based acoustic features with deep belief networks,” in Proceedings of the 2011 IEEE Workshop on Applications of Signal Processing to Audio and Acoustics (WASPAA) , 2011
2011
Cited alongside, same era.
P. Hamel, S. Lemieux, Y. Bengio, and D. Eck, “Temporal pooling and multiscale learning for automatic annotation and ranking of music audio,” in Proceedings of the 12th International Conference on Music Information Retrieval (ISMIR) , 2011
2011
Cited alongside, same era.
Y. Bengio, A. Courville, and P. Vincent, “Representation learning:a review and new perspective,” IEEE Transaction on Pattern Analysis and Machine Intelligennce , vol. 35, no. 8, pp. 1798–1828, 2013
2013
Later among the works it cites.
S. Dieleman and B. Schrauwen, “Multiscale approaches to music audio feature learning,” in Proceedings of the 14th International Conference on Music Information Retrieval (ISMIR) , 2013
2013
Later among the works it cites.
C.-C. M. Yeh, L. Su, and Y.-H. Yang, “Dual-layer bag-of-frames model for music genre classification,” in Proceedings of the 37th International Conference on Acoustics, Speech, and Signal Processing (ICASSP) , 2013
2013
Later among the works it cites.
A. van den Oord, S. Dieleman, and B. Schrauwen, “Deep content-based music recommendation,” in Proceedings of the 27th Conference on Neural Information Processing Systems (NIPS) , 2013
2013
Later among the works it cites.
M. D. Zeiler, M. Ranzato, R. Monga, M. Mao, K. Yang, Q. V. Le, P. Nguyen, A. Senior, V. Vanhoucke, J. Dean, and G. E. Hinton, “On rectified linear units for speech processing,” in Proceedings of the IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2013
2013
Later among the works it cites.
G. Dahl, T. N. Sainath, and G. Hinton, “Improving deep neural networks for lvcsr using rectified linear units and dropout,” in Proceedings of the IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2013
2013
Later among the works it cites.
A. van den Oord, S. Dieleman, and B. Schrauwen, “Transfer learning by supervised pre-training for audio-based music classification,” in Proceedings of the 15th International Conference on Music Information Retrieval (ISMIR) , 2014
2014
Later among the works it cites.
Y. Vaizman, B. McFee, and G. Lanckriet, “Codebook based audio feature representation for music information retrieval,” IEEE Transactions on Acoustics, Speech and Signal Processing , 2014
2014
Later among the works it cites.
L. Su, C.-C. M. Yeh, J.-Y. Liu, J.-C. Wang, and Y.-H. Yang, “A systematic evaluation of the bag-of-frames representation for music information retrieval,” IEEE Transactions on Acoustics, Speech and Signal Processing , 2014
2014
Later among the works it cites.
S. Sigtia and S. Dixon, “Improved music feature learning with deep neural networks,” in Proceedings of the 38th International Conference on Acoustics, Speech, and Signal Processing (ICASSP) , 2014
2014
Later among the works it cites.
S. Dieleman and B. Schrauwen, “End-to-end learning for music audio,” in Proceedings of the 38th International Conference on Acoustics, Speech, and Signal Processing (ICASSP) , 2014
2014
Later among the works it cites.
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov, “Dropout: A simple way to prevent neural networks from overfitting,” Journal of Machine Learning Research , 2014
2014
Later among the works it cites.
D. Yu and L. Deng, Automatic Speech Recognition-A Deep Learning Approach . Springer, 2015
2015
Closest in time.