Fetching the paper…
Reading the bibliography…
This paper introduces a new large-scale music dataset, MusicNet, to serve as a source of supervision and evaluation of machine learning methods for music research.
Automatic segmentation of acoustic musical signals using hidden markov models
C. Raphael · 1999
Earlier work this paper cites.
Alignment of monophonic and polyphonic music to a score
N. Orio and D. Schwarz · 2001
Earlier work this paper cites.
RWC music database: Music genre database and musical instrument sound database
M. Goto, H. Hashiguchi, T. Nishimura, and R. Oka · 2003
Earlier work this paper cites.
Polyphonic audio matching and alignment for music retrieval
N. Hu, R. B. Dannenberg, and G. Tzanetakis · 2003
Earlier work this paper cites.
Improving polyphonic and poly-instrumental music to score alignment
F. Soulez, X. Rodet, and D. Schwarz · 2003
Earlier work this paper cites.
Ground-truth transcriptions of real music from force-aligned midi syntheses
R. J. Turetsky and D. P. W. Ellis · 2003
Earlier work this paper cites.
A discriminative model for polyphonic piano transcription
G. Poliner and D. P. W. Ellis · 2007
Earlier work this paper cites.
Introduction to digital speech processing
L. Rabiner and R. Schafer · 2007
Earlier work this paper cites.
High resolution audio synchronization using chroma features
S. Ewert, M. Müller, and P. Grosche · 2009
Earlier work this paper cites.
Multipitch estimation of piano sounds using a new probabilistic spectral smoothness principle
V. Emiya, R. Badeau, and B. David · 2010
Earlier work this paper cites.
Towards Automatic Extraction of Harmony Information from Music Signals
C. Harte · 2010
Earlier work this paper cites.
Understanding features and distance functions for music sequence alignment
O. Izmirli and R. B. Dannenberg · 2010
Earlier work this paper cites.
Joint multi-pitch detection using harmonic envelope estimation for polyphonic music transcription
E. Benetos and S. Dixon · 2011
Cited alongside, same era.
Multiple fundamental frequency estimation by modeling spectral peaks and non-peak regions
Z. Duan, B. Pardo, and C. Zhang · 2011
Cited alongside, same era.
Learning multi-modal similarity
B. McFee and G. Lanckriet · 2011
Cited alongside, same era.
Moving beyond feature design: Deep architectures and automatic feature learning in music informatics
E. J. Humphrey, J. P. Bello, and Y. LeCun · 2012
Cited alongside, same era.
The million song dataset challenge
B. McFee, T. Bertin-Mahieux, D. P. W. Ellis, and G. Lanckriet · 2012
Cited alongside, same era.
Automatic music transcription: challenges and future directions
E. Benetos, S. Dixon, D. Giannoulis, H. Kirchoff, and A. Klapuri · 2013
mir_eval: A transparent implementation of common mir metrics
C. Raffel, B. McFee, E. J. Humphrey, J. Salamon, O. Nieto, D. Liang, and D. P. W. Ellis · 2014
Later among the works it cites.
An iterative multi range non-negative matrix factorization algorithm for polyphonic music transcription
A. Khlif and V. Sethu · 2015
Later among the works it cites.
librosa: Audio and music signal analysis in python
B. McFee, C. Raffel, D. Liang, D. P. W. Ellis, M. McVicar, E. Battenberg, and O. Nieto · 2015
Later among the works it cites.
Large-scale content-based matching of MIDI and audio files
C. Raffel and D. P. W. Ellis · 2015
Later among the works it cites.
Imagenet large scale visual recognition challenge
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein, A. C. Berg, and L. Fei-Fei · 2015
Later among the works it cites.
Deepbach: a steerable model for bach chorales generation
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Learning optimal features for polyphonic audio-to-score alignment
C. Joder, S. Essid, and G. Richard · 2013
Cited alongside, same era.
Deep content-based music recommendation
A. van den Oord, S. Dieleman, and B. Schrauwen · 2013
Cited alongside, same era.
Unsupervised transcription of piano music
T. Berg-Kirkpatrick, J. Andreas, and D. Klein · 2014
Cited alongside, same era.
End-to-end learning for music audio
S. Dieleman and B. Schrauwen · 2014
Cited alongside, same era.
Metric learning for temporal sequence alignment
D. Garreau, R. Lajugie, S. Arlot, and F. Bach · 2014
Cited alongside, same era.
TensorFlow: Large-scale machine learning on heterogeneous systems
M. Abadi, A. Agarwal, P. Barham, E. Brevdo, Z. Chen, C. Citro, G. Corrado, A. Davis, J. Dean, M. Devin, S. Ghemawat, I. Goodfellow, A. Harp, G. Irving, M. Isard, Y. Jia, R. Jozefowicz, L. Kaiser, M. Kudlur, J. Levenberg, D. Mane, R. Monga, S. Moore, D. Murray, C. Olah, M. Schuster, J. Shlens, B. Steiner, I. Sutskever, K. Talwar, P. Tucker, V. Vanhoucke, V. Vasudevan, F. Viegas, O. Vinyals, P. Warden, M. Wattenberg, M. Wicke, Y. Yu, and X. Zheng
Cited in the paper.
Gaëtan Hadjeres and François Pachet · 2016
Closest in time.
On the potential of simple framewise approaches to piano transcription
R. Kelz, M. Dorfer, F. Korzeniowski, S. Böck, A. Arzt, and G. Widmer · 2016
Closest in time.
Feature learning for chord recognition: the deep chroma extractor
F. Korzeniowsk and G. Widmer · 2016
Closest in time.
A weakly-supervised discriminative model for audio-to-score alignment
R. Lajugie, P. Bojanowski, P. Cuvillier, S. Arlot, and F. Bach · 2016
Closest in time.
Directly modeling voiced and unvoiced components in speech waveforms by neural networks
K. Tokuda and H. Zen · 2016
Closest in time.
WaveNet: A generative model for raw audio
A. van den Oord, S. Dieleman, H. Zen, K. Simonyan, O. Vinyals, A. Graves, N. Kalchbrenner, A. Senior, and K. Kavukcuoglu · 2016
Closest in time.