Fetching the paper…
Reading the bibliography…
We present and release Omnizart, a new Python library that provides a streamlined solution to automatic music transcription (AMT).
O. Gillet and G. Richard, “Enst-drums: an extensive audio-visual database for drum signals processing.” in Proc. ISMIR , 2006, pp. 156–159
2006
Earlier work this paper cites.
C.-L. Hsu and J.-S. R. Jang, “On the improvement of singing voice separation for monaural recordings using the mir-1k dataset,” IEEE TASLP , vol. 18, no. 2, pp. 310–319, 2009
2009
Earlier work this paper cites.
J. Mora, F. Gómez, E. Gómez, F. Escobar-Borrego, and J. M. Díaz-Báñez, “Characterization and melodic similarity of a cappella flamenco cantes,” in Proc. ISMIR , 2010, pp. 9–13
2010
Earlier work this paper cites.
M. Mauch and S. Dixon, “Approximate note transcription for the improved identification of difficult chords,” in Proc. ISMIR , 2010, pp. 135–140
2010
Earlier work this paper cites.
J. A. Burgoyne, J. Wild, and I. Fujinaga, “An expert ground truth set for audio chord recognition and music analysis,” in Proc. ISMIR , 2011, pp. 633–638
2011
Earlier work this paper cites.
E. Molina, A. M. Barbancho-Perez, L. J. Tardón, I. Barbancho-Perez et al. , “Evaluation framework for automatic singing transcription,” in Proc. ISMIR , 2014
2014
Earlier work this paper cites.
L. Su and Y. Yang, “Combining Spectral and Temporal Representations for Multipitch Estimation of Polyphonic Music,” IEEE/ACM TASLP , vol. 23, no. 10, pp. 1600–1612, 2015
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
M. Abadi, P. Barham, J. Chen, Z. Chen, A. Davis, J. Dean, M. Devin, S. Ghemawat, G. Irving, M. Isard et al. , “Tensorflow: A system for large-scale machine learning,” in 12th USENIX symposium on operating systems design and implementation (OSDI 16) , 2016, pp. 265–283
2016
Earlier work this paper cites.
R. Kelz, M. Dorfer, F. Korzeniowski, S. Böck, A. Arzt, and G. Widmer, “On the Potential of Simple Framewise Approaches to Piano Transcription,” in Proc. ISMIR , 2016, pp. 475–481
2016
Cited alongside, same era.
S. Böck, F. Korzeniowski, J. Schlüter, F. Krebs, and G. Widmer, “Madmom: A new python audio and music signal processing library,” in Proc. ACM MM , 2016, pp. 1174–1178
2016
Cited alongside, same era.
J. Thickstun, Z. Harchaoui, and S. M. Kakade, “Learning features of music from scratch,” in Proc. ICLR , 2017
2017
Cited alongside, same era.
C. Southall, C.-W. Wu, A. Lerch, and J. Hockman, “Mdb drums: An annotated subset of medleydb for automatic drum transcription,” 2017
2017
Cited alongside, same era.
Y.-T. Wu, B. Chen, and L. Su, “Automatic Music Transcription Leveraging Generalized Cepstral Features and Deep Learning,” in Proc. ICASSP , 2018, pp. 401–405
L. Su, “Vocal melody extraction using patch-based CNN,” in Proc. ICASSP , 2018, pp. 371–375
2018
Later among the works it cites.
T. Miyato, S. Maeda, M. Koyama, and S. Ishii, “Virtual adversarial training: a regularization method for supervised and semi-supervised learning,” IEEE TPAMI , vol. 41, no. 8, pp. 1979–1993, 2018
2018
Later among the works it cites.
C. Hawthorne, A. Stasyuk, A. Roberts, I. Simon, C. A. Huang, S. Dieleman, E. Elsen, J. H. Engel, and D. Eck, “Enabling Factorized Piano Music Modeling and Generation with the MAESTRO Dataset,” in Proc. ICLR , 2019
2019
Later among the works it cites.
Y. Yamada, M. Iwamura, T. Akiba, and K. Kise, “Shakedrop regularization for deep residual learning,” IEEE Access , vol. 7, pp. 186 126–186 136, 2019
2019
Later among the works it cites.
T.-P. Chen and L. Su, “Harmony transformer: Incorporating chord segmentation into harmony recognition,” in Proc. ISMIR , 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
2018
Cited alongside, same era.
N. Parmar, A. Vaswani, J. Uszkoreit, L. Kaiser, N. Shazeer, A. Ku, and D. Tran, “Image Transformer,” in Proc. ICML , 2018, pp. 4052–4061
2018
Cited alongside, same era.
J. Thickstun, Z. Harchaoui, D. P. Foster, and S. M. Kakade, “Invariances and Data Augmentation for Supervised Music Transcription,” in Proc. ICASSP , 2018, pp. 2241–2245
2018
Cited alongside, same era.
2019
Later among the works it cites.
Y.-T. Wu, B. Chen, and L. Su, “Multi-instrument automatic music transcription with self-attention-based instance segmentation,” IEEE/ACM TASLP , vol. 28, pp. 2796–2809, 2020
2020
Later among the works it cites.
Y.-C. Chuang and L. Su, “Beat and downbeat tracking of symbolic music data using deep recurrent neural networks,” in Proc. APSIPA ASC . IEEE, 2020, pp. 346–352
2020
Later among the works it cites.
I.-C. Wei, C.-W. Wu, and L. Su, “Improving automatic drum transcription using large-scale audio-to-midi aligned data,” in Proc. ICASSP , 2021
2021
Closest in time.