Fetching the paper…
Reading the bibliography…
To compliment the existing set of datasets, we present a small dataset entitled vocadito, consisting of 40 short excerpts of monophonic singing, sung in 7 different languages by singers with varying of levels of training, and recorded on a variety of devices.
I. Godt, “An essay on word painting,” in College Music Symposium , vol. 24, no. 2. JSTOR, 1984, pp. 118–129
1984
Earlier work this paper cites.
K. Kroeger, “Word painting in the music of william billings,” American Music , pp. 41–64, 1988
1988
Earlier work this paper cites.
C.-L. Hsu and J.-S. R. Jang, “On the improvement of singing voice separation for monaural recordings using the mir-1k dataset,” IEEE Transactions on Audio, Speech, and Language Processing , vol. 18, no. 2, pp. 310–319, 2009
2009
Earlier work this paper cites.
J. Mora, F. Gómez, E. Gómez, F. J. Borrego, and J. Díaz-Báñez, “Characterization and similarity in a cappella flamenco cantes.” 01 2010, pp. 351–356
2010
Earlier work this paper cites.
J. Mora, F. Gómez, E. Gómez, F. Escobar, and J. M. Díaz-Báñez, “Melodic characterization and similarity in a cappella flamenco cantes,” in Proc. International Society for Music Information Retrieval Conference (ISMIR) , 2010, pp. 351–356
2010
Earlier work this paper cites.
T. De Clercq and D. Temperley, “A corpus analysis of rock harmony,” Popular Music , vol. 30, no. 1, pp. 47–70, 2011
2011
Earlier work this paper cites.
J. B. L. Smith, J. A. Burgoyne, I. Fujinaga, D. De Roure, and J. S. Downie, “Design and creation of a large-scale database of structural annotations.” in 12th International Society for Music Information Retrieval Conference , ser. ISMIR, 2011
2011
Earlier work this paper cites.
E. Gómez and J. Bonada, “Towards computer-assisted flamenco transcription: An experimental comparison of automatic transcription algorithms as applied to a cappella singing,” vol. 37, no. 2, 2013, pp. 73–90
2013
Earlier work this paper cites.
R. M. Bittner, J. Salamon, M. Tierney, M. Mauch, C. Cannam, and J. P. Bello, “Medleydb: A multitrack dataset for annotation-intensive mir research.” in Proc. International Society for Music Information Retrieval Conference (ISMIR) , 2014, pp. 155–160
2014
Earlier work this paper cites.
E. Molina, A. M. Barbancho-Perez, L. J. Tardon-Garcia, I. Barbancho-Perez et al. , “Evaluation framework for automatic singing transcription,” in Proc. International Society for Music Information Retrieval Conference (ISMIR) , 2014, pp. 567–572
2014
Earlier work this paper cites.
M. Mauch and S. Dixon, “pyin: A fundamental frequency estimator using probabilistic threshold distributions,” in 2014 ieee international conference on acoustics, speech and signal processing (icassp) . IEEE, 2014, pp. 659–663
2014
Earlier work this paper cites.
J. Bosch and E. Gómez, “Melody extraction in symphonic classical music: a comparative study of mutual agreement between humans and algorithms,” in Proc. 9th Conference on Interdisciplinary Musicology–CIM14, Berlin, Germany , 2014
2014
Earlier work this paper cites.
E. Molina, L. J. Tardón, A. M. Barbancho, and I. Barbancho, “Sipth: Singing transcription based on hysteresis defined on the pitch-time curve,” IEEE/ACM Transactions on Audio, Speech, and Language Processing , vol. 23, no. 2, pp. 252–263, 2014
2014
Earlier work this paper cites.
J. Devaney and D. Richardson, “The influence of sung vowels on pitch perception,” The Journal of the Acoustical Society of America , vol. 137, no. 4, pp. 2405–2405, 2015
2015
Cited alongside, same era.
C.-Y. Liang, L. Su, Y.-H. Yang, and H.-M. Lin, “Musical offset detection of pitched instruments: The case of violin,” in Proc. International Society for Music Information Retrieval Conference (ISMIR) , 2015, pp. 281–287
2015
Cited alongside, same era.
T.-S. Chan, T.-C. Yeh, Z.-C. Fan, H.-W. Chen, L. Su, Y.-H. Yang, and R. Jang, “Vocal activity informed singing voice separation with the ikala dataset,” in Proc. IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP) , 2015, pp. 718–722
2015
Cited alongside, same era.
M. Mauch, C. Cannam, R. Bittner, G. Fazekas, J. Salamon, J. Dai, J. Bello, and S. Dixon, “Computer-aided melody note transcription using the Tony software: Accuracy and efficiency,” in Proc. International Conference on Technologies for Music Notation and Representation , 2015
——, “cante100 audio,” Jul. 2018. [Online]. Available: https://doi.org/10.5281/zenodo.1324183
2018
Later among the works it cites.
R. Gong, R. C. Repetto, Y. Yang, and X. Serra, “Jingju a cappella singing dataset part1,” Jul. 2018. [Online]. Available: https://doi.org/10.5281/zenodo.1323561
2018
Later among the works it cites.
J. Wilkins, P. Seetharaman, A. Wahl, and B. Pardo, “Vocalset: A singing voice dataset.” in ISMIR , 2018, pp. 468–474
2018
Later among the works it cites.
A. McLeod and M. Steedman, “Evaluating automatic polyphonic music transcription.” in ISMIR , 2018, pp. 42–49
2018
Later among the works it cites.
J. W. Kim, J. Salamon, P. Li, and J. P. Bello, “Crepe: A convolutional representation for pitch estimation,” in Proc. IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP) , 2018, pp. 161–165
2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2015
Cited alongside, same era.
R. Bittner, J. Wilkins, H. Yip, and J. P. Bello, “Medleydb 2.0: New data and a system for sustainable data collection,” ISMIR Late Breaking and Demo Papers , p. 36, 2016
2016
Cited alongside, same era.
N. Kroher and E. Gómez, “Automatic transcription of flamenco singing from polyphonic music recordings,” IEEE/ACM Transactions on Audio, Speech, and Language Processing , vol. 24, no. 5, pp. 901–913, 2016
2016
Cited alongside, same era.
M. Panteli, R. Bittner, J. P. Bello, and S. Dixon, “Towards the characterization of singing styles in world music,” in 2017 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2017, pp. 636–640
2017
Cited alongside, same era.
E. J. Humphrey, S. Reddy, P. Seetharaman, A. Kumar, R. M. Bittner, A. Demetriou, S. Gulati, A. Jansson, T. Jehan, B. Lehner et al. , “An introduction to signal processing for singing-voice analysis: High notes in the effort to automate the understanding of vocals in music,” IEEE Signal Processing Magazine , vol. 36, no. 1, pp. 82–94, 2018
2018
Cited alongside, same era.
P. Larrouy-Maestri and P. Q. Pfordresher, “Pitch perception in music: Do scoops matter?” Journal of Experimental Psychology: Human Perception and Performance , vol. 44, no. 10, p. 1523, 2018
2018
Cited alongside, same era.
P. Larrouy-Maestri, ““i know it when i hear it” on listeners’ perception of mistuning,” Music & Science , vol. 1, p. 2059204318784582, 2018
2018
Cited alongside, same era.
B. Bozkurt, A. Srinivasamurthy, S. Gulati, and X. Serra, “Saraga: research datasets of indian art music,” May 2018. [Online]. Available: https://doi.org/10.5281/zenodo.4301737
2018
Cited alongside, same era.
G. Meseguer-Brocal, A. Cohen-Hadria, and G. Peeters, “Dali: A large dataset of synchronized audio, lyrics and notes, automatically created using teacher-student machine learning paradigm,” in 19th International Society for Music Information Retrieval Conference , 2018
2018
Cited alongside, same era.
Later among the works it cites.
R. M. Bittner and J. J. Bosch, “Generalized metrics for single-f0 estimation evaluation.” in ISMIR , 2019, pp. 738–745
2019
Later among the works it cites.
——, “Creating dali, a large dataset of synchronized audio, lyrics, and notes,” Transactions of the International Society for Music Information Retrieval , vol. 3, no. 1, 2020
2020
Later among the works it cites.
S. Rosenzweig, H. Cuesta, C. Weiß, F. Scherbaum, E. Gómez, and M. Müller, “Dagstuhl choirset: A multitrack dataset for mir research on choral singing,” Transactions of the International Society for Music Information Retrieval , vol. 3, no. 1, 2020
2020
Later among the works it cites.
G. Meseguer-Brocal and G. Peeters, “Content based singing voice source separation via strong conditioning using aligned phonemes,” in 21st International Society for Music Information Retrieval Conference , 2020
2020
Later among the works it cites.
J.-Y. Hsu and L. Su, “Vocano: transcribing singing vocal notes in polyphonic music using source separation and semi-supervised learning,” Under review , 2020
2020
Later among the works it cites.
M. Fuentes, R. Bittner, M. Miron, G. Plaja, P. Ramoneda, V. Lostanlen, D. Rubinstein, A. Jansson, T. Kell, K. Choi, and et al., “mirdata v.0.3.0,” Jan 2021
2021
Closest in time.
S. Rosenzweig, F. Scherbaum, and M. Müller, “Reliability assessment of singing voice f0-estimates using multiple algorithms,” in ICASSP 2021-2021 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2021, pp. 261–265
2021
Closest in time.