Fetching the paper…
Reading the bibliography…
Vocal pitch is an important high-level feature in music audio processing.
A. De Cheveigné and H. Kawahara, “Yin, a fundamental frequency estimator for speech and music,” in The Journal of the Acoustical Society of America , 2002, pp. 1917–1930
1930
Earlier work this paper cites.
A. Camacho and J. G. Harris, “A sawtooth waveform inspired pitch estimator for speech and music,” in The Journal of the Acoustical Society of America , 2008, pp. 1638–1652
2008
Earlier work this paper cites.
C.-L. Hsu and J.-S. R. Jang, “On the improvement of singing voice separation for monaural recordings using the mir-1k dataset,” IEEE Transactions on Audio, Speech, and Language Processing (TASLP) , pp. 310–319, 2010
2010
Earlier work this paper cites.
J. Carroll, S. Tiaden, and F.-G. Zeng, “Fundamental frequency is critical to speech perception in noise in combined acoustic and electric hearing,” The Journal of the Acoustical Society of America , pp. 2054–2062, 2011
2011
Earlier work this paper cites.
M. Mauch and S. Ewert, “The audio degradation toolbox and its application to robustness evaluation,” in Proceeding of International Society for Music Information Retrieval (ISMIR) , 2013
2013
Earlier work this paper cites.
M. Mauch and S. Dixon, “pyin: A fundamental frequency estimator using probabilistic threshold distributions,” in Proceeding of IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2014, pp. 659–663
2014
Earlier work this paper cites.
M. Dong, J. Wu, and J. Luan, “Vocal pitch extraction in polyphonic music using convolutional residual network.” in Proceeding of INTERSPEECH , 2019, pp. 2010–2014
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
C. Raffel, B. McFee, E. J. Humphrey, J. Salamon, O. Nieto, D. Liang, D. P. Ellis, and C. C. Raffel, “mir_eval: A transparent implementation of common mir metrics,” in Proceeding of International Society for Music Information Retrieval (ISMIR) , 2014
2014
Cited alongside, same era.
2014
Cited alongside, same era.
S. Ioffe and C. Szegedy, “Batch normalization: Accelerating deep network training by reducing internal covariate shift,” in International Conference on International Conference on Machine Learning (ICML) , 2015, p. 448–456
2015
Cited alongside, same era.
B. McFee, C. Raffel, D. Liang, D. P. Ellis, M. McVicar, E. Battenberg, and O. Nieto, “librosa: Audio and music signal analysis in python,” in Proceedings of the 14th python in science conference , 2015, pp. 18–25
2015
Cited alongside, same era.
2019
Later among the works it cites.
M. K. Reddy and K. S. Rao, “Excitation modelling using epoch features for statistical parametric speech synthesis,” Computer Speech & Language , 2020
2020
Later among the works it cites.
R. Hennequin, A. Khlif, F. Voituret, and M. Moussallam, “Spleeter: a fast and efficient music source separation tool with pre-trained models,” Journal of Open Source Software , p. 2154, 2020
2020
Later among the works it cites.
S. Singh, R. Wang, and Y. Qiu, “Deepf0: End-to-end fundamental frequency estimation for music and speech signals,” in Proceeding of IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2021, pp. 61–65
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Salamon, R. M. Bittner, J. Bonada, J. J. Bosch, E. Gómez Gutiérrez, and J. P. Bello, “An analysis/synthesis framework for automatic f0 annotation of multitrack datasets,” in Proceeding of International Society for Music Information Retrieval (ISMIR) , 2017
2017
Cited alongside, same era.
Y. V. S. Murthy and S. G. Koolagudi, “Content-based music information retrieval (cb-mir) and its applications toward the music industry: A review,” ACM Computing Surveys (CSUR) , 2018
2018
Cited alongside, same era.
J. W. Kim, J. Salamon, P. Li, and J. P. Bello, “Crepe: A convolutional representation for pitch estimation,” in Proceeding of IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2018
2018
Cited alongside, same era.
F.-R. Stöter, S. Uhlich, A. Liutkus, and Y. Mitsufuji, “Open-unmix-a reference implementation for music source separation,” Journal of Open Source Software , p. 1667, 2019
2019
Cited alongside, same era.
J.-Y. Wang and J.-S. R. Jang, “On the preparation and validation of a large-scale dataset of singing transcription,” in Proceeding of IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2021, pp. 276–280
2021
Later among the works it cites.
W. Wei, P. Li, Y. Yu, and W. Li, “Harmof0: Logarithmic scale dilated convolution for pitch estimation,” in IEEE International Conference on Multimedia and Expo (ICME) , 2022, pp. 1–6
2022
Later among the works it cites.
S. Kum, J. Lee, K. L. Kim, T. Kim, and J. Nam, “Pseudo-label transfer from frame-level to note-level in a teacher-student framework for singing transcription from polyphonic music,” in Proceeding of IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2022, pp. 796–800
2022
Later among the works it cites.
X. Sun, X. Liang, Q. He, B. Zhu, and Z. Ma, “Gio: A timbre-informed approach for pitch tracking in highly noisy environments,” in International Conference on Multimedia Retrieval (ICMR) , 2022, pp. 480–488
2022
Later among the works it cites.