Fetching the paper…
Reading the bibliography…
Singing techniques are used for expressive vocal performances by employing temporal fluctuations of the timbre, the pitch, and other components of the voice.
Y. Sun, A. K. Wong, and M. S. Kamel, “Classification of imbalanced data: A review,” International journal of pattern recognition and artificial intelligence , vol. 23, no. 04, pp. 687–719, 2009
2009
Earlier work this paper cites.
M. Lin, Q. Chen, and S. Yan, “Network in network,” in International Conference on Learning Representations , 2014. [Online]. Available: https://openreview.net/forum?id=ylE6yojDR5yqX
2014
Earlier work this paper cites.
J. Pons, T. Lidy, and X. Serra, “Experimenting with musically motivated convolutional neural networks,” in 2016 14th International Workshop on Content-Based Multimedia Indexing (CBMI) , 2016, pp. 1–6
2016
Earlier work this paper cites.
C. Huang, Y. Li, C. C. Loy, and X. Tang, “Learning deep representation for imbalanced classification,” in Proceedings of the IEEE conference on computer vision and pattern recognition (CVPR) , 2016, pp. 5375–5384
2016
Earlier work this paper cites.
J. Dai, H. Qi, Y. Xiong, Y. Li, G. Zhang, H. Hu, and Y. Wei, “Deformable convolutional networks,” in Proceedings of the IEEE international conference on computer vision (ICCV) , 2017, pp. 764–773
2017
Earlier work this paper cites.
E. J. Humphrey, S. Reddy, P. Seetharaman, A. Kumar, R. M. Bittner, A. Demetriou, S. Gulati, A. Jansson, T. Jehan, B. Lehner et al. , “An introduction to signal processing for singing-voice analysis: High notes in the effort to automate the understanding of vocals in music,” IEEE Signal Processing Magazine , vol. 36, no. 1, pp. 82–94, 2018
2018
Earlier work this paper cites.
J. Wilkins, P. Seetharaman, A. Wahl, and B. A. Pardo, “Vocalset: A singing voice dataset,” in 19th International Society for Music Information Retrieval Conference, (ISMIR) . International Society for Music Information Retrieval, 2018, pp. 468–474
2018
Earlier work this paper cites.
T. Takahashi, S. Fukayama, and M. Goto, “Instrudive: A music visualization system based on automatically recognized instrumentation.” in 19th International Society for Music Information Retrieval Conference, (ISMIR) , 2018, pp. 561–568
2018
Earlier work this paper cites.
P. Lei and S. Todorovic, “Temporal deformable residual networks for action segmentation in videos,” in Proceedings of the IEEE conference on computer vision and pattern recognition (CVPR) , 2018, pp. 6742–6751
2018
Cited alongside, same era.
D. Mahajan, R. Girshick, V. Ramanathan, K. He, M. Paluri, Y. Li, A. Bharambe, and L. Van Der Maaten, “Exploring the limits of weakly supervised pretraining,” in Proceedings of the European conference on computer vision (ECCV) , 2018, pp. 181–196
2018
Cited alongside, same era.
J. Abeßer and M. Müller, “Fundamental frequency contour classification: A comparison between hand-crafted and cnn-based features,” in Proceedings of the IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP) . IEEE, 2019, pp. 486–490
2019
Cited alongside, same era.
X. Zhu, H. Hu, S. Lin, and J. Dai, “Deformable convnets v2: More deformable, better results,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2019, pp. 9308–9316
K. Papadimitriou and G. Potamianos, “Multimodal sign language recognition via temporal deformable convolutional sequence learning.” in Proceedings of the Interspeech , 2020, pp. 2752–2756
2020
Later among the works it cites.
Y. Zhang, H. Yu, and Z. Ma, “Speaker verification system based on deformable cnn and time-frequency attention,” in 2020 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC) . IEEE, 2020, pp. 1689–1692
2020
Later among the works it cites.
Y. Yamamoto, J. Nam, H. Terasawa, and Y. Hiraga, “Investigating time-frequency representations for audio feature extraction in singing technique classification,” in 2021 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC) . IEEE, 2021, pp. 890–896
2021
Later among the works it cites.
B. O’Connor, S. Dixon, and G. Fazekas, “Zero-shot singing technique conversion,” in Proceedings of 15th In-ternational Symposium on Computer Music Multidisciplinary Research (CMMR) , 2021, pp. 235–244
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
J. Chen, Y. Pan, Y. Li, T. Yao, H. Chao, and T. Mei, “Temporal deformable convolutional encoder-decoder networks for video captioning,” in Proceedings of the AAAI conference on artificial intelligence , vol. 33, no. 01, 2019, pp. 8167–8174
2019
Cited alongside, same era.
Y. Cui, M. Jia, T.-Y. Lin, Y. Song, and S. Belongie, “Class-balanced loss based on effective number of samples,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition (CVPR) , 2019, pp. 9268–9277
2019
Cited alongside, same era.
B. Kang, S. Xie, M. Rohrbach, Z. Yan, A. Gordo, J. Feng, and Y. Kalantidis, “Decoupling representation and classifier for long-tailed recognition,” in International Conference on Learning Representations , 2020. [Online]. Available: https://openreview.net/forum?id=r1gRTCVFvB
2020
Cited alongside, same era.
2021
Later among the works it cites.
L. Xue, N. Constant, A. Roberts, M. Kale, R. Al-Rfou, A. Siddhant, A. Barua, and C. Raffel, “mt5: A massively multilingual pre-trained text-to-text transformer,” in Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies , 2021, pp. 483–498
2021
Later among the works it cites.
Y. Yamamoto, D. Moriyama, J. Nam, and H. Terasawa, “Towards computational analysis of singing technique for music information retrieval : A progress report of building dataset and statistical analysis,” Proceedings of the auditory research meeting , vol. 51, no. 8, pp. 569–572, 2021. [Online]. Available: https://ci.nii.ac.jp/naid/40022767415/
2021
Later among the works it cites.
K. An, Y. Zhang, and Z. Ou, “Deformable TDNN with Adaptive Receptive Fields for Speech Recognition,” in Proceedings of Interspeech , 2021, pp. 2067–2071
2071
Closest in time.