Fetching the paper…
Reading the bibliography…
The third instalment of the VoxCeleb Speaker Recognition Challenge was held in conjunction with Interspeech 2021.
J. Carletta, “Unleashing the killer corpus: experiences in creating the multi-everything ami meeting corpus,” Language Resources and Evaluation , vol. 41, no. 2, pp. 181–190, 2007
2007
Earlier work this paper cites.
S.-C. Yin, R. Rose, and P. Kenny, “Adaptive score normalization for progressive model adaptation in text independent speaker verification,” in Proc. ICASSP . IEEE, 2008, pp. 4857–4860
2008
Earlier work this paper cites.
N. Dehak, R. Dehak, P. Kenny, N. Brümmer, P. Ouellet, and P. Dumouchel, “Support vector machines versus fast scoring in the low-dimensional total variability space for speaker verification,” in Proc. Interspeech , 2009
2009
Earlier work this paper cites.
J. S. Chung and A. Zisserman, “Out of time: automated lip sync in the wild,” in Workshop on Multi-view Lip-reading, ACCV , 2016
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proc. CVPR , 2016
2016
Earlier work this paper cites.
A. Nagrani, J. S. Chung, and A. Zisserman, “VoxCeleb: a large-scale speaker identification dataset,” in Proc. Interspeech , 2017
2017
Earlier work this paper cites.
P. Matejka, O. Novotnỳ, O. Plchot, L. Burget, M. D. Sánchez, and J. Cernockỳ, “Analysis of score normalization in multilingual speaker recognition.” in Proc. Interspeech , 2017, pp. 1567–1571
2017
Earlier work this paper cites.
S. O. Sadjadi, T. Kheyrkhah, A. Tong, C. S. Greenberg, D. A. Reynolds, E. Singer, L. P. Mason, and J. Hernandez-Cordero, “The 2016 nist speaker recognition evaluation.” in Proc. Interspeech , 2017, pp. 1353–1357
2017
Earlier work this paper cites.
J. S. Chung, A. Nagrani, and A. Zisserman, “Voxceleb2: Deep speaker recognition,” in Proc. Interspeech , 2018
2018
Earlier work this paper cites.
T. Afouras, J. S. Chung, and A. Zisserman, “The conversation: Deep audio-visual speech enhancement,” in Proc. Interspeech , 2018
2018
Earlier work this paper cites.
NIST 2018 Speaker Recognition Evaluation Plan , 2018 (accessed 31 July 2020), https://www.nist.gov/system/files/documents/2018/08/17/sre18_eval_plan_2018-05-31_v6.pdf , See Section 3.1
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
Q. Wang, C. Downey, L. Wan, P. A. Mansfield, and I. L. Moreno, “Speaker diarization with lstm,” in Proc. ICASSP . IEEE, 2018, pp. 5239–5243
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
H. Wang, Y. Wang, Z. Zhou, X. Ji, D. Gong, J. Zhou, Z. Li, and W. Liu, “Cosface: Large margin cosine loss for deep face recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2018, pp. 5265–5274
2018
Earlier work this paper cites.
N. Ryant, K. Church, C. Cieri, A. Cristia, J. Du, S. Ganapathy, and M. Liberman, “First dihard challenge evaluation plan,” 2018, tech. Rep. , 2018
2018
Earlier work this paper cites.
2019
Earlier work this paper cites.
A. Dutta and A. Zisserman, “The VIA annotation software for images, audio and video,” in Proceedings of the 27th ACM International Conference on Multimedia , ser. MM ’19. New York, NY, USA: ACM, 2019. [Online]. Available: https://doi.org/10.1145/3343031.3350535
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
H. Yamamoto, K. A. Lee, K. Okabe, and T. Koshinaka, “Speaker augmentation and bandwidth extension for deep speaker embedding.” in Proc. Interspeech , 2019, pp. 406–410
2019
Earlier work this paper cites.
J. Deng, J. Guo, N. Xue, and S. Zafeiriou, “Arcface: Additive angular margin loss for deep face recognition,” in Proc. CVPR , 2019
2019
Earlier work this paper cites.
S. Gao, M.-M. Cheng, K. Zhao, X.-Y. Zhang, M.-H. Yang, and P. H. Torr, “Res2net: A new multi-scale backbone architecture,” IEEE transactions on pattern analysis and machine intelligence , 2019
2019
Cited alongside, same era.
2019
Cited alongside, same era.
2019
Cited alongside, same era.
2019
Cited alongside, same era.
S. O. Sadjadi, C. Greenberg, E. Singer, D. Reynolds, L. Mason, and J. Hernandez-Cordero, “The 2019 nist speaker recognition evaluation cts challenge,” Proc. Speaker Odyssey , 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2020
Cited alongside, same era.
J. S. Chung, J. Huh, A. Nagrani, T. Afouras, and A. Zisserman, “Spot the conversation: speaker diarisation in the wild,” in Proc. Interspeech , 2020
2020
Cited alongside, same era.
A. Nagrani, J. S. Chung, W. Xie, and A. Zisserman, “Voxceleb: Large-scale speaker verification in the wild,” Computer Speech & Language , vol. 60, p. 101027, 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
J. S. Chung, J. Huh, S. Mun, M. Lee, H. S. Heo, S. Choe, C. Ham, S. Jung, B.-J. Lee, and I. Han, “In defence of metric learning for speaker recognition,” Proc. Interspeech , 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2020
Cited alongside, same era.
J. Valk and T. Alumäe, “VoxLingua107: a dataset for spoken language recognition,” in Proc. IEEE SLT Workshop , 2021
2021
Later among the works it cites.
J. Thienpondt, B. Desplanques, and K. Demuynck, “The idlab voxceleb speaker recognition challenge 2021 system description,” 2021
2021
Later among the works it cites.
L. Zhang, H. Zhao, Q. Meng, Y. Chen, M. Liu, and L. Xie, “Beijing zkj-npu speaker verification system for voxceleb speaker recognition challenge 2021,” 2021
2021
Later among the works it cites.
M. Zhao, Y. Ma, M. Liu, and M. Xu, “The speakin system for voxceleb speaker recognition challange 2021,” 2021
2021
Later among the works it cites.
J. Cho, J. Villalba, and N. Dehak, “The jhu submission to voxsrc-21: Track 3,” 2021
2021
Later among the works it cites.
J. Slavíček, A. Swart, M. Klčo, and N. Brümmer, “The phonexia voxceleb speaker recognition challenge 2021 system description,” 2021
2021
Later among the works it cites.
D. Cai and M. Li, “The dku-dukeece system for the self-supervision speaker verification task of the 2021 voxceleb speaker recognition challenge,” 2021
2021
Later among the works it cites.
N. Zheng, N. Li, Y. Zhao, C. Weng, and D. Su, “Tencent speaker diarization system for the voxceleb speaker recognition challenge 2021,” https://www.robots.ox.ac.uk/~vgg/data/voxceleb/data_workshop_2021/reports/Tencent_diarization.pdf , 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
“Webrtc voice activity detector,” 2021 (accessed 31 May 2021), https://github.com/wiseman/py-webrtcvad
2021
Later among the works it cites.
X. Ding, X. Zhang, N. Ma, J. Han, G. Ding, and J. Sun, “Repvgg: Making vgg-style convnets great again,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2021, pp. 13 733–13 742
2021
Later among the works it cites.
2021
Later among the works it cites.
D. Cai, W. Wang, and M. Li, “An iterative framework for self-supervised deep speaker representation learning,” in Proc. ICASSP . IEEE, 2021, pp. 6728–6732
2021
Later among the works it cites.
B. J. Borgström, “Unsupervised bayesian adaptation of plda for speaker verification,” in Proc. Interspeech , 2021, pp. 1039–1043
2021
Later among the works it cites.
D. Raj, L. P. Garcia-Perera, Z. Huang, S. Watanabe, D. Povey, A. Stolcke, and S. Khudanpur, “Dover-lap: A method for combining overlap-aware diarization outputs,” in IEEE Spoken Language Technology Workshop . IEEE, 2021, pp. 881–888
2021
Later among the works it cites.