Fetching the paper…
Reading the bibliography…
In this technical report we describe the IDLAB top-scoring submissions for the VoxCeleb Speaker Recognition Challenge 2020 (VoxSRC-20) in the supervised and unsupervised speaker verification tracks.
J. H. Ward, “Hierarchical grouping to optimize an objective function,”
1963
Earlier work this paper cites.
K. Demuynck, J. Roelens, D. V. Compernolle, and P. Wambacq, “SPRAAK: an open source ”SPeech recognition and automatic annotation kit”,” in
2008
Earlier work this paper cites.
D. Sculley, “Web-scale k-means clustering,” in
2010
Earlier work this paper cites.
S. Cumani, P. Batzu, D. Colibro, C. Vair, P. Laface, and V. Vasilakakis, “Comparison of speaker recognition approaches for real applications.” in
2011
Earlier work this paper cites.
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V. Dubourg, J. Vanderplas, A. Passos, D. Cournapeau, M. Brucher, M. Perrot, and E. Duchesnay, “Scikit-learn: Machine learning in Python,”
2011
Earlier work this paper cites.
D. Müllner, “fastcluster: Fast hierarchical, agglomerative clustering routines for R and Python,”
2013
Earlier work this paper cites.
D. Kingma and J. Ba, “Adam: A method for stochastic optimization,”
2014
Earlier work this paper cites.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: An asr corpus based on public domain audio books,” in
2015
Earlier work this paper cites.
D. Snyder, G. Chen, and D. Povey, “MUSAN: A music, speech, and noise corpus,” 2015
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
A. Nagrani, J. S. Chung, and A. Zisserman, “VoxCeleb: A large-scale speaker identification dataset,” in
2017
Cited alongside, same era.
T. Ko, V. Peddinti, D. Povey, M. L. Seltzer, and S. Khudanpur, “A study on data augmentation of reverberant speech for robust speech recognition,” in
2017
Cited alongside, same era.
L. N. Smith, “Cyclical learning rates for training neural networks,” in
2017
Cited alongside, same era.
J. Hu, L. Shen, and G. Sun, “Squeeze-and-Excitation networks,” in
2018
Cited alongside, same era.
J. S. Chung, A. Nagrani, and A. Zisserman, “VoxCeleb2: Deep speaker recognition,” in
2018
Cited alongside, same era.
S. Gao, M.-M. Cheng, K. Zhao, X. Zhang, M.-H. Yang, and P. H. S. Torr, “Res2Net: A new multi-scale backbone architecture,”
2020
Closest in time.
D. Garcia-Romero, G. Sell, and A. McCree, “Magneto: X-vector magnitude estimation network plus offset for improved speaker recognition,” in
2020
Closest in time.
Y. Zhao, T. Zhou, Z. Chen, and J. Wu, “Improving deep CNN networks with long temporal context for text-independent speaker verification,” in
2020
Closest in time.
J. Deng, J. Guo, T. Liu, M. Gong, and S. Zafeiriou, “Sub-center arcface: Boosting face recognition by large-scale noisy web faces,” in
2020
Closest in time.
J. Thienpondt, B. Desplanques, and K. Demuynck, “Cross-lingual speaker verification with domain-balanced hard prototype mining and language-dependent score normalization,” in
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
H. Zeinali, J. Černocký, and L. Burget, “A multi purpose and large scale speech corpus in Persian and English for speaker and speech recognition: the DeepMine database,” in
2019
Cited alongside, same era.
D. S. Park, W. Chan, Y. Zhang, C.-C. Chiu, B. Zoph, E. D. Cubuk, and Q. V. Le, “SpecAugment: A simple data augmentation method for automatic speech recognition,” in
2019
Cited alongside, same era.
B. Desplanques, J. Thienpondt, and K. Demuynck, “ECAPA-TDNN: Emphasized channel attention, propagation and aggregation in TDNN based speaker verification,” in
2020
Cited alongside, same era.
2020
Closest in time.
2020
Closest in time.
2020
Closest in time.
2020
Closest in time.
K. He, H. Fan, Y. Wu, S. Xie, and R. Girshick, “Momentum contrast for unsupervised visual representation learning,” in
2020
Closest in time.