Fetching the paper…
Reading the bibliography…
In this paper, we present the system submission for the VoxCeleb Speaker Recognition Challenge 2020 (VoxSRC-20) by the DKU-DukeECE team.
A. Janin, D. Baron, J. Edwards, D. Ellis, D. Gelbart, N. Morgan, B. Peskin, T. Pfau, E. Shriberg, A. Stolcke
2003
Earlier work this paper cites.
J. Carletta, S. Ashby, S. Bourban, M. Flynn, M. Guillemot, T. Hain, J. Kadlec, V. Karaiskos, W. Kraaij, M. Kronenthal
2005
Earlier work this paper cites.
U. Von Luxburg, “A tutorial on spectral clustering,”
2007
Earlier work this paper cites.
2013
Earlier work this paper cites.
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov, “Dropout: A Simple Way to Prevent Neural Networks from Overfitting,”
2014
Earlier work this paper cites.
D. Snyder, G. Chen, and D. Povey, “MUSAN: A Music, Speech, and Noise Corpus,”
2015
Earlier work this paper cites.
T. K., V. P., D. P., and S. K., “Audio Augmentation for Speech Recognition,” in
2015
Earlier work this paper cites.
D. P. Kingma and J. Ba, “Adam: A Method for Stochastic Optimization,” in
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep Residual Learning for Image Recognition,” in
2016
Earlier work this paper cites.
A. Nagrani, J. S. Chung, and A. Zisserman, “Voxceleb: A Large-Scale Speaker Identification Dataset,” in
2017
Earlier work this paper cites.
A. Nagrani, J. Chung, and A. Zisserman, “Voxceleb: a large-scale speaker identification dataset,” 2017
2017
Cited alongside, same era.
J. S. Chung, A. Nagrani, and A. Zisserman, “Voxceleb2: Deep Speaker Recognition,” in
2018
Cited alongside, same era.
W. Cai, J. Chen, and M. Li, “Exploring the Encoding Layer and Loss Function in End-to-End Speaker and Language Recognition System,” in
2018
Cited alongside, same era.
J. Park, S. Woo, J. Lee, and I. Kweon, “Bam: Bottleneck attention module,” in
2018
Cited alongside, same era.
K. Okabe, T. Koshinaka, and K. Shinoda, “Attentive statistics pooling for deep speaker embedding,” in
2018
Cited alongside, same era.
Q. Wang, C. Downey, L. Wan, P. A. Mansfield, and I. L. Moreno, “Speaker diarization with lstm,” in
X. Qin, M. Li, H. Bu, R. K. Das, W. Rao, S. Narayanan, and H. Li, “The ffsvc 2020 evaluation plan,”
2020
Closest in time.
2020
Closest in time.
2020
Closest in time.
2020
Closest in time.
D. Cai, W. Cai, and M. Li, “Within-Sample Variability-Invariant Loss for Robust Speaker Recognition Under Noisy Environments,” in
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
M. Diez, L. Burget, and P. Matejka, “Speaker diarization based on bayesian hmm with eigenvoice priors.” in
2018
Cited alongside, same era.
H. Y., K. A. L., K. O., and T. K., “Speaker Augmentation and Bandwidth Extension for Deep Speaker Embedding,” in
2019
Cited alongside, same era.
J. Deng, J. Guo, N. Xue, and S. Zafeiriou, “Arcface: Additive angular margin loss for deep face recognition,” in
2019
Cited alongside, same era.
W. Cai, J. Chen, J. Zhang, and M. Li, “On-the-Fly Data Loader and Utterance-Level Aggregation for Speaker and Language Recognition,”
2020
Cited alongside, same era.
2020
Closest in time.
J. Chung, J. Huh, A. Nagrani, T. Afouras, and A. Zisserman, “Spot the conversation: speaker diarisation in the wild,”
2020
Closest in time.
Q. Lin, W. Cai, L. Yang, J. Wang, J. Zhang, and M. Li, “Dihard ii is still hard: Experimental results and discussions from the dku-lenovo team,” in
2020
Closest in time.
Q. Lin and M. L. Tingle Li, “The dku speech activity detection and speaker identification systems for fearless steps challenge phase-02,” in
2020
Closest in time.
Q. Lin, Y. Hou, and M. Li, “Self-attentive similarity measurement strategies in speaker diarization,” 2020
2020
Closest in time.