Fetching the paper…
Reading the bibliography…
The goal of this work is to investigate the performance of popular speaker recognition models on speech segments from movies, where often actors intentionally disguise their voice to play a character.
A. R. Reich and J. E. Duke, “Effects of selected vocal disguises upon speaker identification by listening,”
1979
Earlier work this paper cites.
A. Hirson and M. Duckworth, “Glottal fry and voice disguise: a case study in forensic phonetics,”
1993
Earlier work this paper cites.
N. Dehak, P. J. Kenny, R. Dehak, P. Dumouchel, and P. Ouellet, “Front-end factor analysis for speaker verification,”
2010
Earlier work this paper cites.
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlicek, Y. Qian, P. Schwarz
2011
Earlier work this paper cites.
D. Garcia-Romero and A. McCree, “Supervised domain adaptation for i-vector based speaker recognition,” in
2014
Earlier work this paper cites.
D. Garcia-Romero, X. Zhang, A. McCree, and D. Povey, “Improving speaker recognition performance in the domain adaptation challenge using deep neural networks,” in
2014
Earlier work this paper cites.
M. McLaren, A. Lawson, L. Ferrer, D. Castan, and M. Graciarena, “The speakers in the wild speaker recognition challenge plan,”
2015
Earlier work this paper cites.
O. M. Parkhi, A. Vedaldi, and A. Zisserman, “Deep face recognition,” in
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in
2016
Earlier work this paper cites.
M. McLaren, L. Ferrer, D. Castan, and A. Lawson, “The speakers in the wild (sitw) speaker recognition database.” 2016
2016
Earlier work this paper cites.
M. Arjovsky, S. Chintala, and L. Bottou, “Wasserstein gan,”
2017
Earlier work this paper cites.
W. Cai, D. Cai, W. Liu, G. Li, and M. Li, “Countermeasures for automatic speaker verification replay spoofing attack: On data augmentation, feature representation, classification and fusion.” in
2017
Earlier work this paper cites.
Z. Chen, Z. Xie, W. Zhang, and X. Xu, “Resnet and model fusion for automatic spoofing detection.” in
2017
Earlier work this paper cites.
P. Liu, X. Qiu, and X. Huang, “Adversarial multi-task learning for text classification,”
2017
Earlier work this paper cites.
P. Matejka, O. Novotnỳ, O. Plchot, and L. Burget, “Analysis of score normalization in multilingual speaker recognition.” 2017
2017
Earlier work this paper cites.
A. Nagrani, J. S. Chung, and A. Zisserman, “VoxCeleb: a large-scale speaker identification dataset,” in
2017
Cited alongside, same era.
E. Tzeng, J. Hoffman, K. Saenko, and T. Darrell, “Adversarial discriminative domain adaptation,” in
2017
Cited alongside, same era.
Y. Zhang, R. Barzilay, and T. Jaakkola, “Aspect-augmented adversarial networks for domain adaptation,”
2017
Cited alongside, same era.
W. Cai, J. Chen, and M. Li, “Exploring the encoding layer and loss function in end-to-end speaker and language recognition system,”
2018
Cited alongside, same era.
Q. Cao, L. Shen, W. Xie, O. M. Parkhi, and A. Zisserman, “VGGFace2: A dataset for recognising faces across pose and age,” in
2018
Cited alongside, same era.
S.-W. Chung, J. S. Chung, and H.-G. Kang, “Perfect match: Improved cross-modal embeddings for audio-visual synchronisation,” in
2019
Later among the works it cites.
R. K. Das, J. Yang, and H. Li, “Long range acoustic features for spoofed speech detection,” in
2019
Later among the works it cites.
J. Rohdin, T. Stafylakis, A. Silnova, H. Zeinali, L. Burget, and O. Plchot, “Speaker verification using end-to-end adversarial language adaptation,” in
2019
Later among the works it cites.
2019
Later among the works it cites.
M. Bain, A. Nagrani, A. Brown, and A. Zisserman, “Condensed movies: Story based retrieval with contextual embeddings,” in
2020
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
J. Hoffman, E. Tzeng, T. Park, J.-Y. Zhu, P. Isola, K. Saenko, A. Efros, and T. Darrell, “Cycada: Cycle-consistent adversarial domain adaptation,” in
2018
Cited alongside, same era.
2018
Cited alongside, same era.
Z. Meng, J. Li, Y. Gong
2018
Cited alongside, same era.
A. Nagrani and A. Zisserman, “From benedict cumberbatch to sherlock holmes: Character identification in tv series without a script,”
2018
Cited alongside, same era.
D. Snyder, D. Garcia-Romero, G. Sell, D. Povey, and S. Khudanpur, “X-vectors: Robust dnn embeddings for speaker recognition,” in
2018
Cited alongside, same era.
L. Wan, Q. Wang, A. Papir, and I. L. Moreno, “Generalized end-to-end loss for speaker verification,” in
2018
Cited alongside, same era.
J. S. Chung, J. Huh, and S. Mun, “Delving into voxceleb: environment invariant speaker recognition,”
2020
Closest in time.
J. S. Chung, J. Huh, S. Mun, M. Lee, H. S. Heo, S. Choe, C. Ham, S. Jung, B.-J. Lee, and I. Han, “In defence of metric learning for speaker recognition,” in
2020
Closest in time.
Y. Fan, J. Kang, L. Li, K. Li, H. Chen, S. Cheng, P. Zhang, Z. Zhou, Y. Cai, and D. Wang, “Cn-celeb: a challenging chinese speaker recognition dataset,” in
2020
Closest in time.
D. Garcia-Romero, A. McCree, D. Snyder, and G. Sell, “Jhu-hltcoe system for the voxsrc speaker recognition challenge,” in
2020
Closest in time.
2020
Closest in time.
2020
Closest in time.
C. Luu, P. Bell, and S. Renals, “Channel adversarial training for speaker verification and diarization,” in
2020
Closest in time.
A. Nagrani, J. S. Chung, S. Albanie, and A. Zisserman, “Disentangled speech embeddings using cross-modal self-supervision,” in
2020
Closest in time.
A. Nagrani, J. S. Chung, J. Huh, A. Brown, E. Coto, W. Xie, M. McLaren, D. A. Reynolds, and A. Zisserman, “VoxSRC 2020: The second voxceleb speaker recognition challenge,” 2020
2020
Closest in time.
A. Nagrani, J. S. Chung, W. Xie, and A. Zisserman, “Voxceleb: Large-scale speaker verification in the wild,”
2020
Closest in time.