Fetching the paper…
Reading the bibliography…
Speaker Recognition and Speaker Identification are challenging tasks with essential applications such as automation, authentication, and security.
J. Garofolo, L. Lamel, W. Fisher, J. Fiscus, and D. Pallett, “Darpa timit acoustic-phonetic continous speech corpus cd-rom. nist speech disc 1-1.1,” NASA STI/Recon Technical Report N , vol. 93, p. 27403, 01 1993
1993
Earlier work this paper cites.
Y. Lecun, L. Bottou, Y. Bengio, and P. Haffner, “Gradient-based learning applied to document recognition,” in Proceedings of the IEEE , 1998, pp. 2278–2324
1998
Earlier work this paper cites.
R. H. Woo, A. Park, and T. J. Hazen, “The mit mobile device speaker verification corpus: Data collection and preliminary experiments,” in 2006 IEEE Odyssey - The Speaker and Language Recognition Workshop , June 2006, pp. 1–6
2006
Earlier work this paper cites.
S. Prince and J. H. Elder, “Probabilistic linear discriminant analysis for inferences about identity,” 2007 IEEE 11th International Conference on Computer Vision , pp. 1–8, 2007
2007
Earlier work this paper cites.
N. Dehak, P. J. Kenny, R. Dehak, P. Dumouchel, and P. Ouellet, “Front end factor analysis for speaker verification,” IEEE Transactions on Audio, Speech and Language Processing , 2010
2010
Earlier work this paper cites.
H. Beigi, Fundamentals of Speaker Recognition . Springer Publishing Company, Incorporated, 2011
2011
Earlier work this paper cites.
A. Kanagasundaram, R. Vogt, D. Dean, S. Sridharan, and M. Mason, “i-vector based speaker recognition on short utterances,” in in Interspeech 2011 , 2011, pp. 2341–2344
2011
Earlier work this paper cites.
P. Matejka, O. Glembek, F. Castaldo, M. J. Alam, O. Plchot, P. Kenny, L. Burget, and J. Cernocky, “Full-covariance ubm and heavy-tailed plda in i-vector speaker verification,” ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings , pp. 4828–4831, 05 2011
2011
Earlier work this paper cites.
S. Cumani, O. Plchot, and P. Laface, “Probabilistic linear discriminant analysis of i-vector posterior distributions.” in ICASSP . IEEE, 2013, pp. 7644–7648. [Online]. Available: http://dblp.uni-trier.de/db/conf/icassp/icassp2013.html#CumaniPL13
2013
Earlier work this paper cites.
R. Travadi, M. Van Segbroeck, and S. Narayanan, “Modified-prior i-Vector Estimation for Language Identification of Short Duration Utterances,” in Proceedings of the Fifteenth Annual Conference of the International Speech Communication Association , Singapore, Sep. 2014, pp. 3037–3041
2014
Earlier work this paper cites.
E. Variani, X. Lei, E. McDermott, I. Lopez-Moreno, and J. Gonzalez-Dominguez, “Deep neural networks for small footprint text-dependent speaker verification,” in ICASSP . IEEE, 2014, pp. 4052–4056
2014
Earlier work this paper cites.
Y. Lei, N. Scheffer, L. Ferrer, and M. McLaren, “A novel scheme for speaker recognition using a phonetically-aware deep neural network,” ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings , pp. 1695–1699, 05 2014
2014
Cited alongside, same era.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollar, and C. L. Zitnick, “Microsoft coco: Common objects in context,” in Computer Vision – ECCV 2014 , D. Fleet, T. Pajdla, B. Schiele, and T. Tuytelaars, Eds. Cham: Springer International Publishing, 2014, pp. 740–755
2014
Cited alongside, same era.
S. Ren, K. He, R. Girshick, and J. Sun, “Faster r-cnn: Towards real-time object detection with region proposal networks,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 39, 06 2015
2015
Cited alongside, same era.
S. H. Ghalehjegh and R. C. Rose, “Deep bottleneck features for i-vector based text-independent speaker verification,” in 2015 IEEE Workshop on Automatic Speech Recognition and Understanding (ASRU) , 2015, pp. 555–560
J. Redmon and A. Farhadi, “Yolo9000: Better, faster, stronger,” in 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2017, pp. 6517–6525
2017
Later among the works it cites.
A. Nagrani, J. S. Chung, and A. Zisserman, “Voxceleb: a large-scale speaker identification dataset,” in INTERSPEECH , 2017
2017
Later among the works it cites.
F. Wang, J. Cheng, W. Liu, and H. Liu, “Additive margin softmax for face verification,” IEEE Signal Processing Letters , vol. 25, no. 7, pp. 926–930, July 2018
2018
Later among the works it cites.
J.-W. Jung, H.-S. Heo, I.-H. Yang, H.-J. Shim, and H.-J. Yu, “Avoiding speaker overfitting in end-to-end dnns using raw waveform for text-independent speaker verification,” in Annual Conference of the International Speech Communication Association , 09 2018, pp. 3583–3587
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2015
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , June 2016, pp. 770–778
2016
Cited alongside, same era.
W. Liu, D. Anguelov, D. Erhan, C. Szegedy, S. Reed, C.-Y. Fu, and A. C. Berg, “Ssd: Single shot multibox detector,” Lecture Notes in Computer Science , p. 21–37, 2016
2016
Cited alongside, same era.
2017
Cited alongside, same era.
F. Wang, M. Jiang, C. Qian, S. Yang, C. Li, H. Zhang, X. Wang, and X. Tang, “Residual attention network for image classification,” 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , Jul 2017. [Online]. Available: http://dx.doi.org/10.1109/CVPR.2017.683
2017
Cited alongside, same era.
D. Snyder, D. Garcia-Romero, D. Povey, and S. Khudanpur, “Deep neural network embeddings for text-independent speaker verification.” in INTERSPEECH , F. Lacerda, Ed. ISCA, 2017, pp. 999–1003. [Online]. Available: http://dblp.uni-trier.de/db/conf/interspeech/interspeech2017.html#SnyderGPK17
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2018
Later among the works it cites.
M. Ravanelli and Y. Bengio, “Speaker recognition from raw waveform with sincnet,” in 2018 IEEE Spoken Language Technology Workshop (SLT) , Dec 2018, pp. 1021–1028
2018
Later among the works it cites.
D. Snyder, D. Garcia-Romero, G. Sell, D. Povey, and S. Khudanpur, “X-vectors: Robust dnn embeddings for speaker recognition,” 2018 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , pp. 5329–5333, 2018
2018
Later among the works it cites.
M. Sandler, A. Howard, M. Zhu, A. Zhmoginov, and L. Chen, “Mobilenetv2: Inverted residuals and linear bottlenecks,” in 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2018, pp. 4510–4520
2018
Later among the works it cites.
2018
Later among the works it cites.
J. A. Chagas Nunes, D. Macêdo, and C. Zanchettin, “Additive margin sincnet for speaker recognition,” in 2019 International Joint Conference on Neural Networks (IJCNN) , July 2019, pp. 1–5
2019
Later among the works it cites.