Fetching the paper…
Reading the bibliography…
We investigate deep neural network performance in the textindependent speaker recognition task.
“Learning a similarity metric discriminatively, with application to face verification,”
Sumit Chopra, Raia Hadsell, and Yann LeCun, · 2005
Earlier work this paper cites.
“Front-end factor analysis for speaker verification,”
Najim Dehak, Patrick J Kenny, Réda Dehak, Pierre Dumouchel, and Pierre Ouellet, · 2011
Earlier work this paper cites.
“Cosine similarity metric learning for face verification,”
Hieu V. Nguyen and Li Bai, · 2011
Earlier work this paper cites.
“The kaldi speech recognition toolkit,”
Daniel Povey, Arnab Ghoshal, Gilles Boulianne, Lukas Burget, Ondrej Glembek, Nagendra Goel, Mirko Hannemann, Petr Motlicek, Yanmin Qian, Petr Schwarz, et al., · 2011
Earlier work this paper cites.
“Imagenet classification with deep convolutional neural networks,”
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton, · 2012
Earlier work this paper cites.
“Auto-encoding variational bayes,”
Diederik P Kingma and Max Welling, · 2013
Earlier work this paper cites.
“A novel scheme for speaker recognition using a phonetically-aware deep neural network,”
Yun Lei, Nicolas Scheffer, Luciana Ferrer, and Mitchell McLaren, · 2014
Earlier work this paper cites.
“Deep neural networks for small footprint text-dependent speaker verification,”
Ehsan Variani, Xin Lei, Erik McDermott, Ignacio Lopez Moreno, and Javier Gonzalez-Dominguez, · 2014
Earlier work this paper cites.
“Parallel training of dnns with natural gradient and parameter averaging,”
Daniel Povey, Xiaohui Zhang, and Sanjeev Khudanpur, · 2014
Earlier work this paper cites.
“Non-linear PLDA for i-vector speaker verification,”
Sergey Novoselov, Timur Pekhovsky, Oleg Kudashev, Valentin S Mendelev, and Alexey Prudnikov, · 2015
Earlier work this paper cites.
“Going deeper with convolutions,”
Christian Szegedy, Wei Liu, Yangqing Jia, Pierre Sermanet, Scott Reed, Dragomir Anguelov, Dumitru Erhan, Vincent Vanhoucke, Andrew Rabinovich, et al., · 2015
Cited alongside, same era.
“A time delay neural network architecture for efficient modeling of long temporal contexts,”
Vijayaditya Peddinti, Daniel Povey, and Sanjeev Khudanpur, · 2015
Cited alongside, same era.
“Advances in deep neural network approaches to speaker recognition,”
Mitchell McLaren, Yun Lei, and Luciana Ferrer, · 2015
Cited alongside, same era.
“Deep metric learning using triplet network,”
Elad Hoffer and Nir Ailon, · 2015
Cited alongside, same era.
“Facenet: A unified embedding for face recognition and clustering,”
Florian Schroff, Dmitry Kalenichenko, and James Philbin, · 2015
Cited alongside, same era.
“Delving deep into rectifiers: Surpassing human-level performance on imagenet classification,”
“The ibm 2016 speaker recognition system,”
Seyed Omid Sadjadi, Sriram Ganapathy, and Jason W Pelecanos, · 2016
Later among the works it cites.
“https://www.nist.gov/itl/iad/mig/speaker-recognition-evaluation-2016,”
NIST speaker recognition evaluation 2016, · 2016
Later among the works it cites.
“A discriminative feature learning approach for deep face recognition,”
Yandong Wen, Kaipeng Zhang, Zhifeng Li, and Yu Qiao, · 2016
Later among the works it cites.
“Deep residual learning for image recognition,”
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun, · 2016
Later among the works it cites.
“The 2016 speakers in the wild speaker recognition evaluation.,”
Mitchell McLaren, Luciana Ferrer, Diego Castan, and Aaron Lawson, · 2016
Later among the works it cites.
“A speaker recognition system for the sitw challenge.,”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun, · 2015
Cited alongside, same era.
“Usage of DNN in speaker recognition: advantages and problems,”
Oleg Kudashev, Sergey Novoselov, Timur Pekhovsky, Konstantin Simonchik, and Galina Lavrentyeva, · 2016
Cited alongside, same era.
“End-to-end attention based text-dependent speaker verification,”
Shi-Xiong Zhang, Zhuo Chen, Yong Zhao, Jinyu Li, and Yifan Gong, · 2016
Cited alongside, same era.
“End-to-end text-dependent speaker verification,”
Georg Heigold, Ignacio Moreno, Samy Bengio, and Noam Shazeer, · 2016
Cited alongside, same era.
“Deep neural network based text-dependent speaker recognition: Preliminary results,”
Gautam Bhattacharya, Jahangir Alam, Themos Stafylakis, and Patrick Kenny, · 2016
Cited alongside, same era.
“X-vectors: Robust dnn embeddings for speaker recognition,”
David Snyder, Daniel Garcia-Romero, Gregory Sell, Daniel Povey, and Sanjeev Khudanpur,
Cited in the paper.
Oleg Kudashev, Sergey Novoselov, Konstantin Simonchik, and Alexander Kozlov, · 2016
Later among the works it cites.
“Sphereface: Deep hypersphere embedding for face recognition,”
Weiyang Liu, Yandong Wen, Zhiding Yu, Ming Li, Bhiksha Raj, and Le Song, · 2017
Later among the works it cites.
“Deep neural network embeddings for text-independent speaker verification,”
David Snyder, Daniel Garcia-Romero, Daniel Povey, and Sanjeev Khudanpur, · 2017
Later among the works it cites.
“Audio replay attack detection with deep learning frameworks,”
Galina Lavrentyeva, Sergey Novoselov, Egor Malykh, Alexander Kozlov, Oleg Kudashev, and Vadim Shchemelinin, · 2017
Later among the works it cites.
“Automatic differentiation in PyTorch,”
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer, · 2017
Later among the works it cites.