Fetching the paper…
Reading the bibliography…
Deep neural network based speaker embeddings, such as x-vectors, have been shown to perform well in text-independent speaker recognition/verification tasks.
“Probabilistic linear discriminant analysis for inferences about identity,”
Simon JD Prince and James H Elder, · 2007
Earlier work this paper cites.
“Nist 2008 speaker recognition evaluation: Performance across telephone and room microphone channels,”
Alvin F Martin and Craig S Greenberg, · 2009
Earlier work this paper cites.
“Front-end factor analysis for speaker verification,”
Najim Dehak, Patrick J Kenny, Réda Dehak, Pierre Dumouchel, and Pierre Ouellet, · 2010
Earlier work this paper cites.
“The nist 2010 speaker recognition evaluation,”
Alvin F Martin and Craig S Greenberg, · 2010
Earlier work this paper cites.
“The Kaldi speech recognition toolkit,”
Daniel Povey, Arnab Ghoshal, Gilles Boulianne, Lukas Burget, Ondrej Glembek, Nagendra Goel, Mirko Hannemann, Petr Motlicek, Yanmin Qian, Petr Schwarz, et al., · 2011
Earlier work this paper cites.
“Rsr2015: Database for text-dependent speaker verification using multiple pass-phrases,”
Anthony Larcher, Kong Aik Lee, Bin Ma, and Haizhou Li, · 2012
Earlier work this paper cites.
“Speaker adaptation of neural network acoustic models using i-vectors,”
George Saon, Hagen Soltau, David Nahamoo, and Michael Picheny, · 2013
Earlier work this paper cites.
“A novel scheme for speaker recognition using a phonetically-aware deep neural network,”
Yun Lei, Nicolas Scheffer, Luciana Ferrer, and Mitchell McLaren, · 2014
Earlier work this paper cites.
“Deep neural networks for small footprint text-dependent speaker verification,”
Ehsan Variani, Xin Lei, Erik McDermott, Ignacio Lopez Moreno, and Javier Gonzalez-Dominguez, · 2014
Earlier work this paper cites.
“Advances in deep neural network approaches to speaker recognition,”
Mitchell McLaren, Yun Lei, and Luciana Ferrer, · 2015
Cited alongside, same era.
“The reddots data collection for speaker recognition,”
Kong Aik Lee, Anthony Larcher, Guangsen Wang, Patrick Kenny, Niko Brümmer, David van Leeuwen, Hagai Aronowitz, Marcel Kockmann, Carlos Vaquero, Bin Ma, et al., · 2015
Cited alongside, same era.
“Time delay deep neural network-based universal background models for speaker recognition,”
David Snyder, Daniel Garcia-Romero, and Daniel Povey, · 2015
Cited alongside, same era.
“Musan: A music, speech, and noise corpus,”
David Snyder, Guoguo Chen, and Daniel Povey, · 2015
Cited alongside, same era.
“Adam: A method for stochastic optimization,”
Diederik P. Kingma and Jimmy Ba, · 2015
Cited alongside, same era.
“What does the speaker embedding encode?,”
Shuai Wang, Yanmin Qian, and Kai Yu, · 2017
Later among the works it cites.
“Automatic differentiation in PyTorch,”
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer, · 2017
Later among the works it cites.
“A study on data augmentation of reverberant speech for robust speech recognition,”
Tom Ko, Vijayaditya Peddinti, Daniel Povey, Michael L Seltzer, and Sanjeev Khudanpur, · 2017
Later among the works it cites.
“X-vectors: Robust dnn embeddings for speaker recognition,”
David Snyder, Daniel Garcia-Romero, Gregory Sell, Daniel Povey, and Sanjeev Khudanpur, · 2018
Later among the works it cites.
“What you can cram into a single $&!#* vector: Probing sentence embeddings for linguistic properties,”
Alexis Conneau, German Kruszewski, Guillaume Lample, Loïc Barrault, and Marco Baroni, · 2018
Later among the works it cites.
“Voxceleb2: Deep speaker recognition,”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
David Snyder, Pegah Ghahremani, Daniel Povey, Daniel Garcia-Romero, Yishay Carmiel, and Sanjeev Khudanpur, · 2016
Cited alongside, same era.
“i-vector/hmm based text-dependent speaker verification system for reddots challenge.,”
Hossein Zeinali, Hossein Sameti, Lukas Burget, Jan Cernockỳ, Nooshin Maghsoodi, and Pavel Matejka, · 2016
Cited alongside, same era.
“New release of mixer-6: Improved validity for phonetic study of speaker variation and identification,”
Eleanor Chodroff, Matthew Maciejewski, Jan Trmal, Sanjeev Khudanpur, and John Godfrey, · 2016
Cited alongside, same era.
“Deep neural network embeddings for text-independent speaker verification.,”
David Snyder, Daniel Garcia-Romero, Daniel Povey, and Sanjeev Khudanpur, · 2017
Cited alongside, same era.
J. S. Chung, A. Nagrani, and A. Zisserman, · 2018
Later among the works it cites.
“Disentangling Style Factors from Speaker Representations,”
Jennifer Williams and Simon King, · 2019
Closest in time.
“BERT: Pre-training of deep bidirectional transformers for language understanding,”
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova, · 2019
Closest in time.