Fetching the paper…
Reading the bibliography…
In this paper we describe the recent advancements made in the IBM i-vector speaker recognition system for conversational speech.
S. Ganapathy, S. Thomas, D. Dimitriadis, and S. Rennie, “Investigating factor analysis features for deep neural networks in noisy speech recognition,” in
1902
Earlier work this paper cites.
K. Fukunaga and J. Mantock, “Nonparametric discriminant analysis,”
1983
Earlier work this paper cites.
K. Fukunaga,
1990
Earlier work this paper cites.
S. J. Young, J. J. Odell, and P. C. Woodland, “Tree-based state tying for high accuracy acoustic modelling,” in
1994
Earlier work this paper cites.
V. Digalakis, D. Rtischev, and L. Neumeyer, “Speaker adaptation using constrained estimation of Gaussian mixtures,”
1995
Earlier work this paper cites.
M. J. F. Gales, “Maximum likelihood linear transformations for HMM-based speech recognition,”
1998
Earlier work this paper cites.
C. Cieri, D. Miller, and K. Walker, “The Fisher corpus: A resource for the next generations of speech-to-text,” in
2004
Earlier work this paper cites.
P. Kenny, G. Boulianne, P. Ouellet, and P. Dumouchel, “Joint factor analysis versus eigenchannels in speaker recognition,”
2007
Earlier work this paper cites.
S. J. Prince and J. H. Elder, “Probabilistic linear discriminant analysis for inferences about identity,” in
2007
Cited alongside, same era.
C. Cieri, L. Corson, D. Graff, and K. Walker, “Resources for new research directions in speaker recognition: The Mixer 3, 4 and 5 corpora,” in
2007
Cited alongside, same era.
NIST, “The NIST Year 2008 Speaker Recognition Evaluation Plan,”
2008
Cited alongside, same era.
M. K. Omar and J. Pelecanos, “Training universal background models for speaker recognition,” in
2010
Cited alongside, same era.
P. Kenny, “Bayesian speaker verification with heavy tailed priors,” in
2010
Cited alongside, same era.
N. Dehak, P. Kenny, R. Dehak, P. Dumouchel, and P. Ouellet, “Front-end factor analysis for speaker verification,”
2011
Later among the works it cites.
D. Garcia-Romero and C. Y. Espy-Wilson, “Analysis of i-vector length normalization in speaker recognition systems.” in
2011
Later among the works it cites.
S. O. Sadjadi and J. H. L. Hansen, “Unsupervised speech activity detection using voicing measures and perceptual spectral flux,”
2013
Later among the works it cites.
Y. Lei, N. Scheffer, L. Ferrer, and M. McLaren, “A novel scheme for speaker recognition using a phonetically-aware deep neural network,” in
2014
Later among the works it cites.
S. O. Sadjadi, J. W. Pelecanos, and W. Zhu, “Nearest neighbor discriminant analysis for robust speaker recognition,” in
2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2010
Cited alongside, same era.
NIST, “The NIST Year 2010 Speaker Recognition Evaluation Plan,”
2010
Cited alongside, same era.
H. Soltau, G. Saon, and B. Kingsbury, “The IBM Attila speech recognition toolkit,” in
2010
Cited alongside, same era.
D. Snyder, D. Garcia-Romero, and D. Povey, “Time delay deep neural network-based universal background models for speaker recognition,” in
2015
Later among the works it cites.
S. O. Sadjadi, S. Ganapathy, and J. W. Pelecanos, “Nearest neighbor discriminant analysis for language recognition,” in
2015
Later among the works it cites.
G. Saon, H. K. Kuo, S. Rennie, and M. Picheny, “The IBM 2015 English conversational telephone speech recognition system,” in
2015
Later among the works it cites.