Fetching the paper…
Reading the bibliography…
We propose SpeakerNet - a new neural architecture for speaker recognition and speaker verification tasks.
“Speaker verification using adapted gaussian mixture models,”
D.A. Reynolds, T.F. Quatieri, and R.B. Dunn, · 2000
Earlier work this paper cites.
“Front-end factor analysis for speaker verification,”
N. Dehak, P.J. Kenny, R. Dehak, P. Dumouchel, and P. Ouellet, · 2010
Earlier work this paper cites.
“Improving neural networks by preventing co-adaptation of feature detectors,”
G.E. Hinton, N. Srivastava, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov, · 2012
Earlier work this paper cites.
“Maxout networks,”
I. Goodfellow, D. Warde-Farley, M. Mirza, A. Courville, and Y. Bengio, · 2013
Earlier work this paper cites.
“Deep neural networks for small footprint text-dependent speaker verification,”
E. Variani, X. Lei, E. McDermott, I. Moreno, and J. Gonzalez-Dominguez, · 2014
Earlier work this paper cites.
“Speaker recognition by machines and humans: A tutorial review,”
J.H.L. Hansen and T. Hasan, · 2015
Earlier work this paper cites.
“End-to-end text-dependent speaker verification,”
G. Heigold, I. Moreno, S. Bengio, and N. Shazeer, · 2016
Earlier work this paper cites.
“Deep Voice 2: Multi-speaker neural text-to-speech,”
A. Gibiansky, S. Arik, G. Diamos, J. Miller, K. Peng, W. Ping, J. Raiman, and Y. Zhou, · 2017
Earlier work this paper cites.
“Deep neural network embeddings for text-independent speaker verification.,”
D. Snyder, D. Garcia-Romero, D. Povey, and S. Khudanpur, · 2017
Cited alongside, same era.
“TristouNet: triplet loss for speaker turn embedding,”
H. Bredin, · 2017
Cited alongside, same era.
“Domain and speaker adaptation for cortana speech recognition,”
Y. Zhao, J. Li, S. Zhang, L. Chen, and Y. Gong, · 2018
Cited alongside, same era.
“VoxCeleb2: Deep speaker recognition,”
J.S. Chung, A. Nagrani, and A. Zisserman, · 2018
Cited alongside, same era.
“X-vectors: Robust dnn embeddings for speaker recognition,”
D. Snyder, D. Garcia-Romero, G. Sell, D. Povey, and S. Khudanpur, · 2018
Cited alongside, same era.
“Attention-based models for text-dependent speaker verification,”
F.A. Chowdhury, Q. Wang, I.L. Moreno, and L. Wan, · 2018
Cited alongside, same era.
“Self-attentive speaker embeddings for text-independent speaker verification,”
Y. Zhu, T. Ko, D. Snyder, B. K. Mak, and D. Povey, · 2018
Later among the works it cites.
“Speaker verification using convolutional neural networks,”
H. Salehghaffari, · 2018
Later among the works it cites.
“Generalized end-to-end loss for speaker verification,”
L. Wan, Q. Wang, A. Papir, and I.L. Moreno, · 2018
Later among the works it cites.
“BUT system description to voxceleb speaker recognition challenge 2019,”
H. Zeinali, S. Wang, A. Silnova, P. Matějka, and O. Plchot, · 2019
Later among the works it cites.
“H-vectors: Utterance-level speaker embedding using a hierarchical attention model,”
Y. Shi, Q. Huang, and T. Hain, · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“Attention mechanism in speaker recognition: What does it learn in deep speaker embedding?,”
Q. Wang, K. Okabe, K.A. Lee, H. Yamamoto, and T. Koshinaka, · 2018
Cited alongside, same era.
“Attentive statistics pooling for deep speaker embedding,”
K. Okabe, T. Koshinaka, and K. Shinoda, · 2018
Cited alongside, same era.
“Arcface: Additive angular margin loss for deep face recognition,”
Jiankang Deng, Jia Guo, Niannan Xue, and Stefanos Zafeiriou, · 2019
Later among the works it cites.
“QuartzNet: Deep automatic speech recognition with 1D time-channel separable convolutions,”
S. Kriman, S. Beliaev, B. Ginsburg, J. Huang, O. Kuchaiev, V. Lavrukhin, R. Leary, J. Li, and Y. Zhang, · 2020
Closest in time.
“A comparison of metric learning loss functions for end-to-end speaker verification,”
J.M. Coria, H. Bredin, S. Ghannay, and S. Rosset, · 2020
Closest in time.