Fetching the paper…
Reading the bibliography…
The computing power of mobile devices limits the end-user applications in terms of storage size, processing, memory and energy consumption.
J. S. Garofolo, L. F. Lamel, W. M. Fisher, J. G. Fiscus, and D. S. Pallett, “Darpa timit acoustic-phonetic continous speech corpus cd-rom. nist speech disc 1-1.1,”
1993
Earlier work this paper cites.
R. H. Woo, A. Park, and T. J. Hazen, “The mit mobile device speaker verification corpus: data collection and preliminary experiments,” in
2006
Earlier work this paper cites.
E. Variani, X. Lei, E. McDermott, I. L. Moreno, and J. Gonzalez-Dominguez, “Deep neural networks for small footprint text-dependent speaker verification,” in
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
L. Sifre and S. Mallat, “Rigid-motion scattering for image classification,”
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,”
2014
Earlier work this paper cites.
W. Chan, N. Jaitly, Q. V. Le, and O. Vinyals, “Listen, attend and spell,”
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
V. Sindhwani, T. Sainath, and S. Kumar, “Structured transforms for small-footprint deep learning,” in
2015
Earlier work this paper cites.
Z. Yang, M. Moczulski, M. Denil, N. de Freitas, A. Smola, L. Song, and Z. Wang, “Deep fried convnets,” in
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
P. Safari, O. Ghahabi, and J. Hernando, “From features to speaker vectors by means of restricted boltzmann machine adaptation,” in
2016
Earlier work this paper cites.
S.-X. Zhang, Z. Chen, Y. Zhao, J. Li, and Y. Gong, “End-to-end attention based text-dependent speaker verification,” in
2016
Earlier work this paper cites.
2016
Cited alongside, same era.
J. L. Ba, J. R. Kiros, and G. E. Hinton, “Layer normalization,”
2016
Cited alongside, same era.
D. Snyder, D. Garcia-Romero, D. Povey, and S. Khudanpur, “Deep neural network embeddings for text-independent speaker verification,” in
2017
Cited alongside, same era.
G. Bhattacharya, M. J. Alam, and P. Kenny, “Deep speaker embeddings for short-duration speaker verification,” in
2017
Cited alongside, same era.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin, “Attention is all you need,” in
2017
2018
Later among the works it cites.
C.-C. Chiu, T. N. Sainath, Y. Wu, and
2018
Later among the works it cites.
M. Dehghani, S. Gouws, O. Vinyals, J. Uszkoreit, and L. Kaiser, “Universal transformers,”
2018
Later among the works it cites.
Y. Zhu, T. Ko, D. Snyder, B. Mak, and D. Povey, “Self-attentive speaker embeddings for text-independent speaker verification,” in
2018
Later among the works it cites.
W. Cai, J. Chen, and M. Li, “Exploring the encoding layer and loss function in end-to-end speaker and language recognition system,” in
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
M. Wang, B. Liu, and H. Foroosh, “Factorized convolutional neural networks,” in
2017
Cited alongside, same era.
A. Nagrani, J. S. Chung, and A. Zisserman, “Voxceleb: a large-scale speaker identification dataset,” in
2017
Cited alongside, same era.
D. Snyder, D. Garcia-Romero, G. Sell, D. Povey, and S. Khudanpur, “X-vectors: Robust dnn embeddings for speaker recognition,” in
2018
Cited alongside, same era.
O. Ghahabi and J. Hernando, “Restricted boltzmann machines for vector representation of speech in speaker recognition,”
2018
Cited alongside, same era.
M. Sandler, A. Howard, M. Zhu, A. Zhmoginov, and L.-C. Chen, “Mobilenetv2: Inverted residuals and linear bottlenecks,” in
2018
Later among the works it cites.
J. S. Chung, A. Nagrani, and A. Zisserman, “Voxceleb2: Deep speaker recognition,” in
2018
Later among the works it cites.
K. Okabe, T. Koshinaka, and K. Shinoda, “Attentive statistics pooling for deep speaker embedding,”
2018
Later among the works it cites.
F. Wang, J. Cheng, W. Liu, and H. Liu, “Additive margin softmax for face verification,”
2018
Later among the works it cites.
M. India, P. Safari, and J. Hernando, “Self Multi-Head Attention for Speaker Recognition,” in
2019
Later among the works it cites.
H. Zeinali, L. Burget, J. Rohdin, T. Stafylakis, and J. H. Cernocky, “How to improve your speaker embeddings extractor in generic toolkits,” in
2019
Later among the works it cites.
B. McFee, M. McVicar, S. Balke, and
2019
Later among the works it cites.