Fetching the paper…
Reading the bibliography…
Speaker recognition systems based on deep speaker embeddings have achieved significant performance in controlled conditions according to the results obtained for early NIST SRE (Speaker Recognition Evaluation) datasets.
“Image method for efficiently simulating small-room acoustics,”
Jont B. Allen and David A. Berkley, · 1979
Earlier work this paper cites.
“Bayesian speaker verification with heavy-tailed priors,”
Patrick Kenny, · 2010
Earlier work this paper cites.
“Speech dereverberation based on variance-normalized delayed linear prediction,”
Tomohiro Nakatani, Takuya Yoshioka, Keisuke Kinoshita, Masato Miyoshi, and Biing-Hwang Juang, · 2010
Earlier work this paper cites.
“Cosine similarity metric learning for face verification,”
Hieu V. Nguyen and Li Bai, · 2010
Earlier work this paper cites.
“The VOiCES from a distance challenge 2019 evaluation plan,”
Mahesh K. Nandwana, Julien van Hout, Mitchell McLaren, Colleen Richey, Aaron Lawson, and Maria A. Barrios, · 2014
Earlier work this paper cites.
“U-net: Convolutional networks for biomedical image segmentation,”
Olaf Ronneberger, Philipp Fischer, and Thomas Brox, · 2015
Earlier work this paper cites.
“Deep residual learning for image recognition,”
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun, · 2016
Earlier work this paper cites.
“V-net: Fully convolutional neural networks for volumetric medical image segmentation,”
Fausto Milletari, Nassir Navab, and Seyed-Ahmad Ahmadi, · 2016
Earlier work this paper cites.
“The speakers in the wild (SITW) speaker recognition database,”
Mitchell McLaren, Luciana Ferrer, Diego Castan, and Aaron Lawson, · 2016
Earlier work this paper cites.
“Deep neural network embeddings for text-independent speaker verification,”
David Snyder, Daniel Garcia-Romero, Daniel Povey, and Sanjeev Khudanpur, · 2017
Earlier work this paper cites.
“SphereFace: Deep hypersphere embedding for face recognition,”
Weiyang Liu, Yandong Wen, Zhiding Yu, Ming Li, Bhiksha Raj, and Le Song, · 2017
Earlier work this paper cites.
“Audio replay attack detection with deep learning frameworks,”
Galina Lavrentyeva et all., · 2017
Earlier work this paper cites.
“Nuance–Politecnico di Torino’s 2016 NIST speaker recognition evaluation system,”
Daniel Colibro, Claudio Vair, Emanuele Dalmasso, Kevin Farrell, Gennady Karvitsky, Sandro Cumani, and Pietro Laface, · 2017
Cited alongside, same era.
“VoxCeleb: A large-scale speaker identification dataset,”
Arsha Nagrani, Joon S. Chung, and Andrew Zisserman, · 2017
Cited alongside, same era.
“Voices obscured in complex environmental settings (VOiCES) corpus,”
Colleen Richey et al., · 2018
Cited alongside, same era.
“On deep speaker embeddings for text-independent speaker recognition,”
Sergey Novoselov, Andrey Shulipa, Ivan Kremnev, Alexander Kozlov, and Vadim Shchemelinin, · 2018
Cited alongside, same era.
“A novel learnable dictionary encoding layer for end-to-end language identification,”
Weicheng Cai, Zexin Cai, Xiang Zhang, Xiaoqi Wang, and Ming Li, · 2018
Cited alongside, same era.
“Additive margin softmax for face verification,”
“Utterance-level aggregation for speaker recognition in the wild,”
Weidi Xie, Arsha Nagrani, Joon S. Chung, and Andrew Zisserman, · 2019
Later among the works it cites.
“A deep neural network for short-segment speaker recognition,”
Amirhossein Hajavi and Ali Etemad, · 2019
Later among the works it cites.
“STC speaker recognition systems for the VOiCES from a distance challenge,”
Sergey Novoselov, Alexey Gusev, Artem Ivanov, Timur Pekhovsky, Andrey Shulipa, Galina Lavrentyeva, Vladimir Volokhov, and Alexander Kozlov, · 2019
Later among the works it cites.
“The DKU system for the speaker recognition task of the 2019 VOiCES from a distance challenge,”
Danwei Cai, Xiaoyi Qin, Weicheng Cai, and Ming Li, · 2019
Later among the works it cites.
“Analysis of BUT submission in far-field scenarios of VOiCES 2019 challenge,”
Pavel Matejka et all., · 2019
Later among the works it cites.
“The JHU speaker recognition system for the VOiCES 2019 challenge,”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Feng Wang, Weiyang Liu, Haijun Liu, and Jian Cheng, · 2018
Cited alongside, same era.
“Squeeze-and-excitation networks,”
Jie Hu, Li Shen, and Gang Sun, · 2018
Cited alongside, same era.
“Triplet loss based cosine similarity metric learning for text-independent speaker recognition,”
Sergey Novoselov, Vadim Shchemelinin, Andrey Shulipa, Alexander Kozlov, and Ivan Kremnev, · 2018
Cited alongside, same era.
“Speaker verification in mismatched conditions with frustratingly easy domain adaptation,”
Jahangir M. Alam, Gautam Bhattacharya, and Patrick Kenny, · 2018
Cited alongside, same era.
“VoxCeleb2: Deep speaker recognition,”
Joon S. Chung, Arsha Nagrani, and Andrew Zisserman, · 2018
Cited alongside, same era.
“BUT system description to VoxCeleb speaker recognition challenge 2019,”
Hossein Zeinali, Shuai Wang, Anna Silnova, Pavel Matějka, and Oldřich Plchot, · 2019
Cited alongside, same era.
“Speaker recognition for multi-speaker conversations using x-vectors,”
David Snyder et al., · 2019
Cited alongside, same era.
David Snyder, Jesús Villalba, Nanxin Chen, Daniel Povey, Gregory Sell, Najim Dehak, and Sanjeev Khudanpur, · 2019
Later among the works it cites.
“ArcFace: Additive angular margin loss for deep face recognition,”
Jiankang Deng, Jia Guo, Niannan Xue, and Stefanos Zafeiriou, · 2019
Later among the works it cites.
“Softmax dissection: Towards understanding intra- and inter-class objective for embedding learning,”
Lanqing He, Zhongdao Wang, Yali Li, and Shengjin Wang, · 2019
Later among the works it cites.
“The STC ASR system for the VOiCES from a distance challenge 2019,”
Ivan Medennikov et al., · 2019
Later among the works it cites.
“State-of-the-art speaker recognition with neural network embeddings in NIST SRE18 and speakers in the wild evaluations,”
Jesús Villalba et al., · 2020
Closest in time.
“VoxCeleb: Large-scale speaker verification in the wild,”
Arsha Nagrani, Joon S. Chung, Weidi Xie, and Andrew Zisserman, · 2020
Closest in time.