Fetching the paper…
Reading the bibliography…
In this paper, we present ECAPA2, a novel hybrid neural network architecture and training strategy to produce robust speaker embeddings.
“Musan: A music, speech, and noise corpus,”
David Snyder, Guoguo Chen, and Daniel Povey, · 2015
Earlier work this paper cites.
“Adam: A method for stochastic optimization,”
Diederik Kingma and Jimmy Ba, · 2015
Earlier work this paper cites.
“Understanding the effective receptive field in deep convolutional neural networks,”
Wenjie Luo, Yujia Li, Raquel Urtasun, and Richard Zemel, · 2016
Earlier work this paper cites.
“VoxCeleb: A large-scale speaker identification dataset,”
Arsha Nagrani, Joon Son Chung, and Andrew Zisserman, · 2017
Earlier work this paper cites.
“Axiomatic attribution for deep networks,”
Mukund Sundararajan, Ankur Taly, and Qiqi Yan, · 2017
Earlier work this paper cites.
“A study on data augmentation of reverberant speech for robust speech recognition,”
Tom Ko, Vijayaditya Peddinti, Daniel Povey, Michael L. Seltzer, and Sanjeev Khudanpur, · 2017
Earlier work this paper cites.
“Cyclical learning rates for training neural networks,”
Leslie N. Smith, · 2017
Earlier work this paper cites.
“Analysis of score normalization in multilingual speaker recognition,”
Pavel Matejka, Ondřej Novotný, Oldřich Plchot, Lukas Burget, Mireia Diez, and Jan Černocký, · 2017
Earlier work this paper cites.
“VoxCeleb2: Deep speaker recognition,”
Joon Son Chung, Arsha Nagrani, and Andrew Zisserman, · 2018
Earlier work this paper cites.
“X-vectors: Robust dnn embeddings for speaker recognition,”
David Snyder, Daniel Garcia-Romero, Gregory Sell, Daniel Povey, and Sanjeev Khudanpur, · 2018
Cited alongside, same era.
“Computing receptive fields of convolutional neural networks,”
Andre Araujo, Wade Norris, and Jack Sim, · 2019
Cited alongside, same era.
“How important is a neuron,”
Kedar Dhamdhere, Mukund Sundararajan, and Qiqi Yan, · 2019
Cited alongside, same era.
“Res2Net: A new multi-scale backbone architecture,”
Shanghua Gao, Ming-Ming Cheng, Kai Zhao, Xinyu Zhang, Ming-Hsuan Yang, and Philip H. S. Torr, · 2019
Cited alongside, same era.
“Arcface: Additive angular margin loss for deep face recognition,”
Jiankang Deng, Jia Guo, Niannan Xue, and Stefanos Zafeiriou, · 2019
Cited alongside, same era.
“Short utterance compensation in speaker verification via cosine-based teacher-student learning of speaker embeddings,”
“The NetEase Games system description for text-dependent sub-challenge of SDSVC 2020,”
Zhuxin Chen, Duisheng Chen, Hanyu Ding, and Yue Lin, · 2020
Later among the works it cites.
“Sub-center ArcFace: Boosting face recognition by large-scale noisy web faces,”
Jiankang Deng, Jia Guo, Tongliang Liu, Mingming Gong, and Stefanos Zafeiriou, · 2020
Later among the works it cites.
“Integrating frequency translational invariance in TDNNs and frequency positional information in 2D ResNets to enhance speaker verification,”
Jenthe Thienpondt, Brecht Desplanques, and Kris Demuynck, · 2021
Later among the works it cites.
“The idlab voxsrc-20 submission: Large margin fine-tuning and quality-aware score calibration in dnn based speaker verification,”
Jenthe Thienpondt, Brecht Desplanques, and Kris Demuynck, · 2021
Later among the works it cites.
“The speakin system for voxceleb speaker recognition challange 2021,” 2021
Miao Zhao, Yufeng Ma, Min Liu, and Minqiang Xu, · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jee-Weon Jung, Hee-Soo Heo, Hye-Jin Shim, and Ha-Jin Yu, · 2019
Cited alongside, same era.
“Speaker Augmentation and Bandwidth Extension for Deep Speaker Embedding,”
Hitoshi Yamamoto, Kong Aik Lee, Koji Okabe, and Takafumi Koshinaka, · 2019
Cited alongside, same era.
“SpecAugment: A simple data augmentation method for automatic speech recognition,”
Daniel S. Park, William Chan, Yu Zhang, Chung-Cheng Chiu, Barret Zoph, Ekin D. Cubuk, and Quoc V. Le, · 2019
Cited alongside, same era.
“ECAPA-TDNN: Emphasized channel attention, propagation and aggregation in TDNN based speaker verification,”
Brecht Desplanques, Jenthe Thienpondt, and Kris Demuynck, · 2020
Cited alongside, same era.
“Investigation on Deep Speaker Embedding Extraction Methods for Multi-Genre Speaker Verification,”
Woo Hyun Kang and Jahangir Alam, · 2022
Later among the works it cites.
“ID R&D system description to voxceleb speaker recognition challenge 2022,”
Rostislav Makarov, Nikita Torgashov, Alexander Alenin, Ivan Yakovlev, and Anton Okhotnikov, · 2022
Later among the works it cites.
“Tackling the score shift in cross-lingual speaker verification by exploiting language information,”
Jenthe Thienpondt, Brecht Desplanques, and Kris Demuynck, · 2022
Later among the works it cites.
“Margin-mixup: A method for robust speaker verification in multi-speaker audio,”
Jenthe Thienpondt, Nilesh Madhu, and Kris Demuynck, · 2023
Later among the works it cites.