Fetching the paper…
Reading the bibliography…
Speaker Verification still suffers from the challenge of generalization to novel adverse environments.
“Robust signal-to-noise ratio estimation based on waveform amplitude distribution analysis,”
Chanwoo Kim and Richard M Stern, · 2008
Earlier work this paper cites.
“The diverse environments multi-channel acoustic noise database (demand): A database of multichannel environmental noise recordings,”
Joachim Thiemann, Nobutaka Ito, and Emmanuel Vincent, · 2013
Earlier work this paper cites.
“Musan: A music, speech, and noise corpus,”
David Snyder, Guoguo Chen, and Daniel Povey, · 2015
Earlier work this paper cites.
“Perceptual losses for real-time style transfer and super-resolution,”
Justin Johnson, Alexandre Alahi, and Li Fei-Fei, · 2016
Earlier work this paper cites.
“The speakers in the wild (sitw) speaker recognition database.,”
Mitchell McLaren, Luciana Ferrer, Diego Castan, et al., · 2016
Earlier work this paper cites.
“Homebank: An online repository of daylong child-centered audio recordings,”
Mark VanDam, Anne S Warlaumont, Elika Bergelson, et al., · 2016
Earlier work this paper cites.
“Hearing in a shoe-box: binaural source position and wall absorption estimation using virtually supervised learning,”
Saurabh Kataria, Clément Gaultier, and Antoine Deleforge, · 2017
Earlier work this paper cites.
“Segan: Speech enhancement generative adversarial network,”
Santiago Pascual, Antonio Bonafonte, and Joan Serra, · 2017
Earlier work this paper cites.
Daniel Michelsanti and Zheng-Hua Tan, · 2017
Cited alongside, same era.
“Swish: a self-gated activation function,”
Prajit Ramachandran, Barret Zoph, and Quoc V Le, · 2017
Cited alongside, same era.
“Speech denoising with deep feature losses,”
Francois G Germain, Qifeng Chen, and Vladlen Koltun, · 2018
Cited alongside, same era.
“Squeeze-and-excitation networks,”
Jie Hu, Li Shen, and Gang Sun, · 2018
Cited alongside, same era.
“The jhu-mit system description for nist sre18,”
Jesús Villalba, Nanxin Chen, David Snyder, et al., · 2018
Cited alongside, same era.
“Voiceid loss: Speech enhancement for speaker verification,”
Suwon Shon, Hao Tang, and James Glass, · 2019
Closest in time.
“Multi-plda diarization on children’s speech,”
Jiamin Xie, Leibny Paola García-Perera, Daniel Povey, et al., · 2019
Closest in time.
“Low-Resource Domain Adaptation for Speaker Recognition Using Cycle-GANs,”
Phani Sankar Nidadavolu, Saurabh Kataria, Jesús Villalba, et al., · 2019
Closest in time.
“Unsupervised feature enhancement for speaker verification,”
Phani Sankar Nidadavolu, Saurabh Kataria, Jesús Villalba, Paola Garcia-Perera, and Najim Dehak, · 2019
Closest in time.
“Cycle-gans for domain adaptation of acoustic features for speaker recognition,”
Phani Sankar Nidadavolu, Jesús Villalba, and Najim Dehak, · 2019
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Joon Son Chung, Arsha Nagrani, and Andrew Zisserman, · 2018
Cited alongside, same era.
“The voices from a distance challenge 2019 evaluation plan,”
Mahesh Kumar Nandwana, Julien Van Hout, Mitchell McLaren, et al., · 2019
Cited alongside, same era.
“State-of-the-art speaker recognition with neural network embeddings in nist sre18 and speakers in the wild evaluations,”
Jesús Villalba, Nanxin Chen, David Snyder, et al., · 2019
Cited alongside, same era.
“The jhu speaker recognition system for the voices 2019 challenge,”
David Snyder, Jesús Villalba, Nanxin Chen, et al., · 2019
Closest in time.
“Speaker detection in the wild: lessons learned from jsalt 2019.,”
Paola Garcia, Jesús Villalba, Hérve Bredin, et al., · 2020
Closest in time.