Fetching the paper…
Reading the bibliography…
The task of making speaker verification systems robust to adverse scenarios remain a challenging and an active area of research.
Jesús Villalba, Nanxin Chen, et al., · 1910
Earlier work this paper cites.
“The ami meeting corpus,”
Iain McCowan, Jean Carletta, et al., · 2005
Earlier work this paper cites.
“Robust signal-to-noise ratio estimation based on waveform amplitude distribution analysis,”
Chanwoo Kim and Richard M Stern, · 2008
Earlier work this paper cites.
“Speech dereverberation based on variance-normalized delayed linear prediction,”
Tomohiro Nakatani et al., · 2010
Earlier work this paper cites.
“Generalization of multi-channel linear prediction methods for blind mimo impulse response shortening,”
Takuya Yoshioka and Tomohiro Nakatani, · 2012
Earlier work this paper cites.
“Generative adversarial nets,”
Ian Goodfellow, Pouget-Abadie, et al., · 2014
Earlier work this paper cites.
“The third chime speech separation and recognition challenge: Dataset, task and baselines,”
Jon Barker et al., · 2015
Earlier work this paper cites.
“The speakers in the wild (sitw) speaker recognition database.,”
Mitchell McLaren, Luciana Ferrer, Diego Castan, and Aaron Lawson, · 2016
Earlier work this paper cites.
“Unpaired image-to-image translation using cycle-consistent adversarial networks,”
Jun-Yan Zhu, Park, et al., · 2017
Cited alongside, same era.
Daniel Michelsanti and Zheng-Hua Tan, · 2017
Cited alongside, same era.
“Least squares generative adversarial networks,”
Xudong Mao, Li, et al., · 2017
Cited alongside, same era.
“Voxceleb: a large-scale speaker identification dataset,”
Arsha Nagrani, Joon Son Chung, and Andrew Zisserman, · 2017
Cited alongside, same era.
“X-vectors: Robust dnn embeddings for speaker recognition,”
David Snyder et al., · 2018
Cited alongside, same era.
“State-of-the-art speaker recognition with neural network embeddings in nist sre18 and speakers in the wild evaluations,”
Jesús Villalba, Nanxin Chen, et al., · 2019
Closest in time.
“Multi-plda diarization on children’s speech,”
Jiamin Xie, Leibny Paola García-Perera, et al., · 2019
Closest in time.
“Voiceid loss: Speech enhancement for speaker verification,”
Suwon Shon, Hao Tang, and James Glass, · 2019
Closest in time.
“Low-Resource Domain Adaptation for Speaker Recognition Using Cycle-GANs,”
Phani Sankar Nidadavolu, Saurabh Kataria, Jesús Villalba, and Najim Dehak, · 2019
Closest in time.
“Cycle-gans for domain adaptation of acoustic features for speaker recognition,”
Phani Sankar Nidadavolu, Jesús Villalba, and Najim Dehak, · 2019
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zhong Meng, Jinyu Li, Yifan Gong, et al., · 2018
Cited alongside, same era.
“Voxceleb2: Deep speaker recognition,”
Joon Son Chung, Arsha Nagrani, and Andrew Zisserman, · 2018
Cited alongside, same era.
“The voices from a distance challenge 2019 evaluation plan,”
Mahesh Kumar Nandwana, Julien Van Hout, et al., · 2019
Cited alongside, same era.
Heiga Zen et al., · 2019
Closest in time.
“Ldc2019e60, distant microphone conversational speech in noisy environments,” Private communication in support of the 2019 JHU/CLSP Summer Workshop, 2019
Diego Castán et al., · 2019
Closest in time.
“Speaker detection in the wild: lessons learned from jsalt 2019.,”
Paola Garcia et al., · 2020
Closest in time.