Fetching the paper…
Reading the bibliography…
Current speaker recognition technology provides great performance with the x-vector approach.
“Application-independent evaluation of speaker detection,”
Niko Brümmer and Johan Du Preez, · 2006
Earlier work this paper cites.
“Robust signal-to-noise ratio estimation based on waveform amplitude distribution analysis,”
Chanwoo Kim and Richard M Stern, · 2008
Earlier work this paper cites.
“Agnitio’s speaker recognition system for evalita 2009,”
Niko Brümmer and Albert Strasheim, · 2009
Earlier work this paper cites.
“Speech dereverberation based on variance-normalized delayed linear prediction,”
Tomohiro Nakatani, Takuya Yoshioka, Keisuke Kinoshita, Masato Miyoshi, and Biing-Hwang Juang, · 2010
Earlier work this paper cites.
“The speaker partitioning problem.,”
Niko Brümmer and Edward De Villiers, · 2010
Earlier work this paper cites.
“The kaldi speech recognition toolkit,”
Daniel Povey, Arnab Ghoshal, Gilles Boulianne, Lukas Burget, Ondrej Glembek, Nagendra Goel, Mirko Hannemann, Petr Motlicek, Yanmin Qian, Petr Schwarz, et al., · 2011
Earlier work this paper cites.
Takuya Yoshioka and Tomohiro Nakatani, · 2012
Earlier work this paper cites.
“Generative adversarial nets,”
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio, · 2014
Earlier work this paper cites.
“Musan: A music, speech, and noise corpus,”
David Snyder, Guoguo Chen, and Daniel Povey, · 2015
Earlier work this paper cites.
“Unsupervised representation learning with deep convolutional generative adversarial networks,”
Alec Radford, Luke Metz, and Soumith Chintala, · 2015
Cited alongside, same era.
“The speakers in the wild (sitw) speaker recognition database.,”
Mitchell McLaren, Luciana Ferrer, Diego Castan, and Aaron Lawson, · 2016
Cited alongside, same era.
“Voxceleb: a large-scale speaker identification dataset,”
Arsha Nagrani, Joon Son Chung, and Andrew Zisserman, · 2017
Cited alongside, same era.
“Unpaired image-to-image translation using cycle-consistent adversarial networks,”
Jun-Yan Zhu, Taesung Park, Phillip Isola, and Alexei A Efros, · 2017
Cited alongside, same era.
“Cross-domain speech recognition using nonparallel corpora with cycle-consistent adversarial networks,”
Masato Mimura, Shinsuke Sakai, and Tatsuya Kawahara, · 2017
Cited alongside, same era.
“High-quality nonparallel voice conversion based on cycle-consistent adversarial network,”
Fuming Fang, Junichi Yamagishi, Isao Echizen, and Jaime Lorenzo-Trueba, · 2018
Later among the works it cites.
“Stargan-vc: Non-parallel many-to-many voice conversion with star generative adversarial networks,”
Hirokazu Kameoka, Takuhiro Kaneko, Kou Tanaka, and Nobukatsu Hojo, · 2018
Later among the works it cites.
“A multi-discriminator cyclegan for unsupervised non-parallel speech domain adaptation,”
Ehsan Hosseini-Asl, Yingbo Zhou, Caiming Xiong, and Richard Socher, · 2018
Later among the works it cites.
“Voxceleb2: Deep speaker recognition,”
Joon Son Chung, Arsha Nagrani, and Andrew Zisserman, · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“Parallel-data-free voice conversion using cycle-consistent adversarial networks,”
Takuhiro Kaneko and Hirokazu Kameoka, · 2017
Cited alongside, same era.
“Least squares generative adversarial networks,”
Xudong Mao, Qing Li, Haoran Xie, Raymond YK Lau, Zhen Wang, and Stephen Paul Smolley, · 2017
Cited alongside, same era.
“X-Vectors : Robust DNN Embeddings for Speaker Recognition,”
David Snyder, Daniel Garcia-Romero, Gregory Sell, Daniel Povey, and Sanjeev Khudanpur, · 2018
Cited alongside, same era.
“Cycle-consistent speech enhancement,”
Zhong Meng, Jinyu Li, Yifan Gong, et al., · 2018
Cited alongside, same era.
Colleen Richey, Maria A. Barrios, Zeb Armstrong, Chris Bartels, Horacio Franco, Martin Graciarena, Aaron Lawson, Mahesh Kumar Nandwana, Allen Stauffer, Julien van Hout, Paul Gamble, Jeffrey Hetherly, Cory Stephenson, and Karl Ni, · 2018
Later among the works it cites.
“State-of-the-art speaker recognition for telephone and video speech: the jhu-mit submission for nist sre18,”
Jesús Villalba, Nanxin Chen, David Snyder, Daniel Garcia-Romero, Alan McCree, Gregory Sell, Jonas Borgstrom, Fred Richardson, Suwon Shon, François Grondin, et al., · 2019
Closest in time.
“Cycle-gans for domain adaptation of acoustic features for speaker recognition,”
Phani Sankar Nidadavolu, Jesús Villalba, and Najim Dehak, · 2019
Closest in time.
“Libritts: A corpus derived from librispeech for text-to-speech,”
Heiga Zen, Viet Dang, Rob Clark, Yu Zhang, Ron J Weiss, Ye Jia, Zhifeng Chen, and Yonghui Wu, · 2019
Closest in time.
“The voices from a distance challenge 2019 evaluation plan,”
Mahesh Kumar Nandwana, Julien Van Hout, Mitchell McLaren, Colleen Richey, Aaron Lawson, and Maria Alejandra Barrios, · 2019
Closest in time.