Fetching the paper…
Reading the bibliography…
ASVspoof 5 is the fifth edition in a series of challenges that promote the study of speech spoofing and deepfake attacks, and the design of detection solutions.
“Application-independent evaluation of speaker detection,”
Niko Brümmer and Johan du Preez, · 2006
Earlier work this paper cites.
“Probabilistic linear discriminant analysis for inferences about identity,”
Simon JD Prince and James H Elder, · 2007
Earlier work this paper cites.
“Librispeech: An ASR corpus based on public domain audio books,”
Vassil Panayotov, Guoguo Chen, Daniel Povey, and Sanjeev Khudanpur, · 2015
Earlier work this paper cites.
“The ASVspoof 2017 challenge: Assessing the limits of replay spoofing attack detection,”
Tomi Kinnunen, Md. Sahidullah, Hector Delgado, Massimiliano Todisco, Nicholas Evans, Junichi Yamagishi, and Kong-Aik Lee, · 2017
Earlier work this paper cites.
“VoxCeleb: A Large-Scale Speaker Identification Dataset,”
Arsha Nagrani, Joon Son Chung, and Andrew Zisserman, · 2017
Earlier work this paper cites.
“Creating new language and voice components for the updated marytts text-to-speech synthesis platform,”
Ingmar Steiner and Sébastien Le Maguer, · 2018
Earlier work this paper cites.
“Natural TTS synthesis by conditioning wavenet on Mel spectrogram predictions,”
Jonathan Shen, Ruoming Pang, Ron J Weiss, Mike Schuster, Navdeep Jaitly, Zongheng Yang, Zhifeng Chen, Yu Zhang, Yuxuan Wang, Rj Skerrv-Ryan, et al., · 2018
Earlier work this paper cites.
“ASVspoof 2019: future horizons in spoofed and fake audio detection,”
Massimiliano Todisco, Xin Wang, Ville Vestman, Md. Sahidullah, Hector Delgado, Andreas Nautsch, Junichi Yamagishi, Nicholas Evans, Tomi H Kinnunen, and Kong Aik Lee, · 2019
Earlier work this paper cites.
“CSTR VCTK Corpus: English Multi-speaker Corpus for CSTR Voice Cloning Toolkit (version 0.92),” 2019
Junichi Yamagishi, Christophe Veaux, and Kirsten MacDonald, · 2019
Earlier work this paper cites.
“MLS: A Large-Scale Multilingual Dataset for Speech Research,”
Vineel Pratap, Qiantong Xu, Anuroop Sriram, Gabriel Synnaeve, and Ronan Collobert, · 2020
Earlier work this paper cites.
NIST 2020 CTS Speaker Recognition ChallengeEvaluation Plan
NIST, · 2020
Earlier work this paper cites.
“Tandem assessment of spoofing countermeasures and automatic speaker verification: Fundamentals,”
Tomi Kinnunen, Héctor Delgado, Nicholas Evans, et al., · 2020
Earlier work this paper cites.
“Glow-TTS: A generative flow for text-to-speech via monotonic alignment search,”
Jaehyeon Kim, Sungwon Kim, Jungil Kong, and Sungroh Yoon, · 2020
Earlier work this paper cites.
“HiFi-GAN: Generative adversarial networks for efficient and high fidelity speech synthesis,”
Jungil Kong, Jaehyeon Kim, and Jaekyoung Bae, · 2020
Earlier work this paper cites.
“Voice Conversion Using Speech-to-Speech Neuro-Style Transfer,”
Ehab A. AlBadawy and Siwei Lyu, · 2020
Earlier work this paper cites.
“ECAPA-TDNN: Emphasized Channel Attention, Propagation and Aggregation in TDNN Based Speaker Verification,”
Brecht Desplanques, Jenthe Thienpondt, and Kris Demuynck, · 2020
Cited alongside, same era.
“Wav2vec 2.0: A framework for self-supervised learning of speech representations,”
Alexei Baevski, Yuhao Zhou, Abdelrahman Mohamed, and Michael Auli, · 2020
Cited alongside, same era.
“ASVspoof 2021: Accelerating progress in spoofed and deepfake speech detection,”
Junichi Yamagishi, Xin Wang, Massimiliano Todisco, Md Sahidullah, Jose Patino, Andreas Nautsch, Xuechen Liu, Kong Aik Lee, Tomi Kinnunen, Nicholas Evans, and Hector Delgado, · 2021
Cited alongside, same era.
“Grad-TTS: A diffusion probabilistic model for text-to-speech,”
Vadim Popov, Ivan Vovk, Vladimir Gogoryan, Tasnima Sadekova, and Mikhail Kudinov, · 2021
Cited alongside, same era.
“Fastpitch: Parallel text-to-speech with pitch prediction,”
Adrian Łańcucki, · 2021
Cited alongside, same era.
“ASVspoof 2021: Towards Spoofed and Deepfake Speech Detection in the Wild,”
Xuechen Liu, Xin Wang, Md Sahidullah, Jose Patino, Hector Delgado, Tomi Kinnunen, Massimiliano Todisco, Junichi Yamagishi, Nicholas Evans, Andreas Nautsch, and Kong Aik Lee, · 2023
Later among the works it cites.
“Malafide: a novel adversarial convolutive noise attack against deepfake and spoofing detection systems,”
Michele Panariello, Wanying Ge, Hemlata Tak, Massimiliano Todisco, and Nicholas Evans, · 2023
Later among the works it cites.
Cheng Gong, Xin Wang, Erica Cooper, Dan Wells, Longbiao Wang, Jianwu Dang, Korin Richmond, and Junichi Yamagishi, · 2023
Later among the works it cites.
“Exact Prosody Cloning in Zero-Shot Multispeaker Text-to-Speech,”
Florian Lux, Julia Koch, and Ngoc Thang Vu, · 2023
Later among the works it cites.
“High fidelity neural audio compression,”
Alexandre Défossez, Jade Copet, Gabriel Synnaeve, and Yossi Adi, · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,”
Jaehyeon Kim, Jungil Kong, and Juhee Son, · 2021
Cited alongside, same era.
“StarGANv2-VC: A diverse, unsupervised, non-parallel framework for natural-sounding voice conversion,”
Yinghao Aaron Li, Ali Zare, and Nima Mesgarani, · 2021
Cited alongside, same era.
“End-to-end anti-spoofing with RawNet2,”
Hemlata Tak, Jose Patino, Massimiliano Todisco, Andreas Nautsch, Nicholas Evans, and Anthony Larcher, · 2021
Cited alongside, same era.
“SASV 2022: The First Spoofing-Aware Speaker Verification Challenge,”
Jee-weon Jung, Hemlata Tak, Hye-jin Shim, Hee-Soo Heo, Bong-Jin Lee, Soo-Whan Chung, Ha-Jin Yu, Nicholas Evans, and Tomi Kinnunen, · 2022
Cited alongside, same era.
“Yourtts: Towards zero-shot multi-speaker TTS and zero-shot voice conversion for everyone,”
Edresson Casanova, Julian Weber, Christopher D Shulby, Arnaldo Candido Junior, Eren Gölge, and Moacir A Ponti, · 2022
Cited alongside, same era.
“BigVGAN: A universal neural vocoder with large-scale training,”
Sang-gil Lee, Wei Ping, Boris Ginsburg, Bryan Catanzaro, and Sungroh Yoon, · 2022
Cited alongside, same era.
“Low-resource multilingual and zero-shot multispeaker TTS,”
Florian Lux, Julia Koch, and Ngoc Thang Vu, · 2022
Cited alongside, same era.
Later among the works it cites.
“Towards single integrated spoofing-aware speaker verification embeddings,”
Sung Hwan Mun, Hye-jin Shim, Hemlata Tak, Xin Wang, Xuechen Liu, Md Sahidullah, Myeonghun Jeong, Min Hyun Han, Massimiliano Todisco, Kong Aik Lee, et al., · 2023
Later among the works it cites.
“Spoofed training data for speech spoofing countermeasure can be efficiently created using neural vocoders,”
Xin Wang and Junichi Yamagishi, · 2023
Later among the works it cites.
“a-DCF: an architecture agnostic metric with application to spoofing-robust speaker verification,”
Hye-jin Shim, Jee-weon Jung, Tomi Kinnunen, et al., · 2024
Closest in time.
“t-EER: Parameter-free tandem evaluation of countermeasures and biometric comparators,”
Tomi H. Kinnunen, Kong Aik Lee, Hemlata Tak, et al., · 2024
Closest in time.
“Malacopula: Adversarial automatic speaker verification attacks using a neural-based generalised hammerstein model,”
Massimiliano Todisco, Michele Panariello, Xin Wang, Hector Delgado, Kong-Aik Lee, and Nicholas Evans, · 2024
Closest in time.
“XTTS: A massively multilingual zero-shot text-to-speech model,”
Edresson Casanova, Kelly Davis, Eren Gölge, Görkem Göknar, Iulian Gulea, Logan Hart, Aya Aljafari, Joshua Meyer, Reuben Morais, Samuel Olayemi, et al., · 2024
Closest in time.
“ASVspoof 5 evaluation plan (phase 2),”
Hector Delgado, Nicholas Evans, Jee-weon Jung, Tomi Kinnunen, Ivan Kukanov, Kong-Aik Lee, Xuechen Liu, Hye-jin Shim, Md Sahidullah, Hemlata Tak, Massimiliano Todisco, Xin Wang, and Junichi Yamagishi, · 2024
Closest in time.
“Calibration tutorial,” https://github.com/luferrer/CalibrationTutorial, 2024
Luciana Ferrer, · 2024
Closest in time.
“Revisiting and improving scoring fusion for spoofing-aware speaker verification using compositional data analysis,”
Xin Wang, Tomi Kinnunen, Lee Kong Aik, Paul-Gauthier Noe, and Junichi Yamagishi, · 2024
Closest in time.
“ASVspoof 2015: the first automatic speaker verification spoofing and countermeasures challenge,”
Zhizheng Wu, Tomi Kinnunen, Nicholas Evans, Junichi Yamagishi, Cemal Hanilçi, Md. Sahidullah, and Aleksandr Sizov, · 2041
Closest in time.