Fetching the paper…
Reading the bibliography…
A speech spoofing countermeasure (CM) that discriminates between unseen spoofed and bona fide data requires diverse training data.
“A Cross-vocoder Study of Speaker Independent Synthetic Speech Detection Using Phase Information,”
Jon Sanchez, Ibon Saratxaga, Inma Hernaez, Eva Navas, and Daniel Erro, · 2014
Earlier work this paper cites.
“Adam: A method for stochastic optimization,”
Diederik P Kingma and Jimmy Ba, · 2014
Earlier work this paper cites.
“Spoofing and Countermeasures for Speaker Verification: A survey,”
Zhizheng Wu, Nicholas Evans, Tomi Kinnunen, Junichi Yamagishi, Federico Alegre, and Haizhou Li, · 2015
Earlier work this paper cites.
“Joint Speaker Verification and Antispoofing in the i-vector Space,”
Aleksandr Sizov, Elie Khoury, Tomi Kinnunen, Zhizheng Wu, and Sébastien Marcel, · 2015
Earlier work this paper cites.
“VoxCeleb2: Deep speaker recognition,”
Joon Son Chung, Arsha Nagrani, and Andrew Zisserman, · 2018
Earlier work this paper cites.
“WaveGlow: A Flow-based Generative Network for Speech Synthesis,”
Ryan Prenger, Rafael Valle, and Bryan Catanzaro, · 2019
Earlier work this paper cites.
“FairSeq: A fast, extensible toolkit for sequence modeling,”
Myle Ott, Sergey Edunov, Alexei Baevski, Angela Fan, Sam Gross, Nathan Ng, David Grangier, and Michael Auli, · 2019
Earlier work this paper cites.
“ASVspoof 2019: A Large-scale Public Database of Synthesized, Converted and Replayed Speech,”
Xin Wang, Junichi Yamagishi, Massimiliano Todisco, and Others, · 2020
Earlier work this paper cites.
“HiFi-GAN: Generative Adversarial Networks for Efficient and High Fidelity Speech Synthesis,”
Jungil Kong, Jaehyeon Kim, and Jaekyoung Bae, · 2020
Earlier work this paper cites.
“Neural Source-Filter Waveform Models for Statistical Parametric Speech Synthesis,”
Xin Wang, Shinji Takaki, and Junichi Yamagishi, · 2020
Earlier work this paper cites.
“Wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations,”
Alexei Baevski, Yuhao Zhou, Abdelrahman Mohamed, and Michael Auli, · 2020
Cited alongside, same era.
“Speech is Silver, Silence is Golden: What do ASVspoof-trained Models Really Learn?,”
Nicolas Müller, Franziska Dieckmann, Pavel Czempin, Roman Canals, Konstantin Böttinger, and Jennifer Williams, · 2021
Cited alongside, same era.
“The Effect of Silence and Dual-Band Fusion in Anti-Spoofing System,”
Yuxiang Zhang, Wenchao Wang, and Pengyuan Zhang, · 2021
Cited alongside, same era.
“WaveFake: A Data Set to Facilitate Audio DeepFake Detection,”
Joel Frank and Lea Schönherr, · 2021
Cited alongside, same era.
“Unsupervised Cross-Lingual Representation Learning for Speech Recognition,”
Alexis Conneau, Alexei Baevski, Ronan Collobert, Abdelrahman Mohamed, and Michael Auli, · 2021
Cited alongside, same era.
“Self-Supervised Speech Representation Learning: A Review,”
“Does Audio Deepfake Detection Generalize?,”
Nicolas M Müller, Pavel Czempin, Franziska Dieckmann, Adam Froghyar, and Konstantin Böttinger, · 2022
Later among the works it cites.
“RawBoost: A Raw Data Boosting and Augmentation Method applied to Automatic Speaker Verification Anti-Spoofing,”
Hemlata Tak, Madhu R Kamble, Jose Patino, Massimiliano Todisco, and Nicholas W D Evans, · 2022
Later among the works it cites.
“Deepfake audio detection by speaker verification,”
Alessandro Pianese, Davide Cozzolino, Giovanni Poggi, and Luisa Verdoliva, · 2022
Later among the works it cites.
“Spoofed Training Data for Speech Spoofing Countermeasure Can Be Efficiently Created Using Neural Vocoders,”
Xin Wang and Junichi Yamagishi, · 2023
Closest in time.
“AI-Synthesized Voice Detection Using Neural Vocoder Artifacts,”
Chengzhe Sun, Shan Jia, Shuwei Hou, and Siwei Lyu, · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Abdel-Rahman Mohamed, Hung-yi Lee, Lasse Borgholt, Jakob D Havtorn, Joakim Edin, Christian Igel, Katrin Kirchhoff, Shang-Wen Li, Karen Livescu, Lars Maaloe, Tara N Sainath, and Shinji Watanabe, · 2022
Cited alongside, same era.
“Automatic speaker verification spoofing and deepfake detection using wav2vec 2.0 and data augmentation,”
Hemlata Tak, Massimiliano Todisco, Xin Wang, Jee-weon Jung, Junichi Yamagishi, and Nicholas Evans, · 2022
Cited alongside, same era.
“Investigating Self-Supervised Front Ends for Speech Spoofing Countermeasures,”
Xin Wang and Junichi Yamagishi, · 2022
Cited alongside, same era.
“The Vicomtech Audio Deepfake Detection System Based on Wav2vec2 for the 2022 ADD Challenge,”
Juan M Martin-Donas and Aitor Alvarez, · 2022
Cited alongside, same era.
“Towards Single Integrated Spoofing-aware Speaker Verification Embeddings,”
Sung Hwan Mun, Hye-jin Shim, Hemlata Tak, Xin Wang, Xuechen Liu, Md Sahidullah, Myeonghun Jeong, Min Hyun Han, Massimiliano Todisco, Kong Aik Lee, and Others, · 2023
Closest in time.
“Improving Generalization Ability of Countermeasures for New Mismatch Scenario by Combining Multiple Advanced Regularization Terms,”
Chang Zeng, Xin Wang, Xiaoxiao Miao, Erica Cooper, and Junichi Yamagishi, · 2023
Closest in time.
“ASVspoof 2021: Towards Spoofed and Deepfake Speech Detection in the Wild,”
Xuechen Liu, Xin Wang, Md Sahidullah, Jose Patino, Héctor Delgado, Tomi Kinnunen, Massimiliano Todisco, Junichi Yamagishi, Nicholas Evans, Andreas Nautsch, and Kong Aik Lee, · 2023
Closest in time.
“Investigating active-learning-based training data selection for speech spoofing countermeasure,”
Xin Wang and Junichi Yamagishi, · 2023
Closest in time.