Fetching the paper…
Reading the bibliography…
Audio deepfake detection is an emerging topic in the artificial intelligence community.
Asvspoof 2015: the first automatic speaker verification spoofing and countermeasures challenge,
Z. Wu, T. Kinnunen, N. Evans, J. Yamagishi, C. Hanilc¸i, et al., · 2015
Earlier work this paper cites.
Thchs-30: A free chinese speech corpus,
D. Wang, X. Zhang, · 2015
Earlier work this paper cites.
A comparison of features for synthetic speech detection,
M. Sahidullah, T. Kinnunen, C. Hanilçi, · 2015
Earlier work this paper cites.
Deep residual learning for image recognition,
K. He, X. Zhang, S. Ren, J. Sun, · 2016
Earlier work this paper cites.
Towards open set deep networks,
A. Bendale, T. E. Boult, · 2016
Earlier work this paper cites.
The asvspoof 2017 challenge: Assessing the limits of replay spoofing attack detection,
T. Kinnunen, M. Sahidullah, H. Delgado, N. E. M. Todisco, et al., · 2017
Earlier work this paper cites.
Aishell-1: An open-source mandarin speech corpus and a speech recognition baseline,
H. Bu, J. Du, X. Na, B. Wu, H. Zheng, · 2017
Earlier work this paper cites.
The Voice Conversion Challenge 2018: Promoting Development of Parallel and Nonparallel Methods ,
J. Lorenzo-Trueba, J. Yamagishi, T. Toda, D. Saito, F. Villavicencio, T. Kinnunen, Z. Ling, · 2018
Earlier work this paper cites.
Asvspoof 2019: Future horizons in spoofed and fake audio detection,
M. Todisco, X. Wang, V. Vestman, M. Sahidullah, K. Lee, · 2019
Earlier work this paper cites.
An overview of voice conversion and its challenges: From statistical modeling to deep learning,
B. Sisman, J. Yamagishi, S. King, H. Li, · 2020
Cited alongside, same era.
Voice Conversion Challenge 2020 –- Intra-lingual semi-parallel and cross-lingual voice conversion –-,
Z. Yi, W.-C. Huang, X. Tian, J. Yamagishi, R. K. Das, T. Kinnunen, Z.-H. Ling, T. Toda, · 2020
Cited alongside, same era.
Aishell-3: A multi-speaker mandarin tts corpus and the baselines,
Y. Shi, H. Bu, X. Xu, S. Zhang, M. Li, · 2020
Cited alongside, same era.
Recent advances in open set recognition: A survey,
C. Geng, S.-j. Huang, S. Chen, · 2020
Cited alongside, same era.
Light convolutional neural network with feature genuinization for detection of synthetic speech attacks,
Z. Wu, R. K. Das1, J. Yang, H. Li, · 2020
Cited alongside, same era.
Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,
J. Kim, J. Kong, J. Son, · 2021
Later among the works it cites.
Remember the ‘deepfake cheerleader mom’? prosecutors now admit they can’t prove fake-video claims,
D. Harwell, · 2021
Later among the works it cites.
Asvspoof 2021: accelerating progress in spoofed and deepfake speech detection,
J. Yamagishi, X. Wang, M. Todisco, M. Sahidullah, J. Patino, A. Nautsch, X. Liu, K. A. Lee, T. Kinnunen, N. Evans, · 2021
Later among the works it cites.
Half-truth: A partially fake audio detection dataset,
J. Yi, Y. Bai, J. Tao, H. Ma, Z. Tian, C. Wang, T. Wang, R. Fu, · 2021
Later among the works it cites.
Continual learning for fake audio detection,
H. Ma, J. Yi, J. Tao, Y. Bai, Z. Tian, C. Wang, · 2021
Later among the works it cites.
Dfgc 2021: A deepfake game competition,
B. Peng, H. Fan, W. Wang, J. Dong, Y. Li, S. Lyu, Q. Li, Z. Sun, H. Chen, B. Chen, et al., · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
wav2vec 2.0: A framework for self-supervised learning of speech representations,
A. Baevski, Y. Zhou, A. Mohamed, M. Auli, · 2020
Cited alongside, same era.
A survey on neural speech synthesis,
X. Tan, T. Qin, F. Soong, T.-Y. Liu, · 2021
Cited alongside, same era.
Grad-tts: A diffusion probabilistic model for text-to-speech,
V. Popov, I. Vovk, V. Gogoryan, T. Sadekova, M. Kudinov, · 2021
Cited alongside, same era.
An initial investigation for detecting vocoder fingerprints of fake audio,
X. Yan, J. Yi, J. Tao, C. Wang, H. Ma, T. Wang, S. Wang, R. Fu,
Cited in the paper.
System fingerprints detection for deepfake audio: An initial dataset and investigation,
X. Yan, J. Yi, J. Tao, C. Wang, H. Ma, Z. Tian, R. Fu,
Cited in the paper.
Later among the works it cites.
Add 2022: the first audio deep synthesis detection challenge,
J. Yi, R. Fu, J. Tao, S. Nie, H. Ma, C. Wang, T. Wang, Z. Tian, Y. Bai, C. Fan, et al., · 2022
Later among the works it cites.
Dfgc 2022: The second deepfake game competition,
B. Peng, W. Xiang, Y. Jiang, W. Wang, J. Dong, Z. Sun, Z. Lei, S. Lyu, · 2022
Later among the works it cites.