Fetching the paper…
Reading the bibliography…
Audio deepfakes pose significant threats, including impersonation, fraud, and reputation damage.
Constant q cepstral coefficients: A spoofing countermeasure for automatic speaker verification
Massimiliano Todisco, Héctor Delgado, and Nicholas Evans · 2017
Earlier work this paper cites.
Csi-net: Unified body characterization and action recognition
Fei Wang, Jinsong Han, Shiyuan Zhang, Xu He, and Dong Huang · 2018
Earlier work this paper cites.
A light cnn for deep face representation with noisy labels
Xiang Wu, Ran He, Zhenan Sun, and Tieniu Tan · 2018
Earlier work this paper cites.
Jee-weon Jung, Hee-Soo Heo, Ju-ho Kim, Hye-jin Shim, and Ha-Jin Yu · 2019
Earlier work this paper cites.
Sentence-bert: Sentence embeddings using siamese bert-networks
N Reimers · 2019
Earlier work this paper cites.
Asvspoof 2019: Future horizons in spoofed and fake audio detection
Massimiliano Todisco, Xin Wang, Ville Vestman, Md Sahidullah, Héctor Delgado, Andreas Nautsch, Junichi Yamagishi, Nicholas Evans, Tomi Kinnunen, and Kong Aik Lee · 2019
Earlier work this paper cites.
Anti-forensic against double jpeg compression detection using adversarial generative network
Kutub Uddin, Yoonmo Yang, and Byung Tae Oh · 2019
Earlier work this paper cites.
wav2vec 2.0: A framework for self-supervised learning of speech representations
Alexei Baevski, Yuhao Zhou, Abdelrahman Mohamed, and Michael Auli · 2020
Earlier work this paper cites.
Holmes: health online model ensemble serving for deep learning models in intensive care units
Shenda Hong, Yanbo Xu, Alind Khare, Satria Priambada, Kevin Maher, Alaa Aljiffry, Jimeng Sun, and Alexey Tumanov · 2020
Earlier work this paper cites.
Fairseq s2t: Fast speech-to-text modeling with fairseq
Changhan Wang, Yun Tang, Xutai Ma, Anne Wu, Sravya Popuri, Dmytro Okhonko, and Juan Pino · 2020
Earlier work this paper cites.
Asvspoof 2019: A large-scale public database of synthesized, converted and replayed speech
Xin Wang, Junichi Yamagishi, Massimiliano Todisco, Héctor Delgado, Andreas Nautsch, Nicholas Evans, Md Sahidullah, Ville Vestman, Tomi Kinnunen, Kong Aik Lee, et al · 2020
Earlier work this paper cites.
Deep4snet: deep learning for fake speech classification
Dora M Ballesteros, Yohanna Rodriguez-Ortega, Diego Renza, and Gonzalo Arce · 2021
Earlier work this paper cites.
Wavefake: A data set to facilitate audio deepfake detection
Joel Frank and Lea Schönherr · 2021
Earlier work this paper cites.
Towards end-to-end synthetic speech detection
Guang Hua, Andrew Beng Jin Teoh, and Haijian Zhang · 2021
Earlier work this paper cites.
Cross-modal speaker verification and recognition: A multilingual perspective
Shah Nawaz, Muhammad Saad Saeed, Pietro Morerio, Arif Mahmood, Ignazio Gallo, Muhammad Haroon Yousaf, and Alessio Del Bue · 2021
Cited alongside, same era.
End-to-end anti-spoofing with rawnet2
Hemlata Tak, Jose Patino, Massimiliano Todisco, Andreas Nautsch, Nicholas Evans, and Anthony Larcher · 2021
Cited alongside, same era.
Analysis of generative adversarial network targeting anti-forensic in jpeg compressed domain
Kutub Uddin, Yoonmo Yang, and Byung Tae Oh · 2021
Cited alongside, same era.
Asvspoof 2021: accelerating progress in spoofed and deepfake speech detection
Junichi Yamagishi, Xin Wang, Massimiliano Todisco, Md Sahidullah, Jose Patino, Andreas Nautsch, Xuechen Liu, Kong Aik Lee, Tomi Kinnunen, Nicholas Evans, et al · 2021
Cited alongside, same era.
Fake speech detection using residual network with transformer encoder
Zhenyu Zhang, Xiaowei Yi, and Xianfeng Zhao · 2021
Cited alongside, same era.
Spoof detection using voice contribution on lfcc features and resnet-34
Khaing Zar Mon, Kasorn Galajit, Candy Olivia Mawalim, Jessada Karnjana, Tsuyoshi Isshiki, and Pakinee Aimmanee · 2023
Later among the works it cites.
Speaker recognition in realistic scenario using multimodal data
Saqlain Hussain Shah, Muhammad Saad Saeed, Shah Nawaz, and Muhammad Haroon Yousaf · 2023
Later among the works it cites.
A robust open-set multi-instance learning for defending adversarial attacks in digital image
Kutub Uddin, Yoonmo Yang, Tae Hyun Jeong, and Byung Tae Oh · 2023
Later among the works it cites.
To-rawnet: improving rawnet with tcn and orthogonal regularization for fake audio detection
Chenglong Wang, Jiangyan Yi, Jianhua Tao, Chuyuan Zhang, Shuai Zhang, Ruibo Fu, and Xun Chen · 2023
Later among the works it cites.
An end-to-end multi-module audio deepfake generation system for add challenge 2023
Sheng Zhao, Qilong Yuan, Yibo Duan, and Zhuoyue Chen · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Deepfake audio detection via mfcc features using machine learning
Ameer Hamza, Abdul Rehman Rehman Javed, Farkhund Iqbal, Natalia Kryvinska, Ahmad S Almadhor, Zunera Jalil, and Rouba Borghol · 2022
Cited alongside, same era.
Aasist: Audio anti-spoofing using integrated spectro-temporal graph attention networks
Jee-weon Jung, Hee-Soo Heo, Hemlata Tak, Hye-jin Shim, Joon Son Chung, Bong-Jin Lee, Ha-Jin Yu, and Nicholas Evans · 2022
Cited alongside, same era.
Defense against adversarial attacks on audio deepfake detection
Piotr Kawa, Marcin Plata, and Piotr Syga · 2022
Cited alongside, same era.
Specrnet: Towards faster and more accessible audio deepfake detection
Piotr Kawa, Marcin Plata, and Piotr Syga · 2022
Cited alongside, same era.
Does audio deepfake detection generalize?
Nicolas M Müller, Pavel Czempin, Franziska Dieckmann, Adam Froghyar, and Konstantin Böttinger · 2022
Cited alongside, same era.
Fully automated end-to-end fake audio detection
Chenglong Wang, Jiangyan Yi, Jianhua Tao, Haiyang Sun, Xun Chen, Zhengkun Tian, Haoxin Ma, Cunhang Fan, and Ruibo Fu · 2022
Cited alongside, same era.
Securing voice biometrics: One-shot learning approach for audio deepfake detection
Awais Khan and Khalid Mahmood Malik · 2023
Cited alongside, same era.
Later among the works it cites.
Rawbmamba: End-to-end bidirectional state space model for audio deepfake detection
Yujie Chen, Jiangyan Yi, Jun Xue, Chenglong Wang, Xiaohui Zhang, Shunbo Dong, Siding Zeng, Jianhua Tao, Lv Zhao, and Cunhang Fan · 2024
Later among the works it cites.
Securing social media against deepfakes using identity, behavioral, and geometric signatures
Muhammad Umar Farooq, Awais Khan, Ijaz Ul Haq, and Khalid Mahmood Malik · 2024
Later among the works it cites.
Frame-to-utterance convergence: A spectra-temporal approach for unified spoofing detection
Awais Khan, Khalid Mahmood Malik, and Shah Nawaz · 2024
Later among the works it cites.
Audio-deepfake detection: Adversarial attacks and countermeasures
Mouna Rabhi, Spiridon Bakiras, and Roberto Di Pietro · 2024
Later among the works it cites.
Counter-act against gan-based attacks: A collaborative learning approach for anti-forensic detection
Kutub Uddin, Tae Hyun Jeong, and Byung Tae Oh · 2024
Later among the works it cites.
Clad: Robust audio deepfake detection against manipulation attacks with contrastive learning
Haolin Wu, Jing Chen, Ruiying Du, Cong Wu, Kun He, Xingcan Shang, Hao Ren, and Guowen Xu · 2024
Later among the works it cites.
I can hear you: Selective robust training for deepfake audio detection
Zirui Zhang, Wei Hao, Aroon Sankoh, William Lin, Emanuel Mendiola-Ortiz, Junfeng Yang, and Chengzhi Mao · 2024
Later among the works it cites.
A transferable anti-forensic attack on forensic cnns using a generative adversarial network
Xinwei Zhao, Chen Chen, and Matthew C Stamm · 2025
Closest in time.