Fetching the paper…
Reading the bibliography…
We present our analysis of a significant data artifact in the official 2019/2021 ASVspoof Challenge Dataset.
Comparison of parametric representations for monosyllabic word recognition in continuously spoken sentences
Steven Davis and Paul Mermelstein · 1980
Earlier work this paper cites.
Calculation of a constant q spectral transform
Judith C Brown · 1991
Earlier work this paper cites.
Mel-cepstral distance measure for objective speech quality assessment
Robert Kubichek · 1993
Earlier work this paper cites.
Detecting converted speech and natural speech for anti-spoofing attack in speaker recognition
Zhizheng Wu, Eng Siong Chng, and Haizhou Li · 2012
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
A comparison of features for synthetic speech detection
Md Sahidullah, Tomi Kinnunen, and Cemal Hanilçi · 2015
Earlier work this paper cites.
Superseded-cstr vctk corpus: English multi-speaker corpus for cstr voice cloning toolkit
Christophe Veaux, Junichi Yamagishi, Kirsten MacDonald, et al · 2017
Earlier work this paper cites.
Montreal forced aligner: Trainable text-speech alignment using kaldi
Michael McAuliffe, Michaela Socolof, Sarah Mihuc, Michael Wagner, and Morgan Sonderegger · 2017
Earlier work this paper cites.
A spoofing benchmark for the 2018 voice conversion challenge: Leveraging from spoofing countermeasures for speech artifact assessment
Tomi Kinnunen, Jaime Lorenzo-Trueba, Junichi Yamagishi, Tomoki Toda, Daisuke Saito, Fernando Villavicencio, and Zhenhua Ling · 2018
Earlier work this paper cites.
The voice conversion challenge 2018: Promoting development of parallel and nonparallel methods
Jaime Lorenzo-Trueba, Junichi Yamagishi, Tomoki Toda, Daisuke Saito, Fernando Villavicencio, Tomi Kinnunen, and Zhenhua Ling · 2018
Cited alongside, same era.
Style tokens: Unsupervised style modeling, control and transfer in end-to-end speech synthesis
Yuxuan Wang, Daisy Stanton, Yu Zhang, RJ-Skerry Ryan, Eric Battenberg, Joel Shor, Ying Xiao, Ye Jia, Fei Ren, and Rif A Saurous · 2018
Cited alongside, same era.
Transfer learning from speaker verification to multispeaker text-to-speech synthesis
Ye Jia, Yu Zhang, Ron J Weiss, Quan Wang, Jonathan Shen, Fei Ren, Zhifeng Chen, Patrick Nguyen, Ruoming Pang, Ignacio Lopez Moreno, et al · 2018
Cited alongside, same era.
Asvspoof 2019: The 3rd automatic speaker verification spoofing and countermeasures challenge database
Junichi Yamagishi, Massimiliano Todisco, Md Sahidullah, Héctor Delgado, Xin Wang, Nicolas Evans, Tomi Kinnunen, Kong Aik Lee, Ville Vestman, and Andreas Nautsch · 2019
Cited alongside, same era.
Recurrent convolutional structures for audio spoof and video deepfake detection
Akash Chintha, Bao Thai, Saniat Javid Sohrawardi, Kartavya Bhatt, Andrea Hickerson, Matthew Wright, and Raymond Ptucha · 2020
Later among the works it cites.
Densely connected convolutional network for audio spoofing detection
Zheng Wang, Sanshuai Cui, Xiangui Kang, Wei Sun, and Zhonghua Li · 2020
Later among the works it cites.
Voice biometric system security: Design and analysis of countermeasures for replay attacks
Bhusan Chettri · 2020
Later among the works it cites.
https://www.forbes.com/sites/jessedamiani/2019/09/03/a-voice-deepfake-was-used-to-scam-a-ceo-out-of-243000
A voice deepfake was used to scam a ceo out of $243,000 · 2021
Closest in time.
https://www.asvspoof.org/index2019.html
Asvspoof 2019: Automatic speaker verification spoofing and countermeasures challenge · 2021
Closest in time.
https://www.asvspoof.org/index2019.html
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Unmasking clever hans predictors and assessing what machines really learn
Sebastian Lapuschkin, Stephan Wäldchen, Alexander Binder, Grégoire Montavon, Wojciech Samek, and Klaus-Robert Müller · 2019
Cited alongside, same era.
Deep residual neural networks for audio spoofing detection
Moustafa Alzantot, Ziqi Wang, and Mani B Srivastava · 2019
Cited alongside, same era.
Ensemble models for spoofing detection in automatic speaker verification
Bhusan Chettri, Daniel Stoller, Veronica Morfi, Marco A Martínez Ramírez, Emmanouil Benetos, and Bob L Sturm · 2019
Cited alongside, same era.
Asvspoof 2019: A large-scale public database of synthesized, converted and replayed speech
Xin Wang, Junichi Yamagishi, Massimiliano Todisco, Héctor Delgado, Andreas Nautsch, Nicholas Evans, Md Sahidullah, Ville Vestman, Tomi Kinnunen, Kong Aik Lee, et al · 2020
Cited alongside, same era.
Wavenet: A generative model for raw audio
Aäron van den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew Senior, and Koray Kavukcuoglu
Cited in the paper.
Asvspoof 2021: Automatic speaker verification spoofing and countermeasures challenge · 2021
Closest in time.
End-to-end anti-spoofing with rawnet2
Hemlata Tak, Jose Patino, Massimiliano Todisco, Andreas Nautsch, Nicholas Evans, and Anthony Larcher · 2021
Closest in time.
librosa/librosa: 0.8.1rc2, May 2021
Brian McFee et Al · 2021
Closest in time.