Fetching the paper…
Reading the bibliography…
Deepfake speech represents a real and growing threat to systems and society.
An Overview on Audio, Signal, Speech, & Language Processing for COVID-19
Gauri Deshpande and Björn Schuller. 2020 · 2005
Earlier work this paper cites.
Monitor alarm fatigue: an integrative review
Maria Cvach. 2012 · 2012
Earlier work this paper cites.
Automatic detection of inhalation breath pauses for improved pause modelling in HMM-TTS
Norbert Braunschweiler and Langzhou Chen. 2013 · 2013
Earlier work this paper cites.
ASVspoof 2015: the first automatic speaker verification spoofing and countermeasures challenge. In Sixteenth annual conference of the international speech communication association
Zhizheng Wu, Tomi Kinnunen, Nicholas Evans, Junichi Yamagishi, Cemal Hanilçi, Md Sahidullah, and Aleksandr Sizov. 2015 · 2015
Earlier work this paper cites.
The ASVspoof 2017 challenge: Assessing the limits of replay spoofing attack detection
Tomi Kinnunen, Md Sahidullah, Héctor Delgado, Massimiliano Todisco, Nicholas Evans, Junichi Yamagishi, and Kong Aik Lee. 2017 · 2017
Earlier work this paper cites.
Can we steal your vocal identity from the Internet?: Initial investigation of cloning Obama’s voice using GAN, WaveNet and low-quality found data. In Speaker and Language Recognition Workshop
Jaime Lorenzo-Trueba, Fuming Fang, Xin Wang, Isao Echizen, Junichi Yamagishi, and Tomi Kinnunen. 2018 · 2018
Earlier work this paper cites.
Tuning model parameters in class-imbalanced learning with precision-recall curve
Guang-Hui Fu, Lun-Zhao Yi, and Jianxin Pan. 2019 · 2019
Earlier work this paper cites.
Deep sensing of breathing signal during conversational speech
Venkata Srikanth Nallanthighal and H Strik. 2019 · 2019
Earlier work this paper cites.
Deepfakes and cheap fakes
Britt Paris and Joan Donovan. 2019 · 2019
Earlier work this paper cites.
Detecting Deep Fakes With Mice : Machines vs Biology. In Black Hat USA
Jonathan Saunders. 2019 · 2019
Earlier work this paper cites.
Robust performance metrics for authentication systems. In Network and Distributed Systems Security (NDSS) Symposium 2019
Shridatt Sugrim, Can Liu, Meghan McLean, and Janne Lindqvist. 2019 · 2019
Earlier work this paper cites.
Casting to corpus: Segmenting and selecting spontaneous dialogue for TTS with a CNN-LSTM speaker-dependent breath detector. In ICASSP 2019-2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 6925–6929
Éva Székely, Gustav Eje Henter, and Joakim Gustafson. 2019 · 2019
Earlier work this paper cites.
ASVspoof 2019: Future horizons in spoofed and fake audio detection
Massimiliano Todisco, Xin Wang, Ville Vestman, Md Sahidullah, Héctor Delgado, Andreas Nautsch, Junichi Yamagishi, Nicholas Evans, Tomi Kinnunen, and Kong Aik Lee. 2019 · 2019
Cited alongside, same era.
Breathing and Speech Planning in Spontaneous Speech Synthesis. In ICASSP 2020 - 2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . 7649–7653
Éva Székely, Gustav Eje Henter, Jonas Beskow, and Joakim Gustafson. 2020 · 2020
Cited alongside, same era.
Take a Breath: Respiratory Sounds Improve Recollection in Synthetic Speech
2021 · 2021
Cited alongside, same era.
Sok: The faults in our asrs: An overview of attacks against automatic speech recognition and speaker identification systems. In 2021 IEEE symposium on security and privacy (SP) . IEEE, 730–747
Hadi Abdullah, Kevin Warren, Vincent Bindschaedler, Nicolas Papernot, and Patrick Traynor. 2021 · 2021
Cited alongside, same era.
Combining automatic speaker verification and prosody analysis for synthetic speech detection
Luigi Attorresi, Davide Salvi, Clara Borrelli, Paolo Bestagini, and Stefano Tubaro. 2022 · 2022
Later among the works it cites.
Who Are You (I Really Wanna Know)? Detecting Audio { \{ DeepFakes } \} Through Vocal Tract Reconstruction. In 31st USENIX Security Symposium (USENIX Security 22) . 2691–2708
Logan Blue, Kevin Warren, Hadi Abdullah, Cassidy Gibson, Luis Vargas, Jessica O’Dell, Kevin Butler, and Patrick Traynor. 2022 · 2022
Later among the works it cites.
SASV 2022: The first spoofing-aware speaker verification challenge
Jee-weon Jung, Hemlata Tak, Hye-jin Shim, Hee-Soo Heo, Bong-Jin Lee, Soo-Whan Chung, Ha-Jin Yu, Nicholas Evans, and Tomi Kinnunen. 2022 · 2022
Later among the works it cites.
On breathing pattern information in synthetic speech. In Proc. Interspeech
Zohreh Mostaani and Mathew Magimai Doss. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Niko Brümmer, Luciana Ferrer, and Albert Swart. 2021 · 2021
Cited alongside, same era.
Asvspoof 2021: Automatic speaker verification spoofing and countermeasures challenge evaluation plan
Héctor Delgado, Nicholas Evans, Tomi Kinnunen, Kong Aik Lee, Xuechen Liu, Andreas Nautsch, Jose Patino, Md Sahidullah, Massimiliano Todisco, Xin Wang, et al · 2021
Cited alongside, same era.
Do deepfakes feel emotions? A semantic approach to detecting deepfakes via emotional inconsistencies. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 1013–1022
Brian Hosler, Davide Salvi, Anthony Murray, Fabio Antonacci, Paolo Bestagini, Stefano Tubaro, and Matthew C Stamm. 2021 · 2021
Cited alongside, same era.
Speech is silver, silence is golden: What do ASVspoof-trained models really learn?
Nicolas M Müller, Franziska Dieckmann, Pavel Czempin, Roman Canals, Konstantin Böttinger, and Jennifer Williams. 2021 · 2021
Cited alongside, same era.
Deep learning architectures for estimating breathing signal and respiratory parameters from speech recordings
Venkata Srikanth Nallanthighal, Zohreh Mostaani, Aki Härmä, Helmer Strik, and Mathew Magimai-Doss. 2021 · 2021
Cited alongside, same era.
ASVspoof 2021: accelerating progress in spoofed and deepfake speech detection
Junichi Yamagishi, Xin Wang, Massimiliano Todisco, Md Sahidullah, Jose Patino, Andreas Nautsch, Xuechen Liu, Kong Aik Lee, Tomi Kinnunen, Nicholas Evans, et al · 2021
Cited alongside, same era.
Dos and don’ts of machine learning in computer security. In 31st USENIX Security Symposium (USENIX Security 22) . 3971–3988
Daniel Arp, Erwin Quiring, Feargus Pendlebury, Alexander Warnecke, Fabio Pierazzi, Christian Wressnegger, Lorenzo Cavallaro, and Konrad Rieck. 2022 · 2022
Cited alongside, same era.
Deep convolutional recurrent neural network for rare acoustic event detection
Shahin Amiriparian and Björn Schuller. [n. d.]
Cited in the paper.
Hemlata Tak, Massimiliano Todisco, Xin Wang, Jee-weon Jung, Junichi Yamagishi, and Nicholas Evans. 2022 · 2022
Later among the works it cites.
Add 2022: the first audio deep synthesis detection challenge. In ICASSP 2022-2022 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 9216–9220
Jiangyan Yi, Ruibo Fu, Jianhua Tao, Shuai Nie, Haoxin Ma, Chenglong Wang, Tao Wang, Zhengkun Tian, Ye Bai, Cunhang Fan, et al · 2022
Later among the works it cites.
BTS-E: Audio Deepfake Detection Using Breathing-Talking-Silence Encoder. In ICASSP 2023-2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 1–5
Thien-Phuc Doan, Long Nguyen-Vu, Souhwan Jung, and Kihun Hong. 2023 · 2023
Later among the works it cites.
"Get in Researchers; We’re Measuring Reproducibility": A Reproducibility Study of Machine Learning Papers in Tier 1 Security Conferences. In Proceedings of the 2023 ACM SIGSAC Conference on Computer and Communications Security . 3433–3459
Daniel Olszewski, Allison Lu, Carson Stillman, Kevin Warren, Cole Kitroser, Alejandro Pascual, Divyajyoti Ukirde, Kevin Butler, and Patrick Traynor. 2023 · 2023
Later among the works it cites.
SoK: The Good, The Bad, and The Unbalanced: Measuring Structural Limitations of Deepfake Datasets. In Proceedings of the USENIX Security Symposium (Security)
Seth Layton, Tyler Tucker, Daniel Olszewski, Kevin Warren, Kevin Butler, and Patrick Traynor. 2024 · 2024
Closest in time.
Better Be Computer or I’m Dumb": A Large-Scale Evaluation of Humans as Audio Deepfake Detectors. In Proceedings of the ACM Conference on Computer and Communications Security (CCS)
Kevin Warren, Tyler Tucker, Anna Crowder, Daniel Olszewski, Allison Lu, Caroline Fedele, Magdalena Pasternak, Seth Layton, Kevin Butler, Carrie Gates, and Patrick Traynor. 2024 · 2024
Closest in time.
The INTERSPEECH 2020 Computational Paralinguistics Challenge: Elderly Emotion, Breathing & Masks. In Interspeech 2020 . ISCA, 2042–2046
Björn W. Schuller, Anton Batliner, Christian Bergler, Eva-Maria Messner, Antonia Hamilton, Shahin Amiriparian, Alice Baird, Georgios Rizos, Maximilian Schmitt, Lukas Stappen, Harald Baumeister, Alexis Deighton MacIntyre, and Simone Hantke. 2020 · 2046
Closest in time.