Fetching the paper…
Reading the bibliography…
Multimodal large language models (MLLMs), which bridge the gap between audio-visual and natural language processing, achieve state-of-the-art performance on several audio-visual tasks.
Roy, N., Hassanieh, H., Roy Choudhury, R.: Backdoor: Making microphones hear inaudible sounds. In: Proceedings of the 15th Annual International Conference on Mobile Systems, Applications, and Services. pp. 2–14 (2017)
2017
Earlier work this paper cites.
Carlini, N., Wagner, D.: Audio adversarial examples: Targeted attacks on speech-to-text. In: 2018 IEEE Security and Privacy Workshops (SPW). pp. 1–7. IEEE (2018)
2018
Earlier work this paper cites.
2019
Earlier work this paper cites.
Esmaeilpour, M., Cardinal, P., Koerich, A.L.: A robust approach for securing audio classification against adversarial attacks. IEEE Transactions on Information Forensics and Security 15
2019
Earlier work this paper cites.
Gao, Y., Xu, C., Wang, D., Chen, S., Ranasinghe, D.C., Nepal, S.: Strip: A defence against trojan attacks on deep neural networks. In: Proceedings of the 35th annual computer security applications conference. pp. 113–125 (2019)
2019
Earlier work this paper cites.
Gong, T., Ramos, A.G.C., Bhattacharya, S., Mathur, A., Kawsar, F.: Audidos: Real-time denial-of-service adversarial attacks on deep audio models. In: 2019 18th IEEE International Conference On Machine Learning And Applications (ICMLA). pp. 978–985. IEEE (2019)
2019
Earlier work this paper cites.
Kwon, H., Kim, Y., Yoon, H., Choi, D.: Selective audio adversarial example in evasion attack on speech recognition system. IEEE Transactions on Information Forensics and Security 15
2019
Earlier work this paper cites.
Kwon, H., Yoon, H., Park, K.W.: Poster: Detecting audio adversarial example through audio modification. In: Proceedings of the 2019 ACM SIGSAC Conference on Computer and Communications Security. pp. 2521–2523 (2019)
2019
Earlier work this paper cites.
Li, J., Qu, S., Li, X., Szurley, J., Kolter, J.Z., Metze, F.: Adversarial music: Real world audio adversary against wake-word detection system. Advances in Neural Information Processing Systems 32
2019
Earlier work this paper cites.
Taori, R., Kamsetty, A., Chu, B., Vemuri, N.: Targeted adversarial examples for black box audio systems. In: Proceedings of the 2019 IEEE security and privacy workshops (SPW). pp. 15–20. IEEE (2019)
2019
Earlier work this paper cites.
Chang, K.H., Huang, P.H., Yu, H., Jin, Y., Wang, T.C.: Audio adversarial examples generation with recurrent neural networks. In: 2020 25th Asia and South Pacific Design Automation Conference (ASP-DAC). pp. 488–493. IEEE (2020)
2020
Earlier work this paper cites.
Du, X., Pun, C.M., Zhang, Z.: A unified framework for detecting audio adversarial examples. In: Proceedings of the 28th ACM International Conference on Multimedia. pp. 3986–3994 (2020)
2020
Earlier work this paper cites.
Kong, Y., Zhang, J.: Adversarial audio: A new information hiding method. In: Proc. Interspeech 2020. pp. 2287–2291 (2020)
2020
Earlier work this paper cites.
Li, Z., Shi, C., Xie, Y., Liu, J., Yuan, B., Chen, Y.: Practical adversarial attacks against speaker recognition systems. In: Proceedings of the 21st international workshop on mobile computing systems and applications. pp. 9–14 (2020)
2020
Earlier work this paper cites.
Liu, X., Wan, K., Ding, Y., Zhang, X., Zhu, Q.: Weighted-sampling audio adversarial example attack. In: Proceedings of the AAAI Conference on Artificial Intelligence. vol. 34, pp. 4908–4915 (2020)
2020
Earlier work this paper cites.
Mao, J., Zhu, S., Liu, J.: An inaudible voice attack to context-based device authentication in smart iot systems. Journal of Systems Architecture 104
2020
Earlier work this paper cites.
Wu, J., Chen, B., Luo, W., Fang, Y.: Audio steganography based on iterative adversarial attacks against convolutional neural networks. IEEE transactions on information forensics and security 15
2020
Earlier work this paper cites.
Du, X., Pun, C.M.: Robust audio patch attacks using physical sample simulation and adversarial patch noise generation. IEEE Transactions on Multimedia 24
2021
Earlier work this paper cites.
Hussain, S., Neekhara, P., Dubnov, S., McAuley, J., Koushanfar, F.: Waveguard: Understanding and mitigating audio adversarial examples. In: 30th USENIX security symposium (USENIX Security 21). pp. 2273–2290 (2021)
2021
Earlier work this paper cites.
Kasher, M., Zhao, M., Greenberg, A., Gulati, D., Kokalj-Filipovic, S., Spasojevic, P.: Inaudible manipulation of voice-enabled devices through backdoor using robust adversarial audio attacks. In: Proceedings of the 3rd ACM Workshop on Wireless Security and Machine Learning. pp. 37–42 (2021)
2021
Earlier work this paper cites.
Ma, P., Petridis, S., Pantic, M.: Detecting adversarial attacks on audiovisual speech recognition. In: Proceedings of the IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). pp. 6403–6407 (2021)
2021
Earlier work this paper cites.
Olivier, R., Raj, B., Shah, M.: High-frequency adversarial defense for speech and audio. In: Proceedings of the IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). pp. 2995–2999 (2021)
2021
Earlier work this paper cites.
Takahashi, N., Inoue, S., Mitsufuji, Y.: Adversarial attacks on audio source separation. In: Proceedings of the IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). pp. 521–525 (2021)
2021
Earlier work this paper cites.
Xie, Y., Li, Z., Shi, C., Liu, J., Chen, Y., Yuan, B.: Enabling fast and universal audio adversarial attack using generative model. In: Proceedings of the AAAI conference on Artificial Intelligence. vol. 35, pp. 14129–14137 (2021)
2021
Earlier work this paper cites.
Zhang, H., Yan, Q., Zhou, P., Liu, X.Y.: Generating robust audio adversarial examples with temporal dependency. In: Proceedings of the Twenty-Ninth International Conference on International Joint Conferences on Artificial Intelligence. pp. 3167–3173 (2021)
2021
Earlier work this paper cites.
Zhang, W., Zhao, S., Liu, L., Li, J., Cheng, X., Zheng, T.F., Hu, X.: Attack on practical speaker verification system using universal adversarial perturbations. In: Proceedings of the IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). pp. 2575–2579 (2021)
2021
Earlier work this paper cites.
Zheng, B., Jiang, P., Wang, Q., Li, Q., Shen, C., Wang, C., Ge, Y., Teng, Q., Zhang, S.: Black-box adversarial attacks on commercial speech platforms with minimal information. In: Proceedings of the 2021 ACM SIGSAC Conference on Computer and Communications Security. pp. 86–107 (2021)
2021
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
Chen, G., Zhao, Z., Song, F., Chen, S., Fan, L., Wang, F., Wang, J.: Towards understanding and mitigating audio adversarial examples for speaker recognition. IEEE Transactions on Dependable and Secure Computing 20
2022
Earlier work this paper cites.
Chen, Q., Chen, M., Lu, L., Yu, J., Chen, Y., Wang, Z., Ba, Z., Lin, F., Ren, K.: Push the limit of adversarial example attack on speaker recognition in physical domain. In: Proceedings of the 20th ACM Conference on Embedded Networked Sensor Systems. pp. 710–724 (2022)
2022
Earlier work this paper cites.
Guo, H., Wang, Y., Ivanov, N., Xiao, L., Yan, Q.: Specpatch: Human-in-the-loop adversarial audio spectrogram patch attack on speech recognition. In: Proceedings of the 2022 ACM SIGSAC conference on computer and communications security. pp. 1353–1366 (2022)
2022
Earlier work this paper cites.
Koffas, S., Xu, J., Conti, M., Picek, S.: Can you hear it? backdoor attacks via ultrasonic triggers. In: Proceedings of the 2022 ACM Workshop on Wireless Security and Machine Learning. pp. 57–62 (2022)
2022
Earlier work this paper cites.
Lan, J., Zhang, R., Yan, Z., Wang, J., Chen, Y., Hou, R.: Adversarial attacks and defenses in speaker recognition systems: A survey. Journal of Systems Architecture 127
2022
Earlier work this paper cites.
Liu, P., Zhang, S., Yao, C., Ye, W., Li, X.: Backdoor attacks against deep neural networks by personalized audio steganography. In: 2022 26th International Conference on Pattern Recognition (ICPR). pp. 68–74. IEEE (2022)
2022
Earlier work this paper cites.
Liu, Q., Zhou, T., Cai, Z., Tang, Y.: Opportunistic backdoor attacks: Exploring human-imperceptible vulnerabilities on speech recognition systems. In: Proceedings of the 30th ACM International Conference on Multimedia. pp. 2390–2398 (2022)
2022
Earlier work this paper cites.
Luo, Y., Tai, J., Jia, X., Zhang, S.: Practical backdoor attack against speaker recognition system. In: International Conference on Information Security Practice and Experience. pp. 468–484. Springer (2022)
2022
Earlier work this paper cites.
Mun, H., Seo, S., Son, B., Yun, J.: Black-box audio adversarial attack using particle swarm optimization. IEEE Access 10
2022
Earlier work this paper cites.
O’Reilly, P., Bugler, A., Bhandari, K., Morrison, M., Pardo, B.: Voiceblock: Privacy through real-time adversarial attacks with audio-to-audio models. Advances in Neural Information Processing Systems 35
2022
Earlier work this paper cites.
Qu, X., Wei, P., Gao, M., Sun, Z., Ong, Y.S., Ma, Z.: Synthesising audio adversarial examples for automatic speech recognition. In: Proceedings of the 28th ACM SIGKDD Conference on Knowledge Discovery and Data Mining. pp. 1430–1440 (2022)
2022
Earlier work this paper cites.
Shi, C., Zhang, T., Li, Z., Phan, H., Zhao, T., Wang, Y., Liu, J., Yuan, B., Chen, Y.: Audio-domain position-independent backdoor attack via unnoticeable triggers. In: Proceedings of the 28th Annual International Conference on Mobile Computing And Networking. pp. 583–595 (2022)
2022
Earlier work this paper cites.
Vadillo, J., Santana, R.: On the human evaluation of universal audio adversarial perturbations. Computers & Security 112
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
Wang, S., Zhang, Z., Zhu, G., Zhang, X., Zhou, Y., Huang, J.: Query-efficient adversarial attack with low perturbation against end-to-end speech recognition systems. IEEE Transactions on Information Forensics and Security 18
2022
Earlier work this paper cites.
Xin, J., Lyu, X., Ma, J.: Natural backdoor attacks on speech recognition models. In: International Conference on Machine Learning for Cyber Security. pp. 597–610. Springer (2022)
2022
Earlier work this paper cites.
Zhu, J., Chen, L., Xu, D., Zhao, W.: Backdoor defence for voice print recognition model based on speech enhancement and weight pruning. IEEE access 10
2022
Earlier work this paper cites.
Chen, H., Zhang, J., Chen, K., Zhang, W., Yu, N.: Model access control based on hidden adversarial examples for automatic speech recognition. IEEE Transactions on Artificial Intelligence 5
2023
Earlier work this paper cites.
Chen, M., Lu, L., Yu, J., Ba, Z., Lin, F., Ren, K.: Advreverb: Rethinking the stealthiness of audio adversarial examples to human perception. IEEE Transactions on Information Forensics and Security 19
2023
Earlier work this paper cites.
Chen, Y.W., Ke, B.H., Chen, B.Z., Chiu, S.R., Tu, C.W., Kuo, J.J.: Knowledge distillation based defense for audio trigger backdoor in federated learning. In: GLOBECOM 2023-2023 IEEE Global Communications Conference. pp. 4271–4276. IEEE (2023)
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
Dou, Z., Hu, X., Yang, H., Liu, Z., Fang, M.: Adversarial attacks to multi-modal models. In: Proceedings of the 1st ACM Workshop on Large AI Systems and Models with Privacy and Safety Analysis. pp. 35–46 (2023)
2023
Cited alongside, same era.
Ge, Y., Zhao, L., Wang, Q., Duan, Y., Du, M.: Advddos: Zero-query adversarial attacks against commercial speech recognition systems. IEEE Transactions on Information Forensics and Security 18
2023
Cited alongside, same era.
Gong, X., Fang, Z., Li, B., Wang, T., Chen, Y., Wang, Q.: Palette: Physically-realizable backdoor attacks against video recognition models. IEEE Transactions on Dependable and Secure Computing 21
2023
Cited alongside, same era.
Guo, F., Sun, Z., Chen, Y., Ju, L.: Towards the universal defense for query-based audio adversarial attacks on speech recognition system. Cybersecurity 6
2023
Cited alongside, same era.
2024
Later among the works it cites.
2024
Later among the works it cites.
Mengara, O.: A backdoor approach with inverted labels using dirty label-flipping attacks. IEEE Access (2024)
2024
Later among the works it cites.
2024
Later among the works it cites.
Miao, Y., Zhu, Y., Yu, L., Zhu, J., Gao, X.S., Dong, Y.: T2vsafetybench: Evaluating the safety of text-to-video generative models. Advances in Neural Information Processing Systems 37
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Guo, H., Chen, X., Guo, J., Xiao, L., Yan, Q.: Masterkey: Practical backdoor attack against speaker verification systems. In: Proceedings of the 29th Annual International Conference on Mobile Computing and Networking. pp. 1–15 (2023)
2023
Cited alongside, same era.
Guo, H., Wang, G., Wang, Y., Chen, B., Yan, Q., Xiao, L.: Phantomsound: Black-box, query-efficient audio adversarial attack via split-second phoneme injection. In: Proceedings of the 26th International Symposium on Research in Attacks, Intrusions and Defenses. pp. 366–380 (2023)
2023
Cited alongside, same era.
Kim, H., Park, J., Lee, J.: Generating transferable adversarial examples for speech classification. Pattern Recognition 137
2023
Cited alongside, same era.
Ko, K., Kim, S., Kwon, H.: Multi-targeted audio adversarial example for use against speech recognition systems. Computers & Security 128
2023
Cited alongside, same era.
Koffas, S., Pajola, L., Picek, S., Conti, M.: Going in style: Audio backdoors through stylistic transformations. In: Proceedings of the IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). pp. 1–5 (2023)
2023
Cited alongside, same era.
Kwon, H., Nam, S.H.: Audio adversarial detection through classification score on speech recognition systems. Computers & Security 126
2023
Cited alongside, same era.
Lee, Y., Chen, K., Meng, G., Lv, P., et al.: Aliasing backdoor attacks on pre-trained models. In: 32nd USENIX Security Symposium (USENIX Security 23). pp. 2707–2724 (2023)
2023
Cited alongside, same era.
Li, H., Jia, P., Li, W., Ma, B., Li, B., Wu, D., Li, H.: Towards efficient universal adversarial attack on audio classification models: A two-step method. In: International Symposium on Emerging Information Security and Applications. pp. 20–37. Springer (2023)
2023
Cited alongside, same era.
2024
Later among the works it cites.
Park, N., Kim, J.: Toward robust asr system against audio adversarial examples using agitated logit. ACM Transactions on Privacy and Security 27
2024
Later among the works it cites.
Qiu, S., You, X., Rong, W., Huang, L., Liang, Y.: Boosting imperceptibility of adversarial attacks for environmental sound classification. In: 2024 IEEE 36th International Conference on Tools with Artificial Intelligence (ICTAI). pp. 790–797. IEEE (2024)
2024
Later among the works it cites.
Rabhi, M., Bakiras, S., Di Pietro, R.: Audio-deepfake detection: Adversarial attacks and countermeasures. Expert Systems with Applications 250
2024
Later among the works it cites.
Schoof, C., Koffas, S., Conti, M., Picek, S.: Emoback: Backdoor attacks against speaker identification using emotional prosody. In: Proceedings of the 2024 Workshop on Artificial Intelligence and Security. pp. 137–148 (2024)
2024
Later among the works it cites.
2024
Later among the works it cites.
Sun, X., Zhang, Y., Tang, X., Bedi, A.S., Bera, A.: Trustnavgpt: Modeling uncertainty to improve trustworthiness of audio-guided llm-based robot navigation. In: 2024 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS). pp. 8794–8801. IEEE (2024)
2024
Later among the works it cites.
Tang, Y., Sun, L., Xu, X.: Silenttrig: An imperceptible backdoor attack against speaker identification with hidden triggers. Pattern Recognition Letters 177
2024
Later among the works it cites.
Wu, Y., Chen, J., Lei, T., Yu, J., Hossain, M.S.: Web 3.0 security: Backdoor attacks in federated learning-based automatic speaker verification systems in the 6g era. Future Generation Computer Systems 160
2024
Later among the works it cites.
Xiao, Y., Yao, W., Li, Z., Yang, J., Wen, W.: Phoneme semantic backdoor attacks with multiple task learning for speech classification task. In: National Conference on Man-Machine Speech Communication. pp. 79–90. Springer (2024)
2024
Later among the works it cites.
Xin, J., Lv, X.: Speechguard: Online defense against backdoor attacks on speech recognition models. In: 2024 International Joint Conference on Neural Networks (IJCNN). pp. 1–8. IEEE (2024)
2024
Later among the works it cites.
Xiong, B., Xing, Z., Wen, W.: Phoneme substitution: A novel approach for backdoor attacks on speech recognition systems. In: 2024 IEEE 36th International Conference on Tools with Artificial Intelligence (ICTAI). pp. 540–547. IEEE (2024)
2024
Later among the works it cites.
Xu, Z., Liu, Y., Deng, G., Li, Y., Picek, S.: A comprehensive study of jailbreak attack versus defense for large language models. In: Findings of the Association for Computational Linguistics ACL 2024. pp. 7432–7449 (2024)
2024
Later among the works it cites.
Yan, B., Lan, J., Yan, Z.: Backdoor attacks against voice recognition systems: A survey. ACM Computing Surveys 57
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Ye, Z., Yan, D., Dong, L., Shen, K.: Breaking speaker recognition with paddingback. In: Proceedings of the IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). pp. 4435–4439 (2024)
2024
Later among the works it cites.
2024
Later among the works it cites.
Yun, Z., Ao, J., Ko, T., Ronen, E., Sharif, M.: Sounding the alarm: Backdooring acoustic foundation models for physically realizable triggers (2024)
2024
Later among the works it cites.
Zhang, M., Ji, S., Cai, H., Dong, H., Zhang, P., Li, Y.: Audio steganography based backdoor attack for speech recognition software. In: 2024 IEEE 48th Annual Computers, Software, and Applications Conference (COMPSAC). pp. 1208–1217. IEEE (2024)
2024
Later among the works it cites.
Zhang, S., Pan, Y., Liu, Q., Yan, Z., Choo, K.K.R., Wang, G.: Backdoor attacks and defenses targeting multi-domain ai models: A comprehensive review. ACM Computing Surveys 57
2024
Later among the works it cites.
Zhang, T., Phan, H., Tang, Z., Shi, C., Wang, Y., Yuan, B., Chen, Y.: Inaudible backdoor attack via stealthy frequency trigger injection in audio spectrogram. In: Proceedings of the 30th Annual International Conference on Mobile Computing and Networking. pp. 31–45 (2024)
2024
Later among the works it cites.
Zhao, S., Gan, L., Tuan, L.A., Fu, J., Lyu, L., Jia, M., Wen, J.: Defending against weight-poisoning backdoor attacks for parameter-efficient fine-tuning. In: Findings of the Association for Computational Linguistics: NAACL 2024. pp. 3421–3438 (2024)
2024
Later among the works it cites.
Zhao, S., Jia, M., Tuan, L.A., Pan, F., Wen, J.: Universal vulnerabilities in large language models: Backdoor attacks for in-context learning. In: Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing. pp. 11507–11522 (2024)
2024
Later among the works it cites.
Zhao, S., Tian, J., Fu, J., Chen, J., Wen, J.: Feamix: Feature mix with memory batch based on self-consistency learning for code generation and code translation. IEEE Transactions on Emerging Topics in Computational Intelligence (2024)
2024
Later among the works it cites.
Zhao, S., Tuan, L.A., Fu, J., Wen, J., Luo, W.: Exploring clean label backdoor attacks and defense in language models. IEEE/ACM Transactions on Audio, Speech, and Language Processing (2024)
2024
Later among the works it cites.
2024
Later among the works it cites.
Chiu, C.W., Huang, L., Li, B., Chen, H.: Do as i say not as i do’: A semi-automated approach for jailbreak prompt attack against multimodal llms. arXiv e-prints pp. arXiv–2502 (2025)
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
Hu, W., Gu, S., Wang, Y., Hong, R.: Videojail: Exploiting video-modality vulnerabilities for jailbreak attacks on multimodal large language models. In: ICLR 2025 Workshop on Building Trust in Language Models and Applications (2025)
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
Keller, L., Glavaš, G.: Speechtaxi: On multilingual semantic speech classification. In: Proceedings of the IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). pp. 1–5 (2025)
2025
Closest in time.
Luong, H.T., Li, H., Zhang, L., Lee, K.A., Chng, E.S.: Llamapartialspoof: An llm-driven fake speech dataset simulating disinformation generation. In: Proceedings of the IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). pp. 1–5 (2025)
2025
Closest in time.
2025
Closest in time.
Xiao, L., Mao, R., Zhao, S., Lin, Q., Jia, Y., He, L., Cambria, E.: Exploring cognitive and aesthetic causality for multimodal aspect-based sentiment analysis. IEEE Transactions on Affective Computing (2025)
2025
Closest in time.
Xu, W., Xu, Y., Zhang, S.: Sample-independent federated learning backdoor attack in speaker recognition. Cluster Computing 28
2025
Closest in time.
Yao, W., Yang, J., He, Y., Liu, J., Wen, W.: Imperceptible rhythm backdoor attacks: Exploring rhythm transformation for embedding undetectable vulnerabilities on speech recognition. Neurocomputing 614
2025
Closest in time.
Zhang, Z., Liang, S., Shimada, D., Xu, C.: Rethinking audio-visual adversarial vulnerability from temporal and modality perspectives. In: The Thirteenth International Conference on Learning Representations (2025)
2025
Closest in time.
Zhao, S., Jia, M., Guo, Z., Gan, L., XU, X., Wu, X., Fu, J., Yichao, F., Pan, F., Luu, A.T.: A survey of recent backdoor attacks and defenses in large language models. Transactions on Machine Learning Research (2025)
2025
Closest in time.
Zhao, S., Xu, X., Xiao, L., Wen, J., Tuan, L.A.: Clean-label backdoor attack and defense: An examination of language model vulnerability. Expert Systems with Applications 265
2025
Closest in time.