Fetching the paper…
Reading the bibliography…
Large Language Models (LLMs) have achieved significantly advanced capabilities in understanding and generating human language text, which have gained increasing popularity over recent years.
K. Papineni, S. Roukos, T. Ward, and W.-J. Zhu, “Bleu: a method for automatic evaluation of machine translation,” in Proceedings of the 40th Annual Meeting of the Association for Computational Linguistics , P. Isabelle, E. Charniak, and D. Lin, Eds. Philadelphia, Pennsylvania, USA: Association for Computational Linguistics, Jul. 2002, pp. 311–318. [Online]. Available: https://aclanthology.org/P02-1040
2002
Earlier work this paper cites.
2004
Earlier work this paper cites.
C.-Y. Lin, “ROUGE: A package for automatic evaluation of summaries,” in Text Summarization Branches Out . Barcelona, Spain: Association for Computational Linguistics, Jul. 2004, pp. 74–81. [Online]. Available: https://aclanthology.org/W04-1013
2004
Earlier work this paper cites.
2005
Earlier work this paper cites.
2006
Earlier work this paper cites.
2010
Earlier work this paper cites.
2011
Earlier work this paper cites.
A. L. Maas, R. E. Daly, P. T. Pham, D. Huang, A. Y. Ng, and C. Potts, “Learning word vectors for sentiment analysis,” in Proceedings of the 49th Annual Meeting of the Association for Computational Linguistics: Human Language Technologies , D. Lin, Y. Matsumoto, and R. Mihalcea, Eds. Portland, Oregon, USA: Association for Computational Linguistics, Jun. 2011, pp. 142–150. [Online]. Available: https://aclanthology.org/P11-1015
2011
Earlier work this paper cites.
R. Socher, A. Perelygin, J. Wu, J. Chuang, C. D. Manning, A. Ng, and C. Potts, “Recursive deep models for semantic compositionality over a sentiment treebank,” in Proceedings of the 2013 Conference on Empirical Methods in Natural Language Processing , D. Yarowsky, T. Baldwin, A. Korhonen, K. Livescu, and S. Bethard, Eds. Seattle, Washington, USA: Association for Computational Linguistics, Oct. 2013, pp. 1631–1642. [Online]. Available: https://aclanthology.org/D13-1170
2013
Earlier work this paper cites.
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
X. Zhang, J. Zhao, and Y. LeCun, “Character-level convolutional networks for text classification,” in Advances in Neural Information Processing Systems , C. Cortes, N. Lawrence, D. Lee, M. Sugiyama, and R. Garnett, Eds., vol. 28. Curran Associates, Inc., 2015. [Online]. Available: https://proceedings.neurips.cc/paper_files/paper/2015/file/250cf8b51c773f3f8dc8b4be867a9a02-Paper.pdf
2015
Earlier work this paper cites.
N. Papernot, P. McDaniel, X. Wu, S. Jha, and A. Swami, “Distillation as a defense to adversarial perturbations against deep neural networks,” in 2016 IEEE Symposium on Security and Privacy (SP) , 2016, pp. 582–597
2016
Earlier work this paper cites.
P. Blanchard, E. M. El Mhamdi, R. Guerraoui, and J. Stainer, “Machine learning with adversaries: Byzantine tolerant gradient descent,” in Advances in Neural Information Processing Systems , I. Guyon, U. V. Luxburg, S. Bengio, H. Wallach, R. Fergus, S. Vishwanathan, and R. Garnett, Eds., vol. 30. Curran Associates, Inc., 2017. [Online]. Available: https://proceedings.neurips.cc/paper_files/paper/2017/file/f4b9ec30ad9f68f89b29639786cb62ef-Paper.pdf
2017
Earlier work this paper cites.
Y. Gu, Q. Zhao, Y. Zhang, and Z. Lin, “Pt-cfi: Transparent backward-edge control flow violation detection using intel processor trace,” in Proceedings of the Seventh ACM on Conference on Data and Application Security and Privacy , 2017, pp. 173–184
2017
Earlier work this paper cites.
N. Papernot, P. McDaniel, A. Sinha, and M. P. Wellman, “Sok: Security and privacy in machine learning,” in 2018 IEEE European Symposium on Security and Privacy (EuroS&P) , 2018
2018
Earlier work this paper cites.
K. Liu, B. Dolan-Gavitt, and S. Garg, “Fine-pruning: Defending against backdooring attacks on deep neural networks,” 05 2018
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
B. Tran, J. Li, and A. Madry, “Spectral signatures in backdoor attacks,” 2018
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
O. de Gibert, N. Perez, A. García-Pablos, and M. Cuadros, “Hate speech dataset from a white supremacy forum,” in Proceedings of the 2nd Workshop on Abusive Language Online (ALW2) , D. Fišer, R. Huang, V. Prabhakaran, R. Voigt, Z. Waseem, and J. Wernimont, Eds. Brussels, Belgium: Association for Computational Linguistics, Oct. 2018, pp. 11–20. [Online]. Available: https://aclanthology.org/W18-5102
2018
Earlier work this paper cites.
2019
Earlier work this paper cites.
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, I. Sutskever et al. , “Language models are unsupervised multitask learners,” OpenAI blog , vol. 1, no. 8, p. 9, 2019
2019
Earlier work this paper cites.
J. Dai, C. Chen, and Y. Li, “A backdoor attack against lstm-based text classification systems,” IEEE Access , vol. 7, pp. 138 872–138 878, 2019
2019
Earlier work this paper cites.
B. Wang, Y. Yao, S. Shan, H. Li, B. Viswanath, H. Zheng, and B. Y. Zhao, “Neural cleanse: Identifying and mitigating backdoor attacks in neural networks,” in 2019 IEEE Symposium on Security and Privacy (SP) , 2019, pp. 707–723
2019
Earlier work this paper cites.
J. Dai and C. Chen, “A backdoor attack against lstm-based text classification systems,” 2019
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Y. Gao, C. Xu, D. Wang, S. Chen, D. C. Ranasinghe, and S. Nepal, “Strip: a defence against trojan attacks on deep neural networks,” in Proceedings of the 35th Annual Computer Security Applications Conference , ser. ACSAC ’19. New York, NY, USA: Association for Computing Machinery, 2019, p. 113–125. [Online]. Available: https://doi.org/10.1145/3359789.3359790
2019
Earlier work this paper cites.
A. Abusnaina, A. Khormali, H. Alasmary, J. Park, A. Anwar, and A. Mohaisen, “Adversarial learning attacks on graph-based iot malware detection systems,” in 2019 IEEE 39th international conference on distributed computing systems (ICDCS) . IEEE, 2019, pp. 1296–1305
2019
Earlier work this paper cites.
H. Chen, C. Fu, J. Zhao, and F. Koushanfar, “Deepinspect: A black-box trojan detection and mitigation framework for deep neural networks,” in Proceedings of the Twenty-Eighth International Joint Conference on Artificial Intelligence, IJCAI-19 . International Joint Conferences on Artificial Intelligence Organization, 7 2019, pp. 4658–4664. [Online]. Available: https://doi.org/10.24963/ijcai.2019/647
2019
Earlier work this paper cites.
Y. Liu, W.-C. Lee, G. Tao, S. Ma, Y. Aafer, and X. Zhang, “Abs: Scanning neural networks for back-doors by artificial brain stimulation,” in Proceedings of the 2019 ACM SIGSAC Conference on Computer and Communications Security , ser. CCS ’19. New York, NY, USA: Association for Computing Machinery, 2019, p. 1265–1282. [Online]. Available: https://doi.org/10.1145/3319535.3363216
2019
Earlier work this paper cites.
Q. Zhao, C. Zuo, G. Pellegrino, and Z. Lin, “Geo-locating drivers: A study of sensitive data leakage in ride-hailing services,” in 26th Annual Network and Distributed System Security Symposium (NDSS 2019) . Internet Society, 2019
2019
Earlier work this paper cites.
T. B. Brown, B. Mann, N. Ryder, M. Subbiah, J. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, S. Agarwal, A. Herbert-Voss, G. Krueger, T. Henighan, R. Child, A. Ramesh, D. M. Ziegler, J. Wu, C. Winter, C. Hesse, M. Chen, E. Sigler, M. Litwin, S. Gray, B. Chess, J. Clark, C. Berner, S. McCandlish, A. Radford, I. Sutskever, and D. Amodei, “Language models are few-shot learners,” 2020
2020
Earlier work this paper cites.
C. Raffel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, and P. J. Liu, “Exploring the limits of transfer learning with a unified text-to-text transformer,” Journal of machine learning research , vol. 21, no. 140, pp. 1–67, 2020
2020
Earlier work this paper cites.
Z. Feng, D. Guo, D. Tang, N. Duan, X. Feng, M. Gong, L. Shou, B. Qin, T. Liu, D. Jiang, and M. Zhou, “CodeBERT: A pre-trained model for programming and natural languages,” in Findings of the Association for Computational Linguistics: EMNLP 2020 , T. Cohn, Y. He, and Y. Liu, Eds. Online: Association for Computational Linguistics, Nov. 2020, pp. 1536–1547. [Online]. Available: https://aclanthology.org/2020.findings-emnlp.139
2020
Earlier work this paper cites.
K. Kurita, P. Michel, and G. Neubig, “Weight poisoning attacks on pre-trained models,” 2020
2020
Earlier work this paper cites.
H. Alasmary, A. Abusnaina, R. Jang, M. Abuhamad, A. Anwar, D. Nyang, and D. Mohaisen, “Soteria: Detecting adversarial examples in control flow graph-based malware classifiers,” in 2020 IEEE 40th International Conference on Distributed Computing Systems (ICDCS) . IEEE, 2020, pp. 888–898
2020
Earlier work this paper cites.
P. Lewis, E. Perez, A. Piktus, F. Petroni, V. Karpukhin, N. Goyal, H. Küttler, M. Lewis, W.-t. Yih, T. Rocktäschel, S. Riedel, and D. Kiela, “Retrieval-augmented generation for knowledge-intensive nlp tasks,” in Advances in Neural Information Processing Systems , H. Larochelle, M. Ranzato, R. Hadsell, M. Balcan, and H. Lin, Eds., vol. 33. Curran Associates, Inc., 2020, pp. 9459–9474. [Online]. Available: https://proceedings.neurips.cc/paper_files/paper/2020/file/6b493230205f780e1bc26945df7481e5-Paper.pdf
2020
Earlier work this paper cites.
B. G. Doan, E. Abbasnejad, and D. C. Ranasinghe, “Februus: Input purification defense against trojan attacks on deep neural network systems,” in Annual Computer Security Applications Conference , ser. ACSAC ’20. ACM, Dec. 2020. [Online]. Available: http://dx.doi.org/10.1145/3427228.3427264
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
Q. Zhao, H. Wen, Z. Lin, D. Xuan, and N. Shroff, “On the accuracy of measured proximity of bluetooth-based contact tracing apps,” in Security and Privacy in Communication Networks: 16th EAI International Conference, SecureComm 2020, Washington, DC, USA, October 21-23, 2020, Proceedings, Part I 16 . Springer, 2020, pp. 49–60
2020
Earlier work this paper cites.
Q. Zhao, C. Zuo, B. Dolan-Gavitt, G. Pellegrino, and Z. Lin, “Automatic uncovering of hidden behaviors from input validation in mobile apps,” in 2020 IEEE Symposium on Security and Privacy (SP) . IEEE, 2020, pp. 1106–1120
2020
Earlier work this paper cites.
2021
Earlier work this paper cites.
B. Wang and A. Komatsuzaki, “Gpt-j-6b: A 6 billion parameter autoregressive language model,” 2021
2021
Earlier work this paper cites.
D. Guo, S. Ren, S. Lu, Z. Feng, D. Tang, S. Liu, L. Zhou, N. Duan, A. Svyatkovskiy, S. Fu, M. Tufano, S. K. Deng, C. Clement, D. Drain, N. Sundaresan, J. Yin, D. Jiang, and M. Zhou, “Graphcodebert: Pre-training code representations with data flow,” 2021
2021
Earlier work this paper cites.
W. U. Ahmad, S. Chakraborty, B. Ray, and K.-W. Chang, “Unified pre-training for program understanding and generation,” 2021
2021
Earlier work this paper cites.
Y. Wang, W. Wang, S. Joty, and S. C. Hoi, “CodeT5: Identifier-aware unified pre-trained encoder-decoder models for code understanding and generation,” in Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing , M.-F. Moens, X. Huang, L. Specia, and S. W.-t. Yih, Eds. Online and Punta Cana, Dominican Republic: Association for Computational Linguistics, Nov. 2021, pp. 8696–8708. [Online]. Available: https://aclanthology.org/2021.emnlp-main.685
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
F. Qi, M. Li, Y. Chen, Z. Zhang, Z. Liu, Y. Wang, and M. Sun, “Hidden killer: Invisible textual backdoor attacks with syntactic trigger,” in Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers) , C. Zong, F. Xia, W. Li, and R. Navigli, Eds. Online: Association for Computational Linguistics, Aug. 2021, pp. 443–453. [Online]. Available: https://aclanthology.org/2021.acl-long.37
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
R. Schuster, C. Song, E. Tromer, and V. Shmatikov, “You autocomplete me: Poisoning vulnerabilities in neural code completion,” in 30th USENIX Security Symposium (USENIX Security 21) . USENIX Association, Aug. 2021, pp. 1559–1575. [Online]. Available: https://www.usenix.org/conference/usenixsecurity21/presentation/schuster
2021
Earlier work this paper cites.
F. Qi, Y. Chen, X. Zhang, M. Li, Z. Liu, and M. Sun, “Mind the style of text! adversarial and backdoor attacks based on text style transfer,” in Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing , M.-F. Moens, X. Huang, L. Specia, and S. W.-t. Yih, Eds. Online and Punta Cana, Dominican Republic: Association for Computational Linguistics, Nov. 2021, pp. 4569–4580. [Online]. Available: https://aclanthology.org/2021.emnlp-main.374
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
X. Zhang, Z. Zhang, S. Ji, and T. Wang, “Trojaning language models for fun and profit,” 2021
2021
Earlier work this paper cites.
K. Chen, Y. Meng, X. Sun, S. Guo, T. Zhang, J. Li, and C. Fan, “Badpre: Task-agnostic backdoor attacks to pre-trained nlp foundation models,” 2021
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
F. Qi, Y. Yao, S. Xu, Z. Liu, and M. Sun, “Turn the combination lock: Learnable textual backdoor attacks via word substitution,” in Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers) , C. Zong, F. Xia, W. Li, and R. Navigli, Eds. Online: Association for Computational Linguistics, Aug. 2021, pp. 4873–4883. [Online]. Available: https://aclanthology.org/2021.acl-long.377
2021
Earlier work this paper cites.
K. Shao, J. Yang, Y. Ai, H. Liu, and Y. Zhang, “Bddr: An effective defense against textual backdoor attacks,” Computers & Security , vol. 110, p. 102433, 2021. [Online]. Available: https://www.sciencedirect.com/science/article/pii/S0167404821002571
2021
Earlier work this paper cites.
W. Yang, Y. Lin, P. Li, J. Zhou, and X. Sun, “RAP: Robustness-Aware Perturbations for defending against backdoor attacks on NLP models,” in Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing , M.-F. Moens, X. Huang, L. Specia, and S. W.-t. Yih, Eds. Online and Punta Cana, Dominican Republic: Association for Computational Linguistics, Nov. 2021, pp. 8365–8381. [Online]. Available: https://aclanthology.org/2021.emnlp-main.659
2021
Earlier work this paper cites.
T. Ni, Y. Chen, K. Song, and W. Xu, “A simple and fast human activity recognition system using radio frequency energy harvesting,” in Adjunct Proceedings of the 2021 ACM International Joint Conference on Pervasive and Ubiquitous Computing and Proceedings of the 2021 ACM International Symposium on Wearable Computers , 2021, pp. 666–671
2021
Earlier work this paper cites.
A. Abusnaina, Y. Wu, S. Arora, Y. Wang, F. Wang, H. Yang, and D. Mohaisen, “Adversarial example detection using latent neighborhood graph,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 7687–7696
2021
Earlier work this paper cites.
Y. Li, X. Lyu, N. Koren, L. Lyu, B. Li, and X. Ma, “Anti-backdoor learning: Training clean models on poisoned data,” in Advances in Neural Information Processing Systems , M. Ranzato, A. Beygelzimer, Y. Dauphin, P. Liang, and J. W. Vaughan, Eds., vol. 34. Curran Associates, Inc., 2021, pp. 14 900–14 912. [Online]. Available: https://proceedings.neurips.cc/paper_files/paper/2021/file/7d38b1e9bd793d3f45e0e212a729a93c-Paper.pdf
2021
Earlier work this paper cites.
C. Chen and J. Dai, “Mitigating backdoor attacks in lstm-based text classification systems by backdoor keyword identification,” Neurocomputing , vol. 452, pp. 253–262, 2021
2021
Earlier work this paper cites.
D. Wu and Y. Wang, “Adversarial neuron pruning purifies backdoored deep models,” in Advances in Neural Information Processing Systems , M. Ranzato, A. Beygelzimer, Y. Dauphin, P. Liang, and J. W. Vaughan, Eds., vol. 34. Curran Associates, Inc., 2021, pp. 16 913–16 925. [Online]. Available: https://proceedings.neurips.cc/paper_files/paper/2021/file/8cbe9ce23f42628c98f80fa0fac8b19a-Paper.pdf
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
W. Yang, Y. Lin, P. Li, J. Zhou, and X. Sun, “Rethinking stealthiness of backdoor attack against nlp models,” in Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers) , 2021, pp. 5543–5557
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
J. Xu, D. Ju, M. Li, Y.-L. Boureau, J. Weston, and E. Dinan, “Bot-adversarial dialogue for safe conversational agents,” in Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies , 2021, pp. 2950–2968
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2022
Earlier work this paper cites.
K. Y. Yoo and N. Kwak, “Backdoor attacks in federated learning by rare embeddings and gradient ensembling,” in Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing , Y. Goldberg, Z. Kozareva, and Y. Zhang, Eds. Abu Dhabi, United Arab Emirates: Association for Computational Linguistics, Dec. 2022, pp. 72–88. [Online]. Available: https://aclanthology.org/2022.emnlp-main.6
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
W. Du, Y. Zhao, B. Li, G. Liu, and S. Wang, “Ppt: Backdoor attacks on pre-trained models via poisoned prompt tuning,” in Proceedings of the Thirty-First International Joint Conference on Artificial Intelligence, IJCAI-22 , L. D. Raedt, Ed. International Joint Conferences on Artificial Intelligence Organization, 7 2022, pp. 680–686, main Track. [Online]. Available: https://doi.org/10.24963/ijcai.2022/96
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
Y. Gao, Y. Kim, B. G. Doan, Z. Zhang, G. Zhang, S. Nepal, D. C. Ranasinghe, and H. Kim, “Design and evaluation of a multi-domain trojan detection method on deep neural networks,” IEEE Transactions on Dependable and Secure Computing , 2022
2022
Earlier work this paper cites.
T. D. Nguyen, P. Rieger, R. De Viti, H. Chen, B. B. Brandenburg, H. Yalame, H. Möllering, H. Fereidooni, S. Marchal, M. Miettinen et al. , “ { \{ FLAME } \} : Taming backdoors in federated learning,” in 31st USENIX Security Symposium (USENIX Security 22) , 2022, pp. 1415–1432
2022
Earlier work this paper cites.
X. Chen, Y. Dong, Z. Sun, S. Zhai, Q. Shen, and Z. Wu, “Kallima: A clean-label framework for textual backdoor attacks,” in European Symposium on Research in Computer Security . Springer, 2022, pp. 447–466
2022
Cited alongside, same era.
L. Gan, J. Li, T. Zhang, X. Li, Y. Meng, F. Wu, Y. Yang, S. Guo, and C. Fan, “Triggerless backdoor attack for nlp tasks with clean labels,” 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Y. Chen, T. Ni, W. Xu, and T. Gu, “Swipepass: Acoustic-based second-factor user authentication for smartphones,” Proceedings of the ACM on Interactive, Mobile, Wearable and Ubiquitous Technologies , vol. 6, no. 3, pp. 1–25, 2022
2022
Cited alongside, same era.
M. Omar and D. Mohaisen, “Making adversarially-trained language models forget with model retraining: A case study on hate speech detection,” in Companion Proceedings of the Web Conference 2022 , 2022, pp. 887–893
2022
Cited alongside, same era.
M. Omar, S. Choi, D. Nyang, and D. Mohaisen, “Quantifying the performance of adversarial training on language models with distribution shifts,” in Proceedings of the 1st Workshop on Cybersecurity and Social Sciences , 2022, pp. 3–9
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
B. Zhu, Y. Qin, G. Cui, Y. Chen, W. Zhao, C. Fu, Y. Deng, Z. Liu, J. Wang, W. Wu, M. Sun, and M. Gu, “Moderate-fitting as a natural backdoor defender for pre-trained language models,” in Advances in Neural Information Processing Systems , A. H. Oh, A. Agarwal, D. Belgrave, and K. Cho, Eds., 2022. [Online]. Available: https://openreview.net/forum?id=C7cv9fh8m-b
2022
Cited alongside, same era.
J. Guan, Z. Tu, R. He, and D. Tao, “Few-shot backdoor defense using shapley estimation,” in 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2022, pp. 13 348–13 357
2022
Cited alongside, same era.
H. Wang, J. Hong, A. Zhang, J. Zhou, and Z. Wang, “Trap and replace: Defending backdoor attacks by trapping them into an easy-to-replace subnetwork,” in Advances in Neural Information Processing Systems , S. Koyejo, S. Mohamed, A. Agarwal, D. Belgrave, K. Cho, and A. Oh, Eds., vol. 35. Curran Associates, Inc., 2022, pp. 36 026–36 039. [Online]. Available: https://proceedings.neurips.cc/paper_files/paper/2022/file/ea06e6e9e80f1c3d382317fff67041ac-Paper-Conference.pdf
2022
Cited alongside, same era.
Y. Li, T. Li, K. Chen, J. Zhang, S. Liu, W. Wang, T. Zhang, and Y. Liu, “Badedit: Backdooring large language models by model editing,” 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
S. Yan, S. Wang, Y. Duan, H. Hong, K. Lee, D. Kim, and Y. Hong, “An llm-assisted easy-to-trigger backdoor attack on code completion models: Injecting disguised vulnerabilities against strong detection,” 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
H. Huang, Z. Zhao, M. Backes, Y. Shen, and Y. Zhang, “Composite backdoor attacks against large language models,” in Findings of the Association for Computational Linguistics: NAACL 2024 , K. Duh, H. Gomez, and S. Bethard, Eds. Mexico City, Mexico: Association for Computational Linguistics, Jun. 2024, pp. 1459–1472. [Online]. Available: https://aclanthology.org/2024.findings-naacl.94
2024
Later among the works it cites.
J. Yan, V. Yadav, S. Li, L. Chen, Z. Tang, H. Wang, V. Srinivasan, X. Ren, and H. Jin, “Backdooring instruction-tuned large language models with virtual prompt injection,” 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
H. Wang, R. Zhong, J. Wen, and J. Steinhardt, “Adaptivebackdoor: Backdoored language model agents that detect human overseers,” in ICML 2024 Workshop on Foundation Models in the Wild , 2024. [Online]. Available: https://openreview.net/forum?id=RredrFZ4tQ
2024
Later among the works it cites.
B. Chen, N. Ivanov, G. Wang, and Q. Yan, “Multi-turn hidden backdoor in large language model-powered chatbot models,” in Proceedings of the 19th ACM Asia Conference on Computer and Communications Security , ser. ASIA CCS ’24. New York, NY, USA: Association for Computing Machinery, 2024, p. 1316–1330. [Online]. Available: https://doi.org/10.1145/3634737.3656289
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
D. Liu and S. Zhang, “Alanca: Active learning guided adversarial attacks for code comprehension on diverse pre-trained and large language models,” in 2024 IEEE International Conference on Software Analysis, Evolution and Reengineering (SANER) , 2024, pp. 602–613
2024
Later among the works it cites.
H. Aghakhani, W. Dai, A. Manoel, X. Fernandes, A. Kharkar, C. Kruegel, G. Vigna, D. Evans, B. Zorn, and R. Sim, “Trojanpuzzle: Covertly poisoning code-suggestion models,” 2024
2024
Later among the works it cites.
R. Zhang, H. Li, R. Wen, W. Jiang, Y. Zhang, M. Backes, Y. Shen, and Y. Zhang, “Instruction backdoor attacks against customized LLMs,” in 33rd USENIX Security Symposium (USENIX Security 24) . Philadelphia, PA: USENIX Association, Aug. 2024, pp. 1849–1866. [Online]. Available: https://www.usenix.org/conference/usenixsecurity24/presentation/zhang-rui
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
D. Lu, T. Pang, C. Du, Q. Liu, X. Yang, and M. Lin, “Test-time backdoor attacks on multimodal large language models,” 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Y. Wen, N. Jain, J. Kirchenbauer, M. Goldblum, J. Geiping, and T. Goldstein, “Hard prompts made easy: Gradient-based discrete optimization for prompt tuning and discovery,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
P. Ren, C. Zuo, X. Liu, W. Diao, Q. Zhao, and S. Guo, “Demistify: Identifying on-device machine learning models stealing and reuse vulnerabilities in mobile apps,” in Proceedings of the 46th IEEE/ACM International Conference on Software Engineering , 2024, pp. 1–13
2024
Later among the works it cites.
X. Meng, L. Wang, S. Guo, L. Ju, and Q. Zhao, “Ava: Inconspicuous attribute variation-based adversarial attack bypassing deepfake detection,” in 2024 IEEE Symposium on Security and Privacy (SP) . IEEE, 2024, pp. 74–90
2024
Later among the works it cites.
H. Liu, Y. Zhou, Y. Yang, Q. Zhao, T. Zhang, and T. Xiang, “Stealthiness assessment of adversarial perturbation: From a visual perspective,” IEEE Transactions on Information Forensics and Security , 2024
2024
Later among the works it cites.
P. Guo, F. Liu, X. Lin, Q. Zhao, and Q. Zhang, “L-autoda: Large language models for automatically evolving decision-based adversarial attacks,” in Proceedings of the Genetic and Evolutionary Computation Conference Companion , 2024, pp. 1846–1854
2024
Later among the works it cites.
T. Ni, “Sensor security in virtual reality: Exploration and mitigation,” in Proceedings of the 22nd Annual International Conference on Mobile Systems, Applications and Services , 2024, pp. 758–759
2024
Later among the works it cites.
T. Ni, Z. Sun, M. Han, Y. Xie, G. Lan, Z. Li, T. Gu, and W. Xu, “Rehsense: Towards battery-free wireless sensing via radio frequency energy harvesting,” in Proceedings of the Twenty-Fifth International Symposium on Theory, Algorithmic Foundations, and Protocol Design for Mobile Networks and Mobile Computing , 2024, pp. 211–220
2024
Later among the works it cites.
2024
Later among the works it cites.
H. Li, Y. Chen, Z. Zheng, Q. Hu, C. Chan, H. Liu, and Y. Song, “Backdoor removal for generative large language models,” 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Z. Sha, X. He, P. Berrang, M. Humbert, and Y. Zhang, “Fine-tuning is all you need to mitigate backdoor attacks,” 2024. [Online]. Available: https://openreview.net/forum?id=ywGSgEmOYb
2024
Later among the works it cites.
2024
Later among the works it cites.
C. Chen, H. Hong, T. Xiang, and M. Xie, “Anti-backdoor model: A novel algorithm to remove backdoors in a non-invasive way,” IEEE Transactions on Information Forensics and Security , vol. 19, pp. 7420–7434, 2024
2024
Later among the works it cites.
R. Bie, J. Jiang, H. Xie, Y. Guo, Y. Miao, and X. Jia, “Mitigating backdoor attacks in pre-trained encoders via self-supervised knowledge distillation,” IEEE Transactions on Services Computing , vol. 17, no. 5, pp. 2613–2625, 2024
2024
Later among the works it cites.
Y. Liu, X. Xu, Z. Hou, and Y. Yu, “Causality based front-door defense against backdoor attack on language models,” in Proceedings of the 41st International Conference on Machine Learning , ser. Proceedings of Machine Learning Research, R. Salakhutdinov, Z. Kolter, K. Heller, A. Weller, N. Oliver, J. Scarlett, and F. Berkenkamp, Eds., vol. 235. PMLR, 21–27 Jul 2024, pp. 32 239–32 252. [Online]. Available: https://proceedings.mlr.press/v235/liu24bu.html
2024
Later among the works it cites.
2024
Later among the works it cites.
C. Wei, W. Meng, Z. Zhang, M. Chen, M. Zhao, W. Fang, L. Wang, Z. Zhang, and W. Chen, “Lmsanitator: Defending prompt-tuning against task-agnostic backdoors,” in Proceedings 2024 Network and Distributed System Security Symposium , ser. NDSS 2024. Internet Society, 2024. [Online]. Available: http://dx.doi.org/10.14722/ndss.2024.23238
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
S. Choi and D. Mohaisen, “Attributing chatgpt-generated source codes,” IEEE Transactions on Dependable and Secure Computing , 2025
2025
Closest in time.