Detecting Backdoor Attacks on Deep Neural Networks by Activation Clustering
Original
Bryant Chen, Wilka Carvalho, Nathalie Baracaldo, Heiko Ludwig, Benjamin Edwards, Taesung Lee, Ian Molloy, and Biplav Srivastava · 2018
Later among the works it cites.
Manipulating machine learning: Poisoning attacks and countermeasures for regression learning
Matthew Jagielski, Alina Oprea, Battista Biggio, Chang Liu, Cristina Nita-Rotaru, and Bo Li · 2018
Later among the works it cites.
Adversarial Malware Binaries: Evading Deep Learning for Malware Detection in Executables
Bojan Kolosnjaji, Ambra Demontis, Battista Biggio, Davide Maiorca, Giorgio Giacinto, Claudia Eckert, and Fabio Roli · 2018
Later among the works it cites.
Deep Convolutional Malware Classifiers Can Learn from Raw Executables and Labels Only
Marek Krčál, Ondřej Švec, Martin Bálek, and Otakar Jašek · 2018
Later among the works it cites.
Fine-Pruning: Defending Against Backdooring Attacks on Deep Neural Networks
Original
Kang Liu, Brendan Dolan-Gavitt, and Siddharth Garg · 2018
Later among the works it cites.
Trojaning Attack on Neural Networks
Yingqi Liu, Shiqing Ma, Yousra Aafer, Wen-Chuan Lee, Juan Zhai, Weihang Wang, and Xiangyu Zhang · 2018
Later among the works it cites.
Malrec: Compact full-trace malware recording for retrospective deep analysis
Giorgio Severi, Tim Leek, and Brendan Dolan-Gavitt · 2018
Later among the works it cites.
Poison Frogs! Targeted Clean-Label Poisoning Attacks on Neural Networks
Ali Shafahi, W. Ronny Huang, Mahyar Najibi, Octavian Suciu, Christoph Studer, Tudor Dumitras, and Tom Goldstein · 2018
Later among the works it cites.
Machine Learning Aided Static Malware Analysis: A Survey and Tutorial
Andrii Shalaginov, Sergii Banin, Ali Dehghantanha, and Katrin Franke · 2018
Later among the works it cites.
When Does Machine Learning FAIL? Generalized Transferability for Evasion and Poisoning Attacks
Octavian Suciu, Radu Ma, Tudor Dumitras, and Hal Daume Iii · 2018
Later among the works it cites.
Spectral signatures in backdoor attacks
Brandon Tran, Jerry Li, and Aleksander Mądry · 2018
Later among the works it cites.
https://skylightcyber.com/2019/07/18/cylance-i-kill-you/
Skylight Cyber | Cylance, I Kill You! · 2019
Later among the works it cites.
On Evaluating Adversarial Robustness
Original
Nicholas Carlini, Anish Athalye, Nicolas Papernot, Wieland Brendel, Jonas Rauber, Dimitris Tsipras, Ian Goodfellow, Aleksander Madry, and Alexey Kurakin · 2019
Later among the works it cites.
Why do adversarial attacks transfer? explaining transferability of evasion and poisoning attacks
Ambra Demontis, Marco Melis, Maura Pintor, Matthew Jagielski, Battista Biggio, Alina Oprea, Cristina Nita-Rotaru, and Fabio Roli · 2019
Later among the works it cites.
Exploring Adversarial Examples in Malware Detection
Octavian Suciu, Scott E. Coull, and Jeffrey Johns · 2019
Later among the works it cites.
Clean-Label Backdoor Attacks
Alexander Turner, Dimitris Tsipras, and Aleksander Mądry · 2019
Later among the works it cites.
Neural Cleanse: Identifying and Mitigating Backdoor Attacks in Neural Networks
Bolun Wang, Yuanshun Yao, Shawn Shan, Huiying Li, Bimal Viswanath, Haitao Zheng, and Ben Y. Zhao · 2019
Later among the works it cites.
Adversarial machine learning reading list
Nicholas Carlini · 2020
Closest in time.
Adversarial machine learning–industry perspectives
Ram Shankar Siva Kumar, Magnus Nyström, John Lambert, Andrew Marshall, Mario Goertzel, Andi Comissoneru, Matt Swann, and Sharon Xia · 2020
Closest in time.
From local explanations to global understanding with explainable AI for trees
Scott M. Lundberg, Gabriel Erion, Hugh Chen, Alex DeGrave, Jordan M. Prutkin, Bala Nair, Ronit Katz, Jonathan Himmelfarb, Nisha Bansal, and Su-In Lee · 2020
Closest in time.
Intriguing Properties of Adversarial ML Attacks in the Problem Space
Fabio Pierazzi, Feargus Pendlebury, Jacopo Cortellazzi, and Lorenzo Cavallaro · 2020
Closest in time.