Fetching the paper…
Reading the bibliography…
This paper provides a unified view to explain different adversarial attacks and defense methods, i.e.
The use of multiple measurements in taxonomic problems
Ronald A Fisher · 1936
Earlier work this paper cites.
A value for n-person games
Lloyd S Shapley · 1953
Earlier work this paper cites.
Probabilistic values for games
Robert J Weber · 1988
Earlier work this paper cites.
An axiomatic approach to the concept of interaction among players in cooperative games
Michel Grabisch and Marc Roubens · 1999
Earlier work this paper cites.
Detecting statistical interactions with additive groves of trees
Daria Sorokina, Rich Caruana, Mirek Riedewald, and Daniel Fink · 2008
Earlier work this paper cites.
A unified approach to interpreting and boosting adversarial transferability
Xin Wang, Jie Ren, Shuyun Lin, Xiangming Zhu, Yisen Wang, and Quanshi Zhang · 2010
Earlier work this paper cites.
Intriguing properties of neural networks
Christian Szegedy, Wojciech Zaremba, Ilya Sutskever, Joan Bruna, Dumitru Erhan, Ian Goodfellow, and Rob Fergus · 2013
Earlier work this paper cites.
Explaining and harnessing adversarial examples
Ian J Goodfellow, Jonathon Shlens, and Christian Szegedy · 2014
Earlier work this paper cites.
Dropout: a simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov · 2014
Earlier work this paper cites.
3d shapenets for 2.5 d object recognition and next-best-view prediction
Zhirong Wu, Shuran Song, Aditya Khosla, Xiaoou Tang, and Jianxiong Xiao · 2014
Earlier work this paper cites.
Imagenet large scale visual recognition challenge
Olga Russakovsky, Jia Deng, Hao Su, Jonathan Krause, Sanjeev Satheesh, Sean Ma, Zhiheng Huang, Andrej Karpathy, Aditya Khosla, Michael Bernstein, Alexander C. Berg, and Li Fei-Fei · 2015
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Adversarial examples in the physical world
Alexey Kurakin, Ian Goodfellow, and Samy Bengio · 2016
Earlier work this paper cites.
The limitations of deep learning in adversarial settings
Nicolas Papernot, Patrick McDaniel, Somesh Jha, Matt Fredrikson, Z Berkay Celik, and Ananthram Swami · 2016
Earlier work this paper cites.
Sergey Zagoruyko and Nikos Komodakis · 2016
Earlier work this paper cites.
Towards evaluating the robustness of neural networks
Nicholas Carlini and David Wagner · 2017
Earlier work this paper cites.
Zoo: Zeroth order optimization based black-box attacks to deep neural networks without training substitute models
Pin-Yu Chen, Huan Zhang, Yash Sharma, Jinfeng Yi, and Cho-Jui Hsieh · 2017
Earlier work this paper cites.
Houdini: Fooling deep structured prediction models
Moustapha Cisse, Yossi Adi, Natalia Neverova, and Joseph Keshet · 2017
Earlier work this paper cites.
Keeping the bad guys out: Protecting and vaccinating deep learning with jpeg compression
Nilaksh Das, Madhuri Shanbhogue, Shang-Tse Chen, Fred Hohman, Li Chen, Michael E Kounavis, and Duen Horng Chau · 2017
Earlier work this paper cites.
Improved regularization of convolutional neural networks with cutout
Terrance DeVries and Graham W Taylor · 2017
Earlier work this paper cites.
Towards interpretable deep neural networks by leveraging adversarial examples
Yinpeng Dong, Hang Su, Jun Zhu, and Fan Bao · 2017
Earlier work this paper cites.
Deepcloak: Masking deep neural network models for robustness against adversarial samples
Ji Gao, Beilun Wang, Zeming Lin, Weilin Xu, and Yanjun Qi · 2017
Earlier work this paper cites.
Formal guarantees on the robustness of a classifier against adversarial manipulation
Matthias Hein and Maksym Andriushchenko · 2017
Earlier work this paper cites.
Densely connected convolutional networks
Gao Huang, Zhuang Liu, Laurens Van Der Maaten, and Kilian Q Weinberger · 2017
Earlier work this paper cites.
Delving into transferable adversarial examples and black-box attacks
Yanpei Liu, Xinyun Chen, Chang Liu, and Dawn Song · 2017
Earlier work this paper cites.
Magnet: a two-pronged defense against adversarial examples
Dongyu Meng and Hao Chen · 2017
Cited alongside, same era.
Biologically inspired protection of deep networks from adversarial attacks
Aran Nayebi and Surya Ganguli · 2017
Cited alongside, same era.
Practical black-box attacks against machine learning
Nicolas Papernot, Patrick McDaniel, Ian Goodfellow, Somesh Jha, Z Berkay Celik, and Ananthram Swami · 2017
Cited alongside, same era.
Pointnet: Deep learning on point sets for 3d classification and segmentation
Charles R Qi, Hao Su, Kaichun Mo, and Leonidas J Guibas · 2017
Cited alongside, same era.
Axiomatic attribution for deep networks
Mukund Sundararajan, Ankur Taly, and Qiqi Yan · 2017
Cited alongside, same era.
Obfuscated gradients give a false sense of security: Circumventing defenses to adversarial examples
On relating explanations and adversarial examples
Alexey Ignatiev, Nina Narodytska, and Joao Marques-Silva · 2019
Later among the works it cites.
Adversarial examples are not bugs, they are features
Andrew Ilyas, Shibani Santurkar, Dimitris Tsipras, Logan Engstrom, Brandon Tran, and Aleksander Madry · 2019
Later among the works it cites.
Towards hierarchical importance attribution: Explaining compositional semantics for neural sequence models
Xisen Jin, Zhongyu Wei, Junyi Du, Xiangyang Xue, and Xiang Ren · 2019
Later among the works it cites.
On the convergence and robustness of adversarial training
Yisen Wang, Xingjun Ma, James Bailey, Jinfeng Yi, Bowen Zhou, and Quanquan Gu · 2019
Later among the works it cites.
Feature denoising for improving adversarial robustness
Cihang Xie, Yuxin Wu, Laurens van der Maaten, Alan L Yuille, and Kaiming He · 2019
Later among the works it cites.
Interpreting adversarial examples by activation promotion and suppression
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Anish Athalye, Nicholas Carlini, and David Wagner · 2018
Cited alongside, same era.
Practical black-box attacks on deep neural networks using efficient query mechanisms
Arjun Nitin Bhagoji, Warren He, Bo Li, and Dawn Song · 2018
Cited alongside, same era.
Analysis of classifiers’ robustness to adversarial perturbations
Alhussein Fawzi, Omar Fawzi, and Pascal Frossard · 2018
Cited alongside, same era.
The relationship between high-dimensional geometry and adversarial examples
Justin Gilmer, Luke Metz, Fartash Faghri, Samuel S Schoenholz, Maithra Raghu, Martin Wattenberg, Ian Goodfellow, and G Brain · 2018
Cited alongside, same era.
Black-box adversarial attacks with limited queries and information
Andrew Ilyas, Logan Engstrom, Anish Athalye, and Jessy Lin · 2018
Cited alongside, same era.
Consistent individualized feature attribution for tree ensembles
Scott M Lundberg, Gabriel G Erion, and Su-In Lee · 2018
Cited alongside, same era.
Characterizing adversarial subspaces using local intrinsic dimensionality
Xingjun Ma, Bo Li, Yisen Wang, Sarah M Erfani, Sudanthi Wijewickrema, Grant Schoenebeck, Dawn Song, Michael E Houle, and James Bailey · 2018
Cited alongside, same era.
Kaidi Xu, Sijia Liu, Gaoyuan Zhang, Mengshu Sun, Pu Zhao, Quanfu Fan, Chuang Gan, and Xue Lin · 2019
Later among the works it cites.
ML-LOO: detecting adversarial examples with feature attribution
Puyudi Yang, Jianbo Chen, Cho-Jui Hsieh, Jane-Ling Wang, and Michael I. Jordan · 2019
Later among the works it cites.
A fourier perspective on model robustness in computer vision
Dong Yin, Raphael Gontijo Lopes, Jonathon Shlens, Ekin D Cubuk, and Justin Gilmer · 2019
Later among the works it cites.
Theoretically principled trade-off between robustness and accuracy
Hongyang Zhang, Yaodong Yu, Jiantao Jiao, Eric Xing, Laurent El Ghaoui, and Michael Jordan · 2019
Later among the works it cites.
Interpreting adversarially trained convolutional neural networks
Tianyuan Zhang and Zhanxing Zhu · 2019
Later among the works it cites.
Improving query efficiency of black-box adversarial attack
Yang Bai, Yuyuan Zeng, Yong Jiang, Yisen Wang, Shu-Tao Xia, and Weiwei Guo · 2020
Later among the works it cites.
Proper network interpretability helps adversarial robustness in classification
Akhilan Boopathy, Sijia Liu, Gaoyuan Zhang, Cynthia Liu, Pin-Yu Chen, Shiyu Chang, and Luca Daniel · 2020
Later among the works it cites.
Concise explanations of neural networks using adversarial training
Prasad Chalasani, Jiefeng Chen, Amrita Roy Chowdhury, Xi Wu, and Somesh Jha · 2020
Later among the works it cites.
Understanding global feature contributions with additive importance measures
Ian Covert, Scott Lundberg, and Su-In Lee · 2020
Later among the works it cites.
Explaining explanations: Axiomatic feature interactions for deep networks
Joseph D Janizek, Pascal Sturmfels, and Su-In Lee · 2020
Later among the works it cites.
A singular value perspective on model robustness
Malhar Jere, Maghav Kumar, and Farinaz Koushanfar · 2020
Later among the works it cites.
A game theoretic analysis of additive adversarial attacks and defenses
Ambar Pal and Rene Vidal · 2020
Later among the works it cites.
Do adversarially robust imagenet models transfer better?
Hadi Salman, Andrew Ilyas, Logan Engstrom, Ashish Kapoor, and Aleksander Madry · 2020
Later among the works it cites.
The shapley taylor interaction index
Mukund Sundararajan, Kedar Dhamdhere, and Ashish Agarwal · 2020
Later among the works it cites.
Game-theoretic interactions of different orders
Hao Zhang, Xu Cheng, Yiting Chen, and Quanshi Zhang · 2020
Later among the works it cites.
Improving adversarial robustness via channel-wise activation suppressing
Yang Bai, Yuyuan Zeng, Yong Jiang, Shu-Tao Xia, Xingjun Ma, and Yisen Wang · 2021
Closest in time.
Spectraldefense: Detecting adversarial attacks on cnns in the fourier domain
Paula Harder, Franz-Josef Pfreundt, Margret Keuper, and Janis Keuper · 2021
Closest in time.
Learning baseline values for shapley values
Jie Ren, Zhanpeng Zhou, Qirui Chen, and Quanshi Zhang · 2021
Closest in time.
Analysis and applications of class-wise robustness in adversarial training
Qi Tian, Kun Kuang, Kelu Jiang, Fei Wu, and Yisen Wang · 2021
Closest in time.
Interpreting attributions and interactions of adversarial attacks
Xin Wang, Shuyun Lin, Hao Zhang, Yufei Zhu, and Quanshi Zhang · 2021
Closest in time.