Fetching the paper…
Reading the bibliography…
One of the remarkable properties of robust computer vision models is that their input-gradients are often aligned with human perception, referred to in the literature as perceptually-aligned gradients (PAGs).
Improving generalization performance using double backpropagation
Harris Drucker and Yann Le Cun · 1992
Earlier work this paper cites.
Training with noise is equivalent to tikhonov regularization
Chris M Bishop · 1995
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Alex Krizhevsky, Geoffrey Hinton, et al · 2009
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
A connection between score matching and denoising autoencoders
Pascal Vincent · 2011
Earlier work this paper cites.
The MNIST database of handwritten digit images for machine learning research
Li Deng · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton · 2012
Earlier work this paper cites.
Deep inside convolutional networks: Visualising image classification models and saliency maps
Karen Simonyan, Andrea Vedaldi, and Andrew Zisserman · 2014
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Smoothgrad: removing noise by adding noise
Daniel Smilkov, Nikhil Thorat, Been Kim, Fernanda Viégas, and Martin Wattenberg · 2017
Earlier work this paper cites.
Axiomatic attribution for deep networks
Mukund Sundararajan, Ankur Taly, and Qiqi Yan · 2017
Earlier work this paper cites.
Towards deep learning models resistant to adversarial attacks
Aleksander Madry, Aleksandar Makelov, Ludwig Schmidt, Dimitris Tsipras, and Adrian Vladu · 2018
Earlier work this paper cites.
Knowledge transfer with jacobian matching
Suraj Srinivas and François Fleuret · 2018
Cited alongside, same era.
Sanity checks for saliency maps
Julius Adebayo, Justin Gilmer, Michael Muelly, Ian Goodfellow, Moritz Hardt, and Been Kim · 2018
Cited alongside, same era.
The riemannian geometry of deep generative models
Hang Shao, Abhishek Kumar, and P Thomas Fletcher · 2018
Cited alongside, same era.
The unreasonable effectiveness of deep features as a perceptual metric
Richard Zhang, Phillip Isola, Alexei A Efros, Eli Shechtman, and Oliver Wang · 2018
Cited alongside, same era.
Robustness may be at odds with accuracy
Dimitris Tsipras, Shibani Santurkar, Logan Engstrom, Alexander Turner, and Aleksander Madry · 2019
Cited alongside, same era.
Image synthesis with a single (robust) classifier
Shibani Santurkar, Andrew Ilyas, Dimitris Tsipras, Logan Engstrom, Brandon Tran, and Aleksander Madry · 2019
Cited alongside, same era.
Robustbench: a standardized adversarial robustness benchmark
Francesco Croce, Maksym Andriushchenko, Vikash Sehwag, Edoardo Debenedetti, Nicolas Flammarion, Mung Chiang, Prateek Mittal, and Matthias Hein · 2020
Later among the works it cites.
Do input gradients highlight discriminative features?
Harshay Shah, Prateek Jain, and Praneeth Netrapalli · 2021
Later among the works it cites.
Rethinking the role of gradient-based attribution methods for model interpretability
Suraj Srinivas and Francois Fleuret · 2021
Later among the works it cites.
Have we learned to explain?: How interpretability methods can learn to encode predictions in their interpretations
Neil Jethani, Mukund Sudarshan, Yindalon Aphinyanaphongs, and Rajesh Ranganath · 2021
Later among the works it cites.
Towards understanding the generative capability of adversarially robust classifiers
Yao Zhu, Jiacheng Ma, Jiacheng Sun, Zewei Chen, Rongxin Jiang, Yaowu Chen, and Zhenguo Li · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Are perceptually-aligned gradients a general property of robust classifiers?
Simran Kaur, Jeremy Cohen, and Zachary C Lipton · 2019
Cited alongside, same era.
Certified adversarial robustness via randomized smoothing
Jeremy Cohen, Elan Rosenfeld, and Zico Kolter · 2019
Cited alongside, same era.
On the benefits of models with perceptually-aligned gradients
Gunjan Aggarwal, Abhishek Sinha, Nupur Kumari, and Mayank Singh · 2020
Cited alongside, same era.
Do adversarially robust imagenet models transfer better?
Hadi Salman, Andrew Ilyas, Logan Engstrom, Ashish Kapoor, and Aleksander Madry · 2020
Cited alongside, same era.
Your classifier is secretly an energy based model and you should treat it like one
Will Grathwohl, Kuan-Chieh Wang, Jörn-Henrik Jacobsen, David Duvenaud, Mohammad Norouzi, and Kevin Swersky · 2020
Cited alongside, same era.
Fairwashing explanations with off-manifold detergent
Christopher Anders, Plamen Pasliev, Ann-Kathrin Dombrowski, Klaus-Robert Müller, and Pan Kessel
Cited in the paper.
Elucidating the design space of diffusion-based generative models
Tero Karras, Miika Aittala, Timo Aila, and Samuli Laine · 2022
Later among the works it cites.
Which explanation should i choose? a function approximation perspective to characterizing post hoc explanations
Tessa Han, Suraj Srinivas, and Himabindu Lakkaraju · 2022
Later among the works it cites.
Enhancing diffusion-based image synthesis with robust classifier guidance
Bahjat Kawar, Roy Ganz, and Michael Elad · 2023
Closest in time.
Classifier robustness enhancement via test-time transformation
Tsachi Blau, Roy Ganz, Chaim Baskin, Michael Elad, and Alex Bronstein · 2023
Closest in time.
Do perceptually aligned gradients imply adversarial robustness?
Roy Ganz, Bahjat Kawar, and Michael Elad · 2023
Closest in time.
The manifold hypothesis for gradient-based explanations
Sebastian Bordt, Uddeshya Upadhyay, Zeynep Akata, and Ulrike von Luxburg · 2023
Closest in time.