Scaling provable adversarial defenses
Wong, E., Schmidt, F., Metzen, J. H., and Kolter, J. Z · 2018
Later among the works it cites.
Feature squeezing: Detecting adversarial examples in deep neural networks
Xu, W., Evans, D., and Qi, Y · 2018
Later among the works it cites.
Certified adversarial robustness via randomized smoothing
Original
Cohen, J. M., Rosenfeld, E., and Kolter, J. Z · 2019
Later among the works it cites.
Simple black-box adversarial attacks
Original
Guo, C., Gardner, J. R., You, Y., Gordon Wilson, A., and Q. Weinberger, K · 2019
Later among the works it cites.
Ciidefence: Defeating adversarial attacks by fusing class-specific image inpainting and image denoising
Gupta, P. and Rahtu, E · 2019
Later among the works it cites.
Functional adversarial attacks
Laidlaw, C. and Feizi, S · 2019
Later among the works it cites.
Tight certificates of adversarial robustness for randomly smoothed classifiers
Lee, G.-H., Yuan, Y., Chang, S., and Jaakkola, T · 2019
Later among the works it cites.
Certified adversarial robustness with additive noise
Li, B., Chen, C., Wang, W., and Carin, L · 2019
Later among the works it cites.
Wasserstein adversarial examples via projected sinkhorn iterations
Original
Wong, E., Schmidt, F. R., and Kolter, J. Z · 2019
Later among the works it cites.
Feature denoising for improving adversarial robustness
Xie, C., Wu, Y., Maaten, L. v. d., Yuille, A. L., and He, K · 2019
Later among the works it cites.
A framework for robustness certification of smoothed classifiers using f-divergences
Dvijotham, K. D., Hayes, J., Balle, B., Kolter, Z., Qin, C., Gyorgy, A., Xiao, K., Gowal, S., and Kohli, P · 2020
Closest in time.
Improved image wasserstein attacks and defenses
Original
Hu, J. E., Swaminathan, A., Salman, H., and Yang, G · 2020
Closest in time.
Randomized smoothing of all shapes and sizes, 2020
Yang, G., Duan, T., Hu, J. E., Salman, H., Razenshteyn, I., and Li, J · 2020
Closest in time.
Macer: Attack-free and scalable robust training via maximizing certified radius
Zhai, R., Dan, C., He, D., Zhang, H., Gong, B., Ravikumar, P., Hsieh, C.-J., and Wang, L · 2020
Closest in time.