Fetching the paper…
Reading the bibliography…
In recent years, the Vision Transformer (ViT) model has gradually become mainstream in various computer vision tasks, and the robustness of the model has received increasing attention.
H. Robbins and S. Monro, “A stochastic approximation method,” The annals of mathematical statistics , pp. 400–407, 1951
1951
Earlier work this paper cites.
A. Krizhevsky, “Learning multiple layers of features from tiny images,” 2009
2009
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in NeurIPS , 2012, pp. 1106–1114
2012
Earlier work this paper cites.
C. Szegedy, W. Zaremba, I. Sutskever, J. Bruna, D. Erhan, I. J. Goodfellow, and R. Fergus, “Intriguing properties of neural networks,” in ICLR , 2014
2014
Earlier work this paper cites.
K. Simonyan and A. Zisserman, “Very deep convolutional networks for large-scale image recognition,” in ICLR , 2015
2015
Earlier work this paper cites.
L. Chen, G. Papandreou, I. Kokkinos, K. Murphy, and A. L. Yuille, “Semantic image segmentation with deep convolutional nets and fully connected crfs,” in ICLR , 2015
2015
Earlier work this paper cites.
I. J. Goodfellow, J. Shlens, and C. Szegedy, “Explaining and harnessing adversarial examples,” in ICLR , 2015
2015
Earlier work this paper cites.
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. S. Bernstein, A. C. Berg, and L. Fei-Fei, “Imagenet large scale visual recognition challenge,” IJCV , vol. 115, no. 3, pp. 211–252, 2015
2015
Earlier work this paper cites.
Y. Le and X. S. Yang, “Tiny imagenet visual recognition challenge,” 2015
2015
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in CVPR , 2016, pp. 770–778
2016
Earlier work this paper cites.
C. Szegedy, V. Vanhoucke, S. Ioffe, J. Shlens, and Z. Wojna, “Rethinking the inception architecture for computer vision,” in CVPR , 2016, pp. 2818–2826
2016
Earlier work this paper cites.
W. Liu, Y. Wen, Z. Yu, M. Li, B. Raj, and L. Song, “Sphereface: Deep hypersphere embedding for face recognition,” in CVPR , 2017, pp. 6738–6746
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
A. Kurakin, I. J. Goodfellow, and S. Bengio, “Adversarial machine learning at scale,” in ICLR , 2017
2017
Earlier work this paper cites.
N. Carlini and D. A. Wagner, “Towards evaluating the robustness of neural networks,” in 2017 IEEE Symposium on Security and Privacy, SP 2017, San Jose, CA, USA, May 22-26, 2017 , 2017, pp. 39–57
2017
Earlier work this paper cites.
G. Huang, Z. Liu, L. van der Maaten, and K. Q. Weinberger, “Densely connected convolutional networks,” in CVPR , 2017, pp. 2261–2269
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin, “Attention is all you need,” in NeurIPS , 2017, pp. 5998–6008
2017
Earlier work this paper cites.
P. Chen, H. Zhang, Y. Sharma, J. Yi, and C. Hsieh, “ZOO: zeroth order optimization based black-box attacks to deep neural networks without training substitute models,” in Proceedings of the 10th ACM Workshop on Artificial Intelligence and Security, AISec@CCS 2017, Dallas, TX, USA, November 3, 2017 . ACM, 2017, pp. 15–26
2017
Earlier work this paper cites.
H. Wang, Y. Wang, Z. Zhou, X. Ji, D. Gong, J. Zhou, Z. Li, and W. Liu, “Cosface: Large margin cosine loss for deep face recognition,” in CVPR , 2018, pp. 5265–5274
2018
Earlier work this paper cites.
——, “Deeplab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected crfs,” IEEE TPAMI , vol. 40, no. 4, pp. 834–848, 2018
2018
Earlier work this paper cites.
Y. Dong, F. Liao, T. Pang, H. Su, J. Zhu, X. Hu, and J. Li, “Boosting adversarial attacks with momentum,” in CVPR , 2018, pp. 9185–9193
2018
Earlier work this paper cites.
A. Madry, A. Makelov, L. Schmidt, D. Tsipras, and A. Vladu, “Towards deep learning models resistant to adversarial attacks,” in ICLR , 2018
2018
Earlier work this paper cites.
J. Deng, J. Guo, N. Xue, and S. Zafeiriou, “Arcface: Additive angular margin loss for deep face recognition,” in CVPR , 2019, pp. 4690–4699
2019
Earlier work this paper cites.
H. Zhang, Y. Yu, J. Jiao, E. P. Xing, L. E. Ghaoui, and M. I. Jordan, “Theoretically principled trade-off between robustness and accuracy,” in ICML , vol. 97, 2019, pp. 7472–7482
2019
Earlier work this paper cites.
A. Shafahi, M. Najibi, A. Ghiasi, Z. Xu, J. P. Dickerson, C. Studer, L. S. Davis, G. Taylor, and T. Goldstein, “Adversarial training for free!” in NeurIPS , 2019, pp. 3353–3364
2019
Cited alongside, same era.
N. Houlsby, A. Giurgiu, S. Jastrzebski, B. Morrone, Q. de Laroussilhe, A. Gesmundo, M. Attariyan, and S. Gelly, “Parameter-efficient transfer learning for NLP,” in ICML , vol. 97, 2019, pp. 2790–2799
2019
Cited alongside, same era.
A. Athalye, N. Carlini, and D. A. Wagner, “Obfuscated gradients give a false sense of security: Circumventing defenses to adversarial examples,” in ICML , ser. Proceedings of Machine Learning Research, vol. 80. PMLR, 2018, pp. 274–283
2019
Cited alongside, same era.
N. Carion, F. Massa, G. Synnaeve, N. Usunier, A. Kirillov, and S. Zagoruyko, “End-to-end object detection with transformers,” in ECCV , vol. 12346, 2020, pp. 213–229
2020
Cited alongside, same era.
Y. Fu, S. Zhang, S. Wu, C. Wan, and Y. Lin, “Patch-fool: Are vision transformers always robust against adversarial perturbations?” in ICLR , 2022
2022
Later among the works it cites.
W. Nie, B. Guo, Y. Huang, C. Xiao, A. Vahdat, and A. Anandkumar, “Diffusion models for adversarial purification,” in ICML , vol. 162, 2022, pp. 16 805–16 827
2022
Later among the works it cites.
B. Wu, J. Gu, Z. Li, D. Cai, X. He, and W. Liu, “Towards efficient adversarial training on vision transformers,” in ECCV , vol. 13673, 2022, pp. 307–325
2022
Later among the works it cites.
E. J. Hu, Y. Shen, P. Wallis, Z. Allen-Zhu, Y. Li, S. Wang, L. Wang, and W. Chen, “Lora: Low-rank adaptation of large language models,” in ICLR , 2022
2022
Later among the works it cites.
Y. Mao, L. Mathias, R. Hou, A. Almahairi, H. Ma, J. Han, S. Yih, and M. Khabsa, “Unipelt: A unified framework for parameter-efficient language model tuning,” in ACL , 2022, pp. 6253–6264
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. Li, S. Zhu, S. Paul, A. K. Roy-Chowdhury, C. Song, S. V. Krishnamurthy, A. Swami, and K. S. Chan, “Connecting the dots: Detecting adversarial perturbations using context inconsistency,” in ECCV , vol. 12368, 2020, pp. 396–413
2020
Cited alongside, same era.
F. Croce and M. Hein, “Provable robustness against all adversarial $l_p$-perturbations for $p\geq 1$,” in ICLR , 2020
2020
Cited alongside, same era.
E. Wong, L. Rice, and J. Z. Kolter, “Fast is better than free: Revisiting adversarial training,” in ICLR , 2020
2020
Cited alongside, same era.
F. Croce and M. Hein, “Reliable evaluation of adversarial robustness with an ensemble of diverse parameter-free attacks,” in ICML , vol. 119, 2020, pp. 2206–2216
2020
Cited alongside, same era.
F. Yang, H. Yang, J. Fu, H. Lu, and B. Guo, “Learning texture transformer network for image super-resolution,” in CVPR , 2020, pp. 5790–5799
2020
Cited alongside, same era.
Y. Wang, D. Zou, J. Yi, J. Bailey, X. Ma, and Q. Gu, “Improving adversarial robustness requires revisiting misclassified examples,” in ICLR , 2020
2020
Cited alongside, same era.
J. Howard and S. Gugger, “Fastai: A layered API for deep learning,” Inf. , vol. 11, no. 2, p. 108, 2020
2020
Cited alongside, same era.
A. Dosovitskiy, L. Beyer, A. Kolesnikov, D. Weissenborn, X. Zhai, T. Unterthiner, M. Dehghani, M. Minderer, G. Heigold, S. Gelly, J. Uszkoreit, and N. Houlsby, “An image is worth 16x16 words: Transformers for image recognition at scale,” in ICLR , 2021
2021
Cited alongside, same era.
2022
Later among the works it cites.
W. Yu, M. Luo, P. Zhou, C. Si, Y. Zhou, X. Wang, J. Feng, and S. Yan, “Metaformer is actually what you need for vision,” in CVPR , 2022, pp. 10 809–10 819
2022
Later among the works it cites.
Y. Wang, T. Ye, L. Cao, W. Huang, F. Sun, F. He, and D. Tao, “Bridged transformer for vision and point cloud 3d object detection,” in CVPR , 2022, pp. 12 104–12 113
2022
Later among the works it cites.
B. Cheng, I. Misra, A. G. Schwing, A. Kirillov, and R. Girdhar, “Masked-attention mask transformer for universal image segmentation,” in CVPR , 2022, pp. 1280–1289
2022
Later among the works it cites.
2022
Later among the works it cites.
Y. Mo, D. Wu, Y. Wang, Y. Guo, and Y. Wang, “When adversarial training meets vision transformers: Recipes from training to architecture,” in NeurIPS , 2022
2022
Later among the works it cites.
M. Jia, L. Tang, B. Chen, C. Cardie, S. J. Belongie, B. Hariharan, and S. Lim, “Visual prompt tuning,” in ECCV , vol. 13693, 2022, pp. 709–727
2022
Later among the works it cites.
Y. Sung, J. Cho, and M. Bansal, “VL-ADAPTER: parameter-efficient transfer learning for vision-and-language tasks,” in CVPR , 2022, pp. 5217–5227
2022
Later among the works it cites.
K. Zhu, X. Hu, J. Wang, X. Xie, and G. Yang, “Improving generalization of adversarial training via robust critical fine-tuning,” in ICCV , 2023, pp. 4401–4411
2023
Later among the works it cites.
Z. Liu, Y. Xu, X. Ji, and A. B. Chan, “TWINS: A fine-tuning framework for improved transferability of adversarial robustness and generalization,” in CVPR , 2023, pp. 16 436–16 446
2023
Later among the works it cites.
Q. Zhang, M. Chen, A. Bukharin, P. He, Y. Cheng, W. Chen, and T. Zhao, “Adaptive budget allocation for parameter-efficient fine-tuning,” in ICLR , 2023
2023
Later among the works it cites.
H. Wang, X. Yang, J. Chang, D. Jin, J. Sun, S. Zhang, X. Luo, and Q. Tian, “Parameter-efficient tuning of large-scale multimodal foundation model,” in NeurIPS , 2023
2023
Later among the works it cites.
G. Wu, W. Zheng, Y. Lu, and Q. Tian, “PSLT: A light-weight vision transformer with ladder self-attention and progressive shift,” IEEE TPAMI , vol. 45, no. 9, pp. 11 120–11 135, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
S. Rebuffi, F. Croce, and S. Gowal, “Revisiting adapters with adversarial training,” in ICLR , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
S. Wang, J. Zhang, Z. Yuan, and S. Shan, “Pre-trained model guided fine-tuning for zero-shot adversarial robustness,” in CVPR , 2024, pp. 24 502–24 511
2024
Closest in time.
D. J. Kopiczko, T. Blankevoort, and Y. M. Asano, “Vera: Vector-based random matrix adaptation,” in ICLR , 2024
2024
Closest in time.