Fetching the paper…
Reading the bibliography…
Adversarial attacks, particularly \textbf{targeted} transfer-based attacks, can be used to assess the adversarial robustness of large visual-language models (VLMs), allowing for a more thorough examination of potential security flaws before deployment.
Z. Wang, A. C. Bovik, H. R. Sheikh, and E. P. Simoncelli, “Image quality assessment: from error visibility to structural similarity,” IEEE transactions on image processing , vol. 13, no. 4, pp. 600–612, 2004
2004
Earlier work this paper cites.
A. Hyvärinen and P. Dayan, “Estimation of non-normalized statistical models by score matching.” Journal of Machine Learning Research , vol. 6, no. 4, 2005
2005
Earlier work this paper cites.
A. Mittal, A. K. Moorthy, and A. C. Bovik, “Blind/referenceless image spatial quality evaluator,” in 2011 conference record of the forty fifth asilomar conference on signals, systems and computers (ASILOMAR) . IEEE, 2011, pp. 723–727
2011
Earlier work this paper cites.
2016
Earlier work this paper cites.
B. Zhou, A. Khosla, A. Lapedriza, A. Oliva, and A. Torralba, “Learning deep features for discriminative localization,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 2921–2929
2016
Earlier work this paper cites.
R. R. Selvaraju, M. Cogswell, A. Das, R. Vedantam, D. Parikh, and D. Batra, “Grad-cam: Visual explanations from deep networks via gradient-based localization,” in Proceedings of the IEEE International Conference on Computer Vision , 2017, pp. 618–626
2017
Earlier work this paper cites.
M. Heusel, H. Ramsauer, T. Unterthiner, B. Nessler, and S. Hochreiter, “Gans trained by a two time-scale update rule converge to a local nash equilibrium,” in Advances in Neural Information Processing Systems , vol. 30, 2017, pp. 6629–6640
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
Y. Song, R. Shu, N. Kushman, and S. Ermon, “Constructing unrestricted adversarial examples with generative models,” in Advances in Neural Information Processing Systems , vol. 31, 2018, pp. 8322–8333
2018
Earlier work this paper cites.
Z. Zhao, D. Dua, and S. Singh, “Generating natural adversarial examples,” in International Conference on Learning Representations , 2018. [Online]. Available: https://openreview.net/forum?id=H1BLjgZCb
2018
Earlier work this paper cites.
A. Madry, A. Makelov, L. Schmidt, D. Tsipras, and A. Vladu, “Towards deep learning models resistant to adversarial attacks,” in International Conference on Learning Representations , 2018. [Online]. Available: https://openreview.net/forum?id=rJzIBfZAb
2018
Earlier work this paper cites.
Y. Dong, F. Liao, T. Pang, H. Su, J. Zhu, X. Hu, and J. Li, “Boosting adversarial attacks with momentum,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 9185–9193
2018
Earlier work this paper cites.
X. Xu, X. Chen, C. Liu, A. Rohrbach, T. Darrell, and D. Song, “Fooling vision and language models despite localization and attention mechanism,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 4951–4961
2018
Earlier work this paper cites.
R. Zhang, P. Isola, A. A. Efros, E. Shechtman, and O. Wang, “The unreasonable effectiveness of deep features as a perceptual metric,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 586–595
2018
Earlier work this paper cites.
A. Chattopadhay, A. Sarkar, P. Howlader, and V. N. Balasubramanian, “Grad-cam++: Generalized gradient-based visual explanations for deep convolutional networks,” in 2018 IEEE winter conference on applications of computer vision (WACV) . IEEE, 2018, pp. 839–847
2018
Earlier work this paper cites.
Y. Xu, B. Wu, F. Shen, Y. Fan, Y. Zhang, H. T. Shen, and W. Liu, “Exact adversarial attack to image captioning via structured output learning with latent variables,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2019, pp. 4135–4144
2019
Earlier work this paper cites.
Y. Song and S. Ermon, “Generative modeling by estimating gradients of the data distribution,” Advances in neural information processing systems , vol. 32, 2019
2019
Earlier work this paper cites.
B. Sun, N.-h. Tsai, F. Liu, R. Yu, and H. Su, “Adversarial defense by stratified convolutional sparse coding,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2019, pp. 11 447–11 456
2019
Earlier work this paper cites.
A. S. Shamsabadi, R. Sanchez-Matilla, and A. Cavallaro, “Colorfool: Semantic adversarial colorization,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 1151–1160
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell et al. , “Language models are few-shot learners,” Advances in neural information processing systems , vol. 33, pp. 1877–1901, 2020
2020
Earlier work this paper cites.
C. Raffel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, and P. J. Liu, “Exploring the limits of transfer learning with a unified text-to-text transformer,” The Journal of Machine Learning Research , vol. 21, no. 1, pp. 5485–5551, 2020
2020
Earlier work this paper cites.
J. Ho, A. Jain, and P. Abbeel, “Denoising diffusion probabilistic models,” Advances in Neural Information Processing Systems , vol. 33, pp. 6840–6851, 2020
2020
Earlier work this paper cites.
J. Song, C. Meng, and S. Ermon, “Denoising diffusion implicit models,” in International Conference on Learning Representations , 2021. [Online]. Available: https://openreview.net/forum?id=St1giarCHLP
2021
Earlier work this paper cites.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark et al. , “Learning transferable visual models from natural language supervision,” in International Conference on Machine Learning . PMLR, 2021, pp. 8748–8763
2021
Earlier work this paper cites.
P.-T. Jiang, C.-B. Zhang, Q. Hou, M.-M. Cheng, and Y. Wei, “Layercam: Exploring hierarchical class activation maps for localization,” IEEE Transactions on Image Processing , vol. 30, pp. 5875–5888, 2021
2021
Cited alongside, same era.
J. Li, D. Li, C. Xiong, and S. Hoi, “Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,” in International Conference on Machine Learning . PMLR, 2022, pp. 12 888–12 900
2022
Cited alongside, same era.
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer, “High-resolution image synthesis with latent diffusion models,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 10 684–10 695
2022
Cited alongside, same era.
N. Aafaq, N. Akhtar, W. Liu, M. Shah, and A. Mian, “Language model agnostic gray-box adversarial attack on image captioning,” IEEE Transactions on Information Forensics and Security , vol. 18, pp. 626–638, 2022
2022
W.-L. Chiang, Z. Li, Z. Lin, Y. Sheng, Z. Wu, H. Zhang, L. Zheng, S. Zhuang, Y. Zhuang, J. E. Gonzalez et al. , “Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality,” See https://vicuna. lmsys. org (accessed 14 April 2023) , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
J. Wu, W. Gan, Z. Chen, S. Wan, and S. Y. Philip, “Multimodal large language models: A survey,” in 2023 IEEE International Conference on Big Data (BigData) . IEEE, 2023, pp. 2247–2256
2023
Later among the works it cites.
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Y. Xiong, J. Lin, M. Zhang, J. E. Hopcroft, and K. He, “Stochastic variance reduced ensemble adversarial attack for boosting the adversarial transferability,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, pp. 14 983–14 992
2022
Cited alongside, same era.
Y. Long, Q. Zhang, B. Zeng, L. Gao, X. Liu, J. Zhang, and J. Song, “Frequency domain model augmentation for adversarial attack,” in European Conference on Computer Vision . Springer, 2022, pp. 549–566
2022
Cited alongside, same era.
J.-B. Alayrac, J. Donahue, P. Luc, A. Miech, I. Barr, Y. Hasson, K. Lenc, A. Mensch, K. Millican, M. Reynolds et al. , “Flamingo: a visual language model for few-shot learning,” Advances in neural information processing systems , vol. 35, pp. 23 716–23 736, 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
C.-H. Ho and N. Vasconcelos, “Disco: Adversarial defense with local implicit functions,” Advances in Neural Information Processing Systems , vol. 35, pp. 23 818–23 837, 2022
2022
Cited alongside, same era.
W. Nie, B. Guo, Y. Huang, C. Xiao, A. Vahdat, and A. Anandkumar, “Diffusion models for adversarial purification,” in International Conference on Machine Learning , 2022, pp. 1–23
2022
Cited alongside, same era.
J. Li, D. Li, S. Savarese, and S. Hoi, “BLIP-2: bootstrapping language-image pre-training with frozen image encoders and large language models,” in International Conference on Machine Learning , 2023, pp. 1–13
2023
Cited alongside, same era.
H. Liu, C. Li, Q. Wu, and Y. J. Lee, “Visual instruction tuning,” in Thirty-seventh Conference on Neural Information Processing Systems , 2023. [Online]. Available: https://openreview.net/forum?id=w0H2xGHlkw
2023
Cited alongside, same era.
S. Han, C. Lin, C. Shen, Q. Wang, and X. Guan, “Interpreting adversarial examples in deep learning: A review,” ACM Computing Surveys , vol. 55, no. 14s, pp. 1–38, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Z. Cai, Y. Tan, and M. S. Asif, “Ensemble-based blackbox attacks on dense prediction,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 4045–4055
2023
Later among the works it cites.
Y. Zhu and Y. Jiang, “A non-global disturbance targeted adversarial example algorithm combined with c&w and grad-cam,” Neural Computing and Applications , vol. 35, no. 29, pp. 21 633–21 644, 2023
2023
Later among the works it cites.
S. H. Vemprala, R. Bonatti, A. Bucker, and A. Kapoor, “Chatgpt for robotics: Design principles and model abilities,” IEEE Access , 2024
2024
Closest in time.
G. Liao, J. Li, and X. Ye, “Vlm2scene: Self-supervised image-text-lidar learning with foundation models for autonomous driving scene understanding,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 38, no. 4, 2024, pp. 3351–3359
2024
Closest in time.
C. Cui, Y. Ma, X. Cao, W. Ye, Y. Zhou, K. Liang, J. Chen, J. Lu, Z. Yang, K.-D. Liao et al. , “A survey on multimodal large language models for autonomous driving,” in Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision , 2024, pp. 958–979
2024
Closest in time.
P. Nagesh, B. Prabha, S. B. Gole, G. Rao, and N. V. Ramana, “Visual assistance for visually impaired people using image caption and text to speech,” in AIP Conference Proceedings , vol. 2512, no. 1. AIP Publishing, 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
J. Zheng, C. Lin, J. Sun, Z. Zhao, Q. Li, and C. Shen, “Physical 3d adversarial attacks against monocular depth estimation in autonomous driving,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, pp. 24 452–24 461
2024
Closest in time.
X. Jia, Y. Chen, X. Mao, R. Duan, J. Gu, R. Zhang, H. Xue, Y. Liu, and X. Cao, “Revisiting and exploring efficient fast adversarial training via law: Lipschitz regularization and auto weight averaging,” IEEE Transactions on Information Forensics and Security , 2024
2024
Closest in time.
X. Jia, Y. Zhang, X. Wei, B. Wu, K. Ma, J. Wang, and X. Cao, “Improving fast adversarial training with prior-guided knowledge,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 2024
2024
Closest in time.
H. Chen, Y. Zhang, Y. Dong, and J. Zhu, “Rethinking model ensemble in transfer-based adversarial attacks,” in The Twelfth International Conference on Learning Representations , 2024. [Online]. Available: https://openreview.net/forum?id=AcJrSoArlh
2024
Closest in time.
Y. Wang, Y. Wu, S. Wu, X. Liu, W. Zhou, L. Zhu, and C. Zhang, “Boosting the transferability of adversarial attacks with frequency-aware perturbation,” IEEE Transactions on Information Forensics and Security , vol. 19, pp. 6293–6304, 2024
2024
Closest in time.
J. C. Costa, T. Roxo, H. Proença, and P. R. Inácio, “How deep learning sees the world: A survey on adversarial attacks & defenses,” IEEE Access , 2024
2024
Closest in time.
Q. Li, Q. Hu, H. Fan, C. Lin, C. Shen, and L. Wu, “Attention-sa: Exploiting model-approximated data semantics for adversarial attack,” IEEE Transactions on Information Forensics and Security , pp. 1–1, 2024
2024
Closest in time.
D.-T. Peng, J. Dong, M. Zhang, J. Yang, and Z. Wang, “Csfadv: Critical semantic fusion guided least-effort adversarial example attacks,” IEEE Transactions on Information Forensics and Security , vol. 19, pp. 5940–5955, 2024
2024
Closest in time.
X. Jia, Y. Zhang, X. Wei, B. Wu, K. Ma, J. Wang, and X. Cao, “Improving fast adversarial training with prior-guided knowledge,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 2024
2024
Closest in time.
M. Yoshida, H. Namura, and M. Okuda, “Adversarial examples for image cropping: Gradient-based and bayesian-optimized approaches for effective adversarial attack,” IEEE Access , 2024
2024
Closest in time.