2023

How Many Unicorns Are in This Image? A Safety Evaluation Benchmark for Vision LLMs

Tu, Haoqin, Cui, Chenhang, Wang, Zijun et al.

Understand

This work focuses on the potential of Vision LLMs (VLLMs) in visual reasoning.

  • Different from prior studies, we shift our focus from evaluating standard performance to introducing a comprehensive safety evaluation suite, covering both out-of-distribution (OOD) generalization and adversarial robustness.
  • For the OOD evaluation, we present two novel VQA datasets, each with one variant, designed to test model performance under challenging conditions.
  • In exploring adversarial robustness, we propose a straightforward attack strategy for misleading VLLMs to produce visual-unrelated responses.

Reading the bibliography…