Fetching the paper…

What Are We Measuring When We Evaluate Large Vision-Language Models? An Analysis of Latent Factors and Biases · Around