2024

CMMU: A Benchmark for Chinese Multi-modal Multi-type Question Understanding and Reasoning

He, Zheqi, Wu, Xinya, Zhou, Pengfei et al.

Understand

Multi-modal large language models(MLLMs) have achieved remarkable progress and demonstrated powerful knowledge comprehension and reasoning abilities.

  • However, the mastery of domain-specific knowledge, which is essential for evaluating the intelligence of MLLMs, continues to be a challenge.
  • Current multi-modal benchmarks for domain-specific knowledge concentrate on multiple-choice questions and are predominantly available in English, which imposes limitations on the comprehensiveness of the evaluation.
  • To this end, we introduce CMMU, a novel benchmark for multi-modal and multi-type question understanding and reasoning in Chinese.

Reading the bibliography…