2023

Taqyim: Evaluating Arabic NLP Tasks Using ChatGPT Models

Alyafeai, Zaid, Alshaibani, Maged S., AlKhamissi, Badr et al.

Understand

Large language models (LLMs) have demonstrated impressive performance on various downstream tasks without requiring fine-tuning, including ChatGPT, a chat-based model built on top of LLMs such as GPT-3.5 and GPT-4.

  • Despite having a lower training proportion compared to English, these models also exhibit remarkable capabilities in other languages.
  • In this study, we assess the performance of GPT-3.5 and GPT-4 models on seven distinct Arabic NLP tasks: sentiment analysis, translation, transliteration, paraphrasing, part of speech tagging, summarization, and diacritization.
  • Our findings reveal that GPT-4 outperforms GPT-3.5 on five out of the seven tasks.

Reading the bibliography…