MARATTO

article · Procedia Computer Science

Evaluating Large Language Models for Arabic Sentiment Analysis: A Comparative Study Using Retrieval-Augmented Generation

20245 citationsOpen accessNile University

Abstract

Sentiment Analysis (SA) is a crucial task in natural language processing. There are numerous studies devoted to this field specially with emerging use of encoder transformers and generative Large Language Models (LLMs). Transformer based models like BERT are often used for sentiment analysis. However, this article examines the performance of the latest generative generative LLMs in Arabic Sentiment Analysis (ASA) using Retrieval-Augmented Generation (RAG) architecture. We evaluated these models using the ASAD, ArSarcasm-v2, and SemEval datasets. Our experimental studies revealed challenges due to dataset imbalances and misclassified neutral labels, which impacted the effectiveness of fine-tuning. By removing the neutral class, significant improvements in model performance were observed across all datasets, with F1-scores increasing by 17%, 18%, and 18% on ASAD, ArSarcasm-v2, and SemEval, respectively.

Research topics

  • Sentiment Analysis and Opinion Mining
  • Topic Modeling
  • Advanced Text Analysis Techniques

Sustainable Development Goals

Read the original research

This page summarises published work. The authoritative version sits with the publisher.

DOI: 10.1016/j.procs.2024.10.210

Is something wrong with this record? Report it or request removal.

Discussion

Discuss this research

Have you built on this work, tried to replicate it, or seen it applied in practice? Share what you know. Verified researchers and MARATTO™ domain experts can open a discussion, and any member can reply. Contributions are reviewed before they appear.

No discussion yet. Open the first thread.