MARATTO

article · Procedia Computer Science

Text Correction for Modern Standard Arabic

20241 citationOpen accessNile University

Abstract

Arabic poses a unique challenge for Natural Language Processing tasks due to its morphological complexity, rich vocabulary, and syntactic flexibility. Mistakes written in Arabic are common even among fluent speakers, and slight mistakes can obstruct a word's meaning. Our project investigates Large Language Models (LLMs)’ capabilities for detecting and automatically correcting syntax and semantic errors in Arabic text. Our project includes an overview of Qatar Arabic Language Bank (QALB), a shared task on automatic correction of Arabic text, which focuses on correcting errors in Arabic text produced by native speakers. We used the QALB dataset for training and evaluation, achieving a WER score of 0.2203 and a GLEU of 0.5956.

Research topics

  • Natural Language Processing Techniques
  • Handwritten Text Recognition Techniques
  • Topic Modeling

Sustainable Development Goals

Read the original research

This page summarises published work. The authoritative version sits with the publisher.

DOI: 10.1016/j.procs.2024.10.211

Is something wrong with this record? Report it or request removal.

Discussion

Discuss this research

Have you built on this work, tried to replicate it, or seen it applied in practice? Share what you know. Verified researchers and MARATTO™ domain experts can open a discussion, and any member can reply. Contributions are reviewed before they appear.

No discussion yet. Open the first thread.