ArSentBERT: fine-tuned bidirectional encoder representations from transformers model for Arabic sentiment classification

Mohamed Fawzy Abdelfattah,Mohamed Abo Rizka,Mohamed Waleed Fakhr

doi:10.11591/eei.v12i2.3914

Abstract

Sentiment analysis in the Arabic language is challenging because of its linguistic complexity. Arabic is complex in words, paragraphs, and sentence structure. Moreover, most Arabic documents contain multiple dialects, writing alphabets, and styles (e.g., Franco-Arab). Nevertheless, fine-tuned bidirectional encoder representations from transformers (BERT) models can provide a reasonable prediction accuracy for Arabic sentiment classification tasks. This paper presents a fine-tuning approach for BERT models for classifying Arabic sentiments. It uses Arabic BERT pre-trained models and tokenizers and includes three stages. The first stage is text preprocessing and data cleaning. The second stage uses transfer-learning of the pre-trained models’ weights and trains all encoder layers. The third stage uses a fully connected layer and a drop-out layer for classification. We tested our fine-tuned models on five different datasets that contain reviews in Arabic with different dialects and compared the results to 11 state-of-the-art models. The experiment results show that our models provide better prediction accuracy than our competitors. We show that the choice of the pre-trained BERT model and the tokenizer type improves the accuracy of Arabic sentiment classification.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: Bulletin of Electrical Engineering and Informatics	Publication Date: Apr 1, 2023
Citations: 6	License type: CC BY-SA 4.0

R Discovery Prime

R Discovery Prime

ArSentBERT: fine-tuned bidirectional encoder representations from transformers model for Arabic sentiment classification

Abstract

Talk to us

Similar Papers

More From: Bulletin of Electrical Engineering and Informatics

Lead the way for us

Similar Papers

Bert model fine-tuning for text classification in knee OA radiology reports
L Chen ... V Pedoia
Osteoarthritis and Cartilage | VOL. 28
L Chen, et. al.L Chen ... V Pedoia
01 Apr 2020
Osteoarthritis and Cartilage | VOL. 28

Engineering Document Summarization Using Sentence Representations Generated by Bidirectional Language Model
Yan Jin ... Yunjian Qiu
-
Yan Jin, et. al.Yan Jin ... Yunjian Qiu
17 Aug 2021
17 Aug 2021

Contextual semantic embeddings based on fine-tuned AraBERT model for Arabic text multi-class categorization
Fatima-Zahra El-Alami ... Noureddine En Nahnahi
Journal of King Saud University - Computer and Information Sciences | VOL. 34
Fatima-Zahra El-Alami, et. al.Fatima-Zahra El-Alami ... Noureddine En Nahnahi
18 Feb 2021
Journal of King Saud University - Computer and Information Sciences | VOL. 34

Exploring the performance and explainability of fine-tuned BERT models for neuroradiology protocol assignment
Salmonn Talebi ... Mohammad R K Mofrad
BMC Medical Informatics and Decision Making | VOL. 24
Salmonn Talebi, et. al.Salmonn Talebi ... Mohammad R K Mofrad
07 Feb 2024
BMC Medical Informatics and Decision Making | VOL. 24

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

ArSentBERT: fine-tuned bidirectional encoder representations from transformers model for Arabic sentiment classification

Abstract

Talk to us

Similar Papers

More From: Bulletin of Electrical Engineering and Informatics