Enhancing Arabic Dialect Detection on Social Media: A Hybrid Model with an Attention Mechanism

被引:2
|
作者
Yafooz, Wael M. S. [1 ]
机构
[1] Taibah Univ, Coll Comp Sci & Engn, Comp Sci Dept, Medina 42353, Saudi Arabia
关键词
Arabic text identification; LSTM; BiLSM; deep learning; social media; IDENTIFICATION;
D O I
10.3390/info15060316
中图分类号
TP [自动化技术、计算机技术];
学科分类号
0812 ;
摘要
Recently, the widespread use of social media and easy access to the Internet have brought about a significant transformation in the type of textual data available on the Web. This change is particularly evident in Arabic language usage, as the growing number of users from diverse domains has led to a considerable influx of Arabic text in various dialects, each characterized by differences in morphology, syntax, vocabulary, and pronunciation. Consequently, researchers in language recognition and natural language processing have become increasingly interested in identifying Arabic dialects. Numerous methods have been proposed to recognize this informal data, owing to its crucial implications for several applications, such as sentiment analysis, topic modeling, text summarization, and machine translation. However, Arabic dialect identification is a significant challenge due to the vast diversity of the Arabic language in its dialects. This study introduces a novel hybrid machine and deep learning model, incorporating an attention mechanism for detecting and classifying Arabic dialects. Several experiments were conducted using a novel dataset that collected information from user-generated comments from Twitter of Arabic dialects, namely, Egyptian, Gulf, Jordanian, and Yemeni, to evaluate the effectiveness of the proposed model. The dataset comprises 34,905 rows extracted from Twitter, representing an unbalanced data distribution. The data annotation was performed by native speakers proficient in each dialect. The results demonstrate that the proposed model outperforms the performance of long short-term memory, bidirectional long short-term memory, and logistic regression models in dialect classification using different word representations as follows: term frequency-inverse document frequency, Word2Vec, and global vector for word representation.
引用
收藏
页数:20
相关论文
共 50 条
  • [1] Arabic dialect identification in social media: A hybrid model with transformer models and BiLSTM
    Alsuwaylimi, Amjad A.
    HELIYON, 2024, 10 (17)
  • [2] A hybrid depression detection model and correlation analysis for social media based on attention mechanism
    Liu, Jiacheng
    Chen, Wanzhen
    Wang, Liangxu
    Ding, Fangyikuang
    INTERNATIONAL JOURNAL OF MACHINE LEARNING AND CYBERNETICS, 2024, 15 (07) : 2631 - 2642
  • [3] A Bi-GRU with attention and CapsNet hybrid model for cyberbullying detection on social media
    Kumar, Akshi
    Sachdeva, Nitin
    WORLD WIDE WEB-INTERNET AND WEB INFORMATION SYSTEMS, 2022, 25 (04): : 1537 - 1550
  • [4] A Bi-GRU with attention and CapsNet hybrid model for cyberbullying detection on social media
    Akshi Kumar
    Nitin Sachdeva
    World Wide Web, 2022, 25 : 1537 - 1550
  • [5] Advancing arabic dialect detection with hybrid stacked transformer models
    Saleh, Hager
    Almohimeed, Abdulaziz
    Hassan, Rasha
    Ibrahim, Mandour M.
    Alsamhi, Saeed Hamood
    Hassan, Moatamad Refaat
    Mostafa, Sherif
    FRONTIERS IN HUMAN NEUROSCIENCE, 2025, 19
  • [6] Arabic Event Detection in Social Media
    Alsaedi, Nasser
    Burnap, Pete
    COMPUTATIONAL LINGUISTICS AND INTELLIGENT TEXT PROCESSING (CICLING 2015), PT I, 2015, 9041 : 384 - 401
  • [7] A deep learning approach for the depression detection of social media data with hybrid feature selection and attention mechanism
    Bhuvaneswari, M.
    Prabha, V. Lakshmi
    EXPERT SYSTEMS, 2023, 40 (09)
  • [8] Rumor Detection on Social Media: A Multi-view Model Using Self-attention Mechanism
    Geng, Yue
    Lin, Zheng
    Fu, Peng
    Wang, Weiping
    COMPUTATIONAL SCIENCE - ICCS 2019, PT I, 2019, 11536 : 339 - 352
  • [9] Arabic named entity recognition in social media based on BiLSTM-CRF using an attention mechanism
    Benali, B. Ait
    Mihi, S.
    Mlouk, A. Ait
    El Bazi, I
    Laachfoubi, N.
    JOURNAL OF INTELLIGENT & FUZZY SYSTEMS, 2022, 42 (06) : 5427 - 5436
  • [10] Dynamic graph convolutional networks with attention mechanism for rumor detection on social media
    Choi, Jiho
    Ko, Taewook
    Choi, Younhyuk
    Byun, Hyungho
    Kim, Chong-kwon
    PLOS ONE, 2021, 16 (08):