Arabic Dialects Identification for All Arabic countries
الباحث الأول:
Ahmed Hussein Aliwy
الباحثين الآخرين:
Hawraa Ali Taher
Zena A. Abutiheen
المجلة:
Barcelona, Spain (Online), December 12, 2020
تاريخ النشر:
None
مختصر البحث:
Arabic dialects are among of three main variant of Arabic language (Classical Arabic, modern standard Arabic and dialectal Arabic). It has many variants according to the country, city (provinces) or town. In this paper, several techniques with multi…
Arabic dialects are among of three main variant of Arabic language (Classical Arabic, modern standard Arabic and dialectal Arabic). It has many variants according to the country, city (provinces) or town. In this paper, several techniques with multiple algorithms are applied for Arabic dialects identification starting from removing noise till classification task using all Ar-abic countries as 21 classes. Three types of classifiers (Naïve Bayes, Logistic Regression, and Decision Tree) are combined using voting with two different methodologies. Also clustering technique is used for decreasing the noise that result from the existing of MSA tweets in the data set for training phase. The results of f measure were 27.17, 41.34 and 52.38 for first methodology without clustering, second methodology without clustering, and second method-ology with clustering, the used data set is NADI shared task data set.