Arabic Text Mining

The rapid growth of the internet has increased the number of online texts. This led to the rapid growth of the number of online texts in the Arabic language. The enormous amount of text must be organized into classes to make the analysis process and text retrieval easier. Text classification is, therefore, a key component of text mining. There are numerous systems and approaches for categorizing literature in English, European (French, German, Spanish), and Asian (Chinese, Japanese). In contrast, there are relatively few studies on categorizing Arabic literature due to the difficulty of the Arabic language. In this work, a brief explanation of key ideas relevant to Arabic text mining are introduced then a new classification system for the Arabic language is presented using light stemming and Classifier Naïve Bayesian (CNB). Texts from two classes: politics and sports, are included in our corpus. Some texts are added to the system, and the system correctly classified them, demonstrating the effectiveness of the system.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
10/24/2021

Transliterating Kurdish texts in Latin into Persian-Arabic script

Kurdish is written in different scripts. The two most popular scripts ar...
research
03/18/2022

Offensive Language Detection in Under-resourced Algerian Dialectal Arabic Language

This paper addresses the problem of detecting the offensive and abusive ...
research
05/12/2015

A Survey of Arabic Dialogues Understanding for Spontaneous Dialogues and Instant Message

Building dialogues systems interaction has recently gained considerable ...
research
05/18/2022

BFCAI at SemEval-2022 Task 6: Multi-Layer Perceptron for Sarcasm Detection in Arabic Texts

This paper describes the systems submitted to iSarcasm shared task. The ...
research
11/09/2012

NF-SAVO: Neuro-Fuzzy system for Arabic Video OCR

In this paper we propose a robust approach for text extraction and recog...
research
11/11/2020

Assessment of text coherence based on the cohesion estimation

In this paper, a graph-based coherence estimation method based on the co...
research
09/27/2017

A Preliminary Study for Building an Arabic Corpus of Pair Questions-Texts from the Web: AQA-Webcorp

With the development of electronic media and the heterogeneity of Arabic...

Please sign up or login with your details

Forgot password? Click here to reset