A Closer Look at Arabic Text Classification

The world has witnessed an information explosion in the past two decades. Electronic devices are now available in many varieties such as PCs, Laptops, book readers, mobile devices and with relatively affordable prices. This and the ubiquitous use of software applications such as social media and clo...

Full description

Saved in:

Bibliographic Details
Published in	International journal of advanced computer science & applications Vol. 10; no. 11
Main Authors	Abdeen, Mohammad A R, AlBouq, Sami, Elmahalawy, Ahmed, Shehata, Sara
Format	Journal Article
Language	English
Published	West Yorkshire Science and Information (SAI) Organization Limited 2019
Subjects	Algorithms Applications programs Classification Cloud computing Data mining Digitization Electronic devices Machine learning Scientific papers Statistical methods Support vector machines Text categorization
Online Access	Get full text

Cover

Loading…

More Information
Summary:	The world has witnessed an information explosion in the past two decades. Electronic devices are now available in many varieties such as PCs, Laptops, book readers, mobile devices and with relatively affordable prices. This and the ubiquitous use of software applications such as social media and cloud applications, and the increasing trend towards digitalization, the amount of information on the global cloud has surged to an unprecedented level. Therefore, a dire need exists in order to mine this massively large amount of data and produce meaningful information. Text Classification is one of the known and well established data mining techniques that has been used and reported in the literature. Text classification methods include statistical and machine learning algorithms such as Naive Baysian, Support Vector Machines and others have widely been used. Many works have been reported regarding text classification of various languages including English, Chinese, Russian, and many others. Arabic is the fifth most spoken language in the world. There has been many works in the literature for Arabic text classification. However, and to the best of our knowledge, there is no recent work that presents a good, critical and comprehensive survey of the Arabic text classification for the past two decades. The aim of this paper is to present a concise and yet comprehensive review of the Arabic text classification. We have covered over 50 research papers covering the past two decades (2000 - 2019). The main focus of this paper is to address the following issues: 1) The techniques reported in the literature including. 2) New Techniques. 3) Most claimed efficient technique. 4) Datasets used and which ones are most popular. 5) Which feature selection techniques are used? 6) Popular classes/categories used. 7) Effect of stemming techniques on classification results.
ISSN:	2158-107X 2156-5570
DOI:	10.14569/IJACSA.2019.0101189