Missing values imputation in Arabic datasets using enhanced robust association rules

Emran, Nurul Akmar and Draman @ Muda, Azah Kamilah and Thabet Salem, Salem Awsan and Zahriah, Sahri and Ali, Abdulrazzak (2022) Missing values imputation in Arabic datasets using enhanced robust association rules. Indonesian Journal of Electrical Engineering and Computer Science, 28 (2). pp. 1067-1075. ISSN 2502-4752

[img] Text
28012-58448-1-PB PUBLISHED.PDF

Download (457kB)

Abstract

Missing value (MV) is one form of data completeness problem in massive datasets. To deal with missing values, data imputation methods were proposed with the aim to improve the completeness of the datasets concerned. Data imputation's accuracy is a common indicator of a data imputation technique's efficiency. However, the efficiency of data imputation can be affected by the nature of the language in which the dataset is written. To overcome this problem, it is necessary to normalize the data, especially in non-Latin languages such as the Arabic language. This paper proposes a method that will address the challenge inherent in Arabic datasets by extending the enhanced robust association rules (ERAR) method with Arabic detection and correction functions. Iterative and Decision Tree methods were used to evaluate the proposed method in an experiment. Experiment results show that the proposed method offers a higher data imputation accuracy than the Iterative and decision tree methods.

Item Type: Article
Uncontrolled Keywords: Arabic dataset, Association rules, Data imputation, Missing values, Morphology
Divisions: Faculty of Information and Communication Technology
Depositing User: Sabariah Ismail
Date Deposited: 28 Mar 2023 13:46
Last Modified: 28 Mar 2023 13:46
URI: http://eprints.utem.edu.my/id/eprint/26431
Statistic Details: View Download Statistic

Actions (login required)

View Item View Item