Comparative Study of Three Imputation Methods to Treat Missing Values
Journal: INTERNATIONAL JOURNAL OF COMPUTERS & TECHNOLOGY (Vol.11, No. 7)Publication Date: 2013-12-11
Authors : Rahul Singhai;
Page : 2779-2786
Keywords : Knowledge Discovery In database; Data mining; Imputation methods; Sampling. Attribute missing values; Data preprocessing.;
Abstract
One relevant problem in data preprocessing is the presence of missing data that leads the poor quality of patterns, extracted after mining. Imputation is one of the widely used procedures that replace the missing values in a data set by some probable values. The advantage of this approach is that the missing data treatment is independent of the learning algorithm used. This allows the user to select the most suitable imputation method for each situation. This paper analyzes the various imputation methods proposed in the field of statistics with respect to data mining. A comparative analysis of three different imputation approaches which can be used to impute missing attribute values in data mining are given that shows the most promising method. An artificial input data (of numeric type) file of 1000 records is used to investigate the performance of these methods. For testing the significance of these methods Z-test approach were used.
Other Latest Articles
- Humanoid Robot Learning How to track and grip
- Performance Analysis of Malicious nodes on Multi hop Cellular Networks
- Depression Analysis using ECG Signal
- Studying the Effect of Paracetamol Drug on the Conductivity of 0.5M Hydrochloric Acid Solution at Different Temperatures
- Semantic link-based Model for User Recommendation in Online community
Last modified: 2016-06-29 18:39:51