Imputing manufacturing material in data mining
Resource
Journal of Intelligent Manufacturing 19 (1): 109-118
Journal
Journal of Intelligent Manufacturing
Pages
109-118
Date Issued
2008
Date
2008
Author(s)
Yeh, Ruey-Ling
Liu, Ching
Shia, Ben-Chang
Cheng, Yu-Ting
Huwang, Ya-Fang
Abstract
Data plays a vital role as a source of information to organizations, especially in times of information and technology. One encounters a not-so-perfect database from which data is missing, and the results obtained from such a database may provide biased or misleading solutions. Therefore, imputing missing data to a database has been regarded as one of the major steps in data mining. The present research used different methods of data mining to construct imputative models in accordance with different types of missing data. When the missing data is continuous, regression models and Neural Networks are used to build imputative models. For the categorical missing data, the logistic regression model, neural network, C5.0 and CART are employed to construct imputative models. The results showed that the regression model was found to provide the best estimate of continuous missing data; but for categorical missing data, the C5.0 model proved the best method. © 2007 Springer Science+Business Media, LLC.
Type
journal article
File(s)![Thumbnail Image]()
Loading...
Name
10.pdf
Size
402.12 KB
Format
Adobe PDF
Checksum
(MD5):fcdbb377e37739c4d2f1ee3fe038ee15
