A. Alamsyah, T. Fadila
Data mining has been widely used to diagnose diseases from medical data. Classification is a data mining technique that can be used to predict disease. In previous studies, a support vector machine was widely used to obtain high accuracy in predicting hepatitis. In this study, the principal component analysis was applied to the support vector machine. A principal component analysis is used to extract features and reduce the number of features or attributes. Principal component analysis can reduce data dimensions without removing important information from the dataset. The extracted and reduced data are then used to classify the support vector machine. Classification performance measurement is done by using a confusion matrix. Hepatitis prediction accuracy achieved was 93.55%. This result is better than the support vector machine classification results without the application of principal component analysis. © Published under licence by IOP Publishing Ltd.
Department of Computer Science, Faculty of Mathematics and Natural Science, Universitas Negeri Semarang, Semarang, Indonesia