Predicting cardiac autonomic neuropathy category for diabetic data with missing values |
| |
Authors: | Jemal Abawajy Andrei Kelarev Morshed Chowdhury Andrew Stranieri Herbert F. Jelinek |
| |
Affiliation: | 1. School of Information Technology, Deakin University, 221 Burwood Hwy, VIC 3125, Australia;2. School of Science, Information Technology and Engineering, University of Ballarat, P.O. Box 663, Ballarat, VIC 3353, Australia;3. Department of Biomedical Engineering, Khalifa University, Abu Dhabi, UAE |
| |
Abstract: | Cardiovascular autonomic neuropathy (CAN) is a serious and well known complication of diabetes. Previous articles circumvented the problem of missing values in CAN data by deleting all records and fields with missing values and applying classifiers trained on different sets of features that were complete. Most of them also added alternative features to compensate for the deleted ones. Here we introduce and investigate a new method for classifying CAN data with missing values. In contrast to all previous papers, our new method does not delete attributes with missing values, does not use classifiers, and does not add features. Instead it is based on regression and meta-regression combined with the Ewing formula for identifying the classes of CAN. This is the first article using the Ewing formula and regression to classify CAN. We carried out extensive experiments to determine the best combination of regression and meta-regression techniques for classifying CAN data with missing values. The best outcomes have been obtained by the additive regression meta-learner based on M5Rules and combined with the Ewing formula. It has achieved the best accuracy of 99.78% for two classes of CAN, and 98.98% for three classes of CAN. These outcomes are substantially better than previous results obtained in the literature by deleting all missing attributes and applying traditional classifiers to different sets of features without regression. Another advantage of our method is that it does not require practitioners to perform more tests collecting additional alternative features. |
| |
Keywords: | Cardiac autonomic neuropathy Diabetes Missing value imputation Regression learners Meta-regression techniques Ewing formula |
本文献已被 ScienceDirect 等数据库收录! |
|