Enhancing stroke disease classification through machine learning models via a novel voting system by feature selection techniques.

Journal: PloS one
PMID:

Abstract

Heart disease remains a leading cause of mortality and morbidity worldwide, necessitating the development of accurate and reliable predictive models to facilitate early detection and intervention. While state of the art work has focused on various machine learning approaches for predicting heart disease, but they could not able to achieve remarkable accuracy. In response to this need, we applied nine machine learning algorithms XGBoost, logistic regression, decision tree, random forest, k-nearest neighbors (KNN), support vector machine (SVM), gaussian naïve bayes (NB gaussian), adaptive boosting, and linear regression to predict heart disease based on a range of physiological indicators. Our approach involved feature selection techniques to identify the most relevant predictors, aimed at refining the models to enhance both performance and interpretability. The models were trained, incorporating processes such as grid search hyperparameter tuning, and cross-validation to minimize overfitting. Additionally, we have developed a novel voting system with feature selection techniques to advance heart disease classification. Furthermore, we have evaluated the models using key performance metrics including accuracy, precision, recall, F1-score, and the area under the receiver operating characteristic curve (ROC AUC). Among the models, XGBoost demonstrated exceptional performance, achieving 99% accuracy, precision, F1-Score, 98% recall, and 100% ROC AUC. This study offers a promising approach to early heart disease diagnosis and preventive healthcare.

Authors

  • Mahade Hasan
    School of Software, Nanjing University of Information Science and Technology, Nanjing, China.
  • Farhana Yasmin
    Department of Computer Science and Technology, Nanjing University of Information Science and Technology, Nanjing, China.
  • Md Mehedi Hassan
    School of Food and Biological Engineering, Jiangsu University, Zhenjiang 212013, PR China.
  • Xue Yu
    Department of Cardiology, Beijing Hospital, National Center of Gerontology, Institute of Geriatric Medicine, Chinese Academy of Medical Sciences, Beijing, China.
  • Soniya Yeasmin
    Department of Computer Science and Engineering, North Western University, Khulna, Bangladesh.
  • Herat Joshi
    Great River Health Systems, Burlington, IA, United States of America.
  • Sheikh Mohammed Shariful Islam
    Institute for Physical Activity and Nutrition, School of Exercise and Nutrition Sciences, Deakin University, Geelong, VIC, 3220, Australia.