PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
May 17, 2026Biomedical Engineering and Computational Biology0 citationsOpen Access

Leveraging Clinical Data for Early Heart Disease Prediction: A Machine Learning Approach With Interpretability Analysis

View Full Paper
EQEmma QumsiyehQAQassam Al-WirdianNENur Şebnem Ersöz

Key Points

  • The aim is to develop a machine-learning approach to predict heart disease using clinical and demographic data.
  • Used a publicly available dataset for heart disease prediction.
  • Evaluated Logistic Regression, Random Forest, K-Nearest Neighbors, and Decision Trees as classification algorithms.
  • Assessed model performance using accuracy, precision, recall, and AUC-ROC metrics.
  • Random Forest and KNN models showed strong predictive performance after hyperparameter optimization.
  • SHAP values were utilized to analyze feature importance, enhancing model interpretability and transparency.

Abstract

Background: Heart disease remains one of the leading causes of mortality worldwide, highlighting the need for early and accurate diagnosis to support effective prevention and treatment strategies. Methods: This study presents a machine-learning-based approach for predicting heart disease using clinical and demographic data from a publicly available dataset. Four widely used classification algorithms—Logistic Regression, Random Forest, K-Nearest Neighbors (KNN), and Decision Trees—were evaluated to identify the most effective predictive model. The dataset underwent comprehensive preprocessing, including handling missing values, categorical encoding, and feature normalization, to enhance data quality and model robustness. Model performance was assessed using accuracy, precision, recall, and AUC-ROC metrics. Results: Findings show that hyperparameter-optimized models, particularly Random Forest and KNN, demonstrated strong predictive performance. Explainability techniques, specifically SHapley Additive exPlanations (SHAP), were incorporated to improve interpretability, transparency, and clinical trust. SHAP values were used to analyze feature importance and provide explanations for individual predictions. Conclusion: The results underscore the potential of interpretable machine-learning models as valuable tools for early diagnosis, risk stratification, and clinical decision support. Future research should employ larger datasets and investigate real-time predictive applications further to enhance the generalizability and clinical utility of these models.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Qumsiyeh et al. (2026) studied this question.

synapsesocial.com/papers/6a095c3f7880e6d24efe24a3https://doi.org/10.1177/11795972261446822
Ask AI
Helpful
Bookmark
Share
View Full Paper