Research Article

Interpretable Ensemble Learning Approach for Breast Cancer Diagnosis Using SHAP-Based Explainable AI

Authors

  • Mahfuz Islam Khan Jabed Department of Information Technology, Washington University of Science and Technology, Alexandria, Virginia, USA https://orcid.org/0009-0001-2141-2894
  • Muhammad Adeel Manzoor Department of Biostatistics and Epidemiology, Monroe University, Bronx, New York, USA https://orcid.org/0009-0008-5904-5752
  • Fardous Mir Tofa Department of Information Technology, Washington University of Science and Technology, Alexandria, Virginia, USA https://orcid.org/0009-0001-1874-1849
  • Muhammad Haseeb Khan Department of Biostatistics and Epidemiology, Monroe University, Bronx, New York, USA

Abstract

Early and reliable breast cancer diagnosis is clinically important, but machine learning models intended for healthcare use must offer both predictive accuracy and interpretability. This study evaluated a structured machine learning workflow on the Wisconsin Diagnostic Breast Cancer dataset. After preprocessing and label encoding, five machine learning classifiers—Logistic Regression, Decision Tree, Random Forest, Support Vector Machine, and XGBoost—were implemented. Model performance was assessed using accuracy, precision, recall, F1-score, ROC-AUC, Brier score, and stratified cross-validation. SHAP was used to provide both global and local interpretability of the selected tree-based model. On the holdout set, Logistic Regression achieved the best overall discrimination and calibration, with 97.37% accuracy, 95.24% recall, ROC-AUC of 0.9960, and the lowest Brier score of 0.0211. SVM, Random Forest, and XGBoost each achieved 97.37% accuracy with perfect precision, but lower recall (92.86%). Cross-validation results also favored Logistic Regression, which achieved a mean ROC-AUC of 0.9955 ± 0.0038. SHAP analysis of XGBoost identified perimeter_worst, concave points_mean, concave points_worst, and radius_worst as the most influential predictors. The study demonstrates that transparent and comparatively simple machine learning models can achieve highly reliable classification performance on the WDBC benchmark, while SHAP improves interpretability of high-performing ensemble models. Although the results are promising, external clinical validation is required before real-world implementation.

Article information

Journal

Journal of Computer Science and Technology Studies

Volume (Issue)

8 (8)

Pages

244-255

Published

2026-07-30

Downloads

Views

24

Downloads

10

Keywords:

Breast Cancer Diagnosis, Machine Learning, Explainable Artificial Intelligence (XAI), SHAP, Clinical Decision Support, Predictive Healthcare Analytics