ArticleBioMed research international2026
An Integrated Statistical and Machine Learning Approach for Breast Cancer Classification Using Tumor Morphological Features.
Article in BioMed research international, 2026. The graph could read no effect estimate from its abstract, so it casts no vote on the map. Cited by 1 paper.
What it found
Each row is one number read from the abstract, on the scale the paper reported it, with its interval. Left of the dashed line favours the treatment, right favours the comparator. Under each row is the sentence it came from. New to these charts? A ten-minute tutorial.
The abstract states no effect estimate the extractor could read, or names no intervention and outcome on the map, so this paper lights no cell and moves no belief. It is still indexed, cited and linked below.
The trial behind it
Trials whose registry record cites this paper, or whose number appears in the abstract. A trial that started after this paper was published is citing it as background, not reporting it.
Neither the registry nor the abstract names a trial number. If this is a trial report, that itself is worth knowing.
Who cites it
1 citing paper in PubMed.
- An Integrated Statistical and Machine Learning Approach for Breast Cancer Classification Using Tumor Morphological Features.BioMed research international · 2026Article
Corrections and comments
PubMed lists nothing against this paper. Absence here is not a guarantee, only a check that was made.
Authors and funding
6 authors.
Funding
No grant is acknowledged in the PubMed record.
Abstract
introductionBreast cancer is one of the most common cancers and a leading cause of death among women worldwide. Early and accurate classification of tumors as benign or malignant is essential for improving patient outcomes. In recent years, statistical and machine learning methods have been widely used to improve diagnostic accuracy; however, combining these approaches in a single framework remains limited.
methodsA retrospective analysis was conducted using the Breast Cancer Wisconsin Diagnostic Dataset (569 tumor samples). Eight morphological features were analyzed. The dataset was split into training (70%) and testing (30%) sets using stratified random sampling. Logistic regression, random forest, and support vector machine (SVM) models were developed. Hyperparameter tuning was performed using five-fold cross-validation. Model performance was evaluated using accuracy, sensitivity, specificity, precision, F1-score, Matthews correlation coefficient (MCC), and area under the curve (AUC). Logistic regression assumptions and model diagnostics were also assessed.
resultsRadius, texture, smoothness, and concavity were significant predictors of malignancy. Five-fold cross-validation indicated stable model performance. On the test set, logistic regression achieved the highest accuracy (95.3%) and AUC (0.983), followed by SVM (94.1%, AUC = 0.980) and random forest (92.9%, AUC = 0.979). Additional performance metrics, including precision, F1-score, and MCC, also demonstrated strong classification performance. Diagnostic analyses confirmed acceptable model fit and no major violations of logistic regression assumptions. No significant differences in AUC were observed among the models (DeLong test, p > 0.05).
conclusionThe study demonstrates that an integrated statistical and machine learning approach provides a robust and accurate method for breast cancer classification. Logistic regression showed slightly better performance, whereas machine learning models also achieved comparable results. This approach has strong potential to support early detection and clinical decision-making.
Indexed as
Identifiers
What OpenQuestion holds
Registered trials
Read under generation 80e0d062 · epoch 390. Bibliography from PubMed, PubMed Central and OpenAlex; grants from NIH RePORTER; trial links from ClinicalTrials.gov; estimates, votes and beliefs from the OpenQuestion graph.