ArticleScientific reports2026
A data-driven machine learning model for effective diabetes diagnosis.
Article in Scientific reports, 2026. The graph could read no effect estimate from its abstract, so it casts no vote on the map. Not yet cited in PubMed.
What it found
Each row is one number read from the abstract, on the scale the paper reported it, with its interval. Left of the dashed line favours the treatment, right favours the comparator. Under each row is the sentence it came from. New to these charts? A ten-minute tutorial.
The abstract states no effect estimate the extractor could read, or names no intervention and outcome on the map, so this paper lights no cell and moves no belief. It is still indexed, cited and linked below.
The trial behind it
Trials whose registry record cites this paper, or whose number appears in the abstract. A trial that started after this paper was published is citing it as background, not reporting it.
Neither the registry nor the abstract names a trial number. If this is a trial report, that itself is worth knowing.
Who cites it
0 citing papers in PubMed.
No citing paper in PubMed yet.
Corrections and comments
PubMed lists nothing against this paper. Absence here is not a guarantee, only a check that was made.
Authors and funding
5 authors.
Funding
No grant is acknowledged in the PubMed record.
Abstract
Diabetes is a prevalent long-lasting disease marked by high blood glucose due to inadequate insulin secretion and insulin resistance may result in serious life-threatening complications. Diabetes global prevalence has raised by fourfold over the past thirty years, and is the ninth foremost disease leads to death across the globe. Meanwhile, developments in machine learning presents new opportunities for prediction and classification of disease. However, despite of numerous existing models, a need for classifying types of diabetes still remains. The objective of this study is to develop an integrated, data-driven machine learning model for predicting the occurrence and classification of diabetes in an effort to improve on these limitations and investigate the ability to differentiate between types of diabetes through machine learning approach. To evaluate the proposed framework for classification of diabetes and its subtypes, publicly accessible diabetes-related datasets were employed for model development and evaluation. Binary classification is used for detecting occurrence of diabetes, while multiclass classification was employed for subtype classification, namely Prediabetes(PD), Type 1 Diabetes(T1D), Type 2 Diabetes(T2D), and Pancreatogenic (Type 3c-T3cD) Diabetes. K-Nearest Neighbors (KNN), Logistic Regression, Naive Bayes, Random Forest, and XGBoost machine learning algorithms were implemented. The XGBoost demonstrated the highest performance among all models with an accuracy of 0.97. Its feature importance scores validated predictive accuracy and to identify key factors that distinguish types of diabetes. The proposed model is intended to serve as a decision-support system for screening and classification tool using routine clinical data to facilitate early diagnosis and treatment planning.
Indexed as
Identifiers
What OpenQuestion holds
Registered trials
Read under generation 80e0d062 · epoch 390. Bibliography from PubMed, PubMed Central and OpenAlex; grants from NIH RePORTER; trial links from ClinicalTrials.gov; estimates, votes and beliefs from the OpenQuestion graph.