Evidence map›Paper›PMID 42539788›Full record

ArticleFrontiers in cardiovascular medicine2026

Explainable machine learning for hypertension prevalence classification: a cross-sectional study in Hainan Province, China.

Yuewei Wu, Shan Huang, Miaomiao Qi, Yuanyuan Zhang, Mei Lu, Tianfa Li, Yueqiong Kong

Abstract read
In one paragraph

Article in Frontiers in cardiovascular medicine, 2026. The graph could read no effect estimate from its abstract, so it casts no vote on the map. Not yet cited in PubMed.

0numbers the graph read from it
0cells of the map it votes in
0citing papers in PubMed
–field-weighted citation impact
1 · What the graph read from it

What it found

Each row is one number read from the abstract, on the scale the paper reported it, with its interval. Left of the dashed line favours the treatment, right favours the comparator. Under each row is the sentence it came from. New to these charts? A ten-minute tutorial.

The abstract states no effect estimate the extractor could read, or names no intervention and outcome on the map, so this paper lights no cell and moves no belief. It is still indexed, cited and linked below.

2 · The registry

The trial behind it

Trials whose registry record cites this paper, or whose number appears in the abstract. A trial that started after this paper was published is citing it as background, not reporting it.

Neither the registry nor the abstract names a trial number. If this is a trial report, that itself is worth knowing.

3 · Its place in the literature

Who cites it

0 citing papers in PubMed.

No citing paper in PubMed yet.

4 · The record

Corrections and comments

PubMed lists nothing against this paper. Absence here is not a guarantee, only a check that was made.

5 · Who and what money

Authors and funding

7 authors.

Yuewei WuKey Laboratory of Emergency and Trauma of Ministry of Education, Department of Cardiology, The First Affiliated Hospital, Hainan Medical University, Haikou, China.
Shan HuangKey Laboratory of Emergency and Trauma of Ministry of Education, Department of Cardiology, The First Affiliated Hospital, Hainan Medical University, Haikou, China.
Miaomiao QiKey Laboratory of Emergency and Trauma of Ministry of Education, Department of Cardiology, The First Affiliated Hospital, Hainan Medical University, Haikou, China.
Yuanyuan ZhangKey Laboratory of Emergency and Trauma of Ministry of Education, Department of Cardiology, The First Affiliated Hospital, Hainan Medical University, Haikou, China.
Mei LuKey Laboratory of Emergency and Trauma of Ministry of Education, Department of Cardiology, The First Affiliated Hospital, Hainan Medical University, Haikou, China.
Tianfa LiKey Laboratory of Emergency and Trauma of Ministry of Education, Department of Cardiology, The First Affiliated Hospital, Hainan Medical University, Haikou, China.
Yueqiong KongKey Laboratory of Emergency and Trauma of Ministry of Education, Department of Cardiology, The First Affiliated Hospital, Hainan Medical University, Haikou, China.

Funding

No grant is acknowledged in the PubMed record.

6 · The paper itself

Abstract

Objective: This study used explainable machine learning models to classify prevalent hypertension status in the general population. Methods: This cross-sectional study used clinical data collected with questionnaires, physical examinations, blood biochemistry, and routine urine tests from 4,800 permanent residents aged ≥18 years old in Hainan Province, China, from 2021 to 2022. The random forest algorithm was applied to select the most significant features based on importance scores of all variable features. Models for hypertension prevalence classification were created using six machine learning techniques. These models were then compared to select a model with high classification performance based on classification accuracy and the area under curve (AUC) values. Calibration assessment and decision curve analysis were additionally performed to assess clinical applicability. To evaluate and illustrate the best models, the SHapley Additive Explanation (SHAP) values and the Local Interpretable Model-Agnostic Explanations (LIME) algorithms were used. Results: In total, 4,606 permanent residents were included in this study (hypertension, 32.5%). They were randomly split into two groups: a training set (3,224, 70%) and a validation set (1,382, 30%). After the random forest algorithm was applied to score feature importance, the top 10 most important features, including age, smoking, urine microalbumin (UALB), educational level, diabetes, body mass index (BMI), sex, triglyceride (TG), income level, and family history of hypertension, were finally included for model construction. With an AUC of 0.8461, the eXtreme Gradient Boosting (XGBoost) model had the greatest performance. The SHAP values were used to quantify the contribution of each input feature to the XGBoost model and demonstrate the importance ranking of predictors. The LIME algorithm integrated the SHAP values to provide a more compelling explanation for each individual classification result. A web page based on the results of this study was developed to support hypertension prevalence screening in clinical environments. Conclusion: A machine learning model based on clinical variables was developed and verified, which showed superior performance in identifying prevalent hypertension cases in the general population, opening up new possibilities for the rapid population screening and prevention of hypertension.

Indexed as

cardiovascular diseasesChinaexplainabilityhypertensionmachine learning

Identifiers

PMID42539788
PMCPMC13424424

What OpenQuestion holds

Textmetadata
LicenceCC BY
Read underepoch 390

Registered trials

None linked

Read under generation 80e0d062 · epoch 390. Bibliography from PubMed, PubMed Central and OpenAlex; grants from NIH RePORTER; trial links from ClinicalTrials.gov; estimates, votes and beliefs from the OpenQuestion graph.