ArticleCurrent urology2026
Machine learning approaches to optimize the integration of sociodemographic factors for predicting cancer-specific survival among patients with high-risk prostate cancer.
Article in Current urology, 2026. The graph could read no effect estimate from its abstract, so it casts no vote on the map. Not yet cited in PubMed.
What it found
Each row is one number read from the abstract, on the scale the paper reported it, with its interval. Left of the dashed line favours the treatment, right favours the comparator. Under each row is the sentence it came from. New to these charts? A ten-minute tutorial.
The abstract states no effect estimate the extractor could read, or names no intervention and outcome on the map, so this paper lights no cell and moves no belief. It is still indexed, cited and linked below.
The trial behind it
Trials whose registry record cites this paper, or whose number appears in the abstract. A trial that started after this paper was published is citing it as background, not reporting it.
Neither the registry nor the abstract names a trial number. If this is a trial report, that itself is worth knowing.
Who cites it
0 citing papers in PubMed.
No citing paper in PubMed yet.
Corrections and comments
PubMed lists nothing against this paper. Absence here is not a guarantee, only a check that was made.
Authors and funding
9 authors.
Funding
No grant is acknowledged in the PubMed record.
Abstract
Background: Sociodemographic factors influence the outcomes of prostate cancer (PCa); however, they are rarely incorporated into clinical risk prediction models. This study aimed to assess whether machine learning approaches could optimize the integration of sociodemographic variables to improve the prediction of cancer-specific survival among patients with high-risk PCa. Materials and methods: Data from the Surveillance, Epidemiology, and End Results database were retrospectively analyzed to identify patients diagnosed with high-risk PCa from 2010 to 2020. Two random forest models were developed: one using clinical and pathological variables (age, stage, prostate-specific antigen level, Gleason grade, time to treatment, and year of diagnosis) and another incorporating available sociodemographic features (race, income, marital status, region, and urbanicity). Five-fold cross-validation was performed to evaluate the model performance and minimize overfitting. Hyperparameter tuning via a grid search optimized the model structure. Performance was assessed using the area under the receiver operating characteristic curve (AUC), Brier scores, sensitivity, and specificity. Parallel analyses were conducted using the XGBoost software. Clinical utility was evaluated using decision curve analysis. Results: We identified 80,858 patients with high-risk PCa. The clinical-only random forest model (AUC, 0.54) significantly improved with the addition of sociodemographic variables (AUC, 0.72; Conclusions: Incorporating sociodemographic variables into machine learning models significantly improved the prediction of cancer-specific survival in high-risk PCa, supporting their inclusion in risk stratification tools.
Indexed as
Identifiers
What OpenQuestion holds
Registered trials
Read under generation 80e0d062 · epoch 390. Bibliography from PubMed, PubMed Central and OpenAlex; grants from NIH RePORTER; trial links from ClinicalTrials.gov; estimates, votes and beliefs from the OpenQuestion graph.