ArticlePLOS digital health2023
Imbalanced class distribution and performance evaluation metrics: A systematic review of prediction accuracy for determining model performance in healthcare systems.
Article in PLOS digital health, 2023. The graph could read no effect estimate from its abstract, so it casts no vote on the map. It carries an expression of concern. Cited by 24 papers.
What it found
Each row is one number read from the abstract, on the scale the paper reported it, with its interval. Left of the dashed line favours the treatment, right favours the comparator. Under each row is the sentence it came from. New to these charts? A ten-minute tutorial.
The abstract states no effect estimate the extractor could read, or names no intervention and outcome on the map, so this paper lights no cell and moves no belief. It is still indexed, cited and linked below.
The trial behind it
Trials whose registry record cites this paper, or whose number appears in the abstract. A trial that started after this paper was published is citing it as background, not reporting it.
Neither the registry nor the abstract names a trial number. If this is a trial report, that itself is worth knowing.
Who cites it
24 citing papers in PubMed.
- A Commentary on AI-based Prediction Models in Periodontitis Progression.Evidence-based dentistry · 2026Article
- Article
- A physics-informed multi-modal transformer framework for predictive maintenance and remaining useful life estimation of electrical submersible pumps.Scientific reports · 2026Article
- Holotomography and Multivariate Analysis Reveal Donor-Specific Responses to Antioxidant Supplementation During Stallion Sperm Cryopreservation.Antioxidants (Basel, Switzerland) · 2026Article
- Cardiovascular Risk Stratification in Youth-Onset Type 2 Diabetes Using Machine Learning.Journal of Korean medical science · 2026Article
- A Verifiable Framework for Brain Tumor Classification: Combining Vision Transformers, Class-Weighted Learning, and SMT-Based Formal Decision Traces.Diagnostics (Basel, Switzerland) · 2026Article
- Sex-Specific Differences in Intracranial Aneurysm Rupture Presentation and Model Performance: Evidence from a Retrospective Cohort.Annals of biomedical engineering · 2026Article
- Predicting individual differences in digital alcohol intervention effectiveness through multimodal data.NPJ digital medicine · 2026Article
- Interpretable intrusion detection for IoT: a CNN-BiLSTM permutation importance framework for deep feature selection.Frontiers in big data · 2026Article
- Modeling multiscale neural dynamics for EEG-based emotion recognition using an attentive wavelet-transformer framework.Frontiers in computational neuroscience · 2026Article
- Adaptive class-aware feature selection for high-dimensional and imbalanced multi-class network intrusion detection.Frontiers in big data · 2026Article
- Construction of a classification model for dementia among Brazilian adults aged 50 and over.Frontiers in aging neuroscience · 2026Article
- Utilizing whole genome sequencing data for machine learning driven prediction of antibiotic resistance inFrontiers in microbiology · 2026Article
- Training Together, Diagnosing Better: Federated Learning for Collagen VI-Related Dystrophies.ArXiv · 2025Article
- Machine learning to predict food effects during drug development: a comprehensive review.Journal of cheminformatics · 2025Review
- Performance of a Vision-Language Model in Detecting Common Dental Conditions on Panoramic Radiographs Using Different Tooth Numbering Systems.Diagnostics (Basel, Switzerland) · 2025Article
- Expression of Concern: Imbalanced class distribution and performance evaluation metrics: A systematic review of prediction accuracy for determining model performance in healthcare systems.PLOS digital health · 2025Article
- Rethinking Domain-Specific Pretraining by Supervised or Self-Supervised Learning for Chest Radiograph Classification: A Comparative Study Against ImageNet Counterparts in Cold-Start Active Learning.Health care science · 2025Article
- A study on early diagnosis for fracture non-union prediction using deep learning and bone morphometric parameters.Frontiers in medicine · 2025Article
- Can artificial intelligence uncover the bioactive peptides' benefits for human health and knowledge? A narrative review.Frontiers in nutrition · 2025Review
Corrections and comments
- Expression of concern
Authors and funding
4 authors.
Funding
No grant is acknowledged in the PubMed record.
Abstract
Focus on predictive algorithm and its performance evaluation is extensively covered in most research studies to determine best or appropriate predictive model with Optimum prediction solution indicated by prediction accuracy score, precision, recall, f1score etc. Prediction accuracy score from performance evaluation has been used extensively as the main determining metric for performance recommendation. It is one of the most widely used metric for identifying optimal prediction solution irrespective of dataset class distribution context or nature of dataset and output class distribution between the minority and majority variables. The key research question however is the impact of class inequality on prediction accuracy score in such datasets with output class distribution imbalance as compared to balanced accuracy score in the determination of model performance in healthcare and other real-world application systems. Answering this question requires an appraisal of current state of knowledge in both prediction accuracy score and balanced accuracy score use in real-world applications where there is unequal class distribution. Review of related works that highlight the use of imbalanced class distribution datasets with evaluation metrics will assist in contextualizing this systematic review.
Identifiers
What OpenQuestion holds
Registered trials
Read under generation 80e0d062 · epoch 390. Bibliography from PubMed, PubMed Central and OpenAlex; grants from NIH RePORTER; trial links from ClinicalTrials.gov; estimates, votes and beliefs from the OpenQuestion graph.