ArticleBioengineering (Basel, Switzerland)2025
Supervised Learning and Large Language Model Benchmarks on Mental Health Datasets: Cognitive Distortions and Suicidal Risks in Chinese Social Media.
Article in Bioengineering (Basel, Switzerland), 2025. The graph could read no effect estimate from its abstract, so it casts no vote on the map. Cited by 8 papers.
What it found
Each row is one number read from the abstract, on the scale the paper reported it, with its interval. Left of the dashed line favours the treatment, right favours the comparator. Under each row is the sentence it came from. New to these charts? A ten-minute tutorial.
The abstract states no effect estimate the extractor could read, or names no intervention and outcome on the map, so this paper lights no cell and moves no belief. It is still indexed, cited and linked below.
The trial behind it
Trials whose registry record cites this paper, or whose number appears in the abstract. A trial that started after this paper was published is citing it as background, not reporting it.
Neither the registry nor the abstract names a trial number. If this is a trial report, that itself is worth knowing.
Who cites it
8 citing papers in PubMed.
- Can Large Language Models Support University Counseling? Evidence from Perceived Counseling Alliance, Disclosure Willingness, and Risk Recognition.Behavioral sciences (Basel, Switzerland) · 2026Article
- Affective and Cognitive Distortions-Aided Suicide Risk Prediction for Long-Form Speech in Psychological Support Hotlines.Bioengineering (Basel, Switzerland) · 2026Article
- Granularity paradox: how emotion taxonomies shape GPT-5's affective cognition and human-AI alignment.Frontiers in psychology · 2026Article
- Validating the architecture of cognitive distortions in Russian discourse using artificial intelligence and bootstrap analysis.Frontiers in psychology · 2026Article
- Large Language Models in Biomedical and Health Informatics: A Review with Bibliometric Analysis.Journal of healthcare informatics research · 2024Article
- Prompt Engineering Paradigms for Medical Applications: Scoping Review.Journal of medical Internet research · 2024Article
- A Comprehensive Review on Synergy of Multi-Modal Data and AI Technologies in Medical Diagnosis.Bioengineering (Basel, Switzerland) · 2024Review
- Comparative analysis of BERT-based and generative large language models for detecting suicidal ideation: a performance evaluation study.Cadernos de saude publica · 2024Article
Corrections and comments
PubMed lists nothing against this paper. Absence here is not a guarantee, only a check that was made.
Authors and funding
10 authors.
Funding
Abstract
On social media, users often express their personal feelings, which may exhibit cognitive distortions or even suicidal tendencies on certain specific topics. Early recognition of these signs is critical for effective psychological intervention. In this paper, we introduce two novel datasets from Chinese social media: SOS-HL-1K for suicidal risk classification, which contains 1249 posts, and SocialCD-3K, a multi-label classification dataset for cognitive distortion detection that contains 3407 posts. We conduct a comprehensive evaluation using two supervised learning methods and eight large language models (LLMs) on the proposed datasets. From the prompt engineering perspective, we experiment with two types of prompt strategies, including four zero-shot and five few-shot strategies. We also evaluate the performance of the LLMs after fine-tuning on the proposed tasks. Experimental results show a significant performance gap between prompted LLMs and supervised learning. Our best supervised model achieves strong results, with an F1-score of 82.76% for the high-risk class in the suicide task and a micro-averaged F1-score of 76.10% for the cognitive distortion task. Without fine-tuning, the best-performing LLM lags by 6.95 percentage points in the suicide task and a more pronounced 31.53 points in the cognitive distortion task. Fine-tuning substantially narrows this performance gap to 4.31% and 3.14% for the respective tasks. While this research highlights the potential of LLMs in psychological contexts, it also shows that supervised learning remains necessary for more challenging tasks.
Indexed as
Identifiers
What OpenQuestion holds
Registered trials
Read under generation 80e0d062 · epoch 390. Bibliography from PubMed, PubMed Central and OpenAlex; grants from NIH RePORTER; trial links from ClinicalTrials.gov; estimates, votes and beliefs from the OpenQuestion graph.