ArticleJournal of the American Medical Informatics Association : JAMIA2024
Evaluation framework for conversational agents with artificial intelligence in health interventions: a systematic scoping review.
Article in Journal of the American Medical Informatics Association : JAMIA, 2024. The graph could read no effect estimate from its abstract, so it casts no vote on the map. Cited by 22 papers, 1 of them a synthesis that pooled it.
What it found
Each row is one number read from the abstract, on the scale the paper reported it, with its interval. Left of the dashed line favours the treatment, right favours the comparator. Under each row is the sentence it came from. New to these charts? A ten-minute tutorial.
The abstract states no effect estimate the extractor could read, or names no intervention and outcome on the map, so this paper lights no cell and moves no belief. It is still indexed, cited and linked below.
The trial behind it
Trials whose registry record cites this paper, or whose number appears in the abstract. A trial that started after this paper was published is citing it as background, not reporting it.
Neither the registry nor the abstract names a trial number. If this is a trial report, that itself is worth knowing.
Who cites it
22 citing papers in PubMed, 1 synthesis or guideline pooled it.
- AI for IMPACTS Framework for Evaluating the Long-Term Real-World Impacts of AI-Powered Clinician Tools: Systematic Review and Narrative Synthesis.Journal of medical Internet research · 2025Pooled it
- Encouraging arm use in stroke survivors: the impact of smart reminders during a home-based intervention.Journal of neuroengineering and rehabilitation · 2024Trial
- AI-Augmented Prevention Science Needs Community-Engaged Prevention Science: a Framework for Greater Accountability.Prevention science : the official journal of the Society for Prevention Research · 2026Article
- Frameworks, Methodologies, and Tools for Evaluating Large Language Models in Digital Mental Health Interventions: Protocol for a Scoping Review.JMIR research protocols · 2026Article
- What Platform Scores Miss: Multidimensional Evaluation of AI Teaching Agents in Medical Education.JMIR medical education · 2026Article
- Large Language Model-Based Behavioral Activation Chatbot for Young People With Depression Using Artificial Users and Clinical Experts: Mixed Methods Evaluation.JMIR mental health · 2026Article
- Metrics Used for the Evaluation of Chatbots Providing Cancer Genetic Risk Assessment and Education: Systematic Review.JMIR AI · 2026Review
- Beyond the Algorithm: A Critical and Evidence-Based Review of Artificial Intelligence in Chronic Pain Rehabilitation.Cureus · 2026Review
- An evidence gap map of digital health interventions for enhancing patient engagement in healthcare.NPJ digital medicine · 2026Article
- Governance for safe and responsible AI in healthcare organisations: a scoping review of frameworks.NPJ digital medicine · 2026Article
- Eye-Tracked Visual Attention to Anthropomorphic Appearance and Empathic Responses in AI Medical Conversational Agents: Dissociating Trust Gains from Attentional Synergy.Journal of eye movement research · 2026Article
- Automated chain-of-thought evaluation framework for large language model-generated emergency department documentation: a simulation-based study.Clinical and experimental emergency medicine · 2026Article
- Review
- Beyond the Bot: A Dual-Phase Framework for Evaluating AI Chatbot Simulations in Nursing Education.Nursing reports (Pavia, Italy) · 2025Article
- User Engagement with A Multimodal Conversational Agent for Self-Care and Chronic Disease Management: A Retrospective Analysis.Journal of medical systems · 2025Article
- Large language models for the mental health community: framework for translating code to care.The Lancet. Digital health · 2025Review
- Nurses' perspectives on privacy and ethical concerns regarding artificial intelligence adoption in healthcare.Heliyon · 2024Article
- Large language models and generative AI in telehealth: a responsible use lens.Journal of the American Medical Informatics Association : JAMIA · 2024Review
- Celebrating Eta Berner and her influence on biomedical and health informatics.Journal of the American Medical Informatics Association : JAMIA · 2024Article
- Redefining Virtual Assistants in Health Care: The Future With Large Language Models.Journal of medical Internet research · 2024Article
Corrections and comments
PubMed lists nothing against this paper. Absence here is not a guarantee, only a check that was made.
Authors and funding
5 authors.
Funding
Abstract
objectivesConversational agents (CAs) with emerging artificial intelligence present new opportunities to assist in health interventions but are difficult to evaluate, deterring their applications in the real world. We aimed to synthesize existing evidence and knowledge and outline an evaluation framework for CA interventions. MATERIALS AND
methodsWe conducted a systematic scoping review to investigate designs and outcome measures used in the studies that evaluated CAs for health interventions. We then nested the results into an overarching digital health framework proposed by the World Health Organization (WHO).
resultsThe review included 81 studies evaluating CAs in experimental (n = 59), observational (n = 15) trials, and other research designs (n = 7). Most studies (n = 72, 89%) were published in the past 5 years. The proposed CA-evaluation framework includes 4 evaluation stages: (1) feasibility/usability, (2) efficacy, (3) effectiveness, and (4) implementation, aligning with WHO's stepwise evaluation strategy. Across these stages, this article presents the essential evidence of different study designs (n = 8), sample sizes, and main evaluation categories (n = 7) with subcategories (n = 40). The main evaluation categories included (1) functionality, (2) safety and information quality, (3) user experience, (4) clinical and health outcomes, (5) costs and cost benefits, (6) usage, adherence, and uptake, and (7) user characteristics for implementation research. Furthermore, the framework highlighted the essential evaluation areas (potential primary outcomes) and gaps across the evaluation stages. DISCUSSION AND
conclusionThis review presents a new framework with practical design details to support the evaluation of CA interventions in healthcare research. PROTOCOL REGISTRATION: The Open Science Framework (https://osf.io/9hq2v) on March 22, 2021.
Indexed as
Identifiers
What OpenQuestion holds
Registered trials
Read under generation 80e0d062 · epoch 390. Bibliography from PubMed, PubMed Central and OpenAlex; grants from NIH RePORTER; trial links from ClinicalTrials.gov; estimates, votes and beliefs from the OpenQuestion graph.