ArticleJournal of medical Internet research2026
Behavior Change Content and Implementation of Large Language Model-Driven Conversational Agents in Cardiometabolic Care: Scoping Review.
Article in Journal of medical Internet research, 2026. The graph could read no effect estimate from its abstract, so it casts no vote on the map. Not yet cited in PubMed.
What it found
Each row is one number read from the abstract, on the scale the paper reported it, with its interval. Left of the dashed line favours the treatment, right favours the comparator. Under each row is the sentence it came from. New to these charts? A ten-minute tutorial.
The abstract states no effect estimate the extractor could read, or names no intervention and outcome on the map, so this paper lights no cell and moves no belief. It is still indexed, cited and linked below.
The trial behind it
Trials whose registry record cites this paper, or whose number appears in the abstract. A trial that started after this paper was published is citing it as background, not reporting it.
Neither the registry nor the abstract names a trial number. If this is a trial report, that itself is worth knowing.
Who cites it
0 citing papers in PubMed.
No citing paper in PubMed yet.
Corrections and comments
PubMed lists nothing against this paper. Absence here is not a guarantee, only a check that was made.
Authors and funding
6 authors.
Funding
No grant is acknowledged in the PubMed record.
Abstract
Background: Large language models (LLMs) are increasingly embedded in conversational agents for cardiometabolic care. These systems could support self-management, but their behavior change content, delivery mechanisms, and implementation transparency are poorly understood. Objective: This scoping review mapped behavior change techniques (BCTs) used in LLM-driven conversational agents for cardiometabolic prevention and management, described how these techniques are delivered across static, rule-based, and generative mechanisms, examined LLM design, personalization, and safety reporting, and summarized user experience and behavioral or clinical outcomes. Methods: We searched PubMed, Web of Science, Embase, CINAHL, APA PsycInfo, IEEE Xplore, ACM Digital Library, arXiv, ClinicalTrials.gov, and the WHO International Clinical Trials Registry Platform for records published from January 1, 2020, to November 30, 2025. The final search was run on March 25, 2026, using this publication-date limit. Eligible studies reported a patient-facing text- or voice-based cardiometabolic conversational agent using an LLM or other transformer-based generative model. Two reviewers independently screened records and extracted data. BCTs were coded using the Behavior Change Technique Taxonomy v1; selected self-management BCTs were classified as static, rule-based or templated, or generative or context-aware. Empirical human-participant- or evaluator-based studies were appraised with the Mixed Methods Appraisal Tool, and a study-specific checklist assessed LLM implementation reporting transparency. Results: Thirty-eight studies were included; 19 involved empirical human-participant- or evaluator-based assessments, whereas 19 were technical and system-level evaluations, including framework-development, simulated-output, and proof-of-concept studies. Studies were concentrated in 2024-2025. Instruction on how to perform behavior was identified in 30 of 38 (79%) studies, information about health consequences in 27 of 38 (71%) studies, and feedback and monitoring techniques in 19 of 38 (50%) studies. Most agents were positioned as educators or coaches targeting type 2 diabetes, obesity, or related cardiometabolic risk, and GPT-family models embedded in hybrid architectures with retrieval-augmented generation or rule-based components predominated. Generative outputs were used mainly for tailored explanations, risk information, and socioemotional responses, whereas self-monitoring, reminders, and structured interactions were more often rule-based or mixed-mode. Only 13 of 38 (34%) studies fully reported prompts or system messages, and 16 of 38 (42%) studies fully reported safety or oversight mechanisms. User evaluations reported good usability and perceived helpfulness, but behavioral or physiological outcomes were sparse and usually limited to pilot, short-term, or single-case designs. Conclusions: LLM-driven conversational agents for cardiometabolic care are proliferating but remain early-stage and methodologically heterogeneous. Current systems primarily use LLMs as educational and explanatory layers with "synthetic empathy" over rule-based data capture and safety functions, while behavior change content remains dominated by information provision and simple feedback. More rigorous comparative studies with longer follow-up are needed before firm conclusions can be drawn about sustained behavioral or clinical benefit.
Indexed as
Identifiers
What Socratic holds
Registered trials
Read under generation 80e0d062 · epoch 390. Bibliography from PubMed, PubMed Central and OpenAlex; grants from NIH RePORTER; trial links from ClinicalTrials.gov; estimates, votes and beliefs from the Socratic graph.