Generative Artificial Intelligence Chatbots for Motivational Interviewing: A Scoping Review From System Design to Intervention Outcomes

arXiv cs.CL Papers

Summary

A scoping review of generative AI chatbots for motivational interviewing reveals they provide MI-consistent interactions with favorable user perceptions, but evidence for sustained behavioral change is limited.

arXiv:2609.20902v1 Announce Type: new Abstract: Motivational interviewing (MI) is a collaborative approach to elicit autonomous motivation for health behavior change. Generative AI (GenAI) offers new ways to deliver MI via conversational systems, but evidence on their design, assessment, and translation into interventions remains fragmented. This scoping review characterized evidence on GenAI-MI chatbots across system design, safety, MI quality, user perceptions, and intervention outcomes. We conducted a PRISMA-ScR scoping review. Nine datasets were searched for studies published or publicly available from January 1, 2015 to June 2, 2026 that used GenAI to generate MI chatbot responses or counselor utterances. Data were extracted using a predefined framework and synthesized descriptively. Forty-seven reports (48 studies) were included. Twenty (41.7%) focused on system design without direct participant use; 28 (58.3%) involved direct interaction. Most systems were text based and disembodied; 23 (47.9%) incorporated dynamic adaptation. Safety measures were unevenly reported. Among studies with direct use, 21/28 (75.0%) reported informed consent or user education. Thirty (62.5%) assessed MI quality, generally suggesting MI-consistent interactions. User perceptions were favorable, especially empathy, usability, helpfulness, and intention to use, though measures were heterogeneous. Eighteen (37.5%) reported intervention outcomes, mostly after a single session. Positive findings were more consistent for short-term motivation than sustained behavioral or functional change. GenAI-MI chatbots can deliver MI-consistent interactions perceived favorably, but evidence for sustained behavioral or functional change is limited. Future research should strengthen runtime safety monitoring, standardize MI quality assessment, and use longer-term comparative designs with behavioral and functional outcomes.
Original Article
View Cached Full Text

Cached at: 09/21/26, 09:02 AM

# Generative Artificial Intelligence Chatbots for Motivational Interviewing: A Scoping Review From System Design to Intervention Outcomes
Source: [https://arxiv.org/abs/2609.20902](https://arxiv.org/abs/2609.20902)
[View PDF](https://arxiv.org/pdf/2609.20902)

> Abstract:Motivational interviewing \(MI\) is a collaborative approach to elicit autonomous motivation for health behavior change\. Generative AI \(GenAI\) offers new ways to deliver MI via conversational systems, but evidence on their design, assessment, and translation into interventions remains fragmented\. This scoping review characterized evidence on GenAI\-MI chatbots across system design, safety, MI quality, user perceptions, and intervention outcomes\. We conducted a PRISMA\-ScR scoping review\. Nine datasets were searched for studies published or publicly available from January 1, 2015 to June 2, 2026 that used GenAI to generate MI chatbot responses or counselor utterances\. Data were extracted using a predefined framework and synthesized descriptively\. Forty\-seven reports \(48 studies\) were included\. Twenty \(41\.7%\) focused on system design without direct participant use; 28 \(58\.3%\) involved direct interaction\. Most systems were text based and disembodied; 23 \(47\.9%\) incorporated dynamic adaptation\. Safety measures were unevenly reported\. Among studies with direct use, 21/28 \(75\.0%\) reported informed consent or user education\. Thirty \(62\.5%\) assessed MI quality, generally suggesting MI\-consistent interactions\. User perceptions were favorable, especially empathy, usability, helpfulness, and intention to use, though measures were heterogeneous\. Eighteen \(37\.5%\) reported intervention outcomes, mostly after a single session\. Positive findings were more consistent for short\-term motivation than sustained behavioral or functional change\. GenAI\-MI chatbots can deliver MI\-consistent interactions perceived favorably, but evidence for sustained behavioral or functional change is limited\. Future research should strengthen runtime safety monitoring, standardize MI quality assessment, and use longer\-term comparative designs with behavioral and functional outcomes\.

## Submission history

From: Run\-Ze Hu \[[view email](https://arxiv.org/show-email/707abf4e/2609.20902)\] **\[v1\]**Thu, 17 Sep 2026 13:43:09 UTC \(1,797 KB\)

Similar Articles

Creating and Evaluating Personas Using Generative AI: A Scoping Review of 81 Articles

arXiv cs.CL

This scoping review analyzes 81 articles (2022-2025) examining the use of generative AI for creating and evaluating user personas, identifying strengths in reproducibility but critical issues including lack of evaluation in 45% of studies, over-reliance on GPT models (86%), and risks of circularity where the same model generates and evaluates personas.

AI chatbots have failed people in crisis. Can that be fixed?

Ars Technica

The article examines how AI chatbots, especially ChatGPT, have been implicated in harmful mental health incidents and explores expert suggestions for reducing harm through greater transparency and de-anthropomorphization, alongside OpenAI's new partnership with the American Psychological Association.

Dynamic In-Group Persona Generation for Enhancing Human-AI Rapport

arXiv cs.AI

This paper introduces a method for LLM-based chatbots to dynamically generate in-group personas by first identifying a user's primary concern and then creating a synthetic persona that shares that concern. A human-subject study demonstrates significant improvements in perceived rapport and user engagement compared to baseline conditions.