AI paper index

Exploring the Effectiveness of a Psychological Counseling Domain-Specific Chatbot (Nancy) for Improving Adult Mental Health: A Randomized Controlled Trial (Preprint)

2026-08-27

One-line summary

An AI research paper on Exploring the Effectiveness of a Psychological Counseling Domain-Specific Chatbot (Nancy) for Improving Adult Mental Health: A Randomized Controlled Trial (Preprint).

Engineering notes

Engineering notes will be added by the aipentium editorial team.

Chinese explanation / 中文解读

中文解读待补充:本站会优先为大语言模型、生成式AI、ChatGPT相关技术、计算机视觉、深度学习等高价值论文补充中文说明。

Original abstract

BACKGROUND Psychological counseling chatbots may expand access to mental health support, but evidence has largely focused on scripted systems or general-purpose large language models (LLMs). Whether domain-specific psychological counseling LLMs provide greater clinical benefit remains unclear. OBJECTIVE This study evaluated the clinical effectiveness of Nancy, a psychological counseling chatbot built on the domain-specific PsyLLM, relative to ChatGPT, a psychoeducational E-book, and a waitlist control (WLC). METHODS In this 4-arm randomized controlled trial (RCT), 501 Mandarin-speaking adults aged 18 years or older with a Patient Health Questionnaire-9 (PHQ-9) score of 5 or higher or a Generalized Anxiety Disorder-7 (GAD-7) score of 5 or higher were randomized 1:1:1:1 to Nancy, ChatGPT, E-book, or the Waitlist Control (WLC) for 28 days. Primary outcomes were PHQ-9 and GAD-7 scores. Secondary outcomes included affect, perceived empathy, and platform-recorded engagement. Intention-to-treat (ITT) analyses used multiple imputation and baseline-adjusted analysis of covariance; per-protocol (PP) analyses were sensitivity analyses. RESULTS Nancy showed greater PHQ-9 reductions than ChatGPT (adjusted difference −2.13, 95% CI −3.66 to −0.59; Holm-adjusted p=.007; Cohen d=0.44), E-book (−5.81, 95% CI −7.23 to −4.38; p<.001; d=1.10), and WLC (−2.95, 95% CI −4.35 to −1.54; p<.001; d=0.61). GAD-7 reductions were also greater with Nancy than with ChatGPT (−1.58, 95% CI −3.08 to −0.08; Holm-adjusted p=.04; d=0.35), E-book (−3.35, 95% CI −4.70 to −2.00; p<.001; d=0.74), and WLC (−1.90, 95% CI −3.24 to −0.56; p=.01; d=0.44). Per-protocol findings were consistent. Positive and negative affect did not differ significantly between groups. Perceived empathy was similar between Nancy and ChatGPT (35.49 vs 34.92; p=.69; Hedges g=0.07), whereas conversational intensity was higher with Nancy (median 17.21 vs 16.39 turns per recorded use day; p<.001; Cliff δ=0.29). CONCLUSIONS Nancy produced greater reductions in depressive and anxiety symptoms than ChatGPT, psychoeducation, and waitlist controls over 28 days. These findings suggest that domain-specific therapeutic alignment may provide added clinical value beyond general-purpose conversational support. CLINICALTRIAL Chinese Clinical Trial Registry (ChiCTR2600121508); https://www.chictr.org.cn/hvshowproject.html?id=298372&v=1.0

5.0Engineering value
7.0Research novelty
4.0Business relevance

Links and sources

Need this topic turned into a technical roadmap?

aipentium can prepare a custom AI literature review, code map, dataset map, and B2B technology assessment.

Request B2B AI research

Comments

No comments yet. Be the first to share your thoughts on this paper.
Login or register to leave a comment