arXiv AI

Generative AI Purpose-built for Social and Mental Health: A Real-World Pilot

arXiv Computation and Language
Sep 21

Generative Artificial Intelligence Chatbots for Motivational Interviewing: A Scoping Review From System Design to Intervention Outcomes

This scoping review examined 48 studies on generative AI chatbots designed to deliver motivational interviewing (MI). It found that most systems were text‑based and disembodied, with about half incorporating dynamic adaptation, and that safety reporting was inconsistent. While user perceptions were generally positive and many studies reported MI‑consistent interactions, evidence for sustained behavioral or functional change remains limited.

By Runze Hu, Jingqi Kong, Yang Yang, Yihang Yang, Jingyao Liu, Haizhou Tang, Shanghang Zhang, Zheng Liu
arXiv Computation and Language
Sep 18

CounselReflect: Opportunities and Challenges for Designing Tools to Support Self-Reflection on Mental Health and Well-Being Conversations with AI

The paper introduces CounselReflect, a tool that converts counseling quality metrics into a framework for users to reflect on their mental‑health AI conversations. Through interviews with 21 users, the study finds that while most participants rarely reflect on their interactions, they identify specific questions they would like such a tool to address. The findings also reveal that users tend to confirm existing beliefs and focus on familiar dimensions, highlighting the need for reflection tools to expose blind spots and encourage a more comprehensive examination of AI interactions, especially when revisiting emotionally charged exchanges.

By Yahan Li, Chaohao Du, Christopher Chun Kuizon, Zeyang Li, Nimra Ishfaq, Shupeng Cheng, Angelica Yinling Sun, Adam C. Frank, Angel Hsing-Chi Hwang, Ruishan Liu
arXiv Computation and Language
Sep 23

From Pattern Recognizers to Personalized Companions: A Survey of Large Language Models in Mental Health

The article surveys how large language models (LLMs) are being applied to mental health, outlining a three‑phase evolution: Phase I uses LLMs as passive information tools and pattern recognizers for assessment; Phase II employs them as empathetic conversationalists for stateless, in‑the‑moment interactions; Phase III aims to create longitudinal, personalized companions that act as stateful cognitive agents. It systematically reviews core technologies, agent architectures (Profile, Memory, Reasoning, Planning), and the datasets and benchmarks that support this progression, offering a coherent narrative and roadmap for future research. The survey also provides a curated resource list at https://github.com/Emo-gml/Awesome-Mental-Health-LLMs.

By He Hu, Yucheng Zhou, Qianning Wang, Yingjian Zou, Chiyuan Ma, Juzheng Si, Jianzhuang Liu, Zitong Yu, Laizhong Cui, Fei Ma, Qi Tian
OpenAI Blog
Sep 23

Introducing MentalHealthBench

Introducing MentalHealthBench, a new benchmark developed by OpenAI, is designed to evaluate AI responses in realistic mental health conversations. The benchmark is expert-informed, focusing on assessing both helpfulness and safety of the AI’s replies. It aims to provide a standardized way to measure performance in sensitive mental health contexts.

arXiv AI
Aug 24

When Vocabulary Comprehension Fails Clinical Reasoning: Evaluating Therapy Bots' Safety Risks for Generation Alpha

The paper evaluates the safety of conversational AI therapy bots for Generation Alpha, revealing that while these models understand 76‑82% of youth‑specific vocabulary, they correctly assess clinical risk only 64‑72% of the time, creating a significant vocabulary‑comprehension gap. Six failure patterns—such as sarcasm masking, minimization acceptance, and semantic drift—were identified, with compounded errors leading to a 94% miss rate when three or more patterns co‑occur. The authors estimate 146,880 missed crises annually and recommend mandatory human‑in‑the‑loop systems, quarterly youth‑specific validation, transparent performance disclosure, and regulatory oversight for youth‑facing mental health AI.

By Manisha Mehta, Virendra Mehta
OpenAI Blog
Oct 14, 2025

Expert Council on Well-Being and AI

OpenAI’s new Expert Council on Well-Being and AI brings together leading psychologists, clinicians, and researchers to guide how ChatGPT supports emotional health, especially for teens. Learn how their insights are shaping safer, more caring AI experiences.