The paper introduces StratCBT, a new dataset of 9,688 psychological counseling sessions with 256K utterances, each counselor response aligned to one of eight Cognitive Behavioral Therapy (CBT) strategies. It was created by modeling clients’ negative thoughts and generating high‑quality conversations through self‑chat, using realistic sessions for guidance. Experiments show that strategy‑aligned generation improves professional and effective counseling when evaluated with large language model‑simulated clients.
By Zimu Wang, Yiwen Jiang, Xiangyu Zhao, Yaling Shen, Jiahe Liu, Stephanie Fong, Maxmartwell H Cheng, Guilherme C Oliveira, Anh Nguyen, Robert Desimone, Barnaby Nelson, Dominic Dwyer, Zongyuan Ge
arXiv:2607. 08257v1 Announce Type: new Abstract: Large language models (LLMs) have shown strong performance on isolated psychiatric tasks, including dialogue, diagnosis, and treatment planning, yet existing benchmarks rarely simulate complete psychiatric clinical encounters.
By Yuming Yang, Xiao Sun, Yuanwei Zou, Zhengxiao Wu, Yun Chen, Jiang Zhong, Haoyang Zeng, Jingwang Huang, Kaiwen Wei
MACBT is a clinician‑facing AI decision‑support system that models the five‑stage CBT workflow with five collaborative agents and incorporates a CBT‑specific longitudinal memory module (CD Memory) to track cognitive distortions across sessions. The system generates pre‑session pathology reports and prioritises interventions, and was trained on a Chinese CBT dialogue corpus using a Qwen3‑14B backbone with supervised fine‑tuning and preference optimisation. Evaluation by GPT‑4 judges shows MACBT outperforms several existing chatbots in professionalism and clinical authenticity, and the memory‑augmented version improves session quality by 12.6% and achieves a longitudinal mean of 2.29 on continuity, progression, and personalization.
By De Jiang, Shuo Zhang, Weiwei Liao, Jianying Zhang, Chuanhui Yu, Hongen Liao, Kehong Yuan
arXiv:2409.02244v3 Announce Type: replace-cross
Abstract: Large language models (LLMs) are increasingly being used as ad hoc therapists. While prior research has found that LLMs outperform human coun...
By Zainab Iftikhar, Sean Ransom, Amy Xiao, Nicole Nugent, Jeff Huang
arXiv:2509.04183v3 Announce Type: replace-cross
Abstract: The growing demand for scalable psychological counseling highlights the need for high-quality, privacy-compliant data, yet such data remains...
By Aishik Mandal, Tanmoy Chakraborty, Iryna Gurevych
arXiv:2507. 02950v3 Announce Type: replace-cross Abstract: Large language models (LLMs) may support counseling training, yet evidence from Japanese-language interactions and automated quality ratings remains limited.
By Keita Kiuchi, Yoshikazu Fujimoto, Hideyuki Goto, Tomonori Hosokawa, Makoto Nishimura, Yosuke Sato, Izumi Sezai, Tomohiro Inoue
The study examined how four large language models (GPT‑5.5, Gemini 3.5 Flash, Claude Opus 4.8, and Fable 5) scored 18 simulated Japanese‑language AI‑to‑AI counseling sessions compared to ratings from 15 human counseling experts. Each model evaluated every transcript three times on four motivational‑interviewing‑informed dimensions and overall quality, consistently giving higher scores for softening sustain talk and overall quality than the expert panel, though the magnitude varied by model. Run‑to‑run reliability (intraclass correlation coefficients ranging from .33 to .96) did not predict closer alignment with expert judgments, and the models’ ability to discriminate counselor conditions was distinct from both reliability and alignment.
By Keita Kiuchi, Yoshikazu Fujimoto, Hideyuki Got\=o, Tomonori Hosokawa, Makoto Nishimura, Y\=osuke Sat\=o, Izumi Sezai, Tomohiro Inoue
arXiv:2605. 01101v2 Announce Type: replace Abstract: This paper develops Virtual Speech Therapist (VST), an intelligent agent-based platform that streamlines stuttering assessment and delivers customized therapy planning through automated and adaptive AI-driven workflows.
By Shakeel Sheikh, Patrick Marmaroli, MD Sahidullah, Slim Ouni, Fabrice Hirsch, Goncalo Leal, Bjorn W Schuller
arXiv:2606. 30887v1 Announce Type: cross Abstract: Large language models show promise for mental health support, yet therapeutic quality improves only when evaluation functions as an actionable control signal rather than a passive metric.
By Mizanur Rahman, Abeer Badawi, Elahe Rahimi, Laleh Seyyed-Kalantari, Frank Rudzicz, Enamul Hoque, Elham Dolatabadi
arXiv:2510.25384v2 Announce Type: replace
Abstract: Large Language Models (LLMs) are promising tools for synthetic data generation in mental health. However, privacy policies and restrictions forced...
By Doan Nam Long Vu, Rui Tan, Lena Moench, Svenja Jule Francke, Daniel Woiwod, Florian Thomas-Odenthal, Sanna Stroth, Tilo Kircher, Christiane Hermann, Udo Dannlowski, Hamidreza Jamalabadi, Simone Balloccu, Shaoxiong Ji
arXiv:2608. 07495v1 Announce Type: cross Abstract: Effective communication during palliative care discussions is a critical clinical skill, yet training clinicians to manage complex patient emotions remains challenging.
By Yining Wu, Tianshu Du, Jinrui Fang, Chi Zhang, Sonal Admane, Ying Ding
Graph2Counsel is a framework that generates synthetic counseling dialogues by leveraging Client Psychological Graphs (CPGs) to encode the relationships among a client’s thoughts, emotions, and behaviors. The system uses a structured prompting pipeline guided by counselor strategies and explores techniques such as Chain‑of‑Thought and Multi‑Agent Feedback to produce 760 realistic sessions from 76 CPGs. Expert evaluation shows the dataset surpasses previous ones in specificity, counselor competence, authenticity, conversational flow, and safety, and fine‑tuning an open‑source model on it improves performance on several counseling benchmarks.
By Aishik Mandal, Hiba Arnaout, Clarissa W. Ong, Juliet Bockhorst, Kate Sheehan, Rachael Moldow, Tanmoy Chakraborty, Iryna Gurevych