arXiv Machine Learning

Automated Data Enrichment using Confidence-Aware Fine-Grained Debate among Open-Source LLMs for Mental Health and Online Safety

arXiv:2512. 06227v3 Announce Type: replace-cross Abstract: Real-world indicators play an important role in many Natural Language Processing (NLP) applications, such as life events for mental health analysis and risky behaviours for online safety, yet labelling such information is often costly and/or difficult due to its multi-label and dynamic nature.

arXiv AI
Aug 20

Large Language Models in Mental Health: A Systematic Review of Applications, Innovations, and Ethical Challenges

The paper reviews how large language models are applied in mental health, covering areas such as social media analysis, clinical conversational agents, therapy support tools, prompt engineering, and multimodal learning. It synthesizes interdisciplinary studies that use social media posts, electronic medical records, and multimodal inputs to detect depression, assess suicide risk, provide personalized therapy, and generate psychoeducational content. The review also discusses advances in model interpretability, annotation strategies, multimodal fusion techniques, and highlights ethical, sociotechnical, and regulatory challenges while proposing frameworks for safe, equitable, and accountable deployment.

By Yisong Chen, Yifan Gao, Sijing Yu, Chuqing Zhao, Yang Lu
arXiv AI
3d ago

Speech-based Psychological Crisis Assessment using LLMs

The paper presents an LLM-based framework for automatically classifying crisis levels in psychological support hotlines, addressing variability in human judgments and staffing constraints. It introduces a paralinguistic injection method that embeds non‑verbal emotional cues into transcripts, allowing the model to consider acoustic nuances. A reasoning‑enhanced training strategy encourages the model to produce diagnostic reasoning chains, which regularizes and improves classification, achieving a macro F1‑score of 0.802 and accuracy of 0.805 in 5‑fold cross‑validation.

By Terumi Chiba, Yang Luo, Ziyun Cui, Yongsheng Tong, Chao Zhang
Hugging Face Trending Papers
Jul 16

Self-Evolving Human-Centered Framework for Explainable Depression Symptom Annotation

Annotation quality is a major bottleneck in building reliable and explainable artificial intelligence (XAI) systems for mental health research. In depression-related datasets, labels are often assigned without structured evidence, symptom-level justification, or traceable alignment with the criteria of the Diagnostic and Statistical Manual of Mental Disorders, Fifth Edition, Text Revision (DSM-5-TR), limiting both transparency and downstream model interpretability.

arXiv Computation and Language
Sep 17

Structured Claim-Level Discourse Representations for Dense Health Narratives

The paper introduces a structured framework for claim-level discourse analysis in dense health narratives, addressing the limitations of existing topic- or sentiment-based representations. It identifies an average of 13.22 atomic claims per minute in social media health videos and proposes tuples that link each claim to thematic aspects, stance, and multidimensional pragmatic attributes. A benchmark of 1,191 manually annotated claims from 60 videos across four health domains is created, and experiments show that large language models perform well on thematic categorization and stance prediction but struggle with high-dimensional pragmatic profiling, indicating a need for task-specific inference strategies.

By Farnoushsadat Nilizadeh, Elham Pourabbas Vafa, Shirin Nilizadeh, Eduard Dragut
arXiv Computation and Language
Sep 1

Graph2Counsel: Clinically Grounded Synthetic Counseling Dialogue Generation from Client Psychological Graphs

Graph2Counsel is a framework that generates synthetic counseling dialogues by leveraging Client Psychological Graphs (CPGs) to encode the relationships among a client’s thoughts, emotions, and behaviors. The system uses a structured prompting pipeline guided by counselor strategies and explores techniques such as Chain‑of‑Thought and Multi‑Agent Feedback to produce 760 realistic sessions from 76 CPGs. Expert evaluation shows the dataset surpasses previous ones in specificity, counselor competence, authenticity, conversational flow, and safety, and fine‑tuning an open‑source model on it improves performance on several counseling benchmarks.

By Aishik Mandal, Hiba Arnaout, Clarissa W. Ong, Juliet Bockhorst, Kate Sheehan, Rachael Moldow, Tanmoy Chakraborty, Iryna Gurevych