arXiv:2607. 03744v1 Announce Type: new Abstract: Automatic depression detection from clinical interviews typically models the semantic content and acoustic characteristics of participant speech.
By Hanie Kang, Huang-Cheng Chou, Sudarsana Reddy Kadiri, Shrikanth Narayanan
arXiv:2607. 25888v1 Announce Type: cross Abstract: This study identifies new depression biomarkers based on the dynamical properties of tract variables, which represent geometric features describing the configuration of the speech articulators.
By Sahar Altalhi, Tanaya Guha, Alessandro Vinciarelli
MERID is a framework that uses recursive self‑improvement agents to autonomously develop multimodal pipelines for detecting major depressive disorder. It aligns multimodal records with depression targets, jointly modifies representations, fusion, and predictors, and guides revisions through evidence‑guided evolution to validate improvements before inheritance. Experiments on depression benchmarks show MERID outperforms existing multimodal and agent‑based baselines, especially highlighting the importance of acoustic and linguistic cues.
By Lei Liu, Zhaokang Liang, Qingcheng Zeng, Chenda Duan, Lu Mi, Zhen Tan, Tianyu Liu
EviDep is a multimodal evidential regression framework for estimating depression severity from audio–visual recordings, incorporating multi‑scale temporal modeling and shared–private representation learning. It uses frequency‑aware feature extraction to decompose behavioral sequences into multiple frequency bands, refined by scale‑specific experts, and applies disentangled evidential learning to separate cross‑modal shared and modality‑specific information. The model outputs Normal‑Inverse‑Gamma distributions via multi‑branch evidential regression, enabling estimation of depression severity along with aleatoric and epistemic uncertainty, and demonstrates competitive accuracy on several benchmark datasets.
By Fangyuan Liu, Sirui Zhao, Yangsong Zhang, Jinyang Huang, Feng-Qi Cui, Bin Luo, Tong Xu, Enhong Chen
arXiv:2606. 05561v1 Announce Type: cross Abstract: Speech-based mental health screening offers scalable depression detection, yet clinical deployment faces a significant barrier: users' privacy concerns about demographic information exposure.
By Xueyang Wu, Siyuan Liu, Kezhuo Yang, Guang Ling
arXiv:2501. 16106v2 Announce Type: replace Abstract: Recent advances in multimodal depression recognition for clinical interviews (MDRC) have demonstrated the potential of AI systems by integrating textual, acoustic, and facial cues.
By Wenjie Zheng, Qiming Xie, Jianfei Yu, Yang Wang, Lei Cao, Fei Wang, Shijin Wang, Rui Xia, Chengqing Zong