arXiv Computation and Language By Tong Li, Rasiq Hussain, Mehak Gupta, Joshua R. Oltmanns

"Mirror" Large Language Model Evaluations of Depression are Criterion Contaminated

Read the original on arXiv Computation and Language →

The study examines how large language models (LLMs) predict depression scores from language responses. In a "Mirror" setup, participants answered structured diagnostic interviews that the LLMs used to predict scores, yielding near-perfect predictions. In a "Non-Mirror" setup, participants gave life history interviews; the LLMs still achieved outstanding prediction accuracy, and both conditions correlated similarly with PHQ-9 scores, indicating that the Mirror advantage disappears when predicting an independent measure. Topic modeling showed different depression themes across interview types, suggesting Mirror evaluations are more about reliability than validity and that Non-Mirror approaches may enhance clinical relevance.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Computation and Language.

arXiv AI
Sep 3

Interpretable Symptom Vectors for Depression in a Large Language Model

The study investigates how a large language model, Gemma-3-27B-PT, internally represents depressive symptoms. By applying mechanistic interpretability methods to the model’s residual stream, researchers found that symptom groups are geometrically distinct at layer 21, and that projected symptom vectors align with clinician-annotated rankings across mood, somatic, and suicidality dimensions. Additionally, a single depression vector at this layer can differentiate depressive from non-depressive text with an AUC of 0.789, suggesting a potential emotional valence gate for symptom projection.

By Fangyi Zhu, Ajay Subramanian, Allison Constant, Camille Wang, Ravish Gupta, Corey J. Keller
arXiv Computation and Language
Sep 3

Candidate Generation and Definition-Guided Verification for Sentence-Level Depression Symptom Recognition

The paper introduces a two‑stage framework for recognizing depression symptoms at the sentence level. First, a contrastively fine‑tuned sentence encoder generates a symptom candidate for each sentence. Then, a fine‑tuned language model verifies the candidate’s presence or absence by comparing the sentence, its context, and a diagnostic definition, ensuring the model’s judgment aligns with that definition before responding.

By Weiming Li, Catarina Barata, Miguel Constante, Joao Sanches
arXiv AI
Aug 10

Natural Language Processing Psychometrics

arXiv:2608. 07316v1 Announce Type: cross Abstract: Natural Language Processing (NLP) models predicting mental health outcomes rarely specify what they measure: contextual knowledge, emotional content, or syntactic structure.

By Edoardo Sebastiano De Duro, Emma Franchino, Massimo Stella