arXiv AI By Jin Mu, Guanhua Chen

Making Clinical Language Models Auditable: Concept-Guided Fine-Tuning for Robust Prediction

Read the original on arXiv AI →

The paper introduces CAST, a concept-guided artifact suppression tuning framework that uses sparse autoencoders to identify and suppress note-specific artifacts in clinical language models. CAST labels latent features with an LLM-assisted pipeline and ICD‑10 constraints, then fine‑tunes the model while providing post‑hoc per‑concept attributions for auditability. In experiments on MIMIC‑IV discharge‑note mortality prediction, CAST outperforms standard fine‑tuned encoders and competes with strong LLM baselines while offering a feature‑level audit trail of clinical concepts and suppressed artifacts.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
Jun 30

Primary ICD Category Prediction using LLM-based Probing

arXiv:2606. 28798v1 Announce Type: new Abstract: Objective: ICD codes are central to reimbursement, research, and population health surveillance, yet automated coding systems often struggle to integrate diagnostic signals from both clinical narratives and structured electronic health record (EHR) variables.

By Chengyuan Liu, Xinyue Zhang, Yao Li, Guanting Chen
arXiv Machine Learning
Jul 20

LLM4EHR: Aligning Clinical Time Series with Medical Event Sequences via Large Language Models

arXiv:2607. 15447v1 Announce Type: new Abstract: Recent research in clinical machine learning, focusing on outcome predictions in intensive care unit (ICU), has shifted from bespoke supervised models to foundation models, utilising modern representation learning methods.

By Jingteng Li, Alexander Capstick, Louise Rigny, Iona Biggart, Neil J Sebire, Payam Barnaghi
arXiv Computer Vision
Aug 25

An end-to-end-trained vision-language model for native-language prostate pathology report generation

An end-to-end-trained vision-language model generates prostate biopsy reports in native languages, demonstrated in German. The system uses a tokenizer and model trained from scratch and an automated pipeline that splits composite reports into image-text pairs, producing 17,344 pairs from 2,402 cases without manual annotation. Evaluated on clinical attributes, it achieves 96.2% F1 for malignancy detection and 65.2% for Gleason grading, comparable to an FDA-cleared classifier and validated on external cohorts.

By Christian Grashei, Fabian G\"ulhan, Maximilian Legnar, Fabian St\"ogbauer, Cleo-Aron Weis, Carolin Mogler, Peter Sch\"uffler