arXiv Machine Learning By Zihan Ding, Yinan Liu, Tengfei Ma, Rachel Wong, George Leibowitz, Benjamin Littenberg, Xia Zheng, Richard N. Rosenthal, Fusheng Wang

A Comparative Study of Feature Selection Methods for EHR Diagnosis Codes in Opioid Use Disorder Prediction

Read the original on arXiv Machine Learning →

arXiv:2608. 04180v1 Announce Type: new Abstract: Feature selection is a critical step in electronic health record (EHR)-based predictive modeling, where input variables are often high-dimensional, sparse, noisy, and redundant.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv Machine Learning
Sep 14

Patient-Reported Survey Data Improve Prediction of Opioid Use Disorder

The study examined whether adding patient‑reported survey data to electronic health records (EHRs) improves the prediction of a first opioid use disorder (OUD) diagnosis. Using 267,747 All of Us participants, the authors compared EHR‑only models to EHR+survey models across multiple machine‑learning algorithms and look‑back windows. Survey augmentation consistently increased predictive performance, with the best 24‑month LightGBM model’s PR‑AUC rising from 0.6219 to 0.6603, and survey features ranked as the second most important information domain.

By Xiyue Jiang, Zihan Ding, Grace Han, Yinan Liu, Richard N. Rosenthal, Fusheng Wang
arXiv AI
Jun 30

Primary ICD Category Prediction using LLM-based Probing

arXiv:2606. 28798v1 Announce Type: new Abstract: Objective: ICD codes are central to reimbursement, research, and population health surveillance, yet automated coding systems often struggle to integrate diagnostic signals from both clinical narratives and structured electronic health record (EHR) variables.

By Chengyuan Liu, Xinyue Zhang, Yao Li, Guanting Chen