arXiv AI

LLM-MINE: Large Language Model based Alzheimer's Disease and Related Dementias Phenotypes Mining from Clinical Notes

arXiv:2603. 13673v2 Announce Type: replace Abstract: Accurate extraction of Alzheimer's Disease and Related Dementias (ADRD) phenotypes from electronic health records (EHR) is critical for early-stage detection and disease staging.

arXiv Machine Learning
Sep 18

Pretrained Medical Representations for the Practical Screening of Drug Repositioning Candidates

The paper introduces a unified pre‑training framework for medical representations that incorporates hierarchical sub‑token aggregation, partial masking, and cross‑reference mechanisms to better capture the structure of medical codes. The resulting model outperforms existing BERT‑based approaches on pre‑training tasks and downstream clinical predictions, such as dementia onset and hospitalization. An in‑silico drug repositioning study for Alzheimer’s disease demonstrates the framework’s ability to rediscover known drugs and prioritize new hypotheses without external literature, establishing a workflow for hypothesis generation and prioritization based on observational data.

By Yuhei Fujioka, Daitaro Misawa, Shingo Fukuma
arXiv Machine Learning
Jul 28

Dementia Etiology Diagnosis via Collaborative Meta Knowledge Enhancement

arXiv:2607. 22770v1 Announce Type: new Abstract: Although artificial intelligence (AI) has shown promising performance in several medical tasks, accurate dementia etiology diagnosis with AI remains challenging due to complex overlapping symptoms among diseases.

By Siyuan Du, Mengxi Chen, Xinyang Jiang, Zilong Wang, Jiangchao Yao, Dongsheng Li, Ya Zhang, Lili Qiu, Yanfeng Wang
arXiv Computation and Language
Sep 18

Fine-Tuning Models for Biomedical Relation Extraction

The paper introduces pre‑trained models for extracting variant‑phenotype relations from biomedical text, focusing on the SNPPhenA corpus. Fine‑tuning small BERT‑based models, especially DeBERTa, achieves performance close to the current state‑of‑the‑art. Moreover, careful fine‑tuning of Google’s Gemini Pro 1.0 surpasses existing benchmarks on both sentence‑level and abstract‑level relation extraction tasks.

By Claudiu Creanga, Liviu P. Dinu, Daniela Gifu
arXiv AI
Jun 12

AAbAAC: An Annotated Corpus for Autoimmunity Information Extraction

arXiv:2606. 13051v1 Announce Type: new Abstract: Despite advances in information extraction driven by deep learning and large language models, performance gaps remain in highly specialized biomedical fields, where domainspecific complexity poses challenges for generalist models.

By Fabien Maury (Imagine - U1163, HeKA | U1346), Sol\`ene Grosdidier (Imagine - U1163), Maud de Dieuleveult (Imagine - U1163), Adrien Coulet (HeKA | U1346)
arXiv Computation and Language
Sep 22

Knowledge Graph-Augmented Ambient AI for Clinical Note Generation

arXiv:2609.22239v1 Announce Type: new Abstract: Ambient AI is increasingly adopted in healthcare to automatically generate clinical notes from patient-clinician conversations, with the potential to s...

By Jakir Hossain, Yi-Fei Zhao, Hongjian Wang, Minmei Shih, Katie Leigh Mullen, Ahmad P. Tafti, Leming Zhou, Manoj Purohit, William Hogan, Jay Zeng, Elizabeth Skidmore, Yanshan Wang