arXiv:2603. 13673v2 Announce Type: replace Abstract: Accurate extraction of Alzheimer's Disease and Related Dementias (ADRD) phenotypes from electronic health records (EHR) is critical for early-stage detection and disease staging.
By Mingchen Shao, Yuzhang Xie, Carl Yang, Jiaying Lu
arXiv:2609.19071v1 Announce Type: new
Abstract: Extracting SNP-phenotype associations from biomedical literature is vital but challenging. We benchmarked diverse NLP models, including MLMs, hybrid ar...
By Claudiu Creanga, Teodor Marchitan, Liviu P. Dinu
The paper introduces pre‑trained models for extracting variant‑phenotype relations from biomedical text, focusing on the SNPPhenA corpus. Fine‑tuning small BERT‑based models, especially DeBERTa, achieves performance close to the current state‑of‑the‑art. Moreover, careful fine‑tuning of Google’s Gemini Pro 1.0 surpasses existing benchmarks on both sentence‑level and abstract‑level relation extraction tasks.
By Claudiu Creanga, Liviu P. Dinu, Daniela Gifu
arXiv:2609.12544v1 Announce Type: new
Abstract: Clinical de-identification relies on accurately identifying personally identifiable information (PII). However, manually annotated datasets are costly...
By Linh Uyen Le, Christian Hoang, Huy Hoang Ha
arXiv:2606. 00031v1 Announce Type: cross Abstract: Coronary artery disease (CAD) remains one of the leading causes of death globally, highlighting the need for reliable predictive systems to support early diagnosis and risk assessment.
By Jeba Maliha, Md Rafiul Kabir
The paper introduces MiNER, a fine‑tuned biomedical NLP system that uses BioBERT to extract malaria‑related named entities from scientific literature. It builds a large, annotated corpus of malaria articles, preprocesses the text, and applies supervised learning to improve extraction performance. Experiments show that MiNER outperforms other encoding and machine‑learning methods in precision, recall, and accuracy, and the authors release the human‑labeled dataset for further research.
By V. S. Anoop, Devika N