arXiv AI

FOUND-AF: Benchmarking ECG Foundation Models for Atrial Fibrillation Detection

arXiv:2608. 03597v1 Announce Type: new Abstract: Atrial fibrillation (AF) is the most common sustained cardiac arrhythmia and is associated with increased risks of stroke, heart failure, and mortality.

arXiv Machine Learning
Jul 31

ECG-InterpBench: Benchmarking the Interpretability of ECG Foundation Models with Matched-Scale Sparse Autoencoders

arXiv:2607. 27404v1 Announce Type: new Abstract: Existing benchmarks for electrocardiogram foundation models primarily evaluate downstream predictive performance, providing limited insight into whether their internal representations can be faithfully decomposed, clinically interpreted, or reproduced across independent analyses.

By Yixuan Duan, Wei Qiu
arXiv Machine Learning
Sep 14

A Dataset and Benchmarks for Atrial Fibrillation Detection from Electrocardiograms of Intensive Care Unit Patients

The paper presents a new labelled ICU dataset and benchmarks for detecting atrial fibrillation (AF) from electrocardiograms (ECGs). It compares three AI approaches—feature‑based classifiers, deep learning, and ECG foundation models—across Canadian ICU data and the 2021 PhysioNet challenge, finding that ECG foundation models with transfer learning achieve the highest F1 score (0.89). The study demonstrates the feasibility of automated AF monitoring in ICU settings and provides resources for further research.

By Sarah Nassar, Nooshin Maghsoodi, Sophia Mannina, Shamel Addas, Stephanie Sibley, Gabor Fichtinger, David Pichora, David Maslove, Purang Abolmaesumi, Parvin Mousavi
arXiv Machine Learning
Aug 20

Atrial Fibrillation Detection with Arbitrary Leads via a Codebook-Based Reconstruction-Classification Framework

The paper introduces DCGCNet, a dual-codebook graph collaborative network that jointly reconstructs ECG signals and classifies atrial fibrillation. It incorporates a local‑global contrastive module for noise‑invariant feature learning and an adaptive codebook vector quantizer to prevent codebook collapse. The model achieves state‑of‑the‑art intra‑dataset performance and consistently attains AUC > 0.98 across seven cross‑dataset settings, even under realistic noisy conditions.

By Hongtao Li, Jia Wei, Guoyao Li, Yuchen Lei, Guangnian Ma, Jia Xiao, Yuanjun Lai, Shuzhen Lv, Xueqiang Ouyang
arXiv Machine Learning
Aug 4

Automated ECG Interval Measurement and Wave Delineation Using Fast Fourier Convolution ResNet

arXiv:2608. 00058v1 Announce Type: cross Abstract: Accurate measurement of ECG intervals, including PR, QRS duration, and QT/QTc, is central to cardiac diagnosis, yet the published ECG delineation literature evaluates performance almost exclusively as fiducial-point timing errors on small curated databases, rather than as clinical interval accuracy on large unselected cohorts.

By Farhan Adam Mukadam, Harshit Mishra, Nachiket Makwana, Pradyot Tiwari, Subramani Kandasamy, KVS Hari
arXiv AI
Sep 17

NeuroECG: ECGFounder-Based Deep ECG Representation for EEG-Free Neurological Prognostication After Cardiac Arrest

NeuroECG is a deep learning framework that repurposes a pretrained ECG foundation model to predict neurological outcomes after cardiac arrest without using electroencephalography (EEG). The model fine‑tunes the backbone with a gradual unfreezing strategy on single‑channel bedside ECG, aggregates multiple ECG segments via quantile pooling and PCA, and achieves an AUROC of 0.7333 using ECG alone. When combined with static clinical covariates, NeuroECG improves performance to an AUROC of 0.8077 and an AUPRC of 0.8970, demonstrating that bedside ECG can serve as a low‑cost, auxiliary prognostic tool in an EEG‑free setting.

By Jiaju Gao, Yi Zhao, Chenyang Xu, Yuxi Zhou, Hao Wang
arXiv Computation and Language
Sep 1

ECGQuest: Benchmarking and Fine-Tuning Language Models for Electrocardiography

ECGQuest is a new benchmark that evaluates language models on the contextual knowledge required for electrocardiogram interpretation, featuring 10,904 True/False questions derived from 23 ECG references and 2003‑2025 Computing in Cardiology proceedings. The study tested 23 commercial and open‑source models, finding that zero‑shot accuracy ranged from 49.5% to 74.4% and that fine‑tuning with Low‑Rank Adaptation improved all open‑source models by 6.5–14.1%, with the best fine‑tuned model achieving 76.3% accuracy and a five‑model ensemble reaching 78.5%. ECGQuest demonstrates that parameter‑efficient fine‑tuning can enable smaller models to compete with larger commercial ones on ECG‑specific tasks.

By Mohammadsina Hassannia, Matthew A. Reyna, Reza Sameni
arXiv Machine Learning
Sep 10

AF-Mamba: Efficient Long-Term Signal Modeling for Early Prediction of Atrial Fibrillation Onset

AF-Mamba is a deep learning model that predicts atrial fibrillation (AF) onset one hour in advance using long‑term RR intervals. It combines temporal convolutional networks for local feature extraction with Mamba, a state‑space model for long‑range sequence modeling, achieving high sensitivity (0.889) and specificity (0.943) in subject‑wise testing. The model maintains strong performance across unseen datasets, offering a favorable trade‑off between predictive accuracy and computational efficiency for real‑time ambulatory monitoring.

By Yongbin Lee, Ki H. Chon
arXiv Machine Learning
Jul 7

Do ECG Foundation Models Transfer to Rare Cardiac Diseases? Evidence from Brugada Syndrome Detection

arXiv:2607. 03009v1 Announce Type: new Abstract: Background: Foundation models (FMs) trained on large-scale unlabeled physiological data have emerged as a promising paradigm for medical artificial intelligence.

By Beatrice Zanchi, Giuliana Monachino, Alvise Dei Rossi, Luigi Fiorillo, Georgia Sarquella-Brugada, Giulio Conte, Francesca Dalia Faraci
arXiv AI
Aug 7

ECG-LENS: Lead-Aware Clinical Context Enriched ECG Report Generation and Evaluation

arXiv:2608. 05893v1 Announce Type: new Abstract: Electrocardiography (ECG) is one of the most widely used non-invasive tools for diagnosing cardiovascular disease, but transforming multi-lead ECG recordings into reliable clinical reports remains challenging.

By Akanta Das, Tasinul Islam Ahon, Ahmed Mahir Sultan Rumi, Md Mahbubur Rahman, Tausif Amim Shadly, Tanzima Hashem