arXiv Machine Learning

Learning from Scarce Labels: Multi-View Echocardiography for Ejection Fraction Prediction

The paper introduces the first publicly available dataset of over 25,000 parasternal long‑axis (PLAX) echocardiography videos labeled for left ventricular ejection fraction (EF), created through a novel data‑generation strategy that links clinical notes to video data. Using this dataset, the authors train a reproducible PLAX‑based EF model that achieves a mean absolute error (MAE) of 6.86%, comparable to the clinical standard of apical four‑chamber (A4C) methods. They further show that simple late fusion of PLAX and A4C predictions reduces MAE to 6.37%, highlighting the benefit of multi‑view integration, and release the dataset, models, and demos publicly.

Hugging Face Trending Papers
Sep 2

Learning from Scarce Labels: Multi-View Echocardiography for Ejection Fraction Prediction

The paper introduces the first publicly available dataset for predicting left ventricular ejection fraction (EF) from parasternal long-axis (PLAX) echocardiography, comprising over 25,000 labeled videos generated through a novel data‑generation strategy that correlates clinical notes with echocardiographic videos. Using this dataset, the authors train a reproducible PLAX‑EF model that achieves a mean absolute error (MAE) of 6.86%, comparable to the clinical standard of apical four‑chamber (A4C) methods. They further show that combining PLAX and A4C predictions via simple late fusion reduces MAE to 6.37%, highlighting the benefit of multi‑view integration, and they release the dataset, models, and demos for community use.

arXiv Computer Vision
3d ago

Echo-E$^3$Net: Efficient Endocardial Spatio-Temporal Network for Ejection Fraction Estimation

Echo-E$^3$Net is an anatomy‑guided spatio‑temporal neural network designed to estimate left ventricular ejection fraction (LVEF) from ultrasound images. It uses a dual‑phase Endocardial Border Detector to locate end‑diastole and end‑systole landmarks and an Endocardial Feature Aggregator to fuse these landmarks with global deep‑feature descriptors for EF regression. The model achieves competitive accuracy on EchoNet‑Dynamic and EchoNet‑Pediatric datasets while using only 1.55 M parameters and 8.05 GFLOPs, enabling real‑time deployment on limited‑resource devices.

By Moein Heidari, Afshin Bozorgpour, AmirHossein Zarif-Fakharnia, Wenjin Chen, Dorit Merhof, David J. Foran, Jasmine Grewal, Ilker Hacihaliloglu
arXiv AI
Jul 2

EchoRisk: A Multicentre Echocardiography Dataset and Benchmark for Cardio-Oncology

arXiv:2607. 01039v1 Announce Type: cross Abstract: Therapy-induced cardiotoxicity is the leading non-oncological cause of treatment interruption in breast cancer patients, yet early, automated risk stratification from routine cardiac imaging remains an unsolved problem.

By Grigorios Kalliatakis, Georgia Karanasiou, Georgios Manikis, Manolis Tsiknakis, Dimitrios Fotiadis, Dorothea Tsekoura, Kalliopi Keramida, Vasileios Bouratzis, Lampros Lakkas, Katerina Naka, Andri Papakonstantinou, Anastasia Constantinidou, Kostas Marias
Hugging Face Trending Papers
Jul 1

EchoRisk: A Multicentre Echocardiography Dataset and Benchmark for Cardio-Oncology

Therapy-induced cardiotoxicity is the leading non-oncological cause of treatment interruption in breast cancer patients, yet early, automated risk stratification from routine cardiac imaging remains an unsolved problem. We present EchoRisk, the first curated, multicentre, longitudinal echocardiography dataset with explicit cardiotoxicity labels, released as the primary technical reference for the EchoRisk-MICCAI 2026 challenge.

arXiv AI
Jul 16

Anatomically Faithful but Temporally Blind: Auditing Attribution for Left-Ventricular Ejection-Fraction Estimation from Echocardiography

arXiv:2607. 13738v1 Announce Type: cross Abstract: Background and Objective: Deep video models estimate left-ventricular ejection fraction (EF) from echocardiography with near-expert accuracy, and post-hoc attribution (Chefer relevance for transformers, Grad-CAM for CNNs) is increasingly used to certify that models "look at the right place.

By Hyunkyung Han, Min Jung Kim
arXiv Computer Vision
4d ago

SV-Cine: Diagnosis-Conditioned Segmentation of Single Ventricle Physiology via Generative Data Augmentation

The paper introduces SV-Cine, a cardiac MRI segmentation framework tailored for single ventricle physiology (SVP). It combines a generative data augmentation pipeline that creates synthetic 3D cardiac meshes and MRI, with a diagnosis-conditioned adaptation of the CineMA foundation model that uses patient-level diagnostic information to improve segmentation. Evaluations on an internal cohort show high Dice scores for left and right ventricles, outperforming nnU-Net, and demonstrate that incorporating diagnosis priors can adapt a pretrained model to specialized SVP tasks.

By Lila Cunge, Yuehong Liu, Hang Xu, Thomas Coudert, Pierangelo Renella, J Paul Finn, William Hsu, Kim-Lien Nguyen
arXiv AI
Jun 2

CardioLens: Revealing the Clinical Reality Gap of MLLMs via Multi-Sequence Cardiac MRI Evaluations

arXiv:2606. 00123v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have shown strong performance on public medical benchmarks, yet existing evaluations often remain weak proxies for clinical use, relying on isolated inputs and simplified recognition-style tasks.

By Zixian Su, Hongkai Zhang, Fan Gao, Encheng Su, Taiping Qu, Jingwei Guo, Nan Zhang, Hui Wang, Zhen Zhou, Kairui Bo, Yan Chen, Yue Ren, Shuai Li, Lei Xu, Henggui Zhang