arXiv AI

Anatomy-Guided Residual Motion Diffusion for Controllable 4D Cardiac MRI Synthesis

arXiv:2606. 26764v1 Announce Type: cross Abstract: Developing robust artificial intelligence models for 4D (3D + time) medical imaging is constrained by limited annotated data, inter-device domain shifts, and privacy restrictions.

arXiv AI
Jun 16

Temporally Consistent and Controllable Video Generation of 2D Cine CMR via Latent Space Motion Modeling

arXiv:2606. 14759v1 Announce Type: cross Abstract: Cine cardiac magnetic resonance is the gold standard for assessing cardiac function, but the scarcity of public datasets limits the development of advanced data-driven models.

By Yiheng Cao (SyCoIA - IMT Mines Al\`es), Gustavo Andrade-Miranda (SyCoIA - IMT Mines Al\`es), Jiatian Zhang, Guillaume Sall\'e, Xin Gao
arXiv Computer Vision
Sep 14

SV-Cine: Diagnosis-Conditioned Segmentation of Single Ventricle Physiology via Generative Data Augmentation

The paper introduces SV-Cine, a cardiac MRI segmentation framework tailored for single ventricle physiology (SVP). It combines a generative data augmentation pipeline that creates synthetic 3D cardiac meshes and MRI, with a diagnosis-conditioned adaptation of the CineMA foundation model that uses patient-level diagnostic information to improve segmentation. Evaluations on an internal cohort show high Dice scores for left and right ventricles, outperforming nnU-Net, and demonstrate that incorporating diagnosis priors can adapt a pretrained model to specialized SVP tasks.

By Lila Cunge, Yuehong Liu, Hang Xu, Thomas Coudert, Pierangelo Renella, J Paul Finn, William Hsu, Kim-Lien Nguyen
arXiv AI
Jul 8

CONFLUX: A Latent Diffusion Model for 3D Chest-CT Synthesis with RL Post-Training

arXiv:2607. 02998v2 Announce Type: replace-cross Abstract: Controllable generative models of 3D medical images can synthesize volumes with specified clinical attributes, but this demands samples that are simultaneously high-fidelity, natively 3D, and faithful to the requested conditioning.

By Max Van Puyvelde, Halil Ibrahim Gulluk, Wim Van Criekinge, Olivier Gevaert
arXiv AI
Jul 7

CONFLUX: A Latent Diusion Model for 3D Chest-CT Synthesis with RL Post-Training

arXiv:2607. 02998v1 Announce Type: cross Abstract: Controllable generative models of 3D medical images can synthesize volumes with specified clinical attributes, but this demands samples that are simultaneously high-fidelity, natively 3D, and faithful to the requested conditioning.

By Max Van Puyvelde, Halil Ibrahim Gulluk, Wim Van Criekinge, Olivier Gevaert
arXiv Computer Vision
Aug 21

4DLoG: Generative Modeling of Neurodegenerative Brain Anatomy with 4D Longitudinal Diffusion Model

arXiv:2604. 22700v2 Announce Type: replace Abstract: Modeling and predicting neurodegenerative disease progression from medical images remains a major challenge in medical AI, with significant implications for early diagnosis, disease monitoring, and treatment planning.

By Nivetha Jayakumar, Swakshar Deb, Bahram Jafrasteh, Qingyu Zhao, Miaomiao Zhang
arXiv Computer Vision
Sep 2

CMRVision: A Foundation Model for Cardiac MR Image Analysis

CMRVision is a cardiac magnetic resonance (CMR) foundation model trained with DINOv3-style self‑supervised learning on 36 million multi‑center, multi‑sequence CMR images. It outperforms prior natural‑image, medical‑image, supervised, and CMR baselines on multi‑task segmentation (cine, LGE, mapping) and cine view classification, achieving Dice scores of 0.940–0.967 for LV and 0.855–0.905 for myocardium, and a zero‑shot Dice of 0.692 on unseen LGE long‑axis views. The model demonstrates robust cross‑view generalization and highest average accuracy (0.906) for cine view classification.

By Athira J. Jacob, Puneet Sharma, Daniel Rueckert
arXiv AI
Sep 10

Diffusion Model in Latent Space for Medical Image Segmentation Task

The paper introduces MedSegLatDiff, a diffusion-based framework that combines a variational autoencoder (VAE) with a latent diffusion model for medical image segmentation. By compressing images into a low-dimensional latent space, the method reduces noise and speeds up training, while a weighted cross‑entropy loss preserves tiny structures such as small nodules. Evaluated on ISIC‑2018, CVC‑Clinic, and LIDC‑IDRI datasets, MedSegLatDiff achieves state‑of‑the‑art Dice and IoU scores, generates diverse segmentation hypotheses, and produces confidence maps that enhance interpretability and reliability for clinical deployment.

By Ngoc Huynh Trinh, Hai Toan Nguyen, Son Ba Luong, Quoc Long Tran
arXiv Machine Learning
Aug 27

FlowMoDL: Model-Based Deep Learning with Conjugate-Gradient Data Consistency for Highly Accelerated 4D Flow MRI Reconstruction

FlowMoDL is an unrolled neural network designed for highly accelerated 4D flow MRI reconstruction, optimizing both anatomical magnitude and phase-derived velocity accuracy. It alternates a learned (3+1)D spatiotemporal denoiser with conjugate‑gradient data‑consistency updates, using a dual‑pathway conditioning scheme to handle acceleration factors from 10× to 50×. Trained with a deep‑supervision composite loss that penalizes velocity magnitude and angular errors, FlowMoDL outperforms classical and deep‑learning baselines on the multi‑center CMRx4DFlow dataset, achieving superior gradient‑step efficiency and robust convergence across all acceleration factors.

By Tristan Gottwald, Michelle Bruch, Mubashir-Ul Hassan, Fatma Alickovic, Milan Kloiber, Daniel Tenbrinck, Torsten Panholzer, Melanie Schaller, Jana Hutter