The paper introduces MedSegLatDiff, a diffusion-based framework that combines a variational autoencoder (VAE) with a latent diffusion model for medical image segmentation. By compressing images into a low-dimensional latent space, the method reduces noise and speeds up training, while a weighted cross‑entropy loss preserves tiny structures such as small nodules. Evaluated on ISIC‑2018, CVC‑Clinic, and LIDC‑IDRI datasets, MedSegLatDiff achieves state‑of‑the‑art Dice and IoU scores, generates diverse segmentation hypotheses, and produces confidence maps that enhance interpretability and reliability for clinical deployment.
By Ngoc Huynh Trinh, Hai Toan Nguyen, Son Ba Luong, Quoc Long Tran
We introduce TCAM-Diff, a novel 3D medical image generation model that reduces the memory requirements to encode and generate high-resolution 3D data. This model utilizes a decoder-only autoencoder method to learn triplane representation from dense volume and leverages generalization operations to prevent overfitting.
arXiv:2504. 19621v2 Announce Type: replace Abstract: Machine learning (ML) systems for medical imaging have demonstrated remarkable diagnostic capabilities, but their susceptibility to biases poses significant risks, since biases may negatively impact generalization performance.
By Haroui Ma, Francesco Quinzan, Theresa Willem, Stefan Bauer
arXiv:2606. 14759v1 Announce Type: cross Abstract: Cine cardiac magnetic resonance is the gold standard for assessing cardiac function, but the scarcity of public datasets limits the development of advanced data-driven models.
By Yiheng Cao (SyCoIA - IMT Mines Al\`es), Gustavo Andrade-Miranda (SyCoIA - IMT Mines Al\`es), Jiatian Zhang, Guillaume Sall\'e, Xin Gao
arXiv:2607. 02998v2 Announce Type: replace-cross Abstract: Controllable generative models of 3D medical images can synthesize volumes with specified clinical attributes, but this demands samples that are simultaneously high-fidelity, natively 3D, and faithful to the requested conditioning.
By Max Van Puyvelde, Halil Ibrahim Gulluk, Wim Van Criekinge, Olivier Gevaert
arXiv:2607. 05319v1 Announce Type: cross Abstract: We study why diffusion autoencoders can achieve similar image quality while learning substantially different latent structures.
By Rajat Rasal, Avinash Kori, Tian Xia, Ben Glocker