arXiv Computer Vision

Ordinal Diffusion Models for Color Fundus Images

The paper introduces an ordinal latent diffusion model for generating color fundus images that incorporates the ordered structure of diabetic retinopathy (DR) severity, using a scalar disease representation instead of categorical conditioning. Evaluations on the EyePACS dataset show improved visual realism, with reduced Fréchet inception distance for most stages and a higher quadratic weighted κ from 0.79 to 0.87. Interpolation experiments demonstrate the model captures a continuous spectrum of disease progression derived from coarse, ordered labels.

arXiv Machine Learning
Aug 5

Assessment of Conditional Diffusion Model for Synthetic Histopathology Image Generation

arXiv:2608. 03990v1 Announce Type: new Abstract: Synthetic histopathology image generation has emerged as an approach that may address data scarcity in computational pathology, yet current evaluation methodologies may not fully assess synthetic data quality for medical applications.

By Seyed Kahaki, Shijie Li, Weijie Chen, Nicholas Petrick
arXiv Machine Learning
Jul 15

Steering Diffusion Models via Class-Contrastive Influence for Few-Shot Medical Classification

arXiv:2607. 12464v1 Announce Type: cross Abstract: When labeled data are scarce, off-the-shelf diffusion models can augment training sets for few-shot medical image classification, but not all generated samples are equally useful for the downstream task.

By Jeeyung Kim, Erfan Esmaeili, Qiang Qiu
arXiv AI
Aug 3

DualDiT: A Conditional Dual-Output Diffusion Transformer for Joint OCT Image and Segmentation Mask Generation

arXiv:2607. 29337v1 Announce Type: cross Abstract: Background and Objective: Generating realistic medical images with anatomically accurate segmentation masks helps address the shortage of annotated data in medical imaging, particularly in optical coherence tomography (OCT) of mouse eyes, where manual retinal layer delineation is labour-intensive due to tiny structures and required expertise, resulting in scarce datasets.

By Fernando Garc\'ia-Torres, Roc\'io del Amor, Sandra Morales, \'Alvaro Barroso, Peter Heiduschka, Bj\"orn Kemper, Valery Naranjo
arXiv AI
Sep 10

Diffusion Model in Latent Space for Medical Image Segmentation Task

The paper introduces MedSegLatDiff, a diffusion-based framework that combines a variational autoencoder (VAE) with a latent diffusion model for medical image segmentation. By compressing images into a low-dimensional latent space, the method reduces noise and speeds up training, while a weighted cross‑entropy loss preserves tiny structures such as small nodules. Evaluated on ISIC‑2018, CVC‑Clinic, and LIDC‑IDRI datasets, MedSegLatDiff achieves state‑of‑the‑art Dice and IoU scores, generates diverse segmentation hypotheses, and produces confidence maps that enhance interpretability and reliability for clinical deployment.

By Ngoc Huynh Trinh, Hai Toan Nguyen, Son Ba Luong, Quoc Long Tran
arXiv Computer Vision
Sep 15

MedDiME: Efficient Latent Diffusion with Adaptive Masking for Medical Counterfactual Generation

MedDiME is a latent-space, classifier‑guided diffusion framework designed for medical counterfactual image generation. It introduces a gradient‑driven adaptive masking mechanism that works directly in latent space, enabling spatially precise edits while avoiding the high computational and memory costs of pixel‑space methods. Experiments show MedDiME can produce high‑quality counterfactuals up to 40× faster and using 13× less GPU memory than previous diffusion baselines.

By Yan Zeng, Changlu Guo, Anders Nymark Christensen, Morten Rieger Hannemose, Anders Bjorholm Dahl
Hugging Face Trending Papers
Jul 27

Color Fundus Photography Analysis: Co-evolution of Data, Preprocessing, and Modeling toward Multimodal AI

Color Fundus Photography (CFP) is a primary non-invasive imaging modality for large-scale screening of ophthalmic and systemic diseases. Existing surveys mainly summarize task-specific algorithms, datasets, or preprocessing techniques independently, lacking a unified perspective on their co-evolution with modern artificial intelligence.