arXiv AI

Understanding Sources of Demographic Predictability in Brain MRI via Disentangling Anatomy and Contrast

arXiv:2603. 04113v2 Announce Type: replace-cross Abstract: Demographic attributes can be predicted from medical images, raising concerns about bias in clinical AI systems.

arXiv Computer Vision
Sep 3

Progression as Latent Drift: Generative Forecasting of Slow-Evolving Pathologies

The paper introduces Latent Drift, a generative forecasting framework that predicts slow-evolving neurodegenerative disease progression by learning changes in a compressed semantic representation rather than full-resolution anatomy. It addresses two failure modes—identity collapse and continuous interpolation trap—by removing pixel-level identity from the prediction target and applying Finite Scalar Quantization to suppress high-frequency nuisance fluctuations. Experiments on longitudinal 3D brain MRI demonstrate that Latent Drift outperforms diffusion and autoregressive transformer baselines in both generative fidelity and clinically relevant metrics.

By Yuxiang Feng, Juncheng Wang, Chao Xu, Wenlong Hou, Huihan Wang, Yijie Qian, Yang Liu, Baigui Sun, Yong Liu, Shujun Wang
arXiv AI
Jun 10

Tractogram foundation model

arXiv:2606. 09893v1 Announce Type: cross Abstract: Diffusion MRI (dMRI) tractography is the only noninvasive approach for mapping white-matter pathways in the living human brain.

By Guikun Chen, Yuqian Chen, Yijie Li, Yogesh Rathi, Nikos Makris, Fan Zhang, Wenguan Wang, Lauren J. O'Donnell
arXiv Computer Vision
Sep 23

Decoupling Disease, Covariates, and Individual Variability: A Unified Disentanglement Framework for Medical Image Classification

The paper introduces MedIDL, a Medical Imaging Disentanglement Learning framework that separates disease-related features from confounding covariates and individual variability in medical images. It achieves this by projecting image features into three orthogonal latent spaces—disease classification, covariate alignment, and a Gaussian head for individual variation—using specialized disentanglement heads. Across seven diverse imaging datasets, MedIDL surpasses state‑of‑the‑art supervised and self‑supervised methods in classification accuracy, and its latent representations and gradient‑based visualizations align with known clinical patterns.

By Shengjie Zhang, Jinglin Zhang, Zhuangzhuang Jiang, Ziqi Yu, Yipin Zhang, Qi Zhang, Xiang Chen, Haibo Yang, Fei Gao, Longbiao Cui, Yuan Zhou, Xiao-Yong Zhang, Alzheimer's Disease Neuroimaging Initiative
arXiv Computer Vision
Sep 4

Subgroup performance analysis of adaptation strategies for chest X-ray foundation models

The study examines how three parameter‑efficient adaptation methods—linear heads on the raw CLS token, an MLP, and an attention‑pooling module—affect pathology classification accuracy and subgroup fairness when applied to a frozen Rad‑DINO chest X‑ray encoder. Using the MIMIC‑CXR dataset, the authors evaluate eight pathologies across race, sex, and imaging‑view subgroups, finding that attention pooling yields the best overall performance and encodes protected attributes most strongly, yet higher performance does not consistently reduce subgroup disparities. The results show that attribute encoding strength and layer choice do not reliably predict fairness outcomes, indicating that fairness must be assessed directly for each task.

By Dhruv Gupta, Emma A. M. Stanley, Fabio De Sousa Ribeiro, Sujal R. Desai, Ben Glocker