arXiv Machine Learning

Alzheimer's Disease Diagnosis using a Multimodal Approach with 3D MRI and PET

arXiv:2606. 20037v1 Announce Type: new Abstract: Alzheimer's disease (AD) is an irreversible neurodegenerative disorder and a leading cause of death worldwide.

arXiv Computer Vision
Sep 14

A Multimodal Explainable Deep Learning Framework for Alzheimer's Disease Diagnosis using 3D Magnetic Resonance Imaging and Clinical Data

The study presents an explainable multimodal deep‑learning framework that combines a 3D CNN for T1‑weighted MRI with a feedforward network for harmonized clinical and demographic data to diagnose Alzheimer’s disease. Using 6,479 ADNI records and 1,703 OASIS‑3 records, the authors compare various model configurations on three‑way and pairwise diagnostic tasks, finding that performance and explanations vary by task, modality, fusion strategy, and cohort. SHAP and Integrated Gradients consistently highlight the MMSE score as the most influential tabular feature, while CAM‑based explanations differ across model setups and cohorts, indicating that explainability is not a stable property under cohort shift.

By Yusuf Brima, Marcellin Atemkeng, Lakshmana Rao Namamula, Antoine Vacavant
arXiv Computer Vision
Sep 7

A Generalizable Feature Extractor for Alzheimer's-Related Brain MRI Tasks

The study investigates whether a compact, supervised 3D CNN pretrained for brain‑age prediction can act as a reusable foundation model for various Alzheimer's‑related neuroimaging tasks. By freezing the 7.18 million weights and adding only ~1 % of trainable parameters via Low‑Rank Adaptation, the model achieved high performance across six experiments, including dementia classification, MCI progression prediction, amyloid positivity detection, and volume estimation of hippocampal and white matter hypointensities. The results demonstrate that the pretrained brain‑age model generalizes well to new datasets without retraining, offering a data‑efficient alternative to larger networks.

By Reza Rajabli, D. Louis Collins
arXiv AI
Sep 25

M$^2$PFN: End-to-End Disentangled Alignment for Generalizable Multimodal In-Context Learning in Alzheimer's Disease

M$^2$PFN is an end‑to‑end multimodal framework that extends the TabPFN in‑context learning engine to Alzheimer’s disease diagnosis by aligning 3D‑MRI and tabular features in a shared subspace. It performs differentiable inference through TabPFN’s transformer, back‑propagates gradients into the encoders, and incorporates a frozen tabular‑only prediction via a gated shortcut. On the ADNI cohort it achieves 65.55 % macro‑F1 and 82.21 % macro‑AUC, surpassing unimodal and multimodal baselines, and it generalizes to external cohorts without retraining.

By Lujia Zhong, Shuo Huang, Jianwei Zhang, Xinyu Nie, Yonggang Shi
arXiv Computer Vision
Aug 31

3D MRI-Based Alzheimer's Disease Classification Using Multi-Modal 3D CNN with Leakage-Aware Subject-Level Evaluation

The paper presents a multimodal 3D convolutional neural network that classifies Alzheimer’s disease using raw OASIS 1 MRI volumes. It fuses structural T1 images with gray matter, white matter, and cerebrospinal fluid probability maps to capture complementary neuroanatomical information. Evaluated with 5‑fold subject‑level cross‑validation, the model achieves a mean accuracy of 72.34 % and an ROC AUC of 0.7781, with GradCAM visualizations highlighting anatomically relevant regions such as the medial temporal lobe and ventricles.

By Md Sifat, Sania Akter, Akif Islam, Md. Ekramul Hamid, Abu Saleh Musa Miah, Najmul Hassan, Md Abdur Rahim, Jungpil Shin
arXiv Machine Learning
Jul 30

An Attention-Based Framework for Alzheimers Disease Classification Using Resting-State fMRI

arXiv:2607. 26746v1 Announce Type: cross Abstract: Accurate identification of Alzheimers disease (AD) using resting-state functional magnetic resonance imaging (rs-fMRI) remains challenging due to the high dimensionality, noise, and complex inter-regional dependencies inherent in functional brain connectivity, which limit the effectiveness of traditional approaches based on handcrafted connectivity features or conventional machine learning models.

By Harshiddhi Pathak, Gowtham Reddy N, Mrinal Acharya, Manjunatha Mahadevappa
arXiv AI
Sep 17

MINT: Multimodal Imaging-to-Speech Knowledge Transfer for Early Alzheimer's Screening

MINT (Multimodal Imaging-to-Speech Knowledge Transfer) is a three-stage framework that transfers MRI-derived biomarkers to speech representations for early Alzheimer’s screening. An MRI teacher creates a compact embedding space for CN‑versus‑MCI classification, and a residual projection head aligns speech features to this space using a geometric loss, allowing imaging‑free inference. Experiments on ADNI‑4 show that aligned speech matches speech baselines, while multimodal fusion outperforms MRI alone, and ablations highlight dropout regularization and self‑supervised pretraining as key design choices.

By Vrushank Ahire, Yogesh Kumar, Anouck Girard, M. A. Ganaie
arXiv Computer Vision
Sep 14

Learning Sparse Latent Predictive Foundation Model for Multimodal Neuroimaging

The paper introduces Neuro‑JEPA, a sparse multimodal foundation model that learns unified representations of brain MRI across T1w, T2w, and FLAIR sequences using a latent predictive objective and a Mixture‑of‑Experts architecture. It was pretrained on over 1.5 million scans from 428,647 studies and systematically evaluates architectural, masking, objective, and sparsity choices for robust multimodal representation learning. Across 47 tasks from three health systems and 12 public datasets, Neuro‑JEPA consistently outperforms a simple CNN baseline, demonstrating its effectiveness for both clinical and research applications.

By Haoxu Huang, Long Chen, Jingyun Chen, Jinu Hyun, James Ryan Loftus, Kara Melmed, Daniel Orringer, Jennifer Frontera, Seena Dehkharghani, Arjun Masurkar, Narges Razavian
arXiv Machine Learning
Aug 27

Modality Contribution Score - A Per-Patient Framework for Quantifying the Relative Diagnostic Contribution of Structural MRI and Amyloid PET in Alzheimer's Disease

The paper introduces MCNet, a neural network that assigns a Modality Contribution Score (MCS) to each patient, indicating how much structural MRI versus amyloid PET drives the diagnostic decision for Alzheimer’s disease. Across 327 ADNI-3 participants, MCNet achieved strong three‑class staging (AUC = 0.881) and the MCS showed a clear, statistically significant increase in PET dominance from cognitively normal to AD. The method was validated on an independent OASIS‑3 cohort and compared favorably to SHAP, suggesting it can guide personalized imaging and clinical trial decisions.

By Dawa Chyophel Lepcha, Aaliya Ali, Sophie A. Martin, Deepika Koundal, Pierrick Coupe, Shabbir Syed-Abdul