Hugging Face Trending Papers

AT-Attn: Temporal-Aware Cross-Attention for Longitudinal Multimodal Alzheimer's Disease Diagnosis

In longitudinal Alzheimer's disease (AD) diagnosis support, clinical and imaging information is often collected at irregular visits. Integrating these multimodal observations may improve diagnostic assessment, but naive fusion can degrade performance when MRI is noisy or intermittently unavailable.

arXiv Computer Vision
Sep 14

A Multimodal Explainable Deep Learning Framework for Alzheimer's Disease Diagnosis using 3D Magnetic Resonance Imaging and Clinical Data

The study presents an explainable multimodal deep‑learning framework that combines a 3D CNN for T1‑weighted MRI with a feedforward network for harmonized clinical and demographic data to diagnose Alzheimer’s disease. Using 6,479 ADNI records and 1,703 OASIS‑3 records, the authors compare various model configurations on three‑way and pairwise diagnostic tasks, finding that performance and explanations vary by task, modality, fusion strategy, and cohort. SHAP and Integrated Gradients consistently highlight the MMSE score as the most influential tabular feature, while CAM‑based explanations differ across model setups and cohorts, indicating that explainability is not a stable property under cohort shift.

By Yusuf Brima, Marcellin Atemkeng, Lakshmana Rao Namamula, Antoine Vacavant
arXiv AI
Sep 17

MINT: Multimodal Imaging-to-Speech Knowledge Transfer for Early Alzheimer's Screening

MINT (Multimodal Imaging-to-Speech Knowledge Transfer) is a three-stage framework that transfers MRI-derived biomarkers to speech representations for early Alzheimer’s screening. An MRI teacher creates a compact embedding space for CN‑versus‑MCI classification, and a residual projection head aligns speech features to this space using a geometric loss, allowing imaging‑free inference. Experiments on ADNI‑4 show that aligned speech matches speech baselines, while multimodal fusion outperforms MRI alone, and ablations highlight dropout regularization and self‑supervised pretraining as key design choices.

By Vrushank Ahire, Yogesh Kumar, Anouck Girard, M. A. Ganaie
arXiv AI
Sep 18

A Two-Stage Multi-Modal MRI Framework for Lifespan Brain Age Prediction

The paper introduces a two-stage multi‑modal MRI framework that processes different MRI modalities independently before integrating them through late fusion. First, the model estimates a probability distribution over six developmental stages, then it predicts age using probability‑weighted, stage‑specialized experts. Experiments across nine datasets—from fetal to elderly—show the method outperforms existing baselines, reducing mean absolute error by 13% and 78% in in‑domain and out‑of‑domain settings, and multi‑modal integration yields 12‑13% performance gains. Analysis on ADNI clinical groups indicates that the predicted brain age gap could help characterize Alzheimer’s‑related brain aging.

By Dingyi Zhang, Ruiying Liu, Yun Wang