arXiv Machine Learning

A Scaling Study for fMRI Foundation Models

The study investigates how data volume, model size, and training duration affect the performance of fMRI foundation models. Using over 200 datasets and 10,000 GPU‑hours, the authors find that larger models benefit more from additional data, and that at a fixed compute budget, increasing data yields greater gains than enlarging the model. By selecting optimal combinations of data, size, and duration, they produce models that outperform existing fMRI foundation models on out‑of‑distribution tasks while requiring less pretraining compute.

arXiv Computer Vision
Sep 24

A generalizable structural brain MRI foundation model built through dual-priority federated pretraining

BrainFedFM is a structural brain MRI foundation model that was federatively pretrained on 164,707 3‑D scans from 42 sites using a dual‑priority approach that emphasizes informative anatomical regions locally and prioritizes site contributions globally. The model outperformed seven baseline models—including four centralized foundation models—across 20 downstream tasks (classification, regression, segmentation), achieving a mean rank of 1.68 and a 50% performance gain, especially in classification and regression and among underrepresented populations. These results demonstrate the model’s generalizability and show that federated pretraining can effectively develop neuroimaging foundation models without pooling raw images.

By Zhen Yu, Yang Liu, Xiahai Zhuang, Qingchao Chen
arXiv AI
Jun 16

OmniMouse: Scaling properties of multi-modal, multi-task Brain Models on 150B Neural Tokens

arXiv:2604. 18827v2 Announce Type: replace-cross Abstract: Scaling data and artificial neural networks has transformed AI, driving breakthroughs in language and vision.

By Konstantin F. Willeke, Polina Turishcheva, Alex Gilbert, Goirik Chakrabarty, Hasan A. Bedel, Paul G. Fahey, Yongrong Qiu, Marissa A. Weis, Michaela Vystr\v{c}ilov\'a, Taliah Muhammad, Lydia Ntanavara, Rachel E. Froebe, Kayla Ponder, Zheng Huan Tan, Emin Orhan, Erick Cobos, Sophia Sanborn, Katrin Franke, Fabian H. Sinz, Alexander S. Ecker, Andreas S. Tolias
arXiv Machine Learning
Aug 4

EEG-FM-Compass: Progress, Benchmarking, and Future Directions for EEG Foundation Models

arXiv:2601. 17883v3 Announce Type: replace Abstract: Electroencephalography (EEG) foundation models (FMs) have recently emerged as a promising paradigm for brain-computer interfaces, aiming to learn transferable neural representations from large-scale heterogeneous recordings.

By Dingkun Liu, Yuheng Chen, Zhu Chen, Zhenyao Cui, Yaozhi Wen, Jiayu An, Jingwei Luo, Dongrui Wu
arXiv AI
3d ago

Combining General and Domain-Specific Pretext Tasks for Brain MR Image Segmentation

The paper investigates combining a domain‑specific self‑supervised task—voxel‑level brain age prediction—with a general task—image inpainting—to pretrain models for brain MRI segmentation. A multitask pretraining framework jointly optimizes both objectives, yielding representations that outperform single‑task pretraining and training from scratch on three segmentation benchmarks (multiple sclerosis lesions, ischemic stroke lesions, and cortical structures). The study demonstrates that integrating domain‑specific and general self‑supervised tasks benefits the development of generalizable neuroimaging foundation models.

By Tasneem Nasser, Susanne Schmid, Roberto Souza, Naser El-Sheimy