arXiv AI

Language as the Interface: Foundation-Model Contrastive Learning Links Transcriptomes and Electrophysiology

arXiv Machine Learning
Jun 18

Contextualizing Biological Language Models across Modalities via Logit-Space Contrastive Alignment

arXiv:2606. 18703v1 Announce Type: new Abstract: Pretrained biological language models expose per-token probability distributions through masked-token prediction, providing the likelihood interface central to sequence design, variant scoring, and mechanistic interpretation.

By Yanjun Shao, Yundi Chen, Yashvi Patel, Aurelien Pelissier, Mar\'ia Rodr\'iguez Mart\'inez
arXiv AI
6d ago

Neural State Prediction: Obstructing Shortcut Learning in EEG Foundation Models

Neural State Prediction (NSP) is a latent‑predictive framework designed to curb shortcut learning in EEG foundation models. By using a target encoder updated with an exponential moving average, identity residualization, and topology‑separated context, NSP constrains both the prediction target and the available context. Trained on 2.2 million EEG segments, NSP outperforms baselines on 14 datasets in the EEG‑FM‑Bench, achieving 63.94 % macro balanced accuracy.

By Kieren Yu, Ziyang Liu, Chang Huang, Jintai Chen, Kaishun Wu
arXiv Machine Learning
Sep 7

Brain4FMs: A Benchmark of Foundation Models for Electrical Brain Signal

Brain4FMs is a unified benchmark for evaluating brain foundation models (BFMs) on both scalp EEG and intracranial EEG (iEEG). It incorporates 17 representative models and 21 public datasets spanning clinical diagnosis, sleep staging, communication, and affective computing, and offers dataset-aware preprocessing, cross‑subject evaluation, and standardized downstream workflows. The benchmark highlights that no single BFM consistently outperforms others across all tasks, modalities, and adaptation protocols, prompting further exploratory analyses of model‑specific spatial, spectral, and discrete representations.

By Fanqi Shen, Enhong Yang, Jiahe Li, Junru Hong, Xiaoran Pan, Zhizhang Yuan, Meng Li, Yang Yang
arXiv Computer Vision
Sep 14

Learning Sparse Latent Predictive Foundation Model for Multimodal Neuroimaging

The paper introduces Neuro‑JEPA, a sparse multimodal foundation model that learns unified representations of brain MRI across T1w, T2w, and FLAIR sequences using a latent predictive objective and a Mixture‑of‑Experts architecture. It was pretrained on over 1.5 million scans from 428,647 studies and systematically evaluates architectural, masking, objective, and sparsity choices for robust multimodal representation learning. Across 47 tasks from three health systems and 12 public datasets, Neuro‑JEPA consistently outperforms a simple CNN baseline, demonstrating its effectiveness for both clinical and research applications.

By Haoxu Huang, Long Chen, Jingyun Chen, Jinu Hyun, James Ryan Loftus, Kara Melmed, Daniel Orringer, Jennifer Frontera, Seena Dehkharghani, Arjun Masurkar, Narges Razavian