arXiv:2508. 17742v3 Announce Type: replace-cross Abstract: Electroencephalography foundation models (EEG-FMs) have advanced brain signal analysis, but the lack of standardized evaluation benchmarks impedes model comparison and scientific progress.
By Wei Xiong, Jiangtong Li, Jie Li, Kun Zhu, Changjun Jiang
NeuroAtlas is the largest EEG benchmark to date, comprising 42 datasets and 260,000 hours of clinical EEG data across epilepsy, sleep medicine, brain age estimation, and brain‑computer interfaces. The study evaluates foundation models (FMs) for EEG against supervised baselines and generic time‑series FMs, finding that EEG‑specific FMs do not consistently outperform generic ones. It also demonstrates that standard machine‑learning metrics are inadequate for clinical relevance, advocating for task‑specific measures such as event‑level decision quality, hypnogram features, and brain‑age gap.
By Konstantinos Kontras, Trui Osselaer, Stylianos G. Mouslech, Angeliki-Ilektra Karaiskou, Guido Gagliardi, Thomas Strypsteen, Mohammad Hossein Badiei, Anku Rani, Maarten Vanmarcke, Miguel Bhagubai, Chanakya Ekbote, Jaedong Hwang, Christos Chatzichristos, Paul Pu Liang, Maarten De Vos
arXiv:2601. 17883v3 Announce Type: replace Abstract: Electroencephalography (EEG) foundation models (FMs) have recently emerged as a promising paradigm for brain-computer interfaces, aiming to learn transferable neural representations from large-scale heterogeneous recordings.
By Dingkun Liu, Yuheng Chen, Zhu Chen, Zhenyao Cui, Yaozhi Wen, Jiayu An, Jingwei Luo, Dongrui Wu
Electroencephalography (EEG) foundation models increasingly rely on multi-dataset training and evaluation, yet public EEG datasets still lack a shared task specification layer that can turn heterogeneous recordings into reusable benchmark units. Existing standards organize files, metadata, and provenance, but they do not specify EEG tasks under a common language and rulebook, leaving critical task semantics scattered across papers, code, and manual interpretation.
Brain4FMs is a unified benchmark for evaluating brain foundation models (BFMs) on both scalp EEG and intracranial EEG (iEEG). It incorporates 17 representative models and 21 public datasets spanning clinical diagnosis, sleep staging, communication, and affective computing, and offers dataset-aware preprocessing, cross‑subject evaluation, and standardized downstream workflows. The benchmark highlights that no single BFM consistently outperforms others across all tasks, modalities, and adaptation protocols, prompting further exploratory analyses of model‑specific spatial, spectral, and discrete representations.
By Fanqi Shen, Enhong Yang, Jiahe Li, Junru Hong, Xiaoran Pan, Zhizhang Yuan, Meng Li, Yang Yang
arXiv:2608. 04156v1 Announce Type: new Abstract: Electroencephalography (EEG) analysis extends beyond assigning predefined labels to recordings; it requires workflows connecting natural-language instructions, signal processing, quantitative evidence, and scientific interpretation.
By Yangxuan Zhou, Sha Zhao, Yuning Chen, Chen Wu, Jiquan Wang, Shijian Li, Gang Pan
EEG-AS is an instance-level algorithm selection framework designed for EEG foundation models. It characterizes each EEG instance using latent embeddings, handcrafted neurophysiological features, and an anchor foundation model, then learns to reconstruct the behaviors of other foundation models from privileged prediction tokens. During inference, EEG-AS estimates these behaviors without running the full model portfolio, enabling efficient selection among seven EEG foundation models and significantly reducing the performance gap between the single best solver and the oracle upper bound across seven public EEG benchmarks.
By Yunzhen Zhang, Ruoxi Piao, Hasan Onur Keles, Mustafa Misir
iMINDBench is a new benchmark for intracranial electroencephalography (iEEG) neural decoding that evaluates models on fifteen tasks across three naturalistic movie‑watching datasets from multiple institutions. It standardizes preprocessing tracks and evaluation splits to enable consistent comparisons. The study shows that pretrained systems outperform baselines within their tracks, but strong spectral baselines remain competitive, and scaling up supervised data yields only modest or task‑dependent gains.
By Geeling Chau, Saba Hashemi, Yonghyeon Gwon, Eshani Patel, Jan DeWitt, Christopher Wang, Andrii Zahorodnii, Sabera J Talukder, Danny Dongyeop Han, Chun Kee Chung, Maryam M Shanechi, Yisong Yue
BRIDGE-EEG is an efficient multi‑task EEG classification pipeline that leverages self‑supervised pretraining while dramatically reducing model size. It maps heterogeneous EEG recordings to a unified 62‑channel time‑frequency representation, pretrains an SE‑ResNet18 teacher with SimCLR, and distills it into smaller SE‑ResNet8 and SE‑ResNet4 students. The compact models achieve accuracy comparable to or better than larger foundation models on abnormality detection and emotion recognition, and they consume up to three times less energy on edge devices, enabling deployment on wearable hardware.
By Meghna Roy Chowdhury, Chengwei Zhou, Haotian Yu, Gourav Datta, Shreyas Sen
arXiv:2608. 07033v1 Announce Type: new Abstract: This work investigates whether Electroencephalograph (EEG) foundation models (EFMs) can be made faster and locally deployable without sacrificing accuracy.
By Lingwei Li, Yirong Kan, Peng Chen, Xu Cao, Zheng Chen, Yasuhiko Nakashima
arXiv:2608.24727v1 Announce Type: cross
Abstract: EEG foundation models pretrained via self-supervised learning promise transferable representations, but their generalization remains limited, especia...
By Meghal Dani, Stefanie Liebe
arXiv:2606. 02598v1 Announce Type: new Abstract: Accurate and generalizable estimation of cognitive workload from electroencephalography (EEG) is critical for human-centered and safety-critical systems.
By Jacob Wong, Sohan Singh, Prannaya Gupta, Jin Xing Ang, Kritika Johari, U-Xuan Tan