arXiv Machine Learning
Sep 4

RobustSeiz: An Open-Source Framework for Benchmarking the Robustness of EEG Seizure Detection Models

RobustSeiz is an open‑source, model‑agnostic framework designed to benchmark the robustness of EEG seizure detection models under realistic clinical stressors. It standardizes four public scalp‑EEG corpora into BIDS‑EEG trees, applies controlled distribution shifts—including environmental, noise, and adversarial transforms—across predefined hyperparameter grids, and reports comprehensive performance metrics such as sensitivity, precision, F1, false positives per 24 h, onset timing, and predictive agreement. The framework offers a Dockerized GPU pipeline, experiment registry, and both full‑evaluation and research‑subset modes, and demonstrates its utility by evaluating a contemporary detector on TUSZ across the full shift grid.

By Mohammad Mohammadi, Alireza Zarei
arXiv AI
3d ago

NeuroAtlas: Benchmarking Foundation Models for Clinical EEG and Brain-Computer Interfaces

NeuroAtlas is the largest EEG benchmark to date, comprising 42 datasets and 260,000 hours of clinical EEG data across epilepsy, sleep medicine, brain age estimation, and brain‑computer interfaces. The study evaluates foundation models (FMs) for EEG against supervised baselines and generic time‑series FMs, finding that EEG‑specific FMs do not consistently outperform generic ones. It also demonstrates that standard machine‑learning metrics are inadequate for clinical relevance, advocating for task‑specific measures such as event‑level decision quality, hypnogram features, and brain‑age gap.

By Konstantinos Kontras, Trui Osselaer, Stylianos G. Mouslech, Angeliki-Ilektra Karaiskou, Guido Gagliardi, Thomas Strypsteen, Mohammad Hossein Badiei, Anku Rani, Maarten Vanmarcke, Miguel Bhagubai, Chanakya Ekbote, Jaedong Hwang, Christos Chatzichristos, Paul Pu Liang, Maarten De Vos