arXiv Machine Learning

Honest and Reliable Evaluation and Expert Equivalence Testing of Automated Neonatal Seizure Detection

arXiv Machine Learning
Sep 4

RobustSeiz: An Open-Source Framework for Benchmarking the Robustness of EEG Seizure Detection Models

RobustSeiz is an open‑source, model‑agnostic framework designed to benchmark the robustness of EEG seizure detection models under realistic clinical stressors. It standardizes four public scalp‑EEG corpora into BIDS‑EEG trees, applies controlled distribution shifts—including environmental, noise, and adversarial transforms—across predefined hyperparameter grids, and reports comprehensive performance metrics such as sensitivity, precision, F1, false positives per 24 h, onset timing, and predictive agreement. The framework offers a Dockerized GPU pipeline, experiment registry, and both full‑evaluation and research‑subset modes, and demonstrates its utility by evaluating a contemporary detector on TUSZ across the full shift grid.

By Mohammad Mohammadi, Alireza Zarei
arXiv AI
3d ago

NeuroAtlas: Benchmarking Foundation Models for Clinical EEG and Brain-Computer Interfaces

NeuroAtlas is the largest EEG benchmark to date, comprising 42 datasets and 260,000 hours of clinical EEG data across epilepsy, sleep medicine, brain age estimation, and brain‑computer interfaces. The study evaluates foundation models (FMs) for EEG against supervised baselines and generic time‑series FMs, finding that EEG‑specific FMs do not consistently outperform generic ones. It also demonstrates that standard machine‑learning metrics are inadequate for clinical relevance, advocating for task‑specific measures such as event‑level decision quality, hypnogram features, and brain‑age gap.

By Konstantinos Kontras, Trui Osselaer, Stylianos G. Mouslech, Angeliki-Ilektra Karaiskou, Guido Gagliardi, Thomas Strypsteen, Mohammad Hossein Badiei, Anku Rani, Maarten Vanmarcke, Miguel Bhagubai, Chanakya Ekbote, Jaedong Hwang, Christos Chatzichristos, Paul Pu Liang, Maarten De Vos
arXiv Machine Learning
Jun 2

EEG-FuseFormer: A Transformer-Driven Feature Fusion Framework for Seizure Onset Prediction

arXiv:2606. 02166v1 Announce Type: new Abstract: Epilepsy is one of the most common neurological disorders globally, characterized by recurring seizures and significantly impacting the quality of life.

By Vigneshwar Hariharan (National University of Singapore), Chithra Reghuvaran (University College Dublin), Arlene John (University of Twente), Nhat Pham (Cardiff University), Omer Rana (Cardiff University), Deepu John (University College Dublin), Ganesh Neelakanta Iyer (National University of Singapore)
arXiv AI
Aug 14

Personalized Scorer Modeling: A Learning-Based Framework for Deriving Robust Sleep Stage Labels from Multiple Experts

arXiv:2608. 12446v1 Announce Type: cross Abstract: Sleep stage classification is important for the diagnosis and management of sleep disorders, yet most automatic staging studies evaluate models against a single reference hypnogram despite known inter-scorer variability.

By Seyyed Ali Hoseini, Javad Baseri, Hamid Saadatfar, Edris Hoseini Gol, AmirHossein Eshghi
arXiv AI
Jul 14

DiffEEG: A Self-Supervised Denoising Diffusion Model for Learning EEG Generic Representations

arXiv:2607. 11578v1 Announce Type: cross Abstract: Deep learning for EEG-based seizure detection faces critical challenges: severe annotation scarcity and extreme class imbalance, where ictal events comprise less than 10\% of clinical recordings.

By Abdulkader Helwan, Lina Abou-Abbas, Hussein El Amouri, Belkacem Chikhaoui, Khadidja Henni
arXiv Machine Learning
Jun 9

QDSP: An Interpretable Structured Learning Framework for Predicting Death or Cerebral Palsy in Very Low Birth Weight Infants

arXiv:2606. 07606v1 Announce Type: new Abstract: Very low birth weight infants (VLBWI) are at high risk of mortality and severe neurodevelopmental impairment, including cerebral palsy, yet reliable discharge-time prognostic stratification remains challenging in high-dimensional and data-limited clinical settings.

By Ling Wang, Xiaolong Li, Hui Zhou, Jing Shi, Fuhao Zhang, Dapeng Chen, Nan Mu
arXiv AI
Jun 2

CLSP-REQA: A Real-Time Quality-Aware Closed-Loop Seizure Prediction Framework with Mamba-BiLSTM and Confidence-Gated Intervention

arXiv:2606. 00074v1 Announce Type: cross Abstract: Reliable seizure prediction is a prerequisite for closed-loop neurostimulation therapy, yet existing methods rarely account for the variability in EEG signal quality encountered in real-world deployment, and the overwhelming majority adopt non-strict evaluation protocols that overestimate generalisation performance.

By Mufeng Chen, Qi Wu, Bingchao Huang, Xiwen Lai, Zekai Chen, Xinge Ouyang, Quansheng Ren
arXiv Machine Learning
Jul 28

Beyond Local Inspection: Global, Guideline-Grounded Evaluation of Post-hoc XAI Methods for ECG Classification

arXiv:2607. 24035v1 Announce Type: cross Abstract: Explainable AI (XAI) is used to assess whether artificial intelligence models rely on meaningful patterns, yet explanations that appear plausible for individual predictions may systematically misrepresent model behavior.

By Nils Gumpfer, Michael Guckert, Samuel Sossalla, Birgit A{\ss}mus, Jennifer Hannig