ViTexSZ: Heterogeneous Vision-Text Knowledge Distillation for EEG Seizure Detection
Read the original on arXiv Computer Vision →The Flow has not summarised this story yet — read it at arXiv Computer Vision.
The Flow has not summarised this story yet — read it at arXiv Computer Vision.
arXiv:2607. 21384v1 Announce Type: new Abstract: Electroencephalography (EEG) models used for epilepsy are often limited to specific datasets and tasks.
arXiv:2607. 14314v1 Announce Type: new Abstract: Seizure diagnosis from EEG signals is a critical yet persistently challenging task, due to the complicated neural dynamics and the spurious connections in inter-channel modeling.
EEG-to-Report is a browser-based annotation and feature‑text framework that links routine EEG review with the creation of AI‑ready datasets. It ingests multi‑format EEG data, standardizes channels, and provides an interactive viewer with a multimodal annotation layer that combines typed text and transcribed voice notes. For each annotated segment, a feature extraction engine computes standardized spectral, temporal, entropy, Hjorth, connectivity, and spike‑related descriptors, stored alongside clinical descriptions in a portable JSON schema, producing aligned feature‑text pairs for training multimodal EEG‑language models. The framework also includes an auto‑report module that uses an ensemble of convolutional networks and a large language model to draft clinical narratives for neurologist review, thereby streamlining annotation workflows and enabling editable draft reports.
ProCA: Progressive Contrastive Alignment for Robust EEG Visual Decoding introduces a model‑agnostic framework that adaptively aligns EEG signals with visual semantics. It replaces fixed visual or textual anchors with EEG‑aware class‑level contrastive supervision and employs structure‑consistent interpolation to preserve channel‑wise and temporal importance. Across multiple evaluation settings—including subject‑dependent, subject‑independent, strict cross‑subject transfer, and continual adaptation—ProCA delivers significant performance gains, achieving relative Top‑1 improvements ranging from 7.4% to 28.1%.
arXiv:2608. 11656v1 Announce Type: new Abstract: Recent advances in EEG foundation models have demonstrated the potential of large-scale pretraining to enable generalizable neural decoding across subjects, recording environments, and datasets.
The paper introduces TriDim, a reusable block that preserves the three EEG axes—channel, short‑term temporal, and long‑term temporal—by applying feed‑forward transformations and cross‑axis attention. Stacking these blocks yields TriDimEEG, a standalone EEG decoder that outperforms fifteen other models on eight datasets, achieving a 4.3% relative accuracy gain. Replacing Transformer blocks in existing EEG foundation models with TriDim blocks improves downstream accuracy by 7.4% on average while reducing parameters by 17.0% to 47.3%.