The paper introduces Adaptive Anisotropic Attention (AAA), a method that splits self‑attention into temporal and spatial paths for axis‑structured signals like EEG. A learned gate combines the two paths for each token, and the resulting AXON model outperforms dense attention baselines on six EEG tasks and shows transfer to audio spectrograms. The study demonstrates that aligning attention with the natural axes of structured data provides a beneficial inductive bias.
By Mahir Jain, Parshva Runwal, Aditya Ray Mishra, Arvasu Kulkarni, Sandeep Singh, Siddharth Panwar
ProCA: Progressive Contrastive Alignment for Robust EEG Visual Decoding introduces a model‑agnostic framework that adaptively aligns EEG signals with visual semantics. It replaces fixed visual or textual anchors with EEG‑aware class‑level contrastive supervision and employs structure‑consistent interpolation to preserve channel‑wise and temporal importance. Across multiple evaluation settings—including subject‑dependent, subject‑independent, strict cross‑subject transfer, and continual adaptation—ProCA delivers significant performance gains, achieving relative Top‑1 improvements ranging from 7.4% to 28.1%.
By Kanglei Zhou, Chunyan Lan, Dongyang Li, Jun Zhu, Liyuan Wang
arXiv:2608. 02070v1 Announce Type: cross Abstract: Brain-computer interfaces (BCIs) have been widely used in motor rehabilitation, disease diagnosis, and other neural engineering scenarios.
By Zhu Chen, Dingkun Liu, Yuheng Chen, Dongrui Wu
arXiv:2608. 02070v2 Announce Type: replace-cross Abstract: Brain-computer interfaces (BCIs) have been widely used in motor rehabilitation, disease diagnosis, and other neural engineering scenarios.
By Zhu Chen, Dingkun Liu, Yuheng Chen, Dongrui Wu
arXiv:2605.18251v2 Announce Type: replace-cross
Abstract: Self-initiated attention shifts play a critical role in voluntary behavior but are difficult to study due to the absence of explicit temporal...
By Yuwen Zeng, Dengzhe Hou, Zhang Zhang, Sai Sun, Yongsong Huang, Chia-huei Tseng, Satoshi Shioiri
arXiv:2506.20354v3 Announce Type: replace-cross
Abstract: Learning from multi-variate time-series with heterogeneous channel configurations remains a fundamental challenge for deep neural networks, p...
By Francesco Carzaniga, Michael Hersche, Abu Sebastian, Kaspar Schindler, Abbas Rahimi
The paper introduces AFOR, a tensor‑wise adaptive optimizer for EEG decoding that replaces the fixed second‑moment decay coefficient used in Adam/AdamW with a dynamic coefficient estimated online from local gradient state. AFOR combines a Residual‑Alignment Signal Scorer (RASS) to assess gradient quality and an Adaptive Forgetting Controller (AFC) to map this score to a bounded per‑step decay coefficient, with cumulative‑product initialization correction for consistency. In cross‑subject experiments on three EEG benchmarks, AFOR outperforms standard optimizers, improving mean test accuracy by 3.00%, 2.07%, and 4.38% over Adam.
By Hongyu Zhu, Lin Chen, Jing Chen, Yuting Zhou, Mingsheng Shang
arXiv:2601. 17883v3 Announce Type: replace Abstract: Electroencephalography (EEG) foundation models (FMs) have recently emerged as a promising paradigm for brain-computer interfaces, aiming to learn transferable neural representations from large-scale heterogeneous recordings.
By Dingkun Liu, Yuheng Chen, Zhu Chen, Zhenyao Cui, Yaozhi Wen, Jiayu An, Jingwei Luo, Dongrui Wu
arXiv:2603. 19100v2 Announce Type: replace Abstract: Electroencephalography (EEG) enables non-invasive monitoring of brain activity across clinical and neurotechnology applications, yet building foundation models for EEG remains challenging due to differing electrode topologies and computational scalability, as Transformer architectures incur quadratic sequence complexity.
By Dana\'e Broustail, Anna Tegon, Thorir Mar Ingolfsson, Yawei Li, Luca Benini
Robust and accurate neural decoders are integral to neurotechnologies such as brain-computer interfaces and closed-loop experiments. Recent work has shown that tokenizing neural data at the spike level facilitates multi-session pretraining and delivers state-of-the-art decoding performance.
Electroencephalography (EEG) provides a non-invasive window into dynamic brain activity, yet modeling long-horizon EEG sequences remains challenging due to their high temporal complexity, substantial...
arXiv:2606. 01923v1 Announce Type: cross Abstract: Large Language Models (LLMs) frequently exhibit "contextual disregard" when faced with input evidence that conflicts with their internal parametric memory, leading to persistent factual hallucinations.
By Mingkuan Zhao, Yide Gao, Wentao Hu, Suquan Chen, Tianchen Huang, Zhenhua An, Zetao Chang, Xiayu Sun, Yuheng Min