The paper presents a low‑cost baseline for continuous emotion regression that relies on video–time priors when users watch familiar videos. It compares this baseline to a fusion of EEG and fNIRS signals, finding that the video–time prior alone achieves mean absolute errors within 0.05 and 0.32 of the fusion model in internal and external evaluations. Ablation studies show that video identity and within‑video time explain most of the performance, while EEG–fNIRS contributions are smaller and variable across participants and videos.
By Minghao Kong, Jiurun Chen, Ying Gao, Xiangbin Meng, Rongjie Wang
arXiv:2606. 06647v1 Announce Type: new Abstract: Objective.
By Jun-You Lin, Ying Choon Wu, Tzyy-Ping Jung
arXiv:2607. 24519v2 Announce Type: replace Abstract: Pretrained EEG foundation models are proposed for clinical decoding, but whether reported gains transfer across populations or survive negative controls is unclear.
By Marzieh Zare
arXiv:2508. 17742v3 Announce Type: replace-cross Abstract: Electroencephalography foundation models (EEG-FMs) have advanced brain signal analysis, but the lack of standardized evaluation benchmarks impedes model comparison and scientific progress.
By Wei Xiong, Jiangtong Li, Jie Li, Kun Zhu, Changjun Jiang
ZeroMAG is a zero‑shot multimodal adapter generation framework that extends frozen EEG foundation models to heterogeneous multimodal recordings using only unlabeled target data. It constructs a configuration‑invariant adapter and generates adapter weights in a latent space learned from source adapters, without target‑side optimization. Across six held‑out target datasets and three EFM backbones, ZeroMAG improves balanced accuracy by 7.22 percentage points over EEG‑only inference and 4.89 points over direct weight regression, approaching supervised multimodal adaptation.
By Yubo Wang, Jingying Ma, Xinliang Zhou, Yangxuan Zhou, Jiquan Wang, Sha Zhao, Yiyuan Yang, Yi Ding, Ziyu Jia, Chenyu Liu, Cuntai Guan
arXiv:2608. 15999v1 Announce Type: new Abstract: Automatic emotion assessment can benefit from combining neural and behavioral signals, but many multimodal approaches rely on separate, modality-specific feature-extraction pipelines before fusion.
By Stefanos Gkikas, Eric Nichols, Christian Arzate Cruz, Randy Gomez