arXiv AI

A Multi-Scale Temporal Framework with Dynamic Fusion for EEG-Based Emotion Recognition

arXiv:2608. 09088v1 Announce Type: new Abstract: Mixed emotions represent a clinically relevant but still underexplored target for automatic emotion recognition.

arXiv Machine Learning
5d ago

Differential Attention Unlocks Complementary EEG and Speech Fusion for Emotion Recognition

The paper introduces EmoSpeechBrain, a multimodal emotion recognition framework that fuses EEG and speech signals. It employs differential attention in the EEG encoder to cancel shared noise and an attention-based gating adapter to align modalities and weight their contributions. Experiments on PME4 and EAV datasets show up to 12.9% accuracy improvement over other EEG encoders and surpass unimodal baselines by up to 23.1%.

By Philip H. Lee, Shreeram Suresh Chandra, John H. L. Hansen
arXiv Computer Vision
Aug 31

MSCGC-KAN: Multi-scale Causal Graph Convolution and KAN-inspired Analytic-basis Mapping for EEG Emotion Recognition

The paper introduces MSCGC-KAN, a new EEG emotion recognition approach that builds on a pre‑trained CBraMod backbone. It incorporates a structured task head featuring multi‑scale causal graph convolution and Kolmogorov–Arnold feature mapping to better capture multi‑scale emotional dynamics, inter‑channel connectivity, and nonlinear discriminative patterns. Experiments on FACED and SEED‑VII show significant performance gains over a linear baseline, achieving balanced accuracies of 60.66% and 33.27% respectively.

By Haoliang Gong, Qingshan She, Jiale Xu, Yunyuan Gao, Xugang Xi
arXiv AI
Jun 2

UF-AMA: A unified framework for cross-domain emotion recognition via adaptive multimodal alignment

arXiv:2606. 00170v1 Announce Type: cross Abstract: In recent years, emotion recognition based on physiological signals such as electroencephalogram (EEG) has gained considerable attention, as internal physiological data offer greater objectivity and reliability compared to external behavioral data like facial expressions.

By Zheng Wang, Shuo Wang, Junhong Wang
arXiv AI
Sep 7

Enhancing Multimodal Emotion Recognition via Multi-Feature Encoding and Attention-Based Fusion

The paper introduces a multimodal emotion recognition framework that combines audio and visual feature extraction with an attention-based fusion strategy. Audio features include Wav2Vec2 embeddings, MFCCs, and statistical acoustic descriptors, fused via a BiLSTM, while video features are extracted using a ResNet50-BiLSTM architecture. A multi-head attention mechanism fuses these modalities, and experiments on MELD and IEMOCAP show significant accuracy and robustness gains, especially in unbalanced data settings.

By Xu Lin, Ke Wang, Hui Kang, Xinying Wang
arXiv Machine Learning
1d ago

Fusion techniques of time frequency-based images to predict the outcome of rTMS depression therapy

The study introduces two fusion techniques—montage and blending—to combine EEG-derived time‑frequency images for predicting response to repetitive Transcranial Magnetic Stimulation (rTMS) in depression. Using a lightweight custom CNN, the authors evaluated the methods on two datasets (15 and 46 patients) under both segment‑level and subject‑disjoint cross‑validation. While segment‑level validation yielded high accuracies (up to 99.90 % on the primary dataset), subject‑disjoint validation showed poor performance, with AUC values ranging from 0.31 to 0.54 and a best subject‑level result of 0.874 ± 0.183 AUC.

By Wael Korani, Md Fahimul Kabir Chowdhury, Mohammed Aledhari, Reza Rostami, Reza Kazemi