arXiv:2608.28932v1 Announce Type: new
Abstract: Voice products increasingly need affective cues that are present in speech but absent from transcripts. We introduce VocalAffectBench, a public, test-o...
By Models Luc Debaupte, Tyler Baumgartner, Brandon Tai, Candice Fan, Bill Wang, Yi Zhong
The paper evaluates deep learning models for electrocardiogram‑based emotion recognition, focusing on generalization across datasets rather than dataset‑specific performance. It introduces two open‑source tools—ARRC for standardized benchmarking and ARDT for inter‑dataset training—to merge three public AER datasets (CUADS, ASCERTAIN, DREAMER) into a more variable benchmark. Using these tools, the authors compare three prominent deep learning architectures and two CNN baselines with hyperparameter tuning and 10‑fold cross‑validation, revealing trade‑offs between accuracy and model complexity and providing a reproducible benchmark for future research.
By Timothy C Sweeney-Fanelli, Ajan Ahmed, Masudul Imtiaz
arXiv:2606. 10278v1 Announce Type: cross Abstract: Speech Emotion Recognition (SER) aims to identify a speaker's emotional state from audio signals.
By Youcef Soufiane Gheffari, Samiya Silarbi
arXiv:2606. 24941v2 Announce Type: replace-cross Abstract: Reviewing recorded interviews for affective cues such as composure and agitation is slow and subjective, and cloud services that could automate the task require sensitive audio to leave the device.
By Wai Laam Mak, Isibor Kennedy Ihianle, Pedro Machado
arXiv:2606. 03359v1 Announce Type: cross Abstract: Speech emotion recognition is an important component of modern human-computer interaction systems.
By Daniil Krasnoproshin, Maxim Vashkevich
arXiv:2609.39453v1 Announce Type: cross
Abstract: Speech emotion recognition (SER) is the task of assigning emotion labels to utterances. Early systems relied on acoustic features, whereas recent app...
By Hezhao Zhang, Thomas Hain