arXiv Computation and Language

MultiHuSE: A Multimodal Dataset for Humour Styles and Emotions

MultiHuSE is a multimodal dataset featuring 2,407 high‑definition videos of 50 diverse actors delivering 1,463 text samples in four psychological humour styles—affiliative, aggressive, self‑enhancing, and self‑deprecating—plus neutral content. Each text is performed by multiple actors, allowing analysis of expressive diversity, and a subset includes emotion annotations. Baseline experiments show that multimodal fusion improves humour style classification accuracy over unimodal approaches, especially for affiliative humour.

arXiv AI
Jul 17

Team RAS in 11th ABAW Competition: Multimodal Ambivalence Recognition Approach

arXiv:2607. 14702v1 Announce Type: cross Abstract: Automatic recognition of ambivalence and hesitancy is challenging because these states may be expressed through inconsistent linguistic, acoustic, facial, and contextual patterns, while top-performing systems often rely on computationally expensive ensembles.

By Elena Ryumina (St. Petersburg Federal Research Center of the Russian Academy of Sciences), Maxim Markitantov (St. Petersburg Federal Research Center of the Russian Academy of Sciences), Alexandr Axyonov (St. Petersburg Federal Research Center of the Russian Academy of Sciences), Fedor Shchetinin (HSE University, St. Petersburg, Russia), Timur Abdulkadirov (St. Petersburg Federal Research Center of the Russian Academy of Sciences), Dmitry Ryumin (St. Petersburg Federal Research Center of the Russian Academy of Sciences), Alexey Karpov (St. Petersburg Federal Research Center of the Russian Academy of Sciences)