arXiv:2608.28967v1 Announce Type: cross
Abstract: Electroencephalography (EEG)-based robotic control is commonly formulated as a direct classification problem, in which electrical neural signals are...
By Alexandr Plashchinsky
arXiv:2607. 00547v1 Announce Type: cross Abstract: Existing egocentric benchmarks have primarily constructed the egocentric setting from first-person-view data, which makes it difficult to evaluate egocentric perspective itself in isolation.
By Jihyeok Jung (KAIST AI), Jeewu Lee (Sogang University), Sanghyeop Kim (Sogang University), Chanhee Han (Ministry of Science and ICT), Seong Joon Oh (KAIST AI)
arXiv:2607. 02542v1 Announce Type: new Abstract: General-purpose embodied agents must understand multimodal instructions, anticipate how their environment will evolve, and produce precise control actions over extended horizons.
By Yuan Zhang, Jingfei Ni, Guanchen Lu, Shiqi Zhang, Qingshan Xu, Chi Liu, Xin Nie, Wenjie Xu, Lin Gao, Zhiyuan Cheng, Mingxin Zhou, Jiajia Wu, Diyuan Liu, Jia Pan, Chao Ji
arXiv:2511. 17581v3 Announce Type: replace Abstract: Modeling the cognitive and experiential factors of human navigation is central to deepening our understanding of human-environment interaction and to enabling safe social navigation and effective assistive wayfinding.
By Zhiwen Qiu, Ziang Liu, Wenqian Niu, Tapomayukh Bhattacharjee, Saleh Kalantari
arXiv:2609.39378v1 Announce Type: new
Abstract: Real-world embodied tasks, from everyday activities to professional procedures, require agents to act under physical constraints while tracking evolvin...
By Shulin Tian, Junsu Kim, Shuai Liu, Hao Li, Yujiao Shen, Sihan Li, Zhe Yang, Yeongon Kim, Feiyu Li, Jialin Wu, Yichi Zhang, Wenhui Wang, Runmao Yao, Yuhao Dong, Zhaoxi Chen, Fangzhou Hong, Antonino Furnari, Jingkang Yang, Hongyuan Zhu, Ziwei Liu
The paper introduces ‘One Model for All’, a universal pre‑training framework that tackles EEG‑based emotion recognition across diverse datasets and paradigms. It decouples learning into a univariate self‑supervised contrastive pre‑training stage using a Unified Channel Schema, followed by a multivariate fine‑tuning stage that employs an Adaptive Resampling Transformer and a Graph Attention Network to model spatio‑temporal dependencies. Experiments demonstrate state‑of‑the‑art performance on within‑subject benchmarks (SEED 99.27%, DEAP 93.69%, DREAMER 93.93%) and superior cross‑dataset transfer, with ablation studies highlighting the critical role of the GAT module.
By Xiang Li, You Li, Yazhou Zhang
The paper introduces a redundancy-aware fusion framework for EgoExo proficiency estimation, which integrates fine-grained motion cues from egocentric views with spatial context from exocentric views. It identifies multiview redundancy and overfitting as key challenges and proposes two modules—AdaMVS for adaptive view selection and VIB-GB for compressing redundant signals—to address them. Experiments on EgoExo-4D and EgoExo-Fitness show that the method learns to select informative views and fuse them effectively, achieving state‑of‑the‑art results.
By Xu Dong, Wanqing Li, Anthony Adeyemi-Ejeye, Andrew Gilbert
arXiv:2607. 24126v1 Announce Type: cross Abstract: Brain-machine interfaces provide a link between neural activity and external devices, enabling restoration of motor function and advancing human-machine interaction using non-invasive electroencephalography (EEG).
By Sankalp Sunil Turankar, Yogesh Kumar Meena
NextMe-800 is an approximately 800‑hour first‑person video dataset collected from a single volunteer over 126 days, featuring 1 Hz images, gaze, and audio. The data are captioned at five hierarchical abstraction levels—from atomic actions to major activities—enabling personalized action anticipation as an open‑vocabulary K‑step sequence prediction task. The authors also introduce NextAct, a 1,500‑point benchmark that combines NextMe‑800 with the multi‑person EgoLife dataset, and evaluate models using an embedding‑based soft edit distance to assess how well personal behavior can be anticipated across abstraction levels and prediction horizons.
By Zhaoxu Meng, Yiming Sun, Mingyuan Gao, Jiachang Zhang, Zhuhan Dai, Yipeng Du, Zheng Lian, Jian-Qiao Zhu
arXiv:2608. 15999v1 Announce Type: new Abstract: Automatic emotion assessment can benefit from combining neural and behavioral signals, but many multimodal approaches rely on separate, modality-specific feature-extraction pipelines before fusion.
By Stefanos Gkikas, Eric Nichols, Christian Arzate Cruz, Randy Gomez
arXiv:2603. 16970v2 Announce Type: replace-cross Abstract: Multimodal egocentric activity recognition integrates visual and inertial cues for robust first-person behavior understanding.
By Hyejeong Im, Wonseon Lim, Dae-Won Kim
arXiv:2509. 25667v3 Announce Type: replace-cross Abstract: This paper presents an Artificial Intelligence (AI) integrated approach to Brain-Computer Interface (BCI)-based wheelchair development, utilizing a motor imagery right-left-hand movement mechanism for control.
By Bipul Thapa, Biplov Paneru, Bishwash Paneru, Khem Narayan Poudyal