arXiv Computer Vision

An Affordable AI-Integrated Smart Cane for Multimodal Mobility Assistance of Visually Impaired Users

arXiv AI
Jul 3

VisionAId: An Offline-First Multimodal Android Assistant for People with Visual Impairment, Featuring Personalized Object Retrieval

arXiv:2607. 02371v1 Announce Type: cross Abstract: Over 285 million people worldwide live with a visual impairment, for whom everyday tasks such as avoiding obstacles, locating personal belongings, recognizing familiar faces, or handling cash remain persistent obstacles to personal autonomy.

By Cristian-Gabriel Florea, Stelian Sp\^inu
arXiv Machine Learning
Sep 25

Edge AI on Constrained Devices for Binary Sleep-Wake Classification in Dynamic Environments

This paper introduces an Edge AI system that classifies sleep and wake states on constrained devices using a multimodal pipeline on an ESP32‑S3 microcontroller. It fuses inertial head‑movement sensing with visual pose classification, running in parallel under FreeRTOS to meet real‑time constraints. The two‑stage detection achieves 96.5 % accuracy for motion‑based detection and 89 % for pose classification, proving robust binary sleep‑wake classification in mobile scenarios.

By Stefan Reitmann, Lena Oden
arXiv AI
Sep 10

RevalExo: A Functional Daily-Activity Benchmark for Inertial and Visual Locomotion Mode Recognition in Older Adults and Clinical Cohorts

RevalExo is a new benchmark for locomotion mode recognition that focuses on functional daily activities performed by older adults and clinical cohorts. It includes 27 participants from three groups—healthy older adults, stroke survivors, and older adults with probable sarcopenia—recorded with lower-body IMUs and, for a subset, synchronized egocentric video. The dataset offers 10.1 hours of frame‑level annotations across 11 locomotion modes, and the authors evaluate unimodal, multimodal, cross‑population, and cross‑modal recognition challenges, finding that sensor fusion improves performance but transitions and generalization remain difficult.

By Diwas Lamsal, Juha Carlon, Reinhard Claeys, Maxim Yudayev, Louis Flynn, Tom Verstraten, David Beckw\'ee, Eva Swinnen, Mihai B\^ace, Bart Vanrumste, Benjamin Filtjens
arXiv AI
Sep 21

Touvigation: Embodied Adaptive Object Acquisition for Blind and Low-Vision Users in Unfamiliar Indoor Environments

arXiv:2609.21828v1 Announce Type: cross Abstract: Blind and low-vision users often face challenges when locating and physically acquiring objects in unfamiliar indoor environments. Existing vision-la...

By George Xi Wang, Xiangyu Li, Shaoyue Wen, Jiaqian Hu, Junan Xie, Yupeng Wang, Ziyue Shi, Qijun Chen, Maaike Bouwmeester, Yuhua Jin, Jing Qian
arXiv Computer Vision
Aug 26

From Seeing to Acting: Smart Glasses as First-Person Intelligence Platforms

The paper surveys the evolution of smart glasses from simple capture devices to first‑person intelligence platforms that integrate human perception, context, and action. It introduces a unified framework that formalizes data flow, hardware capabilities, and seven foundational capabilities, and presents an L0‑L5 hierarchy for capture to embodied action. The study also maps nine application scenes, proposes a nine‑dimensional deployment framework, and outlines an evidence ladder for evaluation and trustworthiness.

By Jiangning Zhang, Haojun Chen, Yong Liu