arXiv Machine Learning

HALO: A Heterogeneity-Aware Language-Aligned IMU Foundation Model for Open-Set Human Activity Recognition

HALO is a heterogeneity‑aware, language‑aligned foundation model for inertial measurement unit (IMU) based human activity recognition. It uses a two‑stage training process: first, a self‑supervised encoder learns to handle diverse sensor configurations and natural‑language sensor descriptions; second, the encoder is aligned with text embeddings through synonym‑aware contrastive learning, enabling open‑set recognition via cosine similarity. Trained on ten public HAR datasets, HALO outperforms five state‑of‑the‑art baselines across eight metrics while using only ~35 M parameters, and improves zero‑shot open‑set accuracy by 13.7 percentage points over 87 training labels.

Hugging Face Trending Papers
Jun 9

Closing the Modality Gap in Zero-Shot HAR: Contrastive Training and Separability-Optimized Prototypes on IMU Data

Zero-shot learning (ZSL) for inertial measurement unit (IMU)-based human activity recognition (HAR) faces a central challenge: bridging the gap between sensor embeddings and semantic class representations. We systematically evaluate seven configurations combining three inference methods with two training pipelines on the PAMAP2 dataset, using 14 seen and 4 unseen activity classes with subjects 108 and 109 held out for testing.

arXiv AI
2d ago

Coverage-Aware Virtual IMU Augmentation for Low-Resource Human Activity Recognition

The paper introduces a coverage-aware virtual IMU augmentation framework for human activity recognition. It selects diverse and scarce data points in a learned sensor embedding space, generates virtual IMU samples as prompts, ranks them by proximity and label consistency, and incorporates them into training with reliability-based weights. Experiments on public benchmarks demonstrate consistent performance gains over existing baselines, with ablation studies confirming the framework’s effectiveness.

By Jiayuan Gao, Yingwei Zhang, Ziyao Tang, Yuejia Ma, Yuanzhe Chen, Shuchao Song, Boshi Tang
arXiv AI
Jun 4

Gravity-Aware Hierarchical Routing for Lightweight SensorLLM on Human Activity Recognition

arXiv:2606. 04019v1 Announce Type: cross Abstract: Recent studies on sensor-language alignment have shown that two-stage frameworks can improve the semantic modeling ability of wearable-sensor human activity recognition (HAR), where SensorLLM-style methods first perform motion-to-language alignment and then fine-tune the model for downstream tasks.

By Hao Li, Mingrui Zheng, Yasuyuki Tahara, Yuichi Sei
arXiv Computer Vision
Sep 10

3rd Place Solution to Human Motion Challenges in Real-World and Clinical Settings (MoCha) @ECCV2026: Language-Aligned Motion Representations for Domain-Generalizable UPDRS-Gait Severity Estimation

arXiv:2609.10187v1 Announce Type: new Abstract: In this work, we introduce language-aligned motion representations for domain-generalizable UPDRS-Gait severity estimation, aiming to learn semanticall...

By Soojie Kim, Muhammad Munsif, Minkyung Kim, Seungryul Baek