arXiv Machine Learning

Autoencoder Architectures for Athlete Performance Scoring from Wearable Telemetry

arXiv:2606. 28145v1 Announce Type: new Abstract: Wearable devices produce large, high dimensional training logs for everyday runners, and interpretation rather than data collection is now the limiting step.

arXiv AI
Sep 16

SOTER: A Generative Time-Series Foundation Model for Wearable Human Physiological Signals

SOTER is a generative foundation model designed for wearable physiological time‑series data. It integrates cross‑channel coupling, spectrum‑guided expert specialization, and continuous‑time latent evolution, using a spatial feature‑aware backbone, a PSD‑guided mixture‑of‑experts layer, and a neural controlled differential equation decoder. Trained on 226 billion time points from five public datasets, SOTER outperforms baselines in zero‑shot forecasting, classification, and imputation across six benchmarks, and remains robust to additive noise.

By Fangke Chen, Sirry Chen, Wei Chen, Zhongyu Wei
arXiv AI
Aug 26

Evaluating Deep Multivariate Imputation Models on Wearable Device Data

The paper introduces a new evaluation protocol for deep multivariate imputation models on wearable device data, addressing the issue of structured missingness where sensor features drop out together. Using a Garmin smartwatch dataset from an epilepsy patient, the authors generate realistic block-missing patterns from training data and show that matching the training protocol to this distribution reduces BRITS’ mean absolute error by 43%. They also extend BRITS with time‑of‑day encoding and compare it to linear interpolation and SAITS, finding that no single model dominates and that model rankings vary with evaluation design.

By Skye Goodman, Roussel Desmond Nzoyem, Leandro Junges, Peter Kissack, Yasser Qureshi, Amberly Brigden, Jeff Clark, Nawid Keshtmand
arXiv Machine Learning
Jul 31

ECG-InterpBench: Benchmarking the Interpretability of ECG Foundation Models with Matched-Scale Sparse Autoencoders

arXiv:2607. 27404v1 Announce Type: new Abstract: Existing benchmarks for electrocardiogram foundation models primarily evaluate downstream predictive performance, providing limited insight into whether their internal representations can be faithfully decomposed, clinically interpreted, or reproduced across independent analyses.

By Yixuan Duan, Wei Qiu