arXiv AI

Deep Learning-Enhanced Real-Time Wi-Fi Sensing Through Single Transceiver Pair

The paper presents a deep learning‑enhanced Wi‑Fi sensing system that uses only a single transceiver pair to achieve real‑time human pose estimation and localization. By leveraging prior information and temporal correlation as side information, the system reduces estimation error under hardware constraints. Experimental results show an average pose error of 0.2189 m and a localization error of 0.6124 m while running at 42 fps on commodity hardware.

arXiv Machine Learning
Sep 24

Untangling the Geometry and Speed for RF Sensing Spectrograms

The paper introduces a physics‑informed autoencoder that separates reflector speed from sensing geometry in WiFi spectrograms, enabling joint recovery of speed, geometry factor, relative amplitude, and ridge width for each Doppler ridge. It builds a compact parametric representation of spectrograms validated on a large human‑activity dataset and employs a synthetic‑to‑real training framework to avoid real‑data collection. Experiments on synthetic and 31 real WiFi scenarios show the method outperforms existing baselines in accurately extracting speed and geometry information.

By Mert Torun, Darius Cuenca, Yasamin Mostofi
arXiv Machine Learning
Jul 7

Seeing Through WiFi: Lightweight Human Pose Estimation with Dynamic Kernel Attention

arXiv:2607. 03196v1 Announce Type: cross Abstract: WiFi-based human pose estimation (HPE) enables the detection and interpretation of human body positions and movements without the need for wearable devices while preserving individual privacy concerns.

By Toan D. Gian, Van-Dinh Nguyen, Vo Phi Son, Nhan Thanh Nguyen, Dinh Thai Hoang, Diep N. Nguyen, Nguyen Cong Luong, Symeon Chatzinotas
arXiv AI
Sep 7

A Deep Generative Model for Synthesizing Labeled Wireless Signals

The paper introduces Inter-Instance Generative Adversarial Networks (IIns‑GAN), a deep learning approach for synthesizing realistic labeled wireless signals. Unlike traditional environmental‑model based methods, IIns‑GAN adapts to various scenarios and produces signals that closely match the physical characteristics of real measurements. Experiments on public Ultra‑Wideband datasets show that the generated signals improve model training for tasks such as distance estimation and environment identification.

By Yuxiao Li, Keke Hu, Santiago Mazuelas, Yuan Shen
arXiv Machine Learning
Sep 22

WiNeRF: Measurement Constrained Radiance Fields for Actionable Wireless Channel Modeling

WiNeRF is a neural field framework that learns a spatially continuous, complex-valued wireless channel representation from sparse channel state information collected by commodity WiFi devices. It incorporates system constraints such as antenna geometry, limited spatial resolution, and phase uncertainty through a 3D conical wave sampling model, a multi-resolution implicit scene representation, and a differentiable optimization framework. In diverse indoor environments with non‑line‑of‑sight regions, WiNeRF achieves a median prediction SNR of 5.3 dB, outperforming prior neural baselines by 4.9 dB on average, and produces a task‑agnostic channel representation that can be reused in standard signal‑processing pipelines without hardware or protocol changes.

By Saif Ur Rahman, Rafid Umayer Murshed, Anton Dmitriev, Cagri Tanriover, Rahul C. Shah, Elah\'e Soltanaghai
arXiv Computer Vision
Sep 22

CIG-MAE: Cross-Modal Information-Guided Masked Autoencoder for Self-Supervised WiFi Sensing

CIG-MAE is a self‑supervised framework for WiFi‑based human action recognition that uses a cross‑modal masked autoencoder to reconstruct both amplitude and phase of Channel State Information. It introduces an adaptive, information‑guided masking strategy that focuses on high‑density time‑frequency regions and employs a Barlow Twins regularizer to align cross‑modal representations without negative samples. Experiments on three public datasets show that CIG‑MAE outperforms state‑of‑the‑art SSL methods and even surpasses a fully supervised baseline, highlighting its data efficiency, robustness, and generalization.

By Gang Liu, Yanling Hao, Yixuan Zou