arXiv Machine Learning

Scalable Cross-Attention Transformer for Cooperative Multi-AP OFDM Uplink Reception

arXiv:2602. 04728v3 Announce Type: replace-cross Abstract: We propose a cross-attention Transformer for joint decoding of uplink OFDM signals received by multiple coordinated access points.

arXiv AI
Jul 21

Hybrid Mamba-Attention Neural Architecture for Channel Estimation

arXiv:2601. 17108v2 Announce Type: replace-cross Abstract: This paper proposes a hybrid Mamba-attention neural architecture to achieve improved channel estimation for orthogonal frequency-division multiplexing (OFDM) waveforms, particularly for configurations with a large number of subcarriers.

By Dianxin Luan, Chengsi Liang, Jie Huang, Zheng Lin, Kaitao Meng, John Thompson, Cheng-Xiang Wang
arXiv Machine Learning
Aug 7

EqDeepRx: Learning a Scalable and Interference Mitigating MIMO Receiver

arXiv:2602. 11834v2 Announce Type: replace-cross Abstract: While machine learning (ML)-based receiver algorithms have received a great deal of attention in the recent literature, they often suffer from poor scaling with increasing spatial multiplexing order and lack of explainability and generalization.

By Mikko Honkala, Dani Korpi, Elias Raninen, Janne M. J. Huttunen
arXiv Machine Learning
Sep 22

WiNeRF: Measurement Constrained Radiance Fields for Actionable Wireless Channel Modeling

WiNeRF is a neural field framework that learns a spatially continuous, complex-valued wireless channel representation from sparse channel state information collected by commodity WiFi devices. It incorporates system constraints such as antenna geometry, limited spatial resolution, and phase uncertainty through a 3D conical wave sampling model, a multi-resolution implicit scene representation, and a differentiable optimization framework. In diverse indoor environments with non‑line‑of‑sight regions, WiNeRF achieves a median prediction SNR of 5.3 dB, outperforming prior neural baselines by 4.9 dB on average, and produces a task‑agnostic channel representation that can be reused in standard signal‑processing pipelines without hardware or protocol changes.

By Saif Ur Rahman, Rafid Umayer Murshed, Anton Dmitriev, Cagri Tanriover, Rahul C. Shah, Elah\'e Soltanaghai
arXiv Computer Vision
Sep 22

CIG-MAE: Cross-Modal Information-Guided Masked Autoencoder for Self-Supervised WiFi Sensing

CIG-MAE is a self‑supervised framework for WiFi‑based human action recognition that uses a cross‑modal masked autoencoder to reconstruct both amplitude and phase of Channel State Information. It introduces an adaptive, information‑guided masking strategy that focuses on high‑density time‑frequency regions and employs a Barlow Twins regularizer to align cross‑modal representations without negative samples. Experiments on three public datasets show that CIG‑MAE outperforms state‑of‑the‑art SSL methods and even surpasses a fully supervised baseline, highlighting its data efficiency, robustness, and generalization.

By Gang Liu, Yanling Hao, Yixuan Zou