arXiv:2606. 11857v1 Announce Type: cross Abstract: Multi-channel mixed-SNR training improves out-of-distribution (OOD) generalisation of deep learning channel estimators for IEEE 802.
By Simbarashe Aldrin Ngorima, Albert Helberg, Marelie H. Davel
Multi-channel mixed-SNR training improves out-of-distribution (OOD) generalisation of deep learning channel estimators for IEEE 802. 11p vehicular communications, yet the internal mechanism responsible for this remains unexplained.
arXiv:2602. 22277v2 Announce Type: replace Abstract: AI-native architectures are vital for 6G wireless communications.
By Abdul Karim Gizzini, Yahia Medjahdi
FedPGT introduces a progressive gradient transmission scheme for vehicular federated learning over time‑varying channels, where vehicles send high‑magnitude gradient entries according to instantaneous channel conditions. The authors derive a convergence bound showing diminishing returns governed by a power‑law decay, and formulate a stochastic optimization problem that is solved via a Lyapunov drift‑plus‑penalty approach with per‑slot surrogate variables. A low‑complexity resource allocation algorithm is proposed, and experiments on CIFAR‑10 and Argoverse demonstrate a 3.65% accuracy gain and a 12.66% reduction in displacement error compared to state‑of‑the‑art baselines.
By Jintao Yan, Tan Chen, Yuxuan Sun, Sheng Zhou, Zhisheng Niu
arXiv:2606. 24969v1 Announce Type: new Abstract: While the quadratic sequence-length bottleneck of transformers has fueled a resurgence in recurrent models, effectively capturing complex dynamics requires architectures that balance efficient training with highly expressive latent states.
By Klaus Schertler, Xiomara Runge, Andrea Ceni, David Kappel, Claudio Gallicchio
The paper introduces DRIFT, a lightweight framework for joint channel estimation and prediction in low Earth orbit non-terrestrial networks, aiming to reduce pilot overhead by using data-driven processing after the initial slot. DRIFT refines data-aided channel estimates and forecasts future channel responses with low computational cost, offering two variants based on convolutional and LSTM layers. Simulations show up to 12% spectral efficiency gain over conventional pilot-based systems, with under 200k multiply-accumulate operations suitable for on-board satellite implementation.
By Bruno De Filippo, Carla Amatetti, Alessandro Vanelli-Coralli
arXiv:2410. 11687v3 Announce Type: replace-cross Abstract: Linear recurrent networks (LRNNs) offer linear-time sequence modeling, but standard recurrent updates do not directly expose the supervised products needed for in-context gradient descent.
By Yudou Tian, Neeraj Mohan Sushma, Harshvardhan Mestha, Nicolo Colombo, David Kappel, Anand Subramoney
MambaCSP is a hybrid-attention state space model that replaces transformer-based backbones with a linear-time Mamba architecture for channel state prediction. By adding lightweight patch‑mixer attention layers, it captures long‑range dependencies while maintaining hardware efficiency. Experiments on MISO‑OFDM show 9‑12% higher accuracy, 3× faster throughput, 2.6× lower VRAM usage, and 2.9× faster inference compared to LLM‑based methods.
By Aladin Djuhera, Haris Gacanin, Holger Boche
The paper introduces XACT, a framework that learns sparse attribution masks over coefficients from any invertible time‑frequency transform, such as STFT, continuous wavelet transform, and discrete wavelet transform. XACT extends the virtual inspection layer approach to wavelet transforms, enabling Layer‑wise Relevance Propagation (LRP) to generate explanations in these representations. Experiments on synthetic and two real‑world datasets show that XACT produces precise, sparse, and structured explanations, often outperforming baseline methods in highlighting relevant features.
By Theresa Dahl Frehr, Francisco Pelayo, Lukas Raad, Alicia Garc\'ia Sanz, Thea Br\"usch, Tommy Sonne Alstr{\o}m
arXiv:2607. 28665v1 Announce Type: cross Abstract: Automated driving systems (ADSs) are becoming ubiquitous.
By Bidhya Shrestha, Christos Papadopoulos
The paper introduces Gated Recurrent Transformers, a depth‑sharing architecture that brackets a single shared core with fixed prelude and coda blocks and uses a lightweight projection and element‑wise update gate to modulate recurrent updates. This design allows functional specialization across recurrences while reducing memory footprint. Experiments show that, under equal FLOPs or parameter budgets, the recurrent model matches or surpasses deeper GPT‑2 Small baselines, achieving similar or better accuracy with fewer parameters and lower peak decoding memory.
By Amr Hegazy, Amr Alanwar, Mostafa Elhoushi
arXiv:2507. 09627v3 Announce Type: replace-cross Abstract: Next-generation wireless technologies such as 6G aim to meet demanding requirements such as ultra-high data rates, low latency, and enhanced connectivity.
By Muhammad Kamran Saeed, Ashfaq Khokhar, Shakil Ahmed