arXiv:2608.23794v1 Announce Type: new
Abstract: Mixture-of-Experts (MoE) scales language models by routing each input through a small set of independently parameterized experts. We show that copying...
By Elian Iluk, Gil Ben-Artzi
Mixture-of-Experts (MoE) scales language models by routing each input through a small set of independently parameterized experts. We show that copying this design into convolutional networks fails for...
arXiv:2606. 10277v1 Announce Type: new Abstract: Though wireless foundation models (WFMs) have shown strong potential in learning universal channel representations, their adaptation to various downstream tasks remains constrained by existing paradigms.
By Yuxuan Shi, Tingting Yang, Kangning Ma, Liwen Jing, Yuwei Wang, Mengfan Zheng, Li Sun
arXiv:2602. 11834v2 Announce Type: replace-cross Abstract: While machine learning (ML)-based receiver algorithms have received a great deal of attention in the recent literature, they often suffer from poor scaling with increasing spatial multiplexing order and lack of explainability and generalization.
By Mikko Honkala, Dani Korpi, Elias Raninen, Janne M. J. Huttunen
arXiv:2511. 08972v2 Announce Type: replace Abstract: Sparse Mixture-of-Experts (SMoE) models are scalable and computationally efficient, enabling large increases in model capacity with limited inference overhead.
By Duc Anh Nguyen, Huu Binh Ta, Nhuan Le Duc, Tan Minh Nguyen, Toan Tran
arXiv:2607. 11970v1 Announce Type: cross Abstract: We develop an enhanced in-context learning (ICL) framework to improve the performance of pilot-based beamforming in multi-user multiple-input single-output (MU-MISO) systems.
By Yubo Zhang, Xiaodong Wang
arXiv:2606. 11857v1 Announce Type: cross Abstract: Multi-channel mixed-SNR training improves out-of-distribution (OOD) generalisation of deep learning channel estimators for IEEE 802.
By Simbarashe Aldrin Ngorima, Albert Helberg, Marelie H. Davel
The paper introduces DRIFT, a lightweight framework for joint channel estimation and prediction in low Earth orbit non-terrestrial networks, aiming to reduce pilot overhead by using data-driven processing after the initial slot. DRIFT refines data-aided channel estimates and forecasts future channel responses with low computational cost, offering two variants based on convolutional and LSTM layers. Simulations show up to 12% spectral efficiency gain over conventional pilot-based systems, with under 200k multiply-accumulate operations suitable for on-board satellite implementation.
By Bruno De Filippo, Carla Amatetti, Alessandro Vanelli-Coralli
arXiv:2607. 16877v1 Announce Type: cross Abstract: The increasing complexity of next-generation wireless networks has driven the integration of artificial intelligence (AI) into wireless communications.
By Yangjing Wang, Ouya Wang, Shenglong Zhou, Geoffrey Ye Li
arXiv:2606. 06373v1 Announce Type: cross Abstract: Wireless foundation models have emerged as a promising alternative to building separate models for each wireless task.
By Ahmed Mohamed, Ahmed Aboulfotouh, Hatem Abou-Zeid
arXiv:2607. 13494v1 Announce Type: new Abstract: The development of smart transportation systems and the introduction of 6G wireless communication technologies have significantly changed vehicle network topologies.
By S. M. Abtahiul Alam, Niloy Das, Apurba Adhikary, Yu Qiao, Zhu Han, Choong Seon Hong
MetaNet is a support‑set controller that predicts, for each layer of a Mixture‑of‑Experts model, an expert‑retention threshold and a bounded routing bias while keeping the backbone, experts, and router frozen. On DeepSeek‑MoE‑16B‑Chat, MetaNet offers a tunable trade‑off between accuracy and expert activation: a conservative setting activates 3.61 experts on average (40% fewer than a fixed k=6) with comparable MMLU accuracy, whereas an aggressive setting activates only 2.28 experts (62% fewer) with a modest accuracy drop. The MMLU‑trained controller also transfers to C‑Eval, activating 2.90 experts on average (52% fewer than fixed k=6) at 0.386 accuracy.
By Rongfeng Wang, Shichao Weng, Zhiqiang Wang, Xinyu Liu, Yang Yi, Peilong Zhou, Hongwei Tang