arXiv:2609.39074v1 Announce Type: cross
Abstract: Federated learning (FL) on memory-constrained edge devices faces a dilemma: first-order (FO) optimization (i.e., backpropagation) demands substantial...
By Qiyuan Chen, Xian Wu, Yanan Ma, Xianhao Chen
arXiv:2608. 01426v1 Announce Type: new Abstract: Federated learning (FL) enables distributed optimization and learning across decentralized edge devices while preserving data privacy, but its performance is fundamentally constrained by heterogeneous data distributions, limited communication resources, and energy availability.
By Furkan Bagci, Busra Tegin, Mohammad Kazemi, Tolga M. Duman
arXiv:2408. 05886v5 Announce Type: replace Abstract: Heterogeneous system configurations of distributed clients connected to the central server (CS) via a time-varying wireless network pose significant challenges for popular distributed machine learning (ML) algorithms such as federated learning (FL).
By Ferdous Pervej, Minseok Choi, Andreas F. Molisch
arXiv:2606. 06687v1 Announce Type: new Abstract: We investigate cluster formation, involving the number and composition of clusters, in decentralized federated learning (FL) with heterogeneous machine learning (ML) optimizers.
By Su Wang, Mung Chiang, H. Vincent Poor
Ampere is a new split federated learning system that reduces both on‑device computation and device‑server communication while improving accuracy. It trains device and server blocks sequentially with local losses, eliminating gradient transfers, and uses a lightweight auxiliary network to consolidate activations into a single transfer. Experiments on CNNs and Transformers show up to 11.70 pp accuracy gains, 18.6× faster training, 911× less communication, and 14.5× less computation compared to state‑of‑the‑art SFL baselines.
By Zihan Zhang, Leon Wong, Blesson Varghese
DART-FL is a multitask federated learning framework designed for edge devices that must balance online inference and model training under limited resources. It dynamically allocates resources between inference and training based on current inference backlog and service capacity, then distributes remaining training capacity among tasks using a queue‑aware scheduler that adjusts loss weights. Experiments on image classification datasets with synthetic and real workloads show that DART‑FL adapts to bursty inference demand, improving accuracy for high‑demand tasks while preserving overall multitask performance.
By Yiming Xie, Pinrui Yu, Geng Yuan, Xue Lin, Ningfang Mi