arXiv Machine Learning

Assessing the Impacts of Imperfect Datasets on Client Selections in Federated Learning

arXiv:2608. 02250v1 Announce Type: new Abstract: Federated learning (FL) is a popular distributed learning framework where multiple clients perform local training and a server aggregates the locally updated models.

Hugging Face Trending Papers
Aug 10

FedTVD: Balancing Data Quality and Quantity for Robust Federated Learning

Federated Learning (FL) enables collaborative model training across distributed client devices while preserving data privacy. However, FL faces significant challenges due to data heterogeneity, particularly in terms of label distribution skewness and variations in dataset sizes, which can lead to biased model updates and hinder convergence.

arXiv Machine Learning
1d ago

Latent Information Sharing for Accelerating Federated Learning

The paper introduces a latent information sharing scheme for federated learning that mitigates client drift by sharing a small amount of hidden‑layer activations. The authors demonstrate both theoretically and empirically that this approach improves training efficiency while maintaining convergence guarantees and data privacy. Compared to existing methods such as FedProx, SCAFFOLD, FedPVR, FedProto, and SplitFed, the proposed method achieves higher model accuracy within a fixed round budget without adding significant communication overhead.

By Seungjun Lee, Ensieh Khazaei, Dimitrios Hatzinakos, Baturalp Buyukates, Sunwoo Lee
arXiv Machine Learning
Aug 27

Theoretically Principled Federated Learning for Balancing Privacy and Utility

The paper introduces a general learning framework that protects privacy in federated learning by distorting model parameters, enabling a trade‑off between privacy and utility. The algorithm supports arbitrary privacy measurements and delivers personalized utility‑privacy balances for each parameter, client, and communication round. The authors prove that the gap between their algorithm’s utility loss and the optimal loss is sub‑linear in iterations, provide a convergence rate, and demonstrate empirically that their method outperforms baselines under the same privacy budget.

By Xiaojin Zhang, Wenjie Li, Yiming Li, Wei Chen, Shutao Xia, Qiang Yang
arXiv Machine Learning
Sep 25

Federated Learning of AnDE Classifiers

arXiv:2609. 28695v1 Announce Type: new Abstract: This work presents a federated framework for training Averaged $n$-Dependence Estimators (AnDE) in distributed environments.

By Pablo Torrijos, Juan C. Alfaro, Jos\'e A. G\'amez, Jos\'e M. Puerta
arXiv AI
Aug 25

FedCC: Towards Addressing Label Distribution Skews in Distillation-Based Federated Learning

FedCC is a new algorithm for distillation-based federated learning that tackles label distribution skew by allowing clients to mark ambiguous samples as 'unknown' instead of forcing a potentially wrong classification. By adding this extra class and calibrating pseudo-labels on a public dataset, FedCC balances confidence across majority and minority classes. Experiments show that FedCC outperforms existing methods, achieving 67.3% accuracy even when each client has data from only one of ten classes, whereas baselines drop to near-random performance.

By Wenxuan Ye, Onur Ayan, Xueli An, Georg Carle
arXiv Machine Learning
Sep 3

Similarity-Aware Personalized Federated Learning in Heterogeneous Environments

The paper introduces SAPE-FL, a personalization framework for Federated Learning that anchors each client’s model to both a global model and a similarity-weighted peer-averaged model. By applying dynamic, client-specific regularization based on model and output similarity, SAPE-FL balances global knowledge transfer with peer collaboration, filtering out dissimilar clients. The authors provide theoretical convergence guarantees and demonstrate empirically that SAPE-FL outperforms state‑of‑the‑art methods in highly heterogeneous and low‑data scenarios.

By Arun Kumar A V, Sunil Gupta, Dang Ngyuen, Bao Duong, Dat Phan Trong