arXiv AI
Sep 4

FedPS: Federated Preprocessing for structured data via aggregated Statistics

FedPS is a federated preprocessing framework that uses aggregated statistics to address missing values, inconsistent formats, and heterogeneous feature scales in structured data. It employs data-sketching techniques to summarize local datasets efficiently, enabling federated algorithms for feature scaling, encoding, discretization, and missing-value imputation. The framework also extends preprocessing-related models, such as Bayesian Linear Regression, to both horizontal and vertical federated learning settings, offering communication‑efficient and consistent pipelines for practical deployments.

By Xuefeng Xu, Graham Cormode
arXiv Machine Learning
Jun 19

Variational Consensus Monte Carlo for Bayesian Mixture

arXiv:2606. 19643v1 Announce Type: cross Abstract: Motivated by the privacy, sensitivity and sharing limitations of health data, we present a comprehensive pipeline for inference of Bayesian mixture models within a federated learning setting, i.

By Julie Fendler, Francesca L. Crowe, Tom Marshall, Sylvia Richardson, Paul D. W. Kirk
Hugging Face Trending Papers
Jun 1

IntraShuffler: A Privacy Preserving Framework for Heterogeneous DP Federated Learning

Heterogeneous Differential Privacy (HDP) in Federated Learning (FL) allows clients to select individual privacy budgets ($\varepsilon_i$) according to institutional policies and data sensitivity. In practice, many HDP-FL systems employ $\varepsilon$-aware server aggregation to improve model utility by re-weighting client updates according to their declared privacy budgets.