arXiv Machine Learning By Yu Zhu, Jason Teng, Zehang Richard Li

Bayesian Federated Cause-of-Death Classification and Quantification Under Distribution Shift

Read the original on arXiv Machine Learning →

arXiv:2505. 02257v2 Announce Type: replace-cross Abstract: In regions lacking medically certified causes of death, verbal autopsy (VA) is a widely used tool to ascertain the cause of death through interviews with caregivers.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv Machine Learning.

arXiv Machine Learning
Jun 24

Federated Survival Analysis in Healthcare: A Multi-Model Evaluation on Cross-Institutional Heterogeneous Breast Cancer Data

arXiv:2606. 23871v1 Announce Type: new Abstract: Survival analysis is central to clinical decision-making, yet reliable time-to-event models require large, diverse cohorts that are rarely available at a single institution, while privacy regulations restrict the centralization of patient data.

By Natalia Moreno-Blasco, Anusha Ihalapathirana, Pekka Siirtola, Miguel Fernandez-de-Retana
arXiv Statistics ML
Aug 28

An Accurate and Single-Communication Federated Inference Algorithm

The paper introduces a federated inference algorithm that requires only a single communication round between participating centers and a coordinating server. By extending previous second‑order Taylor expansion methods to third‑order expansions, the algorithm more accurately approximates local log‑likelihood functions, especially when local sample sizes are small. Simulation studies based on real data show that this higher‑order approach improves inference accuracy while maintaining privacy, communication efficiency, and scalability for collaborative biomedical and epidemiological research.

By Laura Montagnani, Anthony CC Coolen, Marianne A Jonker
arXiv AI
Sep 4

FedPS: Federated Preprocessing for structured data via aggregated Statistics

FedPS is a federated preprocessing framework that uses aggregated statistics to address missing values, inconsistent formats, and heterogeneous feature scales in structured data. It employs data-sketching techniques to summarize local datasets efficiently, enabling federated algorithms for feature scaling, encoding, discretization, and missing-value imputation. The framework also extends preprocessing-related models, such as Bayesian Linear Regression, to both horizontal and vertical federated learning settings, offering communication‑efficient and consistent pipelines for practical deployments.

By Xuefeng Xu, Graham Cormode