arXiv Statistics ML

Equivalence Between Nested Gibbs Measures and Log-Linear Combinations of Gibbs Measures

The paper investigates three operations on Gibbs probability measures: renormalization, normalized log-linear combination, and nesting (changing the reference measure). It shows that the measures produced by the second and third operations solve related optimization problems and that, for specific parameters, nesting one Gibbs measure into another is equivalent to log-linearly combining them. This equivalence has practical implications, such as enabling a one-shot federated learning system where clients’ locally trained Gibbs algorithms can be combined on a server to match the performance of a centrally trained Gibbs algorithm.

arXiv Machine Learning
Sep 25

Machine Unlearning for Gibbs Supervised Learning Algorithms

The paper introduces a method for exact unlearning of Gibbs supervised learning algorithms via a variational formulation based on empirical risk minimization with relative entropy regularization (ERM‑RER). By maximizing the expected empirical risk over the data to be removed while regularizing with relative entropy to the original algorithm, the resulting solution is a new Gibbs probability measure that matches the distribution of an algorithm retrained from scratch on the remaining data. The approach also provides a general framework for reweighting data points in ERM‑RER, allowing for up‑ or down‑weighting to control generalization error or other objectives.

By Yaiza Bermudez, Samir M. Perlaza, I\~naki Esnaola
arXiv Machine Learning
Jun 9

Dendrograms of Mixing Measures for Softmax-Gated Gaussian Mixture of Experts: Consistency Without Model Sweeps

arXiv:2510. 12744v2 Announce Type: replace-cross Abstract: We develop a unified statistical framework for softmax-gated Gaussian mixture of experts (SGMoE) that addresses three long-standing obstacles in parameter estimation and model selection: (i) non-identifiability of gating parameters up to common translations, (ii) intrinsic gate-expert interactions that induce coupled differential relations in the likelihood, and (iii) the tight numerator-denominator coupling in the softmax-induced conditional density.

By Do Tien Hai, Trung Nguyen Mai, TrungTin Nguyen, Nhat Ho, Binh T. Nguyen, Christopher Drovandi