arXiv Machine Learning

Gradient Boosted Mixed Models: Flexible Estimation of Mean and Variance Components for Clustered Data

arXiv:2511. 00217v2 Announce Type: replace-cross Abstract: We introduce Gradient Boosted Mixed Models (GBMixed), a framework which extends boosting to clustered data by jointly modeling the mean and variance components in a linear mixed model via likelihood-based gradients.

arXiv Machine Learning
Jul 7

Integrating Neural Encoders in Bayesian Generalized Linear Mixed Models for Multimodal Data

arXiv:2607. 04647v1 Announce Type: cross Abstract: Scalable Bayesian inference for generalized linear mixed models (GLMMs) provides uncertainty-aware analysis of correlated longitudinal data, but existing scalable approaches largely assume low-dimensional tabular predictors and do not directly accommodate high-dimensional modalities such as images and text.

By Yuankang Zhao, Youngsoo Baek, Felipe A. Medeiros, Samuel Berchuck, Matthew M. Engelhard
arXiv Machine Learning
Sep 16

Splitting the Difference: Interpretable Causal Forests for Treatment Effect Heterogeneity and Bias

The paper introduces a new algorithm that uses decision trees and random forests to estimate individual treatment effects while providing interpretability. It modifies the standard random forest splitting criterion by combining a heterogeneity-focused criterion with a bias-correction criterion, enabling the model to handle observational studies with varying treatment propensities without separately estimating propensity scores. The resulting tree structure directly reveals which features drive treatment effect differences, and simulation studies show the method matches or surpasses existing approaches in prediction accuracy while improving interpretability.

By Nicolas Alexander Ihlo, Merle Behr
arXiv Machine Learning
Jun 19

Variational Consensus Monte Carlo for Bayesian Mixture

arXiv:2606. 19643v1 Announce Type: cross Abstract: Motivated by the privacy, sensitivity and sharing limitations of health data, we present a comprehensive pipeline for inference of Bayesian mixture models within a federated learning setting, i.

By Julie Fendler, Francesca L. Crowe, Tom Marshall, Sylvia Richardson, Paul D. W. Kirk
arXiv Machine Learning
Sep 1

Prediction-Powered Conditional Inference

arXiv:2603.05575v2 Announce Type: replace-cross Abstract: We study prediction-powered conditional inference in the setting where labeled data are scarce, unlabeled covariates are abundant, and a blac...

By Yang Sui, Jin Zhou, Hua Zhou, Xiaowu Dai
arXiv Machine Learning
Sep 1

Neural ODE enhanced linear mixed effect models for estimating complex association patterns of time-varying covariates with the marker trajectory

The paper introduces Neural ODE-LMM, a hybrid model that integrates a Neural Ordinary Differential Equation into a linear mixed‑effects framework to flexibly learn time‑varying associations between exposures and outcomes. By encoding covariate trajectories into a continuous‑time latent state, the method preserves standard likelihood inference while capturing complex, potentially cumulative effects without pre‑specifying functional forms. Simulations demonstrate accurate recovery of instantaneous and cumulative effects, and application to the 3C cohort uncovers BMI and fasting glucose trajectories linked to cognitive decline.

By Zhe Aurore Li, Quentin Clairon, C\'ecilia Samieri, Rodolphe Thi\'ebaut, M\'elanie Prague, C\'ecile Proust-Lima