arXiv:2609.24086v1 Announce Type: cross
Abstract: Large-scale multipurpose cohort studies and biobanks often omit covariates needed for specific downstream analyses. We study target-population infere...
By Huali Zhao (School of Mathematics and Statistics, Huazhong University of Science and Technology), Ke Deng (Department of Statistics and Data Science, Tsinghua University)
arXiv:2605. 24212v2 Announce Type: replace-cross Abstract: Deploying clinical prediction models across healthcare systems often fails when key training covariates are unavailable at deployment and labeled outcomes are limited in the target domain.
By Siqi Li, Chuan Hong, Ziye Tian, Benjamin Sieu-Hon Leong, Koshi Nakagawa, Hideharu Tanaka, Sang Do Shin, Khuong Quoc Dai, Do Ngoc Son, Marcus Eng Hock Ong, Nan Liu, Molei Liu
arXiv:2502. 11331v4 Announce Type: replace-cross Abstract: The proliferation of data has sparked significant interest in leveraging findings from one study to estimate treatment effects in a different target population without direct outcome observations.
By Seok-Jin Kim, Hongjie Liu, Molei Liu, Kaizheng Wang
arXiv:2412. 18081v3 Announce Type: replace-cross Abstract: We study Heterogeneous Transfer Learning (HTL) for high-dimensional regression with differing feature sets.
By Jae Ho Chang, Massimiliano Russo, Subhadeep Paul
arXiv:2608. 20255v1 Announce Type: cross Abstract: This paper develops a general transfer learning framework for nonparametric regression with data consisting of multiple groups.
By Junpeng Ren, Carlos Misael Madrid Padilla, Yanzhen Chen, Oscar Hernan Madrid Padilla
arXiv:2504. 15388v3 Announce Type: replace-cross Abstract: In the context of multivariate nonparametric regression with missing covariates, we propose Pattern Embedded Neural Networks (PENNs), which can be applied in conjunction with any existing imputation technique.
By Tianyi Ma, Tengyao Wang, Richard J. Samworth
arXiv:2607. 14346v1 Announce Type: new Abstract: Policy learning methods are increasingly used to inform treatment allocation under budget constraints.
By Johnna Sundberg, Rayid Ghani, Eli Ben-Michael, Edward Kennedy
arXiv:2608. 15783v1 Announce Type: cross Abstract: In transfer-learning settings, a model derived from abundant surrogate labels may be deployed in a target population where gold-standard outcomes are unobserved.
By Longtian Shi, Molei Liu, Doudou Zhou
arXiv:2607. 07767v1 Announce Type: cross Abstract: Missing values undermine statistical inference and machine learning pipelines, yet most imputation methods rely on heuristics or restrictive parametric assumptions that ignore the joint data distribution.
By Andrea Basteri, Carlo Ciliberto, Alessandro Rudi
arXiv:2608. 00701v1 Announce Type: cross Abstract: Reweighting source samples to match a target covariate distribution is a standard response to distribution shift when generalizing evidence from one population to another.
By Ying Jin, Ying Jin, Dominik Rothenh\"ausler
arXiv:2609.40051v1 Announce Type: new
Abstract: Estimating causal effects from observational data is central to science and policy, but the effects are not identified when confounders are unmeasured....
By Yonghan Jung
arXiv:2507.23768v2 Announce Type: replace-cross
Abstract: Existing methods for transfer learning struggle to deal with situations where the source datasets are limited and not guaranteed to be well-a...
By Nathan Wycoff, Ali Arab, Lisa O. Singh