arXiv:2608. 01383v1 Announce Type: new Abstract: Masked prediction learns representations by fitting a schedule-weighted family of conditional laws, but it remains unclear when near-optimal conditional prediction pins down the underlying joint law.
By Yichao Cai, Javen Qinfeng Shi
arXiv:2606. 11267v1 Announce Type: new Abstract: Data leakage -- contamination of a model with information unavailable at baseline -- is the dominant reproducibility failure in machine-learning-based science, yet detection tools require training code, external data, or domain expertise.
By Laurence A. Jacobs
arXiv:2608. 05608v1 Announce Type: cross Abstract: Multimodal classification typically assumes all modalities are available, yet real-world inputs are often incomplete.
By Yunping Shi, En Yu, Kairui Guo, Jie Lu
arXiv:2608. 07183v1 Announce Type: new Abstract: Multimodal fusion architectures typically assume all modalities are available at inference, yet sensor failures, acquisition variability, and cost constraints routinely produce incomplete observations.
By Alireza Moayedikia
arXiv:2608.23960v1 Announce Type: cross
Abstract: Missing labels are usually regarded as a source of information loss in classification. We study a semi-supervised setting in which the probability of...
By You-Gan Wang, Jinran Wu, Geoffrey J. McLachlan
arXiv:2607. 27289v1 Announce Type: new Abstract: The promise of multimodal fusion lies in combining complementary sources of evidence, yet more evidence does not always yield a better prediction.
By Yu Chang, Anzhe Cheng, Chenwei Wu, Zhuoran Wang, Jiahao Chen, Tamoghna Chattopadhyay, Sophia I. Thomopoulos, Paul M. Thompson, Liyue Shen, Paul Bogdan
arXiv:2608. 09240v1 Announce Type: cross Abstract: Multimodal federated learning (FL) supports collaborative modeling in privacy-sensitive health-sensing and medical settings, but realistic deployments often exhibit dual-axis modality missingness: clients have different modality sets, and individual samples may contain only subsets of the modalities available locally.
By Adiba Orzikulova, Jaehyun Kwak, Jaemin Shin, Yunqi Guo, Xiaomin Ouyang, Guoliang Xing, Steven Euijong Whang, Sung-Ju Lee
arXiv:2609.23937v1 Announce Type: cross
Abstract: Robust linear fits can resist response contamination yet remain too dense or unstable for useful global explanations. We propose penalized distillati...
By Wooyoung Shin, Seunghwan Park
The paper studies how coarsening a calibrated probability vector—by converting it to a hard label such as an argmax or a confidence threshold—affects the estimation of a treatment effect vector τ in a partially linear regression setting. It shows that the plug‑in estimator converges to a distorted version Δτ, where the distortion operator ΔΔ depends on the regression of the discarded part of the score on the retained part. The authors derive how this distortion drives coverage loss of Wald confidence intervals, provide estimable formulas for the bias and coverage, and demonstrate severe loss in simulations and real‑data audits.
By Marcell T. Kurbucz
arXiv:2603.05575v2 Announce Type: replace-cross
Abstract: We study prediction-powered conditional inference in the setting where labeled data are scarce, unlabeled covariates are abundant, and a blac...
By Yang Sui, Jin Zhou, Hua Zhou, Xiaowu Dai
arXiv:2507. 00260v3 Announce Type: replace-cross Abstract: When predictors are statistically dependent, the appropriate definition of feature importance depends on the operational goal.
By Jin-Hong Du, Kathryn Roeder, Larry Wasserman
arXiv:2602. 04408v3 Announce Type: replace Abstract: We study the Pareto frontier (optimal trade-off) between utility and separation, a fairness criterion requiring predictive independence from sensitive attributes conditional on the true outcome.
By Shizhou Xu