arXiv AI By Yiyang Shen, Weiran Wang

No Data Wasted: A Semi-supervised Generative Model for Incomplete Multi-view Data Integration with Missing Labels

Read the original on arXiv AI →

The paper presents a semi‑supervised generative model for multi‑view learning that handles missing views and missing labels. It combines a likelihood‑based approach for unlabeled data with an information bottleneck (IB) framework for labeled data, incorporating modality‑specific information and cross‑view mutual information maximization to learn a shared latent space. Experiments show improved predictive and generative performance on complex datasets with limited labeled samples.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv Machine Learning
Aug 20

Pretraining Reusable Inference Across Views with Synthetic Task Priors

The paper introduces SIMPLE, a prior‑fitted multi‑view in‑context learner that learns a reusable, task‑conditioned inference procedure instead of a fixed fusion function. By generating synthetic task priors in embedding space, SIMPLE can handle diverse view configurations, class structures, and missingness patterns. Experiments on multi‑view and multi‑omics benchmarks show that a frozen SIMPLE model performs competitively, and lightweight adapter calibration further improves performance across most datasets.

By Jielong Lu, Zhihao Wu, Jiajun Yu, Zhaoliang Chen, Haishuai Wang