arXiv:2410. 01244v2 Announce Type: replace-cross Abstract: We introduce a novel Wasserstein-1 ($W_1$) path-space divergence for stochastic and deterministic dynamics and establish a Wasserstein Uncertainty Propagation (WUP) theorem that bounds the $W_1$ distance between terminal distributions by the proposed divergence, equivalently characterized by a weighted $L^2$ discrepancy between the underlying drifts and the $W_1$ distance between their initial measures.
By Ziyu Chen, Markos A. Katsoulakis, Benjamin J. Zhang
arXiv:2607. 21636v1 Announce Type: new Abstract: Synthetic tabular data is valued for preserving not only each column's marginal distribution but the dependencies between columns -- structure that carries much of the discriminative signal for minority classes in imbalanced domains such as fraud and clinical risk.
By Jie Zhang
The paper introduces Conditional-Independence-Regularized Distributional Autoencoders, a framework for learning low-dimensional representations of mixed-type data that includes both numerical and categorical variables. It uses an energy-score objective for numerical variables, a likelihood objective for categorical variables, and an auxiliary conditional independence regularization term to capture dependencies between variable types. The authors provide theoretical analysis and demonstrate that the method improves categorical distribution recovery, achieves competitive overall conditional distribution recovery, and preserves mixed-type dependence structure on synthetic and real-world datasets.
By Siyuan Tang, Gongjun Xu, Ji Zhu
arXiv:2609.39124v1 Announce Type: new
Abstract: Generative models for tabular data are typically trained separately for each dataset, limiting knowledge transfer and requiring the storage of many spe...
By Mohamed Amine Ketata, Maximilian Schambach, Stephan G\"unnemann
arXiv:2602. 24201v2 Announce Type: replace Abstract: Estimating density ratios between pairs of intractable data distributions is a core problem in probabilistic modeling, enabling principled comparisons of sample likelihoods under different data-generating processes across conditions.
By Egor Antipov, Alessandro Palma, Lorenzo Consoli, Stephan G\"unnemann, Andrea Dittadi, Fabian J. Theis
arXiv:2511. 17812v3 Announce Type: replace-cross Abstract: Flow matching models effectively represent complex distributions, yet estimating expectations of functions of their outputs remains challenging under limited sampling budgets.
By Xinshuang Liu, Runfa Blark Li, Shaoxiu Wei, Truong Nguyen