arXiv:2609.39124v1 Announce Type: new
Abstract: Generative models for tabular data are typically trained separately for each dataset, limiting knowledge transfer and requiring the storage of many spe...
By Mohamed Amine Ketata, Maximilian Schambach, Stephan G\"unnemann
arXiv:2603. 10823v2 Announce Type: replace-cross Abstract: Deep generative models can help with data scarcity and privacy by producing synthetic training data, but they struggle in low-data, imbalanced tabular settings to fully learn the complex data distribution.
By Xiaofeng Lin, Seungbae Kim, Zhuoya Li, Zachary DeSoto, Charles Fleming, Guang Cheng
arXiv:2602. 07875v3 Announce Type: replace Abstract: Generating tabular data under conditions is critical to applications requiring precise control over the generative process.
By Aditya Shankar, Yuandou Wang, Rihan Hai, Lydia Y. Chen
arXiv:2606. 09257v1 Announce Type: cross Abstract: High-Dimensional Low-Sample Size (HDLSS) tabular domains (e.
By Al Zadid Sultan Bin Habib, Md Younus Ahamed, Prashnna Gyawali, Gianfranco Doretto, Donald A. Adjeroh
arXiv:2609.16069v1 Announce Type: cross
Abstract: Synthetic tabular data can match real data distributions while still violating the semantic constraints that govern valid tabular rows. This reveals...
By Yili Wang, Ruxue Shi, Mengnan Du, Hangting Ye, Yi Chang, Xin Wang
arXiv:2603. 11946v2 Announce Type: replace-cross Abstract: Probabilistic circuits (PCs) enable exact and tractable inference but employ data independent mixture weights that limit their ability to capture local geometry of the data manifold.
By Sahil Sidheekh, Sriraam Natarajan