← 返回论文检索
IJCAI-ECAI 2026Main Track

Generating High-Diversity Synthetic Tabular Data via Less-Constrained Prior

Sanghun Park, Jaesung Lim, Jong-June Jeon, Seunghwan An

PDF 由论文原始站点提供,PaperCompass 不保存论文文件。

摘要

Generating high-quality synthetic tabular data, regarding fidelity, diversity, and utility, is crucial for many practical purposes. Recent tabular data synthesis methods, including two-stage generative modeling approaches, have achieved this with nearly perfect fidelity. However, we observe that there remains room for improving diversity, and we show that this limitation arises from an additional constraint imposed on the latent support. Our main contribution is that we effectively eliminate this redundant constraint by directly deriving the objective function from the KL-divergence between the ground-truth density and the generative model used for synthetic sample generation. We empirically demonstrate that our model, which relies on a single prior distribution, significantly improves the quality of the synthetic data, especially in terms of diversity. Our implementation code is available at https://anonymous.4open.science/r/SPT-1EE2/.