Data Science Wire

Measuring the Dependency Gap: Diagnosing Inter-Column Fidelity in Tabular Generative Models

arXiv cs.LG5d4 min read

arXiv:2607.21636v1 Announce Type: new Abstract: Synthetic tabular data is valued for preserving not only each column's marginal distribution but the dependencies between columns -- structure that carries much of the discriminative signal for minority classes in imbalanced domains such as fraud and clinical risk. Yet the metrics most commonly used to certify synthetic tabular data are, we show, largely blind to inter-column dependency: a baseline that models every column independently (and therefore destroys all dependency) is judged indistinguishable from real data by the logistic-regression C

Read the full story at arXiv cs.LG

More in Machine Learning