CoDi: Co-evolving Contrastive Diffusion Models for Mixed-type Tabular Synthesis
Chaejeong Lee, Jayoung Kim, Noseong Park
摘要
With growing attention to tabular data these days, the attempt to apply a synthetic table to various tasks has been expanded toward various scenarios. Owing to the recent advances in generative modeling, fake data generated by tabular data synthesis models become sophisticated and realistic. However, there still exists a difficulty in modeling discrete variables (columns) of tabular data. In this work, we propose to process continuous and discrete variables separately (but being conditioned on each other) by two diffusion models. The two diffusion models are co-evolved during training by reading conditions from each other. In order to further bind the diffusion models, moreover, we introduce a contrastive learning method with a negative sampling method. In our experiments with 11 realworld tabular datasets and 8 baseline methods, we prove the efficacy of the proposed method, called CoDi. Our code is available at https: //github.com/ChaejeongLee/CoDi .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- Mixed-Type Tabular Data Synthesis with Score-based Diffusion in Latent SpaceHengrui Zhang, Jiani Zhang, Zhengyuan Shen, Balasubramaniam Srinivasan 等ICLR 2024 · 被引用 233 次
- TabEBM: A Tabular Data Augmentation Method with Distinct Class-Specific Energy-Based ModelsAndrei Margeloiu, Xiangjian Jiang, Nikola Simidjievski, Mateja JamnikNeurIPS 2024 · 被引用 19 次
- Diffusion Transformers for Tabular Data Time Series GenerationFabrizio Garuti, Enver Sangineto, Simone Luetto, Lorenzo Forni 等ICLR 2025 · 被引用 12 次
- TabStruct: Measuring Structural Fidelity of Tabular DataXiangjian Jiang, Nikola Simidjievski, Mateja JamnikICLR 2026 · 被引用 10 次
- SynCoGen: Synthesizable 3D Molecule Generation via Joint Reaction and Coordinate ModelingAndrei Rekesh, Miruna Cretu, Dmytro Shevchuk, Pietro Lio 等ICLR 2026 · 被引用 6 次
它引用的顶会 Paper9
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Structured Denoising Diffusion Models in Discrete State-SpacesJacob Austin, Daniel D. Johnson, Jonathan Ho, Daniel Tarlow 等NeurIPS 2021 · 被引用 2,256 次
- Revisiting Deep Learning Models for Tabular DataYury Gorishniy, Ivan Rubachev, Valentin Khrulkov, Artem BabenkoNeurIPS 2021 · 被引用 1,847 次
- Argmax Flows and Multinomial Diffusion: Learning Categorical DistributionsEmiel Hoogeboom, Didrik Nielsen, Priyank Jaini, Patrick Forré 等NeurIPS 2021 · 被引用 782 次
- Tackling the Generative Learning Trilemma with Denoising Diffusion GANsZhisheng Xiao, Karsten Kreis, Arash VahdatICLR 2022 · 被引用 726 次
相关 Paper
- A Learnable Discrete-Prior Fusion Autoencoder with Contrastive Learning for Tabular Data SynthesisRongchao Zhang, Yiwei Lou, Dexuan Xu, Yongzhi Cao 等AAAI 2024 · 被引用 14 次
- TabDiff: a Mixed-type Diffusion Model for Tabular Data GenerationJuntong Shi, Minkai Xu, Harper Hua, Hengrui Zhang 等ICLR 2025
- Controllable Tabular Data Synthesis Using Diffusion ModelsTongyu Liu, Ju Fan, Nan Tang, Guoliang Li 等SIGMOD 2024 · 被引用 13 次
- Discrete Contrastive Diffusion for Cross-Modal Music and Image GenerationYe Zhu, Yu Wu, Kyle Olszewski, Jian Ren 等ICLR 2023 · 被引用 10 次
- CG-TGAN: Conditional Generative Adversarial Networks with Graph Neural Networks for Tabular Data SynthesizingSeungcheol Lee, Moohong MinAAAI 2025 · 被引用 4 次
