paper-with-me

Tabular Data Generation

6개 벤치마크 · 논문 112편 · 이 태스크의 논문 보기 →

Benchmarks

Adult Census Income

결과 6개

Diabetes

결과 6개

HELOC

결과 6개

SICK

결과 6개

Travel

결과 6개

Most implemented

Papers

Generating Benchmark Health Data Using a Tabular Diffusion Transformer

2026-08-14 · Hao Yan, Lisa Pilgram, Dan Liu, Linglong Kong 외 arxiv

Cross-Tabular Data Generation (CTDG) seeks to learn a generative model from multiple heterogeneous tables and produce new synthetic tabular datasets. However, existing synthetic tabular data generation methods are largel…

Tabular Data Generation

TDGT: A Tabular Data Generation Toolkit supporting adaptive GPU-accelerated Bayesian mixture models, diffusion-based models, and latent-space generative modeling

2026-06-30 · Vasileios C. Pezoulas, Nikolaos S. Tachos, Eleni Georga, Kostas Marias 외 arxiv

The growing demand for privacy-preserving data sharing has positioned synthetic data generation as a critical component of responsible AI workflows. Despite notable advances in generative modeling, existing solutions oft…

Synthetic Data GenerationTabular Data Generation

PSyGenTAB: A Privacy-Preserving Framework for Synthetic Clinical Tabular Data Generation via Constrained Optimization

2026-06-16 · Arshia Ilaty, Hossein Shirazi, Manasi Chitale, Kedar Hegde 외 arxiv

The development of medical AI is constrained by limited access to high-quality clinical data due to institutional silos and strict privacy regulations such as HIPAA and GDPR. Synthetic data generation offers a potential …

Synthetic Data GenerationTabular Data Generation

BSTabDiff: Block-Subunit Diffusion Priors for High-Dimensional Tabular Data Generation

2026-06-08 · Al Zadid Sultan Bin Habib, Md Younus Ahamed, Prashnna Gyawali, Gianfranco Doretto 외 arxiv

High-Dimensional Low-Sample Size (HDLSS) tabular domains (e.g., omics) are characterized by $n \ll m$, where $n$ = number of samples, and $m$ = number of features. Such domains often exhibit strong local correlation grou…

Tabular Data Generation

Differentially Private Synthetic Data via APIs 4: Tabular Data

2026-06-06 · Toan Tran, Arturs Backurs, Zinan Lin, Victor Reis 외 arxiv

This paper investigates the problem of generating synthetic tabular data with differential privacy (DP) guarantees, enabling data sharing in sensitive domains. Despite extensive study, state-of-the-art methods often focu…

Tabular Data Generation

Hierarchical Synthetic Tabular Data Generation: A Hybrid Top-Down and Bottom-Up Framework

2026-05-27 · Junfeng Nie, Alvin Jin, Xiaohui Chen arxiv

Existing approaches for synthetic tabular data generation are based on either purely generative models or LLMs, both of which struggle with data heterogeneity, logical consistency, rare-event coverage, and robustness in …

Synthetic Data GenerationTabular Data Generation

전체 112편 보기 →