paper-with-me

홈 › Papers

Diversity-Driven Synthesis: Enhancing Dataset Distillation through Directed Weight Adjustment

2024-09-26 · Jiawei Du, Xin Zhang, Juncheng Hu, Wenxin Huang, Joey Tianyi Zhou

The sharp increase in data-related expenses has motivated research into condensing datasets while retaining the most informative features. Dataset distillation has thus recently come to the fore. This paradigm generates synthetic datasets that are representative enough to replace the original dataset in training a neural network. To avoid redundancy in these synthetic datasets, it is crucial that each element contains unique features and remains diverse from others during the synthesis stage. In this paper, we provide a thorough theoretical and empirical analysis of diversity within synthesized datasets. We argue that enhancing diversity can improve the parallelizable yet isolated synthesizing approach. Specifically, we introduce a novel method that employs dynamic and directed weight adjustment techniques to modulate the synthesis process, thereby maximizing the representativeness and diversity of each synthetic instance. Our method ensures that each batch of synthetic data mirrors the characteristics of a large, varying subset of the original dataset. Extensive experiments across multiple datasets, including CIFAR, Tiny-ImageNet, and ImageNet-1K, demonstrate the superior performance of our method, highlighting its effectiveness in producing diverse and representative synthetic datasets with minimal computational expense. Our code is available at https://github.com/AngusDujw/Diversity-Driven-Synthesis.https://github.com/AngusDujw/Diversity-Driven-Synthesis.

📄 PDF Abstract BibTeX arXiv:2409.17612

Code (1)

angusdujw/diversity-driven-synthesis 공식 구현 pytorch

Tasks

Dataset DistillationDiversity

Similar Papers 제목 키워드 기반

DELT: A Simple Diversity-driven EarlyLate Training for Dataset Distillation

2024-11-29 · CVPR 2025 1 · Zhiqiang Shen, Ammar Sherif, Zeyuan Yin, Shitong Shao

Recent advances in dataset distillation have led to solutions in two main directions. The conventional batch-to-batch matching mechanism is ideal for small-scale datasets and includes bi-level optimization methods on mod…

Dataset DistillationDiversity

SVGDreamer: Text Guided SVG Generation with Diffusion Model

2023-12-27 · CVPR 2024 1 · XiMing Xing, Haitao Zhou, Chuang Wang, Jing Zhang 외

Recently, text-guided scalable vector graphics (SVGs) synthesis has shown promise in domains such as iconography and sketch. However, existing text-to-SVG generation methods lack editability and struggle with visual qual…

DiversityVector Graphics

Diversity-Preserved Distribution Matching Distillation for Fast Visual Synthesis

2026-02-03 · Tianhe Wu, Ruibin Li, Lei Zhang, Kede Ma arxiv

Distribution matching distillation (DMD) facilitates few-step image generation by aligning a distilled student with a reference multi-step teacher. In practice, however, optimizing DMD can reduce sample diversity in few-…

Image Generation

Beyond Background Bias: Saliency-Driven Prototype Alignment for Dataset Distillation

2026-07-28 · Yawen Zou, Wenqi Cai, Guang Li, Ling Xiao 외 arxiv

Dataset distillation aims to synthesize compact datasets that can approximate the performance of full-data training while significantly reducing computational and storage costs. However, diffusion-based distillation meth…

Soft Label Pruning and Quantization for Large-Scale Dataset Distillation

2026-04-20 · Xiao Lingao, Yang He arxiv

Large-scale dataset distillation requires storing auxiliary soft labels that can be 30-40x larger on ImageNet-1K and 200x larger on ImageNet-21K than the condensed images, undermining the goal of dataset compression. We …