paper-with-me

홈 › Papers

Dataset Condensation with Contrastive Signals

2022-02-07 · Saehyung Lee, Sanghyuk Chun, Sangwon Jung, Sangdoo Yun, Sungroh Yoon

Recent studies have demonstrated that gradient matching-based dataset synthesis, or dataset condensation (DC), methods can achieve state-of-the-art performance when applied to data-efficient learning tasks. However, in this study, we prove that the existing DC methods can perform worse than the random selection method when task-irrelevant information forms a significant part of the training dataset. We attribute this to the lack of participation of the contrastive signals between the classes resulting from the class-wise gradient matching strategy. To address this problem, we propose Dataset Condensation with Contrastive signals (DCC) by modifying the loss function to enable the DC methods to effectively capture the differences between classes. In addition, we analyze the new loss function in terms of training dynamics by tracking the kernel velocity. Furthermore, we introduce a bi-level warm-up strategy to stabilize the optimization. Our experimental results indicate that while the existing methods are ineffective for fine-grained image classification tasks, the proposed method can successfully generate informative synthetic datasets for the same tasks. Moreover, we demonstrate that the proposed method outperforms the baselines even on benchmark datasets such as SVHN, CIFAR-10, and CIFAR-100. Finally, we demonstrate the high applicability of the proposed method by applying it to continual learning tasks.

📄 PDF Abstract BibTeX arXiv:2202.02916

Code (2)

saehyung-lee/dcc 공식 구현 pytorch
Guang000/Awesome-Dataset-Distillation

Tasks

AttributeContinual LearningDataset CondensationFine-Grained Image Classificationimage-classificationImage Classification

Similar Papers 제목 키워드 기반

Navigating Complexity: Toward Lossless Graph Condensation via Expanding Window Matching

2024-02-07 · Yuchen Zhang, Tianle Zhang, Kai Wang, Ziyao Guo 외

Graph condensation aims to reduce the size of a large-scale graph dataset by synthesizing a compact counterpart without sacrificing the performance of Graph Neural Networks (GNNs) trained on it, which has shed light on r…

Is Less More? Exploring Token Condensation as Training-free Adaptation for CLIP

2024-10-16 · Zixin Wang, Dong Gong, Sen Wang, Zi Huang 외

Contrastive language-image pre-training (CLIP) has shown remarkable generalization ability in image classification. However, CLIP sometimes encounters performance drops on downstream datasets during zero-shot inference. …

image-classificationImage ClassificationTest-time Adaptation

Transferable Graph Condensation from the Causal Perspective

2026-01-29 · Huaming Du, Yijie Huang, Su Yao, Yiying Wang 외 arxiv

The increasing scale of graph datasets has significantly improved the performance of graph representation learning methods, but it has also introduced substantial training challenges. Graph dataset condensation technique…

Graph Representation LearningContrastive Learning

Bottleneck Tokens for Unified Multimodal Retrieval

2026-04-13 · Siyu Sun, Jing Ren, Zhaohe Liao, Dongxiao Mao 외 arxiv

Adapting decoder-only multimodal large language models (MLLMs) for unified multimodal retrieval faces two structural gaps. First, existing methods rely on implicit pooling, which overloads the hidden state of a standard …

Contrastive Graph Condensation: Advancing Data Versatility through Self-Supervised Learning

2024-11-26 · Xinyi Gao, Yayong Li, Tong Chen, Guanhua Ye 외

With the increasing computation of training graph neural networks (GNNs) on large-scale graphs, graph condensation (GC) has emerged as a promising solution to synthesize a compact, substitute graph of the large-scale ori…

Graph GenerationSelf-Supervised Learning