paper-with-me

홈 › Papers

Towards Uncovering the Intrinsic Data Structures for Unsupervised Domain Adaptation using Structurally Regularized Deep Clustering

2020-12-08 · Hui Tang, Xiatian Zhu, Ke Chen, Kui Jia, C. L. Philip Chen

Unsupervised domain adaptation (UDA) is to learn classification models that make predictions for unlabeled data on a target domain, given labeled data on a source domain whose distribution diverges from the target one. Mainstream UDA methods strive to learn domain-aligned features such that classifiers trained on the source features can be readily applied to the target ones. Although impressive results have been achieved, these methods have a potential risk of damaging the intrinsic data structures of target discrimination, raising an issue of generalization particularly for UDA tasks in an inductive setting. To address this issue, we are motivated by a UDA assumption of structural similarity across domains, and propose to directly uncover the intrinsic target discrimination via constrained clustering, where we constrain the clustering solutions using structural source regularization that hinges on the very same assumption. Technically, we propose a hybrid model of Structurally Regularized Deep Clustering, which integrates the regularized discriminative clustering of target data with a generative one, and we thus term our method as H-SRDC. Our hybrid model is based on a deep clustering framework that minimizes the Kullback-Leibler divergence between the distribution of network prediction and an auxiliary one, where we impose structural regularization by learning domain-shared classifier and cluster centroids. By enriching the structural similarity assumption, we are able to extend H-SRDC for a pixel-level UDA task of semantic segmentation. We conduct extensive experiments on seven UDA benchmarks of image classification and semantic segmentation. With no explicit feature alignment, our proposed H-SRDC outperforms all the existing methods under both the inductive and transductive settings. We make our implementation codes publicly available at https://github.com/huitangtang/H-SRDC.

📄 PDF Abstract BibTeX arXiv:2012.04280

Code (2)

huitangtang/H-SRDC 공식 구현 pytorch
huitangtang/SRDCPP 공식 구현 pytorch

Tasks

ClusteringConstrained ClusteringDeep ClusteringDomain Adaptationimage-classificationImage ClassificationSemantic SegmentationUnsupervised Domain Adaptation

Similar Papers 제목 키워드 기반

Unsupervised Feature Selection with Adaptive Structure Learning

2015-04-03 · Liang Du, Yi-Dong Shen

The problem of feature selection has raised considerable interests in the past decade. Traditional unsupervised methods select the features which can faithfully preserve the intrinsic structures of data, where the intrin…

feature selection

Refinement Contrastive Learning of Cell-Gene Associations for Unsupervised Cell Type Identification

2025-12-11 · Liang Peng, Haopeng Liu, Yixuan Ye, Cheng Liu 외 arxiv

Unsupervised cell type identification is crucial for uncovering and characterizing heterogeneous populations in single cell omics studies. Although a range of clustering methods have been developed, most focus exclusivel…

Representation LearningContrastive Learning

Unsupervised Domain Adaptation for Image Classification via Structure-Conditioned Adversarial Learning

2021-03-04 · Hui Wang, Jian Tian, Songyuan Li, Hanbin Zhao 외

Unsupervised domain adaptation (UDA) typically carries out knowledge transfer from a label-rich source domain to an unlabeled target domain by adversarial learning. In principle, existing UDA approaches mainly focus on t…

Domain AdaptationGeneral Classificationimage-classificationImage Classification+2

Uncovering divergent linguistic information in word embeddings with lessons for intrinsic and extrinsic evaluation

2018-09-06 · CONLL 2018 10 · Mikel Artetxe, Gorka Labaka, Iñigo Lopez-Gazpio, Eneko Agirre

Following the recent success of word embeddings, it has been argued that there is no such thing as an ideal representation for words, as different models tend to capture divergent and often mutually incompatible aspects …

Word Embeddings

Uncovering Intrinsic Capabilities: A Paradigm for Data Curation in Vision-Language Models

2025-09-27 · Junjie Li, Ziao Wang, Jianghong Ma, Xiaofeng Zhang arxiv

Large vision-language models (VLMs) achieve strong benchmark performance, but controlling their behavior through instruction tuning remains difficult. Reducing the budget of instruction tuning dataset often causes regres…