DisCont: Self-Supervised Visual Attribute Disentanglement using Context Vectors
Disentangling the underlying feature attributes within an image with no prior supervision is a challenging task. Models that can disentangle attributes well provide greater interpretability and control. In this paper, we propose a self-supervised framework DisCont to disentangle multiple attributes by exploiting the structural inductive biases within images. Motivated by the recent surge in contrastive learning paradigms, our model bridges the gap between self-supervised contrastive learning algorithms and unsupervised disentanglement. We evaluate the efficacy of our approach, both qualitatively and quantitatively, on four benchmark datasets.
Code (1)
Tasks
AttributeContrastive LearningDisentanglementMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Operationalizing Quantized Disentanglement
Recent theoretical work established the unsupervised identifiability of quantized factors under any diffeomorphism. The theory assumes that quantization thresholds correspond to axis-aligned discontinuities in the probab…
Self-supervised Enhancement of Latent Discovery in GANs
Several methods for discovering interpretable directions in the latent space of pre-trained GANs have been proposed. Latent semantics discovered by unsupervised methods are relatively less disentangled than supervised me…
AttributeDisentanglementImage RetrievalRetrievalSPADE: Self-supervised Pretraining for Acoustic DisEntanglement
Self-supervised representation learning approaches have grown in popularity due to the ability to train models on large amounts of unlabeled data and have demonstrated success in diverse fields such as natural language p…
DisentanglementRepresentation LearningRhythmDetach and Adapt: Learning Cross-Domain Disentangled Deep Representation
While representation learning aims to derive interpretable features for describing visual data, representation disentanglement further results in such features so that particular image attributes can be identified and ma…
AttributeDisentanglementDomain AdaptationRepresentation Learning+2U-VAP: User-specified Visual Appearance Personalization via Decoupled Self Augmentation
Concept personalization methods enable large text-to-image models to learn specific subjects (e.g., objects/poses/3D models) and synthesize renditions in new contexts. Given that the image references are highly biased to…
AttributeDisentanglementSentence