paper-with-me

홈 › Papers

Unsupervised Discovery of Semantic Latent Directions in Diffusion Models

2023-02-24 · Yong-Hyun Park, Mingi Kwon, Junghyo Jo, Youngjung Uh

Despite the success of diffusion models (DMs), we still lack a thorough understanding of their latent space. While image editing with GANs builds upon latent space, DMs rely on editing the conditions such as text prompts. We present an unsupervised method to discover interpretable editing directions for the latent variables $\mathbf{x}_t \in \mathcal{X}$ of DMs. Our method adopts Riemannian geometry between $\mathcal{X}$ and the intermediate feature maps $\mathcal{H}$ of the U-Nets to provide a deep understanding over the geometrical structure of $\mathcal{X}$. The discovered semantic latent directions mostly yield disentangled attribute changes, and they are globally consistent across different samples. Furthermore, editing in earlier timesteps edits coarse attributes, while ones in later timesteps focus on high-frequency details. We define the curvedness of a line segment between samples to show that $\mathcal{X}$ is a curved manifold. Experiments on different baselines and datasets demonstrate the effectiveness of our method even on Stable Diffusion. Our source code will be publicly available for the future researchers.

📄 PDF Abstract BibTeX arXiv:2302.12469

Code (0)

등록된 구현이 없습니다.

Tasks

Attribute

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

LatentCLR: A Contrastive Learning Approach for Unsupervised Discovery of Interpretable Directions

2021-04-02 · ICCV 2021 10 · Oğuz Kaan Yüksel, Enis Simsar, Ezgi Gülperi Er, Pinar Yanardag

Recent research has shown that it is possible to find interpretable directions in the latent spaces of pre-trained Generative Adversarial Networks (GANs). These directions enable controllable image generation and support…

Contrastive LearningImage Generation

NoiseCLR: A Contrastive Learning Approach for Unsupervised Discovery of Interpretable Directions in Diffusion Models

2023-12-08 · CVPR 2024 1 · Yusuf Dalva, Pinar Yanardag

Generative models have been very popular in the recent years for their image generation capabilities. GAN-based models are highly regarded for their disentangled latent space, which is a key feature contributing to their…

Contrastive LearningImage Generation

Enabling Local Editing in Diffusion Models by Joint and Individual Component Analysis

2024-08-29 · Theodoros Kouzelis, Manos Plitsis, Mihalis A. Nicolaou, Yannis Panagakis

Recent advances in Diffusion Models (DMs) have led to significant progress in visual synthesis and editing tasks, establishing them as a strong competitor to Generative Adversarial Networks (GANs). However, the latent sp…

DenoisingImage Manipulation

Discovering Class-Specific GAN Controls for Semantic Image Synthesis

2022-12-02 · Edgar Schönfeld, Julio Borges, Vadim Sushko, Bernt Schiele 외

Prior work has extensively studied the latent space structure of GANs for unconditional image synthesis, enabling global editing of generated images by the unsupervised discovery of interpretable latent directions. Howev…

Image Generation

Unsupervised Discovery of Interpretable Directions in the GAN Latent Space

2020-02-10 · ICML 2020 1 · Andrey Voynov, Artem Babenko

The latent spaces of GAN models often have semantically meaningful directions. Moving in these directions corresponds to human-interpretable image transformations, such as zooming or recoloring, enabling a more controlla…

Saliency Detection