paper-with-me

홈 › Papers

Diffusion Model-Based Image Editing: A Survey

2024-02-27 · Yi Huang, Jiancheng Huang, Yifan Liu, Mingfu Yan, Jiaxi Lv, Jianzhuang Liu, Wei Xiong, He Zhang, Shifeng Chen, Liangliang Cao

Denoising diffusion models have emerged as a powerful tool for various image generation and editing tasks, facilitating the synthesis of visual content in an unconditional or input-conditional manner. The core idea behind them is learning to reverse the process of gradually adding noise to images, allowing them to generate high-quality samples from a complex distribution. In this survey, we provide an exhaustive overview of existing methods using diffusion models for image editing, covering both theoretical and practical aspects in the field. We delve into a thorough analysis and categorization of these works from multiple perspectives, including learning strategies, user-input conditions, and the array of specific editing tasks that can be accomplished. In addition, we pay special attention to image inpainting and outpainting, and explore both earlier traditional context-driven and current multimodal conditional methods, offering a comprehensive analysis of their methodologies. To further evaluate the performance of text-guided image editing algorithms, we propose a systematic benchmark, EditEval, featuring an innovative metric, LMM Score. Finally, we address current limitations and envision some potential directions for future research. The accompanying repository is released at https://github.com/SiatMMLab/Awesome-Diffusion-Model-Based-Image-Editing-Methods.

📄 PDF Abstract BibTeX arXiv:2402.17525

Code (1)

siatmmlab/awesome-diffusion-model-based-image-editing-methods 공식 구현 tf

Tasks

DenoisingImage GenerationImage InpaintingmodelSurveytext-guided-image-editing

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…
Inpainting Train a convolutional neural network to generate the contents of an arbitrary image region conditioned on its surroundings.

Similar Papers 제목 키워드 기반

Preference Alignment on Diffusion Model: A Comprehensive Survey for Image Generation and Editing

2025-02-10 · Sihao Wu, Xiaonan Si, Chi Xing, Jianhong Wang 외

The integration of preference alignment with diffusion models (DMs) has emerged as a transformative approach to enhance image generation and editing capabilities. Although integrating diffusion models with preference ali…

Autonomous DrivingImage Generation

Text-to-image Diffusion Models in Generative AI: A Survey

2023-03-14 · Chenshuang Zhang, Chaoning Zhang, Mengchun Zhang, In So Kweon 외

This survey reviews the progress of diffusion models in generating images from text, ~\textit{i.e.} text-to-image diffusion models. As a self-contained work, this survey starts with a brief introduction of how diffusion …

Image GenerationSurveytext-guided-generationtext-guided-image-editing+2

A Survey of Multimodal-Guided Image Editing with Text-to-Image Diffusion Models

2024-06-20 · Xincheng Shuai, Henghui Ding, Xingjun Ma, RongCheng Tu 외

Image editing aims to edit the given synthetic or real image to meet the specific requirements from users. It is widely studied in recent years as a promising and challenging field of Artificial Intelligence Generative C…

Video Editing

A Survey on Video Diffusion Models

2023-10-16 · Zhen Xing, Qijun Feng, Haoran Chen, Qi Dai 외

The recent wave of AI-generated content (AIGC) has witnessed substantial success in computer vision, with the diffusion model playing a crucial role in this achievement. Due to their impressive generative capabilities, d…

Image GenerationSurveyVideo EditingVideo Generation+1

Diffusion Model-Based Video Editing: A Survey

2024-06-26 · Wenhao Sun, Rong-Cheng Tu, Jingyi Liao, DaCheng Tao

The rapid development of diffusion models (DMs) has significantly advanced image and video applications, making "what you want is what you see" a reality. Among these, video editing has gained substantial attention and s…

modelSurveyVideo Editing