paper-with-me

Papers

ParallelEdits: Efficient Multi-object Image Editing

2024-06-03 · Mingzhen Huang, Jialing Cai, Shan Jia, Vishnu Suresh Lokhande, Siwei Lyu

Text-driven image synthesis has made significant advancements with the development of diffusion models, transforming how visual content is generated from text prompts. Despite these advances, text-driven image editing, a key area in computer graphics, faces unique challenges. A major challenge is making simultaneous edits across multiple objects or attributes. Applying these methods sequentially for multi-attribute edits increases computational demands and efficiency losses. In this paper, we address these challenges with significant contributions. Our main contribution is the development of ParallelEdits, a method that seamlessly manages simultaneous edits across multiple attributes. In contrast to previous approaches, ParallelEdits not only preserves the quality of single attribute edits but also significantly improves the performance of multitasking edits. This is achieved through innovative attention distribution mechanism and multi-branch design that operates across several processing heads. Additionally, we introduce the PIE-Bench++ dataset, an expansion of the original PIE-Bench dataset, to better support evaluating image-editing tasks involving multiple objects and attributes simultaneously. This dataset is a benchmark for evaluating text-driven image editing methods in multifaceted scenarios.

📄 PDF Abstract BibTeX arXiv:2406.00985

Code (0)

등록된 구현이 없습니다.

Tasks

AttributeImage GenerationObject

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Referring Image Editing: Object-level Image Editing via Referring Expressions

2024-01-01 · CVPR 2024 1 · Chang Liu, Xiangtai Li, Henghui Ding

Significant advancements have been made in image editing with the recent advance of the Diffusion model. However most of the current methods primarily focus on global or subject-level modifications and often face lim…

Semantic Segmentation

Object-aware Inversion and Reassembly for Image Editing

2023-10-18 · Zhen Yang, Ganggui Ding, Wen Wang, Hao Chen 외

By comparing the original and target prompts, we can obtain numerous editing pairs, each comprising an object and its corresponding editing target. To allow editability while maintaining fidelity to the input image, exis…

BenchmarkingDenoisingObject

MoEdit: On Learning Quantity Perception for Multi-object Image Editing

2025-03-13 · CVPR 2025 1 · Yanfeng Li, Kahou Chan, Yue Sun, ChanTong Lam 외

Multi-object images are prevalent in various real-world scenarios, including augmented reality, advertisement design, and medical imaging. Efficient and precise editing of these images is critical for these applications.…

AttributeImage GenerationObjectStyle Transfer

PAIR Diffusion: A Comprehensive Multimodal Object-Level Image Editor

2024-01-01 · CVPR 2024 1 · Vidit Goel, Elia Peruzzo, Yifan Jiang, Dejia Xu 외

Generative image editing has recently witnessed extremely fast-paced growth. Some works use high-level conditioning such as text while others use low-level conditioning. Nevertheless most of them lack fine-grained co…

Object

PAIR-Diffusion: A Comprehensive Multimodal Object-Level Image Editor

2023-03-30 · Vidit Goel, Elia Peruzzo, Yifan Jiang, Dejia Xu 외

Generative image editing has recently witnessed extremely fast-paced growth. Some works use high-level conditioning such as text, while others use low-level conditioning. Nevertheless, most of them lack fine-grained cont…

Object