paper-with-me

Papers Text-based Image Editing

“Text-based Image Editing” 태그가 달린 논문 54편 · 필터 해제

Proc3D: Procedural 3D Generation and Parametric Editing of 3D Shapes with Large Language Models

2026-01-18 · Fadlullah Raji, Stefano Petrangeli, Matheus Gadelha, Yu Shen 외 arxiv

Generating 3D models has traditionally been a complex task requiring specialized expertise. While recent advances in generative AI have sought to automate this process, existing methods produce non-editable representatio…

Text-based Image Editing3D GenerationPoint Clouds

VENUS: Visual Editing with Noise Inversion Using Scene Graphs

2026-01-12 · Thanh-Nhan Vo, Trong-Thuan Nguyen, Tam V. Nguyen, Minh-Triet Tran arxiv

State-of-the-art text-based image editing models often struggle to balance background preservation with semantic consistency, frequently resulting either in the synthesis of entirely new images or in outputs that fail to…

Text-based Image Editing

FlowDC: Flow-Based Decoupling-Decay for Complex Image Editing

2025-12-12 · Yilei Jiang, Zhen Wang, Yanghao Wang, Jun Yu 외 arxiv

With the surge of pre-trained text-to-image flow matching models, text-based image editing performance has gained remarkable improvement, especially for \underline{simple editing} that only contains a single editing targ…

Text-based Image Editing

3D-Consistent Multi-View Editing by Correspondence Guidance

2025-11-27 · Josef Bengtson, David Nilsson, Dong In Lee, Yaroslava Lochman 외 arxiv

Recent advancements in diffusion and flow models have greatly improved text-based image editing, yet methods that edit images independently often produce geometrically and photometrically inconsistent results across diff…

Text-based Image Editing

Target-aware Image Editing via Cycle-consistent Constraints

2025-10-23 · Yanghao Wang, Zhen Wang, Long Chen arxiv

Recent pre-trained text-to-image flow models have enabled remarkable progress in text-based image editing. Mainstream approaches adopt a corruption-then-restoration paradigm, where the source image is first corrupted int…

Text-based Image EditingImage Reconstruction

InstructUDrag: Joint Text Instructions and Object Dragging for Interactive Image Editing

2025-10-09 · Haoran Yu, Yi Shi arxiv

Text-to-image diffusion models have shown great potential for image editing, with techniques such as text-based and object-dragging methods emerging as key approaches. However, each of these methods has inherent limitati…

Text-based Image EditingImage Reconstruction

Prompt-to-Prompt: Text-Based Image Editing Via Cross-Attention Mechanisms -- The Research of Hyperparameters and Novel Mechanisms to Enhance Existing Frameworks

2025-10-05 · Linn Bieske, Carla Lorente arxiv

Recent advances in image editing have shifted from manual pixel manipulation to employing deep learning methods like stable diffusion models, which now leverage cross-attention mechanisms for text-driven control. This tr…

Text-based Image Editing

Immunizing Images from Text to Image Editing via Adversarial Cross-Attention

2025-09-12 · Matteo Trippodo, Federico Becattini, Lorenzo Seidenari arxiv

Recent advances in text-based image editing have enabled fine-grained manipulation of visual content guided by natural language. However, such methods are susceptible to adversarial attacks. In this work, we propose a no…

Text-based Image Editing

Discrete Noise Inversion for Next-scale Autoregressive Text-based Image Editing

2025-09-02 · Quan Dao, Xiaoxiao He, Ligong Han, Ngan Hoai Nguyen 외 arxiv

Visual autoregressive models (VAR) have recently emerged as a promising class of generative models, achieving performance comparable to diffusion models in text-to-image generation tasks. While conditional generation has…

Text-based Image EditingText-to-Image Generation

NoHumansRequired: Autonomous High-Quality Image Editing Triplet Mining

2025-07-18 · Maksim Kuprashevich, Grigorii Alekseenko, Irina Tolstykh, Georgii Fedorov 외

Recent advances in generative modeling enable image editing assistants that follow natural language instructions without additional user input. Their supervised training requires millions of triplets: original image, ins…

Image EditingText-based Image EditingTriplet

Cora: Correspondence-aware image editing using few step diffusion

2025-05-29 · Amirhossein Almohammadi, Aryan Mikaeili, Sauradip Nag, Negar Hassanpour 외

Image editing is an important task in computer graphics, vision, and VFX, with recent diffusion-based methods achieving fast and high-quality results. However, edits requiring significant structural changes, such as non-…

Image-to-Image TranslationSemantic correspondenceText-based Image EditingTranslation

POEM: Precise Object-level Editing via MLLM control

2025-04-10 · Marco Schouten, Mehmet Onurcan Kaya, Serge Belongie, Dim P. Papadopoulos

Diffusion models have significantly improved text-to-image generation, producing high-quality, realistic images from textual descriptions. Beyond generation, object-level image editing remains a challenging problem, requ…

Image GenerationObjectObject LocalizationText-based Image Editing+2

ILLUME+: Illuminating Unified MLLM with Dual Visual Tokenization and Diffusion Refinement

2025-04-02 · Runhui Huang, Chunwei Wang, Junwei Yang, Guansong Lu 외

We present ILLUME+ that leverages dual visual tokenization and a diffusion decoder to improve both deep semantic understanding and high-fidelity image generation. Existing unified models have struggled to simultaneously …

DecoderImage GenerationImage ReconstructionSuper-Resolution+2

KV-Edit: Training-Free Image Editing for Precise Background Preservation

2025-02-24 · Tianrui Zhu, Shiyi Zhang, Jiawei Shao, Yansong Tang

Background consistency remains a significant challenge in image editing tasks. Despite extensive developments, existing works still face a trade-off between maintaining similarity to the original image and generating con…

Text-based Image Editing

PartEdit: Fine-Grained Image Editing using Pre-Trained Diffusion Models

2025-02-06 · Aleksandar Cvejic, Abdelrahman Eldesokey, Peter Wonka

We present the first text-based image editing approach for object parts based on pre-trained diffusion models. Diffusion-based image editing approaches capitalized on the deep understanding of diffusion models of image s…

ObjectText-based Image Editing

FeedEdit: Text-Based Image Editing with Dynamic Feedback Regulation

2025-01-01 · CVPR 2025 1 · Fengyi Fu, Lei Zhang, Mengqi Huang, Zhendong Mao

Text-based image editing which aims at generating rigid or non-rigid changes to images conditioned on the given text, has recently attracted considerable interest. Previous works mainly follow the multi-step denoisin…

DenoisingText-based Image Editing

FireFlow: Fast Inversion of Rectified Flow for Image Semantic Editing

2024-12-10 · Yingying Deng, Xiangyu He, Changwang Mei, Peisong Wang 외

Though Rectified Flows (ReFlows) with distillation offers a promising way for fast sampling, its fast inversion transforms images back to structured noise for recovery and following editing remains unsolved. This paper i…

Text-based Image Editing

Action-based image editing guided by human instructions

2024-12-05 · Maria Mihaela Trusca, Mingxiao Li, Marie-Francine Moens

Text-based image editing is typically approached as a static task that involves operations such as inserting, deleting, or modifying elements of an input image based on human instructions. Given the static nature of this…

Text-based Image Editing

LoRA of Change: Learning to Generate LoRA for the Editing Instruction from A Single Before-After Image Pair

2024-11-28 · Xue Song, Jiequan Cui, Hanwang Zhang, Jiaxin Shi 외

In this paper, we propose the LoRA of Change (LoC) framework for image editing with visual instructions, i.e., before-after image pairs. Compared to the ambiguities, insufficient specificity, and diverse interpretations …

SpecificityText-based Image Editing

Pathways on the Image Manifold: Image Editing via Video Generation

2024-11-25 · CVPR 2025 1 · Noam Rotstein, Gal Yona, Daniel Silver, Roy Velich 외

Recent advances in image editing, driven by image diffusion models, have shown remarkable progress. However, significant challenges remain, as these models often struggle to follow complex edit instructions accurately an…

Text-based Image EditingVideo Generation
1–20 / 54 다음 →