paper-with-me

홈 › Papers

Point-Driven Interactive Text and Image Layer Editing Using Diffusion Models

2025-04-18 · Zhenyu Yu, Mohd Yamani Idna Idris, Pei Wang, Yuelong Xia

We present DanceText, a training-free framework for multilingual text editing in images, designed to support complex geometric transformations and achieve seamless foreground-background integration. While diffusion-based generative models have shown promise in text-guided image synthesis, they often lack controllability and fail to preserve layout consistency under non-trivial manipulations such as rotation, translation, scaling, and warping. To address these limitations, DanceText introduces a layered editing strategy that separates text from the background, allowing geometric transformations to be performed in a modular and controllable manner. A depth-aware module is further proposed to align appearance and perspective between the transformed text and the reconstructed background, enhancing photorealism and spatial consistency. Importantly, DanceText adopts a fully training-free design by integrating pretrained modules, allowing flexible deployment without task-specific fine-tuning. Extensive experiments on the AnyWord-3M benchmark demonstrate that our method achieves superior performance in visual quality, especially under large-scale and complex transformation scenarios.

📄 PDF Abstract BibTeX arXiv:2504.14108

Code (0)

등록된 구현이 없습니다.

Tasks

Image Generation

Methods 이 논문이 사용한 방법론

ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

DA Wand: Distortion-Aware Selection using Neural Mesh Parameterization

2022-12-13 · CVPR 2023 1 · Richard Liu, Noam Aigerman, Vladimir G. Kim, Rana Hanocka

We present a neural technique for learning to select a local sub-region around a point which can be used for mesh parameterization. The motivation for our framework is driven by interactive workflows used for decaling, t…

Segmentation

iColoriT: Towards Propagating Local Hint to the Right Region in Interactive Colorization by Leveraging Vision Transformer

2022-07-14 · Jooyeol Yun, Sanghyeon Lee, Minho Park, Jaegul Choo

Point-interactive image colorization aims to colorize grayscale images when a user provides the colors for specific locations. It is essential for point-interactive colorization methods to appropriately propagate user-pr…

ColorizationDecoderImage ColorizationPoint-interactive Image Colorization

Hunyuan-GameCraft-2: Instruction-following Interactive Game World Model

2025-11-28 · Junshu Tang, Jiacheng Liu, Jiaqi Li, Longhuang Wu 외 arxiv

Recent advances in generative world models have enabled remarkable progress in creating open-ended game environments, evolving from static scene synthesis toward dynamic, interactive simulation. However, current approach…

SDMatte: Grafting Diffusion Models for Interactive Matting

2025-08-01 · Longfei Huang, Yu Liang, Hao Zhang, Jinwei Chen 외 arxiv

Recent interactive matting methods have shown satisfactory performance in capturing the primary regions of objects, but they fall short in extracting fine-grained details in edge regions. Diffusion models trained on bill…

P3S-Diffusion:A Selective Subject-driven Generation Framework via Point Supervision

2024-12-27 · Junjie Hu, Shuyong Gao, Lingyi Hong, Qishan Wang 외

Recent research in subject-driven generation increasingly emphasizes the importance of selective subject features. Nevertheless, accurately selecting the content in a given reference image still poses challenges, especia…

Image Generation