paper-with-me

Papers

FreeInsert: Personalized Object Insertion with Geometric and Style Control

2025-09-25 · Yuhong Zhang, Han Wang, Yiwen Wang, Rong Xie, Li Song arxiv

Text-to-image diffusion models have made significant progress in image generation, allowing for effortless customized generation. However, existing image editing methods still face certain limitations when dealing with personalized image composition tasks. First, there is the issue of lack of geometric control over the inserted objects. Current methods are confined to 2D space and typically rely on textual instructions, making it challenging to maintain precise geometric control over the objects. Second, there is the challenge of style consistency. Existing methods often overlook the style consistency between the inserted object and the background, resulting in a lack of realism. In addition, the challenge of inserting objects into images without extensive training remains significant. To address these issues, we propose \textit{FreeInsert}, a novel training-free framework that customizes object insertion into arbitrary scenes by leveraging 3D geometric information. Benefiting from the advances in existing 3D generation models, we first convert the 2D object into 3D, perform interactive editing at the 3D level, and then re-render it into a 2D image from a specified view. This process introduces geometric controls such as shape or view. The rendered image, serving as geometric control, is combined with style and content control achieved through diffusion adapters, ultimately producing geometrically controlled, style-consistent edited images via the diffusion model.

📄 PDF Abstract BibTeX arXiv:2509.20756

Code (0)

등록된 구현이 없습니다.

Tasks

Image Generation3D GenerationImage Editing

Similar Papers 제목 키워드 기반

FreeInsert: Disentangled Text-Guided Object Insertion in 3D Gaussian Scene without Spatial Priors

2025-05-02 · Chenxi Li, Weijie Wang, Qiang Li, Bruno Lepri 외

Text-driven object insertion in 3D scenes is an emerging task that enables intuitive scene editing through natural language. However, existing 2D editing-based methods often rely on spatial priors such as 2D masks or 3D …

ObjectSpatial Reasoning

Magic Insert: Style-Aware Drag-and-Drop

2024-07-02 · Nataniel Ruiz, Yuanzhen Li, Neal Wadhwa, Yael Pritch 외

We present Magic Insert, a method for dragging-and-dropping subjects from a user-provided image into a target image of a different style in a physically plausible manner while matching the style of the target image. This…

Domain AdaptationObject

SwapAnything: Enabling Arbitrary Object Swapping in Personalized Visual Editing

2024-04-08 · Jing Gu, Nanxuan Zhao, Wei Xiong, Qing Liu 외

Effective editing of personal content holds a pivotal role in enabling individuals to express their creativity, weaving captivating narratives within their visual stories, and elevate the overall quality and impact of th…

Image GenerationObject

SceneExpander: Text-Guided 3D Scene Expansion via Free-Form View Insertion

2026-03-28 · Zijian He, Renjie Liu, Yihao Wang, Weizhi Zhong 외 arxiv

World building with 3D scene representations is increasingly important for content creation, simulation, and interactive experiences, yet real workflows are inherently iterative: creators repeatedly extend existing scene…

Test-time Adaptation3D ReconstructionStyle Transfer

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework

2026-05-22 · Xiao Cao, Yansong Qu, Xiangzhen, Chang 외 arxiv

Mask-free video object insertion has emerged as a challenging task, requiring harmonious integration of reference objects into source videos. However, existing methods struggle when references exhibit severe stylistic do…

Video GenerationStyle Transfer