paper-with-me

홈 › Papers

Teleportraits: Training-Free People Insertion into Any Scene

2025-10-07 · Jialu Gao, K J Joseph, Fernando De La Torre arxiv

The task of realistically inserting a human from a reference image into a background scene is highly challenging, requiring the model to (1) determine the correct location and poses of the person and (2) perform high-quality personalization conditioned on the background. Previous approaches often treat them as separate problems, overlooking their interconnections, and typically rely on training to achieve high performance. In this work, we introduce a unified training-free pipeline that leverages pre-trained text-to-image diffusion models. We show that diffusion models inherently possess the knowledge to place people in complex scenes without requiring task-specific training. By combining inversion techniques with classifier-free guidance, our method achieves affordance-aware global editing, seamlessly inserting people into scenes. Furthermore, our proposed mask-guided self-attention mechanism ensures high-quality personalization, preserving the subject's identity, clothing, and body features from just a single reference image. To the best of our knowledge, we are the first to perform realistic human insertions into scenes in a training-free manner and achieve state-of-the-art results in diverse composite scene images with excellent identity preservation in backgrounds and subjects.

📄 PDF Abstract BibTeX arXiv:2510.05660

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

DreamInsert: Zero-Shot Image-to-Video Object Insertion from A Single Image

2025-03-13 · Qi Zhao, Zhan Ma, Pan Zhou

Recent developments in generative diffusion models have turned many dreams into realities. For video object insertion, existing methods typically require additional information, such as a reference video or a 3D asset of…

Object

OmniInsert: Mask-Free Video Insertion of Any Reference via Diffusion Transformer Models

2025-09-22 · Jinshu Chen, Xinghui Li, Xu Bai, Tianxiang Ma 외 arxiv

Recent advances in video insertion based on diffusion models are impressive. However, existing methods rely on complex control signals but struggle with subject consistency, limiting their practical applicability. In thi…

Smart-Insertion-V: Photorealistic Video Insertion via a Closed-Loop Feedback Dual-Stream Framework

2026-05-22 · Xiao Cao, Yansong Qu, Xiangzhen, Chang 외 arxiv

Mask-free video object insertion has emerged as a challenging task, requiring harmonious integration of reference objects into source videos. However, existing methods struggle when references exhibit severe stylistic do…

Video GenerationStyle Transfer

CrimEdit: Controllable Editing for Counterfactual Object Removal, Insertion, and Movement

2025-09-28 · Boseong Jeon, Junghyuk Lee, Jimin Park, Kwanyoung Kim 외 arxiv

Recent works on object removal and insertion have enhanced their performance by handling object effects such as shadows and reflections, using diffusion models trained on counterfactual datasets. However, the performance…

Loiter UAV Reinsertion Guidance for Fixed-wing UAV Corridors

2026-05-13 · Pradeep J, Kedarisetty Siddhardha, Ashwini Ratnoo arxiv

This paper considers fixed-wing unmanned aerial vehicle (UAV) corridors comprising a main lane, a circular loiter lane for managing traffic congestion, and transit lanes connecting the two. In particular, we address the …