paper-with-me

홈 › Papers

PLACID: Identity-Preserving Multi-Object Compositing via Video Diffusion with Synthetic Trajectories

2026-01-30 · Gemma Canet Tarrés, Manel Baradad, Francesc Moreno-Noguer, Yumeng Li arxiv

Recent advances in generative AI have dramatically improved photorealistic image synthesis, yet they fall short for studio-level multi-object compositing. This task demands simultaneous (i) near-perfect preservation of each item's identity, (ii) precise background and color fidelity, (iii) layout and design elements control, and (iv) complete, appealing displays showcasing all objects. However, current state-of-the-art models often alter object details, omit or duplicate objects, and produce layouts with incorrect relative sizing or inconsistent item presentations. To bridge this gap, we introduce PLACID, a framework that transforms a collection of object images into an appealing multi-object composite. Our approach makes two main contributions. First, we leverage a pretrained image-to-video (I2V) diffusion model with text control to preserve objects consistency, identities, and background details by exploiting temporal priors from videos. Second, we propose a novel data curation strategy that generates synthetic sequences where randomly placed objects smoothly move to their target positions. This synthetic data aligns with the video model's temporal priors during training. At inference, objects initialized at random positions consistently converge into coherent layouts guided by text, with the final frame serving as the composite image. Extensive quantitative evaluations and user studies demonstrate that PLACID surpasses state-of-the-art methods in multi-object compositing, achieving superior identity, background, and color preservation, with less omitted objects and visually appealing results.

📄 PDF Abstract BibTeX arXiv:2602.00267

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

IMPRINT: Generative Object Compositing by Learning Identity-Preserving Representation

2024-03-15 · CVPR 2024 1 · Yizhi Song, Zhifei Zhang, Zhe Lin, Scott Cohen 외

Generative object compositing emerges as a promising new avenue for compositional image editing. However, the requirement of object identity preservation poses a significant challenge, limiting practical usage of most ex…

Object

Towards Design Compositing

2026-04-16 · Abhinav Mahajan, Abhikhya Tripathy, Sudeeksha Reddy Pala, Vaibhav Methi 외 arxiv

Graphic design creation involves harmoniously assembling multimodal components such as images, text, logos, and other visual assets collected from diverse sources, into a visually-appealing and cohesive design. Recent me…

Chameleon: Style-Content Disentangled Framework for Cross-Domain Object Compositing

2026-05-31 · Sukhun Ko, Soo Ye Kim, Jihyong Oh arxiv

Image compositing aims to seamlessly insert a foreground object into a background image, and recent advances in diffusion models have significantly enhanced the quality, especially when the foreground and background imag…

Contrastive Learning

PlacidDreamer: Advancing Harmony in Text-to-3D Generation

2024-07-19 · Shuo Huang, Shikun Sun, Zixuan Wang, Xiaoyu Qin 외

Recently, text-to-3D generation has attracted significant attention, resulting in notable performance enhancements. Previous methods utilize end-to-end 3D generation models to initialize 3D Gaussians, multi-view diffusio…

3D GenerationText to 3D

CatalogStitch: Dimension-Aware and Occlusion-Preserving Object Compositing for Catalog Image Generation

2026-04-10 · Sanyam Jain, Pragya Kandari, Manit Singhal, He Zhang 외 arxiv

Generative object compositing methods have shown remarkable ability to seamlessly insert objects into scenes. However, when applied to real-world catalog image generation, these methods require tedious manual interventio…

Image Generation