paper-with-me

Papers

AnyScene: Customized Image Synthesis with Composited Foreground

2024-01-01 · CVPR 2024 1 · Ruidong Chen, Lanjun Wang, Weizhi Nie, Yongdong Zhang, An-An Liu

Recent advancements in text-to-image technology have significantly advanced the field of image customization. Among various applications the task of customizing diverse scenes for user-specified composited elements holds great application value but has not been extensively explored. Addressing this gap we propose AnyScene a specialized framework designed to create varied scenes from composited foreground using textual prompts. AnyScene addresses the primary challenges inherent in existing methods particularly scene disharmony due to a lack of foreground semantic understanding and distortion of foreground elements. Specifically we develop a foreground injection module that guides a pre-trained diffusion model to generate cohesive scenes in visual harmony with the provided foreground. To enhance robust generation we implement a layout control strategy that prevents distortions of foreground elements. Furthermore an efficient image blending mechanism seamlessly reintegrates foreground details into the generated scenes producing outputs with overall visual harmony and precise foreground details. In addition we propose a new benchmark and a series of quantitative metrics to evaluate this proposed image customization task. Extensive experimental results demonstrate the effectiveness of AnyScene which confirms its potential in various applications.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Image Generation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

AnyScene: Towards Highly Controllable Driving Scene Generation at Anywhere and Beyond

2026-05-25 · Haiming Zhang, Junfei Zhou, Feng Jiang, Jingzhong Li 외 arxiv

Generating high-fidelity and controllable synthetic data is critical for advancing end-to-end autonomous driving, particularly for addressing the long tail of rare safety-critical scenarios. Existing occupancy-guided met…

Autonomous Driving3D ReconstructionScene GenerationVideo Generation

DALL-E for Detection: Language-driven Compositional Image Synthesis for Object Detection

2022-06-20 · Yunhao Ge, Jiashu Xu, Brian Nlong Zhao, Neel Joshi 외

We propose a new paradigm to automatically generate training data with accurate labels at scale using the text-toimage synthesis frameworks (e.g., DALL-E, Stable Diffusion, etc.). The proposed approach decouples training…

Image CaptioningImage GenerationObjectobject-detection+2

Infusing Definiteness into Randomness: Rethinking Composition Styles for Deep Image Matting

2022-12-27 · Zixuan Ye, Yutong Dai, Chaoyi Hong, Zhiguo Cao 외

We study the composition style in deep image matting, a notion that characterizes a data generation flow on how to exploit limited foregrounds and random backgrounds to form a training dataset. Prior art executes this fl…

Image MattingTriplet

OCONet: Image Extrapolation by Object Completion

2021-06-19 · CVPR 2021 1 · Richard Strong Bowen, Huiwen Chang, Charles Herrmann, Piotr Teterwak 외

Image extrapolation extends an input image beyond the originally-captured field of view. Existing methods struggle to extrapolate images with salient objects in the foreground or are limited to very specific objects …

DecoderObject

Intrinsic Harmonization for Illumination-Aware Compositing

2023-12-06 · Chris Careaga, S. Mahdi H. Miangoleh, Yağız Aksoy

Despite significant advancements in network-based image harmonization techniques, there still exists a domain disparity between typical training pairs and real-world composites encountered during inference. Most existing…

Image HarmonizationImage Relighting