paper-with-me

홈 › Papers

Sketch-Guided Scene Image Generation

2024-07-09 · Tianyu Zhang, Xiaoxuan Xie, Xusheng Du, Haoran Xie

Text-to-image models are showcasing the impressive ability to create high-quality and diverse generative images. Nevertheless, the transition from freehand sketches to complex scene images remains challenging using diffusion models. In this study, we propose a novel sketch-guided scene image generation framework, decomposing the task of scene image scene generation from sketch inputs into object-level cross-domain generation and scene-level image construction. We employ pre-trained diffusion models to convert each single object drawing into an image of the object, inferring additional details while maintaining the sparse sketch structure. In order to maintain the conceptual fidelity of the foreground during scene generation, we invert the visual features of object images into identity embeddings for scene generation. In scene-level image construction, we generate the latent representation of the scene image using the separated background prompts, and then blend the generated foreground objects according to the layout of the sketch input. To ensure the foreground objects' details remain unchanged while naturally composing the scene image, we infer the scene image on the blended latent representation using a global prompt that includes the trained identity tokens. Through qualitative and quantitative experiments, we demonstrate the ability of the proposed approach to generate scene images from hand-drawn sketches surpasses the state-of-the-art approaches.

📄 PDF Abstract BibTeX arXiv:2407.06469

Code (0)

등록된 구현이 없습니다.

Tasks

Image GenerationObjectScene Generation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

MaskSketch: Unpaired Structure-guided Masked Image Generation

2023-02-10 · CVPR 2023 1 · Dina Bashkirova, Jose Lezama, Kihyuk Sohn, Kate Saenko 외

Recent conditional image generation methods produce images of remarkable diversity, fidelity and realism. However, the majority of these methods allow conditioning only on labels or text prompts, which limits their level…

Conditional Image GenerationDiversityImage GenerationImage-to-Image Translation+2

Text-Guided Scene Sketch-to-Photo Synthesis

2023-02-14 · AprilPyone MaungMaung, Makoto Shing, Kentaro Mitsui, Kei Sawada 외

We propose a method for scene-level sketch-to-photo synthesis with text guidance. Although object-level sketch-to-photo synthesis has been widely studied, whole-scene synthesis is still challenging without reference phot…

Self-Supervised Learning

SketchyCOCO: Image Generation from Freehand Scene Sketches

2020-03-05 · CVPR 2020 6 · Chengying Gao, Qi Liu, Qi Xu, Li-Min Wang 외

We introduce the first method for automatic image generation from scene-level freehand sketches. Our model allows for controllable image generation by specifying the synthesis goal via freehand sketches. The key contribu…

AttributeGenerative Adversarial NetworkImage GenerationObject+1

SketchTriplet: Self-Supervised Scenarized Sketch-Text-Image Triplet Generation

2024-05-29 · Zhenbei Wu, Qiang Wang, Jie Yang

The scarcity of free-hand sketch presents a challenging problem. Despite the emergence of some large-scale sketch datasets, these datasets primarily consist of sketches at the single-object level. There continues to be a…

Image GenerationImage RetrievalSketch-Based Image RetrievalTriplet

Sketch2NeRF: Multi-view Sketch-guided Text-to-3D Generation

2024-01-25 · Minglin Chen, Weihao Yuan, Yukun Wang, Zhe Sheng 외

Recently, text-to-3D approaches have achieved high-fidelity 3D content generation using text description. However, the generated objects are stochastic and lack fine-grained control. Sketches provide a cheap approach to …

3D GenerationNeRFText to 3D