paper-with-me

Papers

StyleForge: Enhancing Text-to-Image Synthesis for Any Artistic Styles with Dual Binding

2024-04-08 · Junseo Park, Beomseok Ko, Hyeryung Jang

Recent advancements in text-to-image models, such as Stable Diffusion, have showcased their ability to create visual images from natural language prompts. However, existing methods like DreamBooth struggle with capturing arbitrary art styles due to the abstract and multifaceted nature of stylistic attributes. We introduce Single-StyleForge, a novel approach for personalized text-to-image synthesis across diverse artistic styles. Using approximately 15 to 20 images of the target style, Single-StyleForge establishes a foundational binding of a unique token identifier with a broad range of attributes of the target style. Additionally, auxiliary images are incorporated for dual binding that guides the consistent representation of crucial elements such as people within the target style. Furthermore, we present Multi-StyleForge, which enhances image quality and text alignment by binding multiple tokens to partial style attributes. Experimental evaluations across six distinct artistic styles demonstrate significant improvements in image quality and perceptual fidelity, as measured by FID, KID, and CLIP scores.

📄 PDF Abstract BibTeX arXiv:2404.05256

Code (0)

등록된 구현이 없습니다.

Tasks

Image Generation

Methods 이 논문이 사용한 방법론

CLIP Contrastive Language-Image Pre-training (CLIP), consisting of a simplified version of ConVIRT trained from scratch, is an efficient method of image representation learning…
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Artistic Glyph Image Synthesis via One-Stage Few-Shot Learning

2019-10-11 · Yue Gao, Yuan Guo, Zhouhui Lian, Yingmin Tang 외

Automatic generation of artistic glyph images is a challenging task that attracts many research interests. Previous methods either are specifically designed for shape synthesis or focus on texture transfer. In this paper…

Few-Shot LearningImage Generation

StyleBlend: Enhancing Style-Specific Content Creation in Text-to-Image Diffusion Models

2025-02-13 · Zichong Chen, Shijin Wang, Yang Zhou

Synthesizing visually impressive images that seamlessly align both text prompts and specific artistic styles remains a significant challenge in Text-to-Image (T2I) diffusion models. This paper introduces StyleBlend, a me…

StyleForge: Indoor Furniture Styling by Counterfactual Reasoning in a Hypergraph Field

2026-08-03 · Lingwei Dang, Shishuo Shang, Pan Liu, Jiajia Cheng 외 hf

Fixed-layout indoor furniture styling requires selecting assets that form a coherent room without changing the prescribed furniture categories, positions, orientations, or scales. Existing approaches typically retrieve e…

StyleCLIPDraw: Coupling Content and Style in Text-to-Drawing Synthesis

2021-11-04 · Peter Schaldenbrand, Zhixuan Liu, Jean Oh

Generating images that fit a given text description using machine learning has improved greatly with the release of technologies such as the CLIP image-text encoder model; however, current methods lack artistic control o…

Style Transfer

WordArt Designer: User-Driven Artistic Typography Synthesis using Large Language Models

2023-10-20 · Jun-Yan He, Zhi-Qi Cheng, Chenyang Li, Jingdong Sun 외

This paper introduces WordArt Designer, a user-driven framework for artistic typography synthesis, relying on the Large Language Model (LLM). The system incorporates four key modules: the LLM Engine, SemTypo, StyTypo, an…

Language ModelingLanguage ModellingLarge Language Model