paper-with-me

홈 › Papers

A Style is Worth One Code: Unlocking Code-to-Style Image Generation with Discrete Style Space

2025-11-13 · Huijie Liu, Shuhao Cui, Haoxiang Cao, Shuai Ma, Kai Wu, Guoliang Kang arxiv

Innovative visual stylization is a cornerstone of artistic creation, yet generating novel and consistent visual styles remains a significant challenge. Existing generative approaches typically rely on lengthy textual prompts, reference images, or parameter-efficient fine-tuning to guide style-aware image generation, but often struggle with style consistency, limited creativity, and complex style representations. In this paper, we affirm that a style is worth one numerical code by introducing the novel task, code-to-style image generation, which produces images with novel, consistent visual styles conditioned solely on a numerical style code. To date, this field has only been primarily explored by the industry (e.g., Midjourney), with no open-source research from the academic community. To fill this gap, we propose CoTyle, the first open-source method for this task. Specifically, we first train a discrete style codebook from a collection of images to extract style embeddings. These embeddings serve as conditions for a text-to-image diffusion model (T2I-DM) to generate stylistic images. Subsequently, we train an autoregressive style generator on the discrete style embeddings to model their distribution, allowing the synthesis of novel style embeddings. During inference, a numerical style code is mapped to a unique style embedding by the style generator, and this embedding guides the T2I-DM to generate images in the corresponding style. Unlike existing methods, our method offers unparalleled simplicity and diversity, unlocking a vast space of reproducible styles from minimal input. Extensive experiments validate that CoTyle effectively turns a numerical code into a style controller, demonstrating a style is worth one code.

📄 PDF Abstract BibTeX arXiv:2511.10555

Code (0)

등록된 구현이 없습니다.

Tasks

parameter-efficient fine-tuningImage Generation

Similar Papers 제목 키워드 기반

Unlocking Fine-Grained and Within-Utterance Speaking Style Control in Prompt-Based Text-to-Speech Models

2026-04-09 · Jaehoon Kang, Yejin Lee, Yoonji Park, Kyuhong Shim arxiv

While prompt-based text-to-speech (TTS) models enable natural language-driven speaking style control, they often provide limited fine-grained control and apply a single global style across an utterance. This restricts pr…

StyleMe3D: Stylization with Disentangled Priors by Multiple Encoders on 3D Gaussians

2025-04-21 · Cailin Zhuang, Yaoqi Hu, Xuanyang Zhang, Wei Cheng 외

3D Gaussian Splatting (3DGS) excels in photorealistic scene reconstruction but struggles with stylized scenarios (e.g., cartoons, games) due to fragmented textures, semantic misalignment, and limited adaptability to abst…

3DGSNeRFStyle Transfer

Speech-Worthy Alignment for Japanese SpeechLLMs via Direct Preference Optimization

2026-03-13 · Mengjie Zhao, Lianbo Liu, Yusuke Fujita, Hao Shi 외 arxiv

SpeechLLMs typically combine ASR-trained encoders with text-based LLM backbones, leading them to inherit written-style output patterns unsuitable for text-to-speech synthesis. This mismatch is particularly pronounced in …

Text-To-Speech Synthesis

EchoStyle: Unlocking High-Fidelity Video Stylization with Reverse Data Synthesis

2026-06-24 · Huaqiu Li, Jiahao Wang, Sijia Cai, Hualian Sheng 외 arxiv

While image stylization has been studied extensively, video stylization remains a critical and largely unsolved challenge in the field of intelligent content creation. Existing methods, usually utilizing a reference imag…

Exploiting Style and Attention in Real-World Super-Resolution

2019-12-21 · Xin Ma, Yi Li, Huaibo Huang, Mandi Luo 외

Real-world image super-resolution (SR) is a challenging image translation problem. Low-resolution (LR) images are often generated by various unknown transformations rather than by applying simple bilinear down-sampling o…

Image Super-ResolutionMutual Information EstimationSuper-Resolution