paper-with-me

홈 › Papers

Stencil: Subject-Driven Generation with Context Guidance

2025-09-21 · Gordon Chen, Ziqi Huang, Cheston Tan, Ziwei Liu arxiv

Recent text-to-image diffusion models can generate striking visuals from text prompts, but they often fail to maintain subject consistency across generations and contexts. One major limitation of current fine-tuning approaches is the inherent trade-off between quality and efficiency. Fine-tuning large models improves fidelity but is computationally expensive, while fine-tuning lightweight models improves efficiency but compromises image fidelity. Moreover, fine-tuning pre-trained models on a small set of images of the subject can damage the existing priors, resulting in suboptimal results. To this end, we present Stencil, a novel framework that jointly employs two diffusion models during inference. Stencil efficiently fine-tunes a lightweight model on images of the subject, while a large frozen pre-trained model provides contextual guidance during inference, injecting rich priors to enhance generation with minimal overhead. Stencil excels at generating high-fidelity, novel renditions of the subject in less than a minute, delivering state-of-the-art performance and setting a new benchmark in subject-driven generation.

📄 PDF Abstract BibTeX arXiv:2509.17120

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Training-free Subject-Enhanced Attention Guidance for Compositional Text-to-image Generation

2024-05-11 · Shengyuan Liu, Bo wang, Ye Ma, Te Yang 외

Existing subject-driven text-to-image generation models suffer from tedious fine-tuning steps and struggle to maintain both text-image alignment and subject fidelity. For generating compositional subjects, it often encou…

AttributeImage GenerationText to Image GenerationText-to-Image Generation

Evotype: Towards the Evolution of Type Stencils

2018-06-26 · Tiago Martins, João Correia, Ernesto Costa, Penousal Machado

Typefaces are an essential resource employed by graphic designers. The increasing demand for innovative type design work increases the need for good technological means to assist the designer in the creation of a typefac…

Vocal Bursts Type Prediction

STENCIL-NET: Data-driven solution-adaptive discretization of partial differential equations

2021-01-15 · Suryanarayana Maddu, Dominik Sturm, Bevan L. Cheeseman, Christian L. Müller 외

Numerical methods for approximately solving partial differential equations (PDE) are at the core of scientific computing. Often, this requires high-resolution or adaptive discretization grids to capture relevant spatio-t…

Token-Level Privacy in Large Language Models

2025-03-05 · Re'em Harel, Niv Gilboa, Yuval Pinter

The use of language models as remote services requires transmitting private information to external providers, raising significant privacy concerns. This process not only risks exposing sensitive data to untrusted servic…

Privacy PreservingSemantic SimilaritySemantic Textual Similarity

ImPoster: Text and Frequency Guidance for Subject Driven Action Personalization using Diffusion Models

2024-09-24 · Divya Kothandaraman, Kuldeep Kulkarni, Sumit Shekhar, Balaji Vasan Srinivasan 외

We present ImPoster, a novel algorithm for generating a target image of a 'source' subject performing a 'driving' action. The inputs to our algorithm are a single pair of a source image with the subject that we wish to e…

Denoising