paper-with-me

Papers

Universal Few-Shot Spatial Control for Diffusion Models

2025-09-09 · Kiet T. Nguyen, Chanhyuk Lee, Donggyun Kim, Dong Hoon Lee, Seunghoon Hong arxiv

Spatial conditioning in pretrained text-to-image diffusion models has significantly improved fine-grained control over the structure of generated images. However, existing control adapters exhibit limited adaptability and incur high training costs when encountering novel spatial control conditions that differ substantially from the training tasks. To address this limitation, we propose Universal Few-Shot Control (UFC), a versatile few-shot control adapter capable of generalizing to novel spatial conditions. Given a few image-condition pairs of an unseen task and a query condition, UFC leverages the analogy between query and support conditions to construct task-specific control features, instantiated by a matching mechanism and an update on a small set of task-specific parameters. Experiments on six novel spatial control tasks show that UFC, fine-tuned with only 30 annotated examples of novel tasks, achieves fine-grained control consistent with the spatial conditions. Notably, when fine-tuned with 0.1% of the full training data, UFC achieves competitive performance with the fully supervised baselines in various control tasks. We also show that UFC is applicable agnostically to various diffusion backbones and demonstrate its effectiveness on both UNet and DiT architectures. Code is available at https://github.com/kietngt00/UFC.

📄 PDF Abstract BibTeX arXiv:2509.07530

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Layered Rendering Diffusion Model for Controllable Zero-Shot Image Synthesis

2023-11-30 · Zipeng Qi, Guoxi Huang, Chenyang Liu, Fei Ye

This paper introduces innovative solutions to enhance spatial controllability in diffusion models reliant on text queries. We first introduce vision guidance as a foundational spatial cue within the perturbed distributio…

DenoisingImage Generation

Motion-Zero: Zero-Shot Moving Object Control Framework for Diffusion-Based Video Generation

2024-01-18 · Changgu Chen, Junwei Shu, Gaoqi He, Changbo Wang 외

Recent large-scale pre-trained diffusion models have demonstrated a powerful generative ability to produce high-quality videos from detailed text descriptions. However, exerting control over the motion of objects in vide…

DenoisingPositionVideo Generation

Ctrl-Adapter: An Efficient and Versatile Framework for Adapting Diverse Controls to Any Diffusion Model

2024-04-15 · Han Lin, Jaemin Cho, Abhay Zala, Mohit Bansal

ControlNets are widely used for adding spatial control to text-to-image diffusion models with different conditions, such as depth maps, scribbles/sketches, and human poses. However, when it comes to controllable video ge…

GPUImage GenerationStyle TransferVideo Editing+2

Universal Diffusion-Based Probabilistic Downscaling

2026-02-12 · Roberto Molinaro, Niall Siegenheim, Henry Martin, Mark Frey 외 arxiv

We introduce a universal diffusion-based downscaling framework that lifts deterministic low-resolution weather forecasts into probabilistic high-resolution predictions without any model-specific fine-tuning. A single con…

Weather Forecasting

VidEdit: Zero-Shot and Spatially Aware Text-Driven Video Editing

2023-06-14 · Paul Couairon, Clément Rambour, Jean-Emmanuel Haugeard, Nicolas Thome

Recently, diffusion-based generative models have achieved remarkable success for image generation and edition. However, existing diffusion-based video editing approaches lack the ability to offer precise control over gen…

Image GenerationVideo Editing