paper-with-me

Papers

CTRLorALTer: Conditional LoRAdapter for Efficient 0-Shot Control & Altering of T2I Models

2024-05-13 · Nick Stracke, Stefan Andreas Baumann, Joshua M. Susskind, Miguel Angel Bautista, Björn Ommer

Text-to-image generative models have become a prominent and powerful tool that excels at generating high-resolution realistic images. However, guiding the generative process of these models to consider detailed forms of conditioning reflecting style and/or structure information remains an open problem. In this paper, we present LoRAdapter, an approach that unifies both style and structure conditioning under the same formulation using a novel conditional LoRA block that enables zero-shot control. LoRAdapter is an efficient, powerful, and architecture-agnostic approach to condition text-to-image diffusion models, which enables fine-grained control conditioning during generation and outperforms recent state-of-the-art approaches.

📄 PDF Abstract BibTeX arXiv:2405.07913

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Accelerating Conditional Prompt Learning via Masked Image Modeling for Vision-Language Models

2025-08-07 · Phuoc-Nguyen Bui, Khanh-Binh Nguyen, Hyunseung Choo arxiv

Vision-language models (VLMs) like CLIP excel in zero-shot learning but often require resource-intensive training to adapt to new tasks. Prompt learning techniques, such as CoOp and CoCoOp, offer efficient adaptation but…

Zero-Shot Learning

Where Is My Spot? Few-Shot Image Generation via Latent Subspace Optimization

2023-01-01 · CVPR 2023 1 · Chenxi Zheng, Bangzhen Liu, Huaidong Zhang, Xuemiao Xu 외

Image generation relies on massive training data that can hardly produce diverse images of an unseen category according to a few examples. In this paper, we address this dilemma by projecting sparse few-shot samples …

Image Generation

Controlling Formality in Low-Resource NMT with Domain Adaptation and Re-Ranking: SLT-CDT-UoS at IWSLT2022

2022-05-12 · IWSLT (ACL) 2022 5 · Sebastian T. Vincent, Loïc Barrault, Carolina Scarton

This paper describes the SLT-CDT-UoS group's submission to the first Special Task on Formality Control for Spoken Language Translation, part of the IWSLT 2022 Evaluation Campaign. Our efforts were split between two front…

Domain AdaptationLow Resource NMTNMTRe-Ranking+2

Controlled and Conditional Text to Image Generation with Diffusion Prior

2023-02-23 · Pranav Aggarwal, Hareesh Ravi, Naveen Marri, Sachin Kelkar 외

Denoising Diffusion models have shown remarkable performance in generating diverse, high quality images from text. Numerous techniques have been proposed on top of or in alignment with models like Stable Diffusion and Im…

DecoderDenoisingImage GenerationPrompt Engineering+2

Zero-shot generalization of transformer neural operators to larger domains

2026-06-12 · Armand de Villeroché, Sibo Cheng, Vincent Le Guen, Marc Bocquet 외 arxiv

Transformer-based neural operators have shown remarkable performance for approximating solution operators of partial differential equations on complex geometries. However, existing approaches implicitly assume a fixed do…

Zero-shot Generalization