paper-with-me

Papers

Personalize Anything for Free with Diffusion Transformer

2025-03-16 · Haoran Feng, Zehuan Huang, Lin Li, Hairong Lv, Lu Sheng

Personalized image generation aims to produce images of user-specified concepts while enabling flexible editing. Recent training-free approaches, while exhibit higher computational efficiency than training-based methods, struggle with identity preservation, applicability, and compatibility with diffusion transformers (DiTs). In this paper, we uncover the untapped potential of DiT, where simply replacing denoising tokens with those of a reference subject achieves zero-shot subject reconstruction. This simple yet effective feature injection technique unlocks diverse scenarios, from personalization to image editing. Building upon this observation, we propose \textbf{Personalize Anything}, a training-free framework that achieves personalized image generation in DiT through: 1) timestep-adaptive token replacement that enforces subject consistency via early-stage injection and enhances flexibility through late-stage regularization, and 2) patch perturbation strategies to boost structural diversity. Our method seamlessly supports layout-guided generation, multi-subject personalization, and mask-controlled editing. Evaluations demonstrate state-of-the-art performance in identity preservation and versatility. Our work establishes new insights into DiTs while delivering a practical paradigm for efficient personalization.

📄 PDF Abstract BibTeX arXiv:2503.12590

Code (0)

등록된 구현이 없습니다.

Tasks

Computational EfficiencyDenoisingDiversityImage GenerationPersonalized Image Generation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Personalize Segment Anything Model with One Shot

2023-05-04 · Renrui Zhang, Zhengkai Jiang, Ziyu Guo, Shilin Yan 외

Driven by large-data pre-training, Segment Anything Model (SAM) has been demonstrated as a powerful and promptable framework, revolutionizing the segmentation models. Despite the generality, customizing SAM for specific …

Image GenerationmodelPersonalized SegmentationSegmentation+5

360Anything: Geometry-Free Lifting of Images and Videos to 360°

2026-01-22 · Ziyi Wu, Daniel Watson, Andrea Tagliasacchi, David J. Fleet 외 arxiv

Lifting perspective images and videos to 360° panoramas enables immersive 3D world generation. Existing approaches often rely on explicit geometric alignment between the perspective and the equirectangular projection (ER…

Generate Anything Anywhere in Any Scene

2023-06-29 · Yuheng Li, Haotian Liu, Yangming Wen, Yong Jae Lee

Text-to-image diffusion models have attracted considerable interest due to their wide applicability across diverse fields. However, challenges persist in creating controllable models for personalized object generation. I…

Data AugmentationObject

RigAnything: Template-Free Autoregressive Rigging for Diverse 3D Assets

2025-02-13 · Isabella Liu, Zhan Xu, Wang Yifan, Hao Tan 외

We present RigAnything, a novel autoregressive transformer-based model, which makes 3D assets rig-ready by probabilistically generating joints, skeleton topologies, and assigning skinning weights in a template-free manne…

LoRAShop: Training-Free Multi-Concept Image Generation and Editing with Rectified Flow Transformers

2025-05-29 · Yusuf Dalva, Hidir Yesiltepe, Pinar Yanardag

We introduce LoRAShop, the first framework for multi-concept image editing with LoRA models. LoRAShop builds on a key observation about the feature interaction patterns inside Flux-style diffusion transformers: concept-s…

DenoisingImage GenerationVisual Storytelling