paper-with-me

Papers

DreamID: High-Fidelity and Fast diffusion-based Face Swapping via Triplet ID Group Learning

2025-04-20 · Fulong Ye, Miao Hua, Pengze Zhang, Xinghui Li, Qichao Sun, Songtao Zhao, Qian He, Xinglong Wu

In this paper, we introduce DreamID, a diffusion-based face swapping model that achieves high levels of ID similarity, attribute preservation, image fidelity, and fast inference speed. Unlike the typical face swapping training process, which often relies on implicit supervision and struggles to achieve satisfactory results. DreamID establishes explicit supervision for face swapping by constructing Triplet ID Group data, significantly enhancing identity similarity and attribute preservation. The iterative nature of diffusion models poses challenges for utilizing efficient image-space loss functions, as performing time-consuming multi-step sampling to obtain the generated image during training is impractical. To address this issue, we leverage the accelerated diffusion model SD Turbo, reducing the inference steps to a single iteration, enabling efficient pixel-level end-to-end training with explicit Triplet ID Group supervision. Additionally, we propose an improved diffusion-based model architecture comprising SwapNet, FaceNet, and ID Adapter. This robust architecture fully unlocks the power of the Triplet ID Group explicit supervision. Finally, to further extend our method, we explicitly modify the Triplet ID Group data during training to fine-tune and preserve specific attributes, such as glasses and face shape. Extensive experiments demonstrate that DreamID outperforms state-of-the-art methods in terms of identity similarity, pose and expression preservation, and image fidelity. Overall, DreamID achieves high-quality face swapping results at 512*512 resolution in just 0.6 seconds and performs exceptionally well in challenging scenarios such as complex lighting, large angles, and occlusions.

📄 PDF Abstract BibTeX arXiv:2504.14509

Code (1)

superhero-7/DreamID 공식 구현

Tasks

AttributeFace SwappingTriplet

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…
Adapter 설명 없음

Similar Papers 제목 키워드 기반

DreamID-V:Bridging the Image-to-Video Gap for High-Fidelity Face Swapping via Diffusion Transformer

2026-01-04 · Xu Guo, Fulong Ye, Xinghui Li, Pengqi Tu 외 arxiv

Video Face Swapping (VFS) requires seamlessly injecting a source identity into a target video while meticulously preserving the original pose, expression, lighting, background, and dynamic information. Existing methods s…

Reinforcement LearningFace Swapping

DreamID-Omni: Unified Framework for Controllable Human-Centric Audio-Video Generation

2026-02-12 · Xu Guo, Fulong Ye, Qichao Sun, Liyang Chen 외 arxiv

Recent advancements in foundation models have revolutionized joint audio-video generation. However, existing approaches typically treat human-centric tasks including reference-based audio-video generation (R2AV), video e…

Video Generation

DreamIdentity: Improved Editability for Efficient Face-identity Preserved Image Generation

2023-07-01 · Zhuowei Chen, Shancheng Fang, Wei Liu, Qian He 외

While large-scale pre-trained text-to-image models can synthesize diverse and high-quality human-centric images, an intractable problem is how to preserve the face identity for conditioned face images. Existing methods e…

Image Generation

FastFace: Tuning Identity Preservation in Distilled Diffusion via Guidance and Attention

2025-05-27 · Sergey Karpukhin, Vadim Titov, Andrey Kuznetsov, Aibek Alanov

In latest years plethora of identity-preserving adapters for a personalized generation with diffusion models have been released. Their main disadvantage is that they are dominantly trained jointly with base diffusion mod…

LDPM: Towards undersampled MRI reconstruction with MR-VAE and Latent Diffusion Prior

2024-11-05 · Xingjian Tang, Jingwei Guan, Linge Li, Ran Shi 외

Diffusion models, as powerful generative models, have found a wide range of applications and shown great potential in solving image reconstruction problems. Some works attempted to solve MRI reconstruction with diffusion…

Image ReconstructionMRI Reconstruction