paper-with-me

홈 › Papers

IP-FaceDiff: Identity-Preserving Facial Video Editing with Diffusion

2025-01-13 · Tharun Anand, Aryan Garg, Kaushik Mitra

Facial video editing has become increasingly important for content creators, enabling the manipulation of facial expressions and attributes. However, existing models encounter challenges such as poor editing quality, high computational costs and difficulties in preserving facial identity across diverse edits. Additionally, these models are often constrained to editing predefined facial attributes, limiting their flexibility to diverse editing prompts. To address these challenges, we propose a novel facial video editing framework that leverages the rich latent space of pre-trained text-to-image (T2I) diffusion models and fine-tune them specifically for facial video editing tasks. Our approach introduces a targeted fine-tuning scheme that enables high quality, localized, text-driven edits while ensuring identity preservation across video frames. Additionally, by using pre-trained T2I models during inference, our approach significantly reduces editing time by 80%, while maintaining temporal consistency throughout the video sequence. We evaluate the effectiveness of our approach through extensive testing across a wide range of challenging scenarios, including varying head poses, complex action sequences, and diverse facial expressions. Our method consistently outperforms existing techniques, demonstrating superior performance across a broad set of metrics and benchmarks.

📄 PDF Abstract BibTeX arXiv:2501.07530

Code (0)

등록된 구현이 없습니다.

Tasks

Video Editing

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…
SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

NegFaceDiff: The Power of Negative Context in Identity-Conditioned Diffusion for Synthetic Face Generation

2025-08-13 · Eduarda Caldeira, Naser Damer, Fadi Boutros arxiv

The use of synthetic data as an alternative to authentic datasets in face recognition (FR) development has gained significant attention, addressing privacy, ethical, and practical concerns associated with collecting and …

Face Recognition

Cartoonish sketch-based face editing in videos using identity deformation transfer

2017-03-25 · Long Zhao, Fangda Han, Xi Peng, Xun Zhang 외

We address the problem of using hand-drawn sketches to create exaggerated deformations to faces in videos, such as enlarging the shape or modifying the position of eyes or mouth. This task is formulated as a 3D face mode…

Face Model

A Latent Transformer for Disentangled Face Editing in Images and Videos

2021-06-22 · ICCV 2021 10 · Xu Yao, Alasdair Newson, Yann Gousseau, Pierre Hellier

High quality facial image editing is a challenging problem in the movie post-production industry, requiring a high degree of control and identity preservation. Previous works that attempt to tackle this problem may suffe…

AttributeDisentanglement

DiffMagicFace: Identity Consistent Facial Editing of Real Videos

2026-04-15 · Huanghao Yin, Shenkun Xu, Kanle Shi, Junhai Yong 외 arxiv

Text-conditioned image editing has greatly benefitted from the advancements in Image Diffusion Models. However, extending these techniques to facial video editing introduces challenges in preserving facial identity throu…

Image Editing

PERSE: Personalized 3D Generative Avatars from A Single Portrait

2024-12-30 · CVPR 2025 1 · Hyunsoo Cha, Inhee Lee, Hanbyul Joo

We present PERSE, a method for building an animatable personalized generative avatar from a reference portrait. Our avatar model enables facial attribute editing in a continuous and disentangled latent space to control e…

Attribute