paper-with-me

Papers

Preserving Source Video Realism: High-Fidelity Face Swapping for Cinematic Quality

2025-12-08 · Zekai Luo, Zongze Du, Zhouhang Zhu, Hao Zhong, Muzhi Zhu, Wen Wang, Yuling Xi, Chenchen Jing, Hao Chen, Chunhua Shen arxiv

Video face swapping is crucial in film and entertainment production, where achieving high fidelity and temporal consistency over long and complex video sequences remains a significant challenge. Inspired by recent advances in reference-guided image editing, we explore whether rich visual attributes from source videos can be similarly leveraged to enhance both fidelity and temporal coherence in video face swapping. Building on this insight, this work presents LivingSwap, the first video reference guided face swapping model. Our approach employs keyframes as conditioning signals to inject the target identity, enabling flexible and controllable editing. By combining keyframe conditioning with video reference guidance, the model performs temporal stitching to ensure stable identity preservation and high-fidelity reconstruction across long video sequences. To address the scarcity of data for reference-guided training, we construct a paired face-swapping dataset, Face2Face, and further reverse the data pairs to ensure reliable ground-truth supervision. Extensive experiments demonstrate that our method achieves state-of-the-art results, seamlessly integrating the target identity with the source video's expressions, lighting, and motion, while significantly reducing manual effort in production workflows. Project webpage: https://aim-uofa.github.io/LivingSwap

📄 PDF Abstract BibTeX arXiv:2512.07951

Code (0)

등록된 구현이 없습니다.

Tasks

Face SwappingImage Editing

Similar Papers 제목 키워드 기반

AVE-Compass: Towards Holistic Evaluation for Audio-Video Editing Abilities

2026-07-17 · Yuqing Wen, Yukai Huang, Qianqian Xie, Jiangtao Wu 외 hf

While instruction-based video editing has advanced rapidly, real-world videos contain tightly coupled audio and visual signals, and editing one modality often requires coordinated changes in the other. Existing benchmark…

Instruction Following

TIGER: Taming Identity, Geometry, and Generative Priors for High-Quality Face Video Restoration

2026-06-23 · Yang Zhou, Wenxue Li, Peng Zhang, Yifei Chen 외 arxiv

Face Video Restoration (FVR) aims to recover high-fidelity facial videos from degraded input while preserving identity and semantic consistency across frames. Existing methods often struggle to simultaneously address thr…

Video RestorationVideo Generation

Advancing Open-source World Models

2026-01-28 · Robbyant Team, Zelin Gao, Qiuyu Wang, Yanhong Zeng 외 arxiv

We present LingBot-World, an open-sourced world simulator stemming from video generation. Positioned as a top-tier world model, LingBot-World offers the following features. (1) It maintains high fidelity and robust dynam…

Video Generation

SimInsert: Seamless Video Object Insertion via Regional Sparse Attention Fusion

2026-05-22 · Xinyu Chen, Yuyi Qian, Jiang Lin, Shenyi Wang 외 arxiv

Video object insertion requires ensuring spatio-temporal coherence and interactive realism, extending far beyond simple content placement. However, current approaches are often hindered by a reliance on explicit motion e…

DreamID-V:Bridging the Image-to-Video Gap for High-Fidelity Face Swapping via Diffusion Transformer

2026-01-04 · Xu Guo, Fulong Ye, Xinghui Li, Pengqi Tu 외 arxiv

Video Face Swapping (VFS) requires seamlessly injecting a source identity into a target video while meticulously preserving the original pose, expression, lighting, background, and dynamic information. Existing methods s…

Reinforcement LearningFace Swapping