paper-with-me

홈 › Papers

FaceCam: Portrait Video Camera Control via Scale-Aware Conditioning

2026-03-05 · Weijie Lyu, Ming-Hsuan Yang, Zhixin Shu arxiv

We introduce FaceCam, a system that generates video under customizable camera trajectories for monocular human portrait video input. Recent camera control approaches based on large video-generation models have shown promising progress but often exhibit geometric distortions and visual artifacts on portrait videos due to scale-ambiguous camera representations or 3D reconstruction errors. To overcome these limitations, we propose a face-tailored scale-aware representation for camera transformations that provides deterministic conditioning without relying on 3D priors. We train a video generation model on both multi-view studio captures and in-the-wild monocular videos, and introduce two camera-control data generation strategies: synthetic camera motion and multi-shot stitching, to exploit stationary training cameras while generalizing to dynamic, continuous camera trajectories at inference time. Experiments on Ava-256 dataset and diverse in-the-wild videos demonstrate that FaceCam achieves superior performance in camera controllability, visual quality, identity and motion preservation.

📄 PDF Abstract BibTeX arXiv:2603.05506

Code (0)

등록된 구현이 없습니다.

Tasks

3D ReconstructionVideo Generation

Similar Papers 제목 키워드 기반

FactorPortrait: Controllable Portrait Animation via Disentangled Expression, Pose, and Viewpoint

2025-12-12 · Jiapeng Tang, Kai Li, Chengxiang Yin, Liuhao Ge 외 arxiv

We introduce FactorPortrait, a video diffusion method for controllable portrait animation that enables lifelike synthesis from disentangled control signals of facial expressions, head movement, and camera viewpoints. Giv…

Novel View Synthesis

Pixel Cube: Diffusion-based Portrait Video Relighting Through Realistic Lighting Reproduction

2026-06-01 · Yufan Zhang, Yu Ji, Ayo Ajiboye, Rundi Wu 외 arxiv

We present a diffusion-based method for relighting dynamic portrait videos with photorealism and temporal consistency. Our method is fueled by a hybrid training dataset that consists of real-captured and rendered dynamic…

AniPortraitGAN: Animatable 3D Portrait Generation from 2D Image Collections

2023-09-05 · Yue Wu, Sicheng Xu, Jianfeng Xiang, Fangyun Wei 외

Previous animatable 3D-aware GANs for human generation have primarily focused on either the human head or full body. However, head-only videos are relatively uncommon in real life, and full body generation typically does…

PP-HumanSeg: Connectivity-Aware Portrait Segmentation with a Large-Scale Teleconferencing Video Dataset

2021-12-14 · Lutao Chu, Yi Liu, Zewu Wu, Shiyu Tang 외

As the COVID-19 pandemic rampages across the world, the demands of video conferencing surge. To this end, real-time portrait segmentation becomes a popular feature to replace backgrounds of conferencing participants. Whi…

Portrait SegmentationSegmentationSemantic Segmentation

PV3D: A 3D Generative Model for Portrait Video Generation

2022-12-13 · Zhongcong Xu, Jianfeng Zhang, Jun Hao Liew, Wenqing Zhang 외

Recent advances in generative adversarial networks (GANs) have demonstrated the capabilities of generating stunning photo-realistic portrait images. While some prior works have applied such image GANs to unconditional 2D…

Video Generation