paper-with-me

홈 › Papers

ID-Consistent, Precise Expression Generation with Blendshape-Guided Diffusion

2025-10-06 · Foivos Paraperas Papantoniou, Stefanos Zafeiriou arxiv

Human-centric generative models designed for AI-driven storytelling must bring together two core capabilities: identity consistency and precise control over human performance. While recent diffusion-based approaches have made significant progress in maintaining facial identity, achieving fine-grained expression control without compromising identity remains challenging. In this work, we present a diffusion-based framework that faithfully reimagines any subject under any particular facial expression. Building on an ID-consistent face foundation model, we adopt a compositional design featuring an expression cross-attention module guided by FLAME blendshape parameters for explicit control. Trained on a diverse mixture of image and video data rich in expressive variation, our adapter generalizes beyond basic emotions to subtle micro-expressions and expressive transitions, overlooked by prior works. In addition, a pluggable Reference Adapter enables expression editing in real images by transferring the appearance from a reference frame during synthesis. Extensive quantitative and qualitative evaluations show that our model outperforms existing methods in tailored and identity-consistent expression generation. Code and models can be found at https://github.com/foivospar/Arc2Face.

📄 PDF Abstract BibTeX arXiv:2510.04706

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Geometry-guided Emotion Modulation for Controllable and Photorealistic Emotional Talking Face Generation

2026-08-01 · Chenggong Hu, Shaoyin Ma, Yi Wang, Li Sun 외 arxiv

Audio-driven emotional talking face generation aims to synthesize realistic videos with expressive facial dynamics. However, existing methods struggle to balance controllability and visual fidelity. Although implicit rep…

Talking Face GenerationContinuous Control

ToonifyGB: StyleGAN-based Gaussian Blendshapes for 3D Stylized Head Avatars

2025-05-15 · Rui-Yang Ju, Sheng-Yen Huang, Yi-Ping Hung

The introduction of 3D Gaussian blendshapes has enabled the real-time reconstruction of animatable head avatars from monocular video. Toonify, a StyleGAN-based framework, has become widely used for facial image stylizati…

Image StylizationVideo Generation

RegHead: Non-Humanoid Head Blendshapes via Feed-Forward Registration

2026-07-13 · Jiahao Luo, Hao Zhang, Jianqi Chen, Yijie He 외 arxiv

We present RegHead, a framework for constructing semantic blendshape sets for animatable non-humanoid head avatars. With a fixed expression vocabulary, semantic blendshapes provide a low-dimensional and interpretable ani…

Image Editing

3D Gaussian Blendshapes for Head Avatar Animation

2024-04-30 · Shengjie Ma, Yanlin Weng, Tianjia Shao, Kun Zhou

We introduce 3D Gaussian blendshapes for modeling photorealistic head avatars. Taking a monocular video as input, we learn a base head model of neutral expression, along with a group of expression blendshapes, each of wh…

High-Quality Mesh Blendshape Generation from Face Videos via Neural Inverse Rendering

2024-01-16 · Xin Ming, Jiawei Li, Jingwang Ling, Libo Zhang 외

Readily editable mesh blendshapes have been widely used in animation pipelines, while recent advancements in neural geometry and appearance representations have enabled high-quality inverse rendering. Building upon these…

Inverse Rendering