paper-with-me

Papers

SuperFace: Preference-Aligned Facial Expression Estimation Beyond Pseudo Supervision

2026-05-07 · Zejian Kang, Xuanyang Xu, Wentao Yang, Kai Zheng, Yuanchen Fei, Hongyuan Zou, Hui Shan, Shuo Yang, Xiangru Huang arxiv

Accurate facial estimation is crucial for realistic digital human animation, and ARKit blendshape coefficients offer an interpretable representation by mapping facial motions to semantic animation controls. However, learning high-quality ARKit coefficient prediction remains limited by the absence of reliable ground-truth supervision. Existing methods typically rely on capture software such as Live Link Face to provide pseudo labels, which may contain noisy activations, biased coefficient magnitudes, and missing or inaccurate facial actions. Consequently, models trained with supervised learning tend to reproduce imperfect pseudo labels rather than optimize for perceptual expression fidelity. In this paper, we propose SuperFace, a preference-driven framework that moves ARKit facial expression estimation from pseudo-label imitation toward human-aligned perceptual optimization. Instead of treating software-estimated coefficients as fixed ground truth, SuperFace uses them only as an initialization and further improves coefficient prediction through human preference feedback on rendered facial expressions. By aligning the model with perceptual judgments rather than numerical pseudo labels, SuperFace enables more visually faithful and expressive facial animation. Experiments show that SuperFace improves expression fidelity over Live Link Face supervision, demonstrating the effectiveness of preference-driven optimization for semantic facial action prediction.

📄 PDF Abstract BibTeX arXiv:2605.06179

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Facial Expression Generation Aligned with Human Preference for Natural Dyadic Interaction

2026-03-07 · Xu Chen, Rui Gao, Xinjie Zhang, Haoyu Zhang 외 arxiv

Achieving natural dyadic interaction requires generating facial expressions that are emotionally appropriate and socially aligned with human preference. Human feedback offers a compelling mechanism to guide such alignmen…

Reinforcement Learning

Superior and Pragmatic Talking Face Generation with Teacher-Student Framework

2024-03-26 · Chao Liang, Jianwen Jiang, Tianyun Zhong, Gaojie Lin 외

Talking face generation technology creates talking videos from arbitrary appearance and motion signal, with the "arbitrary" offering ease of use but also introducing challenges in practical applications. Existing methods…

Face GenerationTalking Face Generation

FERGI: Automatic Scoring of User Preferences for Text-to-Image Generation from Spontaneous Facial Expression Reaction

2023-12-05 · Shuangquan Feng, Junhua Ma, Virginia R. de Sa

Researchers have proposed to use data of human preference feedback to fine-tune text-to-image generative models. However, the scalability of human feedback collection has been limited by its reliance on manual annotation…

Image GenerationText to Image GenerationText-to-Image Generation

Hallo4: High-Fidelity Dynamic Portrait Animation via Direct Preference Optimization and Temporal Motion Modulation

2025-05-29 · Jiahao Cui, Yan Chen, Mingwang Xu, Hanlin Shang 외

Generating highly dynamic and photorealistic portrait animations driven by audio and skeletal motion remains challenging due to the need for precise lip synchronization, natural facial expressions, and high-fidelity body…

Portrait AnimationVideo Alignment

3DFlowRenderer: One-shot Face Re-enactment via Dense 3D Facial Flow Estimation

2024-04-23 · Siddharth Nijhawan, Takuya Yashima, Tamaki Kojima

Performing facial expression transfer under one-shot setting has been increasing in popularity among research community with a focus on precise control of expressions. Existing techniques showcase compelling results in p…

Motion Estimation