paper-with-me

Papers

LiftAvatar: Kinematic-Space Completion for Expression-Controlled 3D Gaussian Avatar Animation

2026-03-02 · Hualiang Wei, Shunran Jia, Jialun Liu, Wenhui Li arxiv

We present LiftAvatar, a new paradigm that completes sparse monocular observations in kinematic space (e.g., facial expressions and head pose) and uses the completed signals to drive high-fidelity avatar animation. LiftAvatar is a fine-grained, expression-controllable large-scale video diffusion Transformer that synthesizes high-quality, temporally coherent expression sequences conditioned on single or multiple reference images. The key idea is to lift incomplete input data into a richer kinematic representation, thereby strengthening both reconstruction and animation in downstream 3D avatar pipelines. To this end, we introduce (i) a multi-granularity expression control scheme that combines shading maps with expression coefficients for precise and stable driving, and (ii) a multi-reference conditioning mechanism that aggregates complementary cues from multiple frames, enabling strong 3D consistency and controllability. As a plug-and-play enhancer, LiftAvatar directly addresses the limited expressiveness and reconstruction artifacts of 3D Gaussian Splatting-based avatars caused by sparse kinematic cues in everyday monocular videos. By expanding incomplete observations into diverse pose-expression variations, LiftAvatar also enables effective prior distillation from large-scale video generative models into 3D pipelines, leading to substantial gains. Extensive experiments show that LiftAvatar consistently boosts animation quality and quantitative metrics of state-of-the-art 3D avatar methods, especially under extreme, unseen expressions.

📄 PDF Abstract BibTeX arXiv:2603.02129

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

POCE: Pose-Controllable Expression Editing

2023-04-18 · Rongliang Wu, Yingchen Yu, Fangneng Zhan, Jiahui Zhang 외

Facial expression editing has attracted increasing attention with the advance of deep neural networks in recent years. However, most existing methods suffer from compromised editing fidelity and limited usability as they…

Architecture of a Web-based Predictive Editor for Controlled Natural Language Processing

2014-06-27 · Stephen Guy, Rolf Schwitter

In this paper, we describe the architecture of a web-based predictive text editor being developed for the controlled natural language PENG$^{ASP)$. This controlled language can be used to write non-monotonic specificatio…

Sentence

Kinematically Controllable Cable Robots with Reconfigurable End-effectors

2025-10-26 · Nan Zhang arxiv

To enlarge the translational workspace of cable-driven robots, one common approach is to increase the number of cables. However, this introduces two challenges: (1) cable interference significantly reduces the rotational…

Decoding RobKiNet: Insights into Efficient Training of Robotic Kinematics Informed Neural Network

2025-09-09 · Yanlong Peng, Zhigang Wang, Ziwen He, Pengxu Chang 외 arxiv

In robots task and motion planning (TAMP), it is crucial to sample within the robot's configuration space to meet task-level global constraints and enhance the efficiency of subsequent motion planning. Due to the complex…

Reinforcement LearningMotion Planning

Analytical Solvers for Common Algebraic Equations Arising in Kinematics Problems

2025-08-31 · Hai-Jun Su arxiv

This paper presents analytical solvers for four common types of algebraic equations encountered in robot kinematics: single trigonometric equations, single-angle trigonometric systems, two-angle trigonometric systems, an…