paper-with-me

홈 › Papers

SCAIL: Towards Studio-Grade Character Animation via In-Context Learning of 3D-Consistent Pose Representations

2025-12-05 · Wenhao Yan, Sheng Ye, Zhuoyi Yang, Jiayan Teng, ZhenHui Dong, Kairui Wen, Xiaotao Gu, Yong-Jin Liu, Jie Tang arxiv

Achieving controllable character animation that meets studio-grade standards remains challenging despite recent progress. Existing approaches can transfer motion from a driving video to a reference image, but often fail to preserve structural fidelity and temporal consistency in wild scenarios involving complex motion and cross-identity animations. In this work, we present \textbf{SCAIL} (a framework toward \textbf{S}tudio-grade \textbf{C}haracter \textbf{A}nimation via \textbf{I}n-context \textbf{L}earning), which is designed to address these challenges from two key innovations. First, we propose a novel 3D pose representation, providing a robust and flexible motion signal. Second, we introduce a full-context pose injection mechanism within a diffusion-transformer, enabling effective spatio-temporal reasoning over full motion sequences. To align with studio-grade requirements, we develop a curated data pipeline ensuring both diversity and quality, and establish a comprehensive benchmark for systematic evaluation. Experiments show that \textbf{SCAIL} achieves state-of-the-art performance and advances character animation toward studio-grade controlling. Code and model are available at \href{https://github.com/zai-org/SCAIL}{zai-org/SCAIL}.

📄 PDF Abstract BibTeX arXiv:2512.05905

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

SCAIL-2: Unifying Controlled Character Animation with End-to-end In-Context Conditioning

2026-06-09 · Wenhao Yan, Fengjia Guo, Zhuoyi Yang, Jie Tang arxiv

Controlled character animation requires transferring motion from a driving sequence to a reference character. Prior works heavily rely on intermediate representations, including pose skeletons to represent motion or mask…

AvatarStudio: High-fidelity and Animatable 3D Avatar Creation from Text

2023-11-29 · Jianfeng Zhang, Xuanmeng Zhang, Huichao Zhang, Jun Hao Liew 외

We study the problem of creating high-fidelity and animatable 3D avatars from only textual descriptions. Existing text-to-avatar methods are either limited to static avatars which cannot be animated or struggle to genera…

NeRF

HeadStudio: Text to Animatable Head Avatars with 3D Gaussian Splatting

2024-02-09 · Zhenglin Zhou, Fan Ma, Hehe Fan, Zongxin Yang 외

Creating digital avatars from textual prompts has long been a desirable yet challenging task. Despite the promising results achieved with 2D diffusion priors, current methods struggle to create high-quality and consisten…

Audio2Rig: Artist-oriented deep learning tool for facial animation

2024-05-30 · Bastien Arcelin, Nicolas Chaverou

Creating realistic or stylized facial and lip sync animation is a tedious task. It requires lot of time and skills to sync the lips with audio and convey the right emotion to the character's face. To allow animators to s…

Deep Learning

Artist-Guided Semiautomatic Animation Colorization

2020-06-22 · Harrish Thasarathan, Mehran Ebrahimi

There is a delicate balance between automating repetitive work in creative domains while staying true to an artist's vision. The animation industry regularly outsources large animation workloads to foreign countries wher…

ColorizationLine Art Colorization