paper-with-me

Papers

UniMate: One Unified Model to Animate Diverse Skeletons

2026-09-04 · Linzhan Mou, Jiahui Lei, Zhiyang Dou, Chenyue Cai, Chaoyue Song, Adam Finkelstein, Szymon Rusinkiewicz hf

Recent advances in automatic rigging now deliver animation-ready 3D assets at scale, yet generating the motion to drive them remains a bottleneck. Existing learned animators are topology-constrained: they rely on category-specific templates or require per-skeleton fine-tuning and reference motions at inference. We present UniMate, a unified foundation model that synthesizes articulated motion for arbitrary skeletons from a rigged 3D asset and a text prompt, with no test-time optimization or per-skeleton retraining. UniMate introduces a topology-aware diffusion transformer, which integrates skeletal topology into attention via three mechanisms: (1) a graph-aware attention bias from pairwise joint relations and geodesic distances; (2) a spectral rotary position embedding generalizing RoPE to arbitrary kinematic trees via the graph Laplacian; and (3) a global topological conditioner attention-pooled from the rest-pose skeleton. We also curate UniML3D, 13,006 motion sequences spanning bipedal, quadrupedal, avian, marine, insectoid, serpentine, and articulated rigid objects with unified canonicalization and text pairing. Trained on this dataset, UniMate outperforms state-of-the-art baselines in quality, generalization, and efficiency, and supports zero-shot cross-topology transfer, in-betweening, expansion, and text-guided editing. Our project page is available at https://linzhanmou.com/unimate/.

📄 PDF Abstract BibTeX arXiv:2609.05415

Code (2)

Friedrich-M/UniMate ★ 56
InsomaniacElf/sg-tamil-tts-resources- ★ 1

Similar Papers 제목 키워드 기반

Object Wake-up: 3D Object Rigging from a Single Image

2021-08-05 · Ji Yang, Xinxin Zuo, Sen Wang, Zhenbo Yu 외

Given a single image of a general object such as a chair, could we also restore its articulated 3D shape similar to human modeling, so as to animate its plausible articulations and diverse motions? This is an interesting…

3D Object Reconstruction3D ReconstructionObjectObject Reconstruction

NECromancer: Breathing Life into Skeletons via BVH Animation

2026-02-06 · Mingxi Xu, Qi Wang, Zhengyu Wen, Phong Dao Thien 외 arxiv

Motion tokenization is a key component of generalizable motion models, yet most existing approaches are restricted to species-specific skeletons, limiting their applicability across diverse morphologies. We propose NECro…

Heterogeneous Skeleton-Based Action Representation Learning

2025-01-01 · CVPR 2025 1 · Hongsong Wang, Xiaoyan Ma, Jidong Kuang, Jie Gui

Skeleton-based human action recognition has received widespread attention in recent years due to its diverse range of application scenarios. Due to the different sources of human skeletons, skeleton data naturally ex…

Action RecognitionAction UnderstandingRepresentation LearningTemporal Action Localization

AniGen: Unified $S^3$ Fields for Animatable 3D Asset Generation

2026-04-09 · Yi-Hua Huang, Zi-Xin Zou, Yuting He, Chirui Chang 외 arxiv

Animatable 3D assets, defined as geometry equipped with an articulated skeleton and skinning weights, are fundamental to interactive graphics, embodied agents, and animation production. While recent 3D generative models …

AnimateScene: Camera-controllable Animation in Any Scene

2025-08-08 · Qingyang Liu, Bingjie Gao, Weiheng Huang, Jun Zhang 외 arxiv

Recent advances in 3D scene reconstruction and 4D human animation have broadened adoption, but integrating the two remains difficult. Key challenges include placing humans at plausible locations and scales without interp…