paper-with-me

홈 › Papers

TPC: Test-time Procrustes Calibration for Diffusion-based Human Image Animation

2024-10-31 · Sunjae Yoon, Gwanhyeong Koo, Younghwan Lee, Chang D. Yoo

Human image animation aims to generate a human motion video from the inputs of a reference human image and a target motion video. Current diffusion-based image animation systems exhibit high precision in transferring human identity into targeted motion, yet they still exhibit irregular quality in their outputs. Their optimal precision is achieved only when the physical compositions (i.e., scale and rotation) of the human shapes in the reference image and target pose frame are aligned. In the absence of such alignment, there is a noticeable decline in fidelity and consistency. Especially, in real-world environments, this compositional misalignment commonly occurs, posing significant challenges to the practical usage of current systems. To this end, we propose Test-time Procrustes Calibration (TPC), which enhances the robustness of diffusion-based image animation systems by maintaining optimal performance even when faced with compositional misalignment, effectively addressing real-world scenarios. The TPC provides a calibrated reference image for the diffusion model, enhancing its capability to understand the correspondence between human shapes in the reference and target images. Our method is simple and can be applied to any diffusion-based image animation system in a model-agnostic manner, improving the effectiveness at test time without additional training.

📄 PDF Abstract BibTeX arXiv:2410.24037

Code (0)

등록된 구현이 없습니다.

Tasks

Image Animation

Methods 이 논문이 사용한 방법론

Procrustes Procrustes
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

COMPOT: Calibration-Optimized Matrix Procrustes Orthogonalization for Transformers Compression

2026-02-16 · Denis Makhov, Dmitriy Shopkhoev, Magauiya Zhussip, Ammar Ali 외 arxiv

Post-training compression of Transformer models commonly relies on truncated singular value decomposition (SVD). However, enforcing a single shared subspace can degrade accuracy even at moderate compression. Sparse dicti…

NaviCache: Test-Time Self-Calibration Caching for Video Generation

2026-06-25 · Zheqi Lv, Zhibo Zhu, Jinke Wang, Qi Tian 외 arxiv

Video Diffusion Models (VDMs) is constrained by immense computational costs. While offline calibration-based acceleration suffers from calibration data dependency, prohibitive calibration duration, and susceptibility to …

Video Generation

Limitations of (Procrustes) Alignment in Assessing Multi-Person Human Pose and Shape Estimation

2024-09-25 · Drazic Martin, Pierre Perrault

We delve into the challenges of accurately estimating 3D human pose and shape in video surveillance scenarios. Beginning with the advocacy for metrics like W-MPJPE and W-PVE, which omit the (Procrustes) realignment step,…

On Procrustes Contamination in Machine Learning Applications of Geometric Morphometrics

2026-01-26 · Lloyd Austin Courtenay arxiv

Geometric morphometrics (GMM) is widely used to quantify shape variation, more recently serving as input for machine learning (ML) analyses. Standard practice aligns all specimens via Generalized Procrustes Analysis (GPA…

Deep Soft Procrustes for Markerless Volumetric Sensor Alignment

2020-03-23 · Vladimiros Sterzentsenko, Alexandros Doumanoglou, Spyridon Thermos, Nikolaos Zioulis 외

With the advent of consumer grade depth sensors, low-cost volumetric capture systems are easier to deploy. Their wider adoption though depends on their usability and by extension on the practicality of spatially aligning…

Pose Estimation