paper-with-me

홈 › Papers

VTON-HandFit: Virtual Try-on for Arbitrary Hand Pose Guided by Hand Priors Embedding

2024-08-22 · CVPR 2025 1 · Yujie Liang, Xiaobin Hu, Boyuan Jiang, Donghao Luo, Kai Wu, Wenhui Han, Taisong Jin, Chengjie Wang

Although diffusion-based image virtual try-on has made considerable progress, emerging approaches still struggle to effectively address the issue of hand occlusion (i.e., clothing regions occluded by the hand part), leading to a notable degradation of the try-on performance. To tackle this issue widely existing in real-world scenarios, we propose VTON-HandFit, leveraging the power of hand priors to reconstruct the appearance and structure for hand occlusion cases. Firstly, we tailor a Handpose Aggregation Net using the ControlNet-based structure explicitly and adaptively encoding the global hand and pose priors. Besides, to fully exploit the hand-related structure and appearance information, we propose Hand-feature Disentanglement Embedding module to disentangle the hand priors into the hand structure-parametric and visual-appearance features, and customize a masked cross attention for further decoupled feature embedding. Lastly, we customize a hand-canny constraint loss to better learn the structure edge knowledge from the hand template of model image. VTON-HandFit outperforms the baselines in qualitative and quantitative evaluations on the public dataset and our self-collected hand-occlusion Handfit-3K dataset particularly for the arbitrary hand pose occlusion cases in real-world scenarios. The Code and dataset will be available at \url{https://github.com/VTON-HandFit/VTON-HandFit}.

📄 PDF Abstract BibTeX arXiv:2408.12340

Code (2)

illumara/vton-handfit 공식 구현
vton-handfit/vton-handfit 공식 구현

Tasks

DisentanglementVirtual Try-on

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음

Similar Papers 제목 키워드 기반

PT-VTON: an Image-Based Virtual Try-On Network with Progressive Pose Attention Transfer

2021-11-23 · Hanhan Zhou, Tian Lan, Guru Venkataramani

The virtual try-on system has gained great attention due to its potential to give customers a realistic, personalized product presentation in virtualized settings. In this paper, we present PT-VTON, a novel pose-transfer…

Pose TransferVirtual Try-on

SPG-VTON: Semantic Prediction Guidance for Multi-pose Virtual Try-on

2021-08-03 · Bingwen Hu, Ping Liu, Zhedong Zheng, Mingwu Ren

Image-based virtual try-on is challenging in fitting a target in-shop clothes into a reference person under diverse human poses. Previous works focus on preserving clothing details ( e.g., texture, logos, patterns ) when…

Virtual Try-on

Enhancing Person-to-Person Virtual Try-On with Multi-Garment Virtual Try-Off

2025-04-17 · Riza Velioglu, Petra Bevandic, Robin Chan, Barbara Hammer

Computer vision is transforming fashion through Virtual Try-On (VTON) and Virtual Try-Off (VTOFF). VTON generates images of a person in a specified garment using a target photo and a standardized garment image, while a m…

Garment ReconstructionImage GenerationVirtual Try-OffVirtual Try-on

D$^4$-VTON: Dynamic Semantics Disentangling for Differential Diffusion based Virtual Try-On

2024-07-21 · Zhaotong Yang, Zicheng Jiang, Xinzhe Li, Huiyu Zhou 외

In this paper, we introduce D$^4$-VTON, an innovative solution for image-based virtual try-on. We address challenges from previous studies, such as semantic inconsistencies before and after garment warping, and reliance …

DenoisingVirtual Try-on

OmniVTON: Training-Free Universal Virtual Try-On

2025-07-20 · Zhaotong Yang, Yuhui Li, Shengfeng He, Xinzhe Li 외 arxiv

Image-based Virtual Try-On (VTON) techniques rely on either supervised in-shop approaches, which ensure high fidelity but struggle with cross-domain generalization, or unsupervised in-the-wild methods, which improve adap…

Domain GeneralizationVirtual Try-on