paper-with-me

홈 › Papers

FreeCloth: Free-form Generation Enhances Challenging Clothed Human Modeling

2024-11-29 · CVPR 2025 1 · Hang Ye, Xiaoxuan Ma, Hai Ci, Wentao Zhu, Yizhou Wang

Achieving realistic animated human avatars requires accurate modeling of pose-dependent clothing deformations. Existing learning-based methods heavily rely on the Linear Blend Skinning (LBS) of minimally-clothed human models like SMPL to model deformation. However, they struggle to handle loose clothing, such as long dresses, where the canonicalization process becomes ill-defined when the clothing is far from the body, leading to disjointed and fragmented results. To overcome this limitation, we propose FreeCloth, a novel hybrid framework to model challenging clothed humans. Our core idea is to use dedicated strategies to model different regions, depending on whether they are close to or distant from the body. Specifically, we segment the human body into three categories: unclothed, deformed, and generated. We simply replicate unclothed regions that require no deformation. For deformed regions close to the body, we leverage LBS to handle the deformation. As for the generated regions, which correspond to loose clothing areas, we introduce a novel free-form, part-aware generator to model them, as they are less affected by movements. This free-form generation paradigm brings enhanced flexibility and expressiveness to our hybrid framework, enabling it to capture the intricate geometric details of challenging loose clothing, such as skirts and dresses. Experimental results on the benchmark dataset featuring loose clothing demonstrate that FreeCloth achieves state-of-the-art performance with superior visual fidelity and realism, particularly in the most challenging cases.

📄 PDF Abstract BibTeX arXiv:2411.19942

Code (0)

등록된 구현이 없습니다.

Tasks

Form

Similar Papers 제목 키워드 기반

TabSD: Large Free-Form Table Question Answering with SQL-Based Table Decomposition

2025-02-19 · Yuxiang Wang, Junhao Gan, Jianzhong Qi

Question answering on free-form tables (TableQA) is challenging due to the absence of predefined schemas and the presence of noise in large tables. While Large Language Models (LLMs) have shown promise in TableQA, they s…

Answer GenerationFormQuestion Answering

Free$^2$Guide: Gradient-Free Path Integral Control for Enhancing Text-to-Video Generation with Large Vision-Language Models

2024-11-26 · JaeMin Kim, Bryan S Kim, Jong Chul Ye

Diffusion models have achieved impressive results in generative tasks like text-to-image (T2I) and text-to-video (T2V) synthesis. However, achieving accurate text alignment in T2V generation remains challenging due to th…

Reinforcement Learning (RL)Text-to-Video GenerationVideo Generation

Model-Guided Dual-Role Alignment for High-Fidelity Open-Domain Video-to-Audio Generation

2025-10-28 · Kang Zhang, Trung X. Pham, Suyeon Lee, Axi Niu 외 arxiv

We present MGAudio, a novel flow-based framework for open-domain video-to-audio generation, which introduces model-guided dual-role alignment as a central design principle. Unlike prior approaches that rely on classifier…

Audio Generation

DefectFill: Realistic Defect Generation with Inpainting Diffusion Model for Visual Inspection

2025-03-18 · CVPR 2025 1 · Jaewoo Song, Daemin Park, Kanghyun Baek, Sangyub Lee 외

Developing effective visual inspection models remains challenging due to the scarcity of defect data. While image generation models have been used to synthesize defect images, producing highly realistic defects remains d…

Image Generation

Visually Guided Decoding: Gradient-Free Hard Prompt Inversion with Language Models

2025-05-13 · Donghoon Kim, Minji Bae, Kyuhong Shim, Byonghyo Shim

Text-to-image generative models like DALL-E and Stable Diffusion have revolutionized visual content creation across various applications, including advertising, personalized media, and design prototyping. However, crafti…

Text Generation