paper-with-me

Papers

WordRobe: Text-Guided Generation of Textured 3D Garments

2024-03-26 · Astitva Srivastava, Pranav Manu, Amit Raj, Varun Jampani, Avinash Sharma

In this paper, we tackle a new and challenging problem of text-driven generation of 3D garments with high-quality textures. We propose "WordRobe", a novel framework for the generation of unposed & textured 3D garment meshes from user-friendly text prompts. We achieve this by first learning a latent representation of 3D garments using a novel coarse-to-fine training strategy and a loss for latent disentanglement, promoting better latent interpolation. Subsequently, we align the garment latent space to the CLIP embedding space in a weakly supervised manner, enabling text-driven 3D garment generation and editing. For appearance modeling, we leverage the zero-shot generation capability of ControlNet to synthesize view-consistent texture maps in a single feed-forward inference step, thereby drastically decreasing the generation time as compared to existing methods. We demonstrate superior performance over current SOTAs for learning 3D garment latent space, garment interpolation, and text-driven texture synthesis, supported by quantitative evaluation and qualitative user study. The unposed 3D garment meshes generated using WordRobe can be directly fed to standard cloth simulation & animation pipelines without any post-processing.

📄 PDF Abstract BibTeX arXiv:2403.17541

Code (0)

등록된 구현이 없습니다.

Tasks

Disentanglementtext-guided-generationTexture Synthesis

Methods 이 논문이 사용한 방법론

CLIP Contrastive Language-Image Pre-training (CLIP), consisting of a simplified version of ConVIRT trained from scratch, is an efficient method of image representation learning…
ALIGN In the ALIGN method, visual and language representations are jointly trained from noisy image alt-text data. The image and text encoders are learned via contrastive loss…

Similar Papers 제목 키워드 기반

GarmentSketch: Large-scale Sketch-to-Fashion Benchmark

2026-06-12 · Duong-Duy-Khang Bui, Minh-Tan Pham, Tam V. Nguyen, Minh-Triet Tran 외 arxiv

Fashion sketching is a cornerstone of design workflows, allowing rapid visualization of creative concepts prior to physical prototyping. Yet, progress in sketch-based fashion image synthesis has been hindered by the abse…

Text-to-Image Generation

GETAvatar: Generative Textured Meshes for Animatable Human Avatars

2023-10-04 · ICCV 2023 1 · Xuanmeng Zhang, Jianfeng Zhang, Rohan Chacko, Hongyi Xu 외

We study the problem of 3D-aware full-body human generation, aiming at creating animatable human avatars with high-quality textures and geometries. Generally, two challenges remain in this field: i) existing methods stru…

Image Generation

TAPS3D: Text-Guided 3D Textured Shape Generation from Pseudo Supervision

2023-03-23 · CVPR 2023 1 · Jiacheng Wei, Hao Wang, Jiashi Feng, Guosheng Lin 외

In this paper, we investigate an open research task of generating controllable 3D textured shapes from the given textual descriptions. Previous works either require ground truth caption labeling or extensive optimization…

Diversity

USR: Unsupervised Separated 3D Garment and Human Reconstruction via Geometry and Semantic Consistency

2023-02-21 · Yue Shi, Yuxuan Xiong, Jingyi Chai, Bingbing Ni 외

Dressed people reconstruction from images is a popular task with promising applications in the creative media and game industry. However, most existing methods reconstruct the human body and garments as a whole with the …

3D geometryVirtual Try-on

Garment3DGen: 3D Garment Stylization and Texture Generation

2024-03-27 · Nikolaos Sarafianos, Tuur Stuyck, Xiaoyu Xiang, Yilei Li 외

We introduce Garment3DGen a new method to synthesize 3D garment assets from a base mesh given a single input image as guidance. Our proposed approach allows users to generate 3D textured clothes based on both real and sy…

Image to 3DTexture Synthesis