Garments2Look: A Multi-Reference Dataset for High-Fidelity Outfit-Level Virtual Try-On with Clothing and Accessories
Virtual try-on (VTON) has advanced single-garment visualization, yet real-world fashion centers on full outfits with multiple garments, accessories, fine-grained categories, layering, and diverse styling, remaining beyond current VTON systems. Existing datasets are category-limited and lack outfit diversity. We introduce Garments2Look, the first large-scale multimodal dataset for outfit-level VTON, comprising 80K many-garments-to-one-look pairs across 40 major categories and 300+ fine-grained subcategories. Each pair includes an outfit with 3-12 reference garment images (Average 4.48), a model image wearing the outfit, and detailed item and try-on textual annotations. To balance authenticity and diversity, we propose a synthesis pipeline. It involves heuristically constructing outfit lists before generating try-on results, with the entire process subjected to strict automated filtering and human validation to ensure data quality. To probe task difficulty, we adapt SOTA VTON methods and general-purpose image editing models to establish baselines. Results show current methods struggle to try on complete outfits seamlessly and to infer correct layering and styling, leading to misalignment and artifacts.
Code (0)
등록된 구현이 없습니다.
Tasks
Virtual Try-onImage EditingSimilar Papers 제목 키워드 기반
Virtual Fashion Photo-Shoots: Building a Large-Scale Garment-Lookbook Dataset
Fashion image generation has so far focused on narrow tasks such as virtual try-on, where garments appear in clean studio environments. In contrast, editorial fashion presents garments through dynamic poses, diverse loca…
Image GenerationVirtual Try-onControllable Human Image Generation with Personalized Multi-Garments
We present BootComp, a novel framework based on text-to-image diffusion models for controllable human image generation with multiple reference garments. Here, the main bottleneck is data acquisition for training: collect…
DenoisingImage GenerationVirtual Try-onDressing in Order: Recurrent Person Image Generation for Pose Transfer, Virtual Try-on and Outfit Editing
We propose a flexible person generation framework called Dressing in Order (DiOr), which supports 2D pose transfer, virtual try-on, and several fashion editing tasks. The key to DiOr is a novel recurrent generation pipel…
Fashion SynthesisImage GenerationPose TransferVirtual Try-onImage Based Virtual Try-On Network From Unpaired Data
This paper presents a new image-based virtual try-on approach (Outfit-VITON) that helps visualize how a composition of clothing items selected from various reference images form a cohesive outfit on a person in a query i…
Virtual Try-onA generic method of wearable items virtual try-on
Virtual try-on synthesizes garments for the target bodies in 2D/3D domains. Even though existing virtual try-on methods focus on redressing garments, the virtual try-on hair, shoes and wearable accessories are still unde…
Virtual Try-on