paper-with-me

Papers

Towards Vision-Language-Garment Models For Web Knowledge Garment Understanding and Generation

2025-06-05 · Jan Ackermann, Kiyohiro Nakayama, Guandao Yang, Tong Wu, Gordon Wetzstein

Multimodal foundation models have demonstrated strong generalization, yet their ability to transfer knowledge to specialized domains such as garment generation remains underexplored. We introduce VLG, a vision-language-garment model that synthesizes garments from textual descriptions and visual imagery. Our experiments assess VLG's zero-shot generalization, investigating its ability to transfer web-scale reasoning to unseen garment styles and prompts. Preliminary results indicate promising transfer capabilities, highlighting the potential for multimodal foundation models to adapt effectively to specialized domains like fashion design.

📄 PDF Abstract BibTeX arXiv:2506.05210

Code (0)

등록된 구현이 없습니다.

Tasks

Zero-shot Generalization

Similar Papers 제목 키워드 기반

EditGarment: An Instruction-Based Garment Editing Dataset Constructed with Automated MLLM Synthesis and Semantic-Aware Evaluation

2025-08-05 · Deqiang Yin, Junyi Guo, Huanda Lu, Fangyu Wu 외 arxiv

Instruction-based garment editing enables precise image modifications via natural language, with broad applications in fashion design and customization. Unlike general editing tasks, it requires understanding garment-spe…

GarmentWeaver: Schema-Aware Structured Synthesis for Multimodal Sewing Patterns

2026-08-31 · Yinwen Lu, Weihao Luo, Yueqi Zhong arxiv

Multimodal Sewing pattern generation aims to infer executable sewing patterns from design cues such as sketches and textual descriptions. As an interpretable and simulation-compatible representation, sewing patterns are …

GarmentPile++: Affordance-Driven Cluttered Garments Retrieval with Vision-Language Reasoning

2026-03-04 · Mingleyang Li, Yuran Wang, Yue Chen, Tianxing Chen 외 arxiv

Garment manipulation has attracted increasing attention due to its critical role in home-assistant robotics. However, the majority of existing garment manipulation works assume an initial state consisting of only one gar…

Object Segmentation

SwiftTailor: Efficient 3D Garment Generation with Geometry Image Representation

2026-03-19 · Phuc Pham, Uy Dieu Tran, Binh-Son Hua, Phong Nguyen arxiv

Realistic and efficient 3D garment generation remains a longstanding challenge in computer vision and digital fashion. Existing methods typically rely on large vision- language models to produce serialized representation…

SKT: Integrating State-Aware Keypoint Trajectories with Vision-Language Models for Robotic Garment Manipulation

2024-09-26 · Xin Li, Siyuan Huang, Qiaojun Yu, Zhengkai Jiang 외

Automating garment manipulation poses a significant challenge for assistive robotics due to the diverse and deformable nature of garments. Traditional approaches typically require separate models for each garment type, w…

Keypoint Detection