paper-with-me

홈 › Papers

AnyDesign: Versatile Area Fashion Editing via Mask-Free Diffusion

2024-08-21 · Yunfang Niu, Lingxiang Wu, Dong Yi, Jie Peng, Ning Jiang, Haiying Wu, Jinqiao Wang

Fashion image editing aims to modify a person's appearance based on a given instruction. Existing methods require auxiliary tools like segmenters and keypoint extractors, lacking a flexible and unified framework. Moreover, these methods are limited in the variety of clothing types they can handle, as most datasets focus on people in clean backgrounds and only include generic garments such as tops, pants, and dresses. These limitations restrict their applicability in real-world scenarios. In this paper, we first extend an existing dataset for human generation to include a wider range of apparel and more complex backgrounds. This extended dataset features people wearing diverse items such as tops, pants, dresses, skirts, headwear, scarves, shoes, socks, and bags. Additionally, we propose AnyDesign, a diffusion-based method that enables mask-free editing on versatile areas. Users can simply input a human image along with a corresponding prompt in either text or image format. Our approach incorporates Fashion DiT, equipped with a Fashion-Guidance Attention (FGA) module designed to fuse explicit apparel types and CLIP-encoded apparel features. Both Qualitative and quantitative experiments demonstrate that our method delivers high-quality fashion editing and outperforms contemporary text-guided fashion editing methods.

📄 PDF Abstract BibTeX arXiv:2408.11553

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Softmax The Softmax output function transforms a previous layer's output into a vector of probabilities. It is commonly used for multiclass classification. Given an input vector $x$…
Attention 설명 없음
Focus 설명 없음

Similar Papers 제목 키워드 기반

MADiff: Text-Guided Fashion Image Editing with Mask Prediction and Attention-Enhanced Diffusion

2024-12-28 · Zechao Zhan, Dehong Gao, Jinxia Zhang, Jiale Huang 외

Text-guided image editing model has achieved great success in general domain. However, directly applying these models to the fashion domain may encounter two issues: (1) Inaccurate localization of editing region; (2) Wea…

Large Language Modeltext-guided-image-editing

Fashion Matrix: Editing Photos by Just Talking

2023-07-25 · Zheng Chong, Xujie Zhang, Fuwei Zhao, Zhenyu Xie 외

The utilization of Large Language Models (LLMs) for the construction of AI systems has garnered significant attention across diverse fields. The extension of LLMs to the domain of fashion holds substantial commercial pot…

Semantic Segmentation

DeepFashion2: A Versatile Benchmark for Detection, Pose Estimation, Segmentation and Re-Identification of Clothing Images

2019-01-23 · CVPR 2019 6 · Yuying Ge, Ruimao Zhang, Lingyun Wu, Xiaogang Wang 외

Understanding fashion images has been advanced by benchmarks with rich annotations such as DeepFashion, whose labels include clothing categories, landmarks, and consumer-commercial image pairs. However, DeepFashion has n…

Pose EstimationRetrievalSemantic Segmentation

Pose-Star: Anatomy-Aware Editing for Open-World Fashion Images

2025-07-04 · Yuran Dong, Mang Ye

To advance real-world fashion image editing, we analyze existing two-stage pipelines(mask generation followed by diffusion-based editing)which overly prioritize generator optimization while neglecting mask controllabilit…

Anatomy

Text-Driven Fashion Image Editing with Compositional Concept Learning and Counterfactual Abduction

2025-01-01 · CVPR 2025 1 · Shanshan Huang, Haoxuan Li, Chunyuan Zheng, Mingyuan Ge 외

Fashion image editing is a valuable tool for designers to convey their creative ideas by visualizing design concepts. With the recent advances in text editing methods, significant progress has been made in fashion im…

counterfactualCounterfactual ReasoningDenoising