paper-with-me

Papers

FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training

2024-12-28 · Jiale Huang, Dehong Gao, Jinxia Zhang, Zechao Zhan, Yang Hu, Xin Wang

Large-scale Vision-Language Pre-training (VLP) has demonstrated remarkable success in the general domain. However, in the fashion domain, items are distinguished by fine-grained attributes like texture and material, which are crucial for tasks such as retrieval. Existing models often fail to leverage these fine-grained attributes from both text and image modalities. To address the above issues, we propose a novel approach for the fashion domain, Fine-grained Attributes Enhanced VLP (FashionFAE), which focuses on the detailed characteristics of fashion data. An attribute-emphasized text prediction task is proposed to predict fine-grained attributes of the items. This forces the model to focus on the salient attributes from the text modality. Additionally, a novel attribute-promoted image reconstruction task is proposed, which further enhances the fine-grained ability of the model by leveraging the representative attributes from the image modality. Extensive experiments show that FashionFAE significantly outperforms State-Of-The-Art (SOTA) methods, achieving 2.9% and 5.2% improvements in retrieval on sub-test and full test sets, respectively, and a 1.6% average improvement in recognition tasks.

📄 PDF Abstract BibTeX arXiv:2412.19997

Code (0)

등록된 구현이 없습니다.

Tasks

AttributeImage ReconstructionRetrieval

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

FashionSAP: Symbols and Attributes Prompt for Fine-grained Fashion Vision-Language Pre-training

2023-04-11 · CVPR 2023 1 · Yunpeng Han, Lisai Zhang, Qingcai Chen, Zhijian Chen 외

Fashion vision-language pre-training models have shown efficacy for a wide range of downstream tasks. However, general vision-language pre-training models pay less attention to fine-grained domain features, while these f…

Attribute

Improving the Annotation of DeepFashion Images for Fine-grained Attribute Recognition

2018-07-31 · Roshanak Zakizadeh, Michele Sasdelli, Yu Qian, Eduard Vazquez

DeepFashion is a widely used clothing dataset with 50 categories and more than overall 200k images where each image is annotated with fine-grained attributes. This dataset is often used for clothes recognition and althou…

Attribute

Fashionpedia: Ontology, Segmentation, and an Attribute Localization Dataset

2020-04-26 · ECCV 2020 8 · Menglin Jia, Mengyun Shi, Mikhail Sirotenko, Yin Cui 외

In this work we explore the task of instance segmentation with attribute localization, which unifies instance segmentation (detect and segment each object instance) and fine-grained visual attribute categorization (recog…

AttributeFine-Grained Visual CategorizationFine-Grained Visual RecognitionInstance Segmentation+3

The iMaterialist Fashion Attribute Dataset

2019-06-13 · Sheng Guo, Weilin Huang, Xiao Zhang, Prasanna Srikhanta 외

Large-scale image databases such as ImageNet have significantly advanced image classification and other visual recognition tasks. However much of these datasets are constructed only for single-label and coarse object-lev…

AttributeGeneral Classificationimage-classificationImage Classification+1

DETR-based Layered Clothing Segmentation and Fine-Grained Attribute Recognition

2023-04-17 · Hao Tian, Yu Cao, P. Y. Mok

Clothing segmentation and fine-grained attribute recognition are challenging tasks at the crossing of computer vision and fashion, which segment the entire ensemble clothing instances as well as recognize detailed attrib…

AttributeSegmentation