paper-with-me

홈 › Papers

Pose Guided Attention for Multi-label Fashion Image Classification

2019-11-12 · Beatriz Quintino Ferreira, João P. Costeira, Ricardo G. Sousa, Liang-Yan Gui, João P. Gomes

We propose a compact framework with guided attention for multi-label classification in the fashion domain. Our visual semantic attention model (VSAM) is supervised by automatic pose extraction creating a discriminative feature space. VSAM outperforms the state of the art for an in-house dataset and performs on par with previous works on the DeepFashion dataset, even without using any landmark annotations. Additionally, we show that our semantic attention module brings robustness to large quantities of wrong annotations and provides more interpretable results.

📄 PDF Abstract BibTeX arXiv:1911.05024

Code (0)

등록된 구현이 없습니다.

Tasks

ClassificationGeneral Classificationimage-classificationImage ClassificationMulti-Label ClassificationMUlTI-LABEL-ClASSIFICATION

Similar Papers 제목 키워드 기반

Attribute-Guided Multi-Level Attention Network for Fine-Grained Fashion Retrieval

2022-12-27 · Ling Xiao, Toshihiko Yamasaki

Fine-grained fashion retrieval searches for items that share a similar attribute with the query image. Most existing methods use a pre-trained feature extractor (e.g., ResNet 50) to capture image representations. However…

Attributeimage-classificationImage Classificationobject-detection+2

Training and challenging models for text-guided fashion image retrieval

2022-04-23 · Eric Dodds, Jack Culpepper, Gaurav Srivastava

Retrieving relevant images from a catalog based on a query image together with a modifying caption is a challenging multimodal task that can particularly benefit domains like apparel shopping, where fine details and subt…

AttributeImage RetrievalRetrieval

MADiff: Text-Guided Fashion Image Editing with Mask Prediction and Attention-Enhanced Diffusion

2024-12-28 · Zechao Zhan, Dehong Gao, Jinxia Zhang, Jiale Huang 외

Text-guided image editing model has achieved great success in general domain. However, directly applying these models to the fashion domain may encounter two issues: (1) Inaccurate localization of editing region; (2) Wea…

Large Language Modeltext-guided-image-editing

Deep Attention-guided Hashing

2018-12-04 · Zhan Yang, Osolo Ian Raymond, Wuqing Sun, Jun Long

With the rapid growth of multimedia data (e.g., image, audio and video etc.) on the web, learning-based hashing techniques such as Deep Supervised Hashing (DSH) have proven to be very efficient for large-scale multimedia…

Deep Attention

Attentive Fashion Grammar Network for Fashion Landmark Detection and Clothing Category Classification

2018-06-01 · CVPR 2018 6 · Wenguan Wang, Yuanlu Xu, Jianbing Shen, Song-Chun Zhu

This paper proposes a knowledge-guided fashion network to solve the problem of visual fashion analysis, e.g., fashion landmark localization and clothing category classification. The suggested fashion model is leveraged w…

General Classification