paper-with-me

홈 › Papers

Two-Stream Multi-Task Network for Fashion Recognition

2019-01-29 · Peizhao Li, Yanjing Li, Xiao-Long Jiang, Xian-Tong Zhen

In this paper, we present a two-stream multi-task network for fashion recognition. This task is challenging as fashion clothing always contain multiple attributes, which need to be predicted simultaneously for real-time industrial systems. To handle these challenges, we formulate fashion recognition into a multi-task learning problem, including landmark detection, category and attribute classifications, and solve it with the proposed deep convolutional neural network. We design two knowledge sharing strategies which enable information transfer between tasks and improve the overall performance. The proposed model achieves state-of-the-art results on large-scale fashion dataset comparing to the existing methods, which demonstrates its great effectiveness and superiority for fashion recognition.

📄 PDF Abstract BibTeX arXiv:1901.10172

Code (0)

등록된 구현이 없습니다.

Tasks

AttributeMulti-Task LearningVocal Bursts Valence Prediction

Similar Papers 제목 키워드 기반

Fashionformer: A simple, Effective and Unified Baseline for Human Fashion Segmentation and Recognition

2022-04-10 · Shilin Xu, Xiangtai Li, Jingbo Wang, Guangliang Cheng 외

Human fashion understanding is one crucial computer vision task since it has comprehensive information for real-world applications. This focus on joint human fashion segmentation and attribute recognition. Contrary to th…

AttributeFashion UnderstandingSegmentation

Streaming Multi-talker Speech Recognition with Joint Speaker Identification

2021-04-05 · Liang Lu, Naoyuki Kanda, Jinyu Li, Yifan Gong

In multi-talker scenarios such as meetings and conversations, speech processing systems are usually required to transcribe the audio as well as identify the speakers for downstream applications. Since overlapped speech i…

Speaker Identificationspeech-recognitionSpeech RecognitionSpeech Separation

FashionViL: Fashion-Focused Vision-and-Language Representation Learning

2022-07-17 · Xiao Han, Licheng Yu, Xiatian Zhu, Li Zhang 외

Large-scale Vision-and-Language (V+L) pre-training for representation learning has proven to be effective in boosting various downstream V+L tasks. However, when it comes to the fashion domain, existing V+L methods are i…

Contrastive LearningImage RetrievalRepresentation LearningRetrieval

Texture and Shape Biased Two-Stream Networks for Clothing Classification and Attribute Recognition

2020-06-01 · CVPR 2020 6 · Yuwei Zhang, Peng Zhang, Chun Yuan, Zhi Wang

Clothes category classification and attribute recognition have achieved distinguished success with the development of deep learning. People have found that landmark detection plays a positive role in these tasks. However…

AttributeClassificationGeneral Classification

Self-supervised learning with bi-label masked speech prediction for streaming multi-talker speech recognition

2022-11-10 · Zili Huang, Zhuo Chen, Naoyuki Kanda, Jian Wu 외

Self-supervised learning (SSL), which utilizes the input data itself for representation learning, has achieved state-of-the-art results for various downstream speech tasks. However, most of the previous studies focused o…

Representation LearningSelf-Supervised Learningspeech-recognitionSpeech Recognition