Two-Stream Multi-Task Network for Fashion Recognition
In this paper, we present a two-stream multi-task network for fashion recognition. This task is challenging as fashion clothing always contain multiple attributes, which need to be predicted simultaneously for real-time industrial systems. To handle these challenges, we formulate fashion recognition into a multi-task learning problem, including landmark detection, category and attribute classifications, and solve it with the proposed deep convolutional neural network. We design two knowledge sharing strategies which enable information transfer between tasks and improve the overall performance. The proposed model achieves state-of-the-art results on large-scale fashion dataset comparing to the existing methods, which demonstrates its great effectiveness and superiority for fashion recognition.
Code (0)
등록된 구현이 없습니다.
Tasks
AttributeMulti-Task LearningVocal Bursts Valence PredictionSimilar Papers 제목 키워드 기반
Fashionformer: A simple, Effective and Unified Baseline for Human Fashion Segmentation and Recognition
Human fashion understanding is one crucial computer vision task since it has comprehensive information for real-world applications. This focus on joint human fashion segmentation and attribute recognition. Contrary to th…
AttributeFashion UnderstandingSegmentationStreaming Multi-talker Speech Recognition with Joint Speaker Identification
In multi-talker scenarios such as meetings and conversations, speech processing systems are usually required to transcribe the audio as well as identify the speakers for downstream applications. Since overlapped speech i…
Speaker Identificationspeech-recognitionSpeech RecognitionSpeech SeparationFashionViL: Fashion-Focused Vision-and-Language Representation Learning
Large-scale Vision-and-Language (V+L) pre-training for representation learning has proven to be effective in boosting various downstream V+L tasks. However, when it comes to the fashion domain, existing V+L methods are i…
Contrastive LearningImage RetrievalRepresentation LearningRetrievalTexture and Shape Biased Two-Stream Networks for Clothing Classification and Attribute Recognition
Clothes category classification and attribute recognition have achieved distinguished success with the development of deep learning. People have found that landmark detection plays a positive role in these tasks. However…
AttributeClassificationGeneral ClassificationSelf-supervised learning with bi-label masked speech prediction for streaming multi-talker speech recognition
Self-supervised learning (SSL), which utilizes the input data itself for representation learning, has achieved state-of-the-art results for various downstream speech tasks. However, most of the previous studies focused o…
Representation LearningSelf-Supervised Learningspeech-recognitionSpeech Recognition