Deep CNN-based Multi-task Learning for Open-Set Recognition
We propose a novel deep convolutional neural network (CNN) based multi-task learning approach for open-set visual recognition. We combine a classifier network and a decoder network with a shared feature extractor network within a multi-task learning framework. We show that this approach results in better open-set recognition accuracy. In our approach, reconstruction errors from the decoder network are utilized for open-set rejection. In addition, we model the tail of the reconstruction error distribution from the known classes using the statistical Extreme Value Theory to improve the overall performance. Experiments on multiple image classification datasets are performed and it is shown that this method can perform significantly better than many competitive open set recognition algorithms available in the literature. The code will be made available at: github.com/otkupjnoz/mlosr.
Code (0)
등록된 구현이 없습니다.
Tasks
DecoderGeneral Classificationimage-classificationImage ClassificationMulti-Task LearningOpen Set LearningSimilar Papers 제목 키워드 기반
MOoSE: Multi-Orientation Sharing Experts for Open-set Scene Text Recognition
Open-set text recognition, which aims to address both novel characters and previously seen ones, is one of the rising subtopics in the text recognition field. However, the current open-set text recognition solutions only…
Mixture-of-ExpertsScene Text RecognitionTowards Open World Recognition
With the of advent rich classification models and high computational power visual recognition systems have found many operational applications. Recognition in the real world poses multiple challenges that are not apparen…
Object RecognitionVideo Emotion Open-vocabulary Recognition Based on Multimodal Large Language Model
Multimodal emotion recognition is a task of great concern. However, traditional data sets are based on fixed labels, resulting in models that often focus on main emotions and ignore detailed emotional changes in complex …
Emotion RecognitionLanguage ModelingLanguage ModellingLarge Language Model+2ODN: Opening the Deep Network for Open-set Action Recognition
In recent years, the performance of action recognition has been significantly improved with the help of deep neural networks. Most of the existing action recognition works hold the \textit{closed-set} assumption that all…
Action RecognitionOpen Set Action RecognitionTemporal Action LocalizationTripletOpenPack: A Large-scale Dataset for Recognizing Packaging Works in IoT-enabled Logistic Environments
Unlike human daily activities, existing publicly available sensor datasets for work activity recognition in industrial domains are limited by difficulties in collecting realistic data as close collaboration with industri…
Activity RecognitionHuman Activity Recognition