paper-with-me

Papers

An Efficient Large-scale Semi-supervised Multi-label Classifier Capable of Handling Missing labels

2016-06-18 · Amirhossein Akbarnejad, Mahdieh Soleymani Baghshah

Multi-label classification has received considerable interest in recent years. Multi-label classifiers have to address many problems including: handling large-scale datasets with many instances and a large set of labels, compensating missing label assignments in the training set, considering correlations between labels, as well as exploiting unlabeled data to improve prediction performance. To tackle datasets with a large set of labels, embedding-based methods have been proposed which seek to represent the label assignments in a low-dimensional space. Many state-of-the-art embedding-based methods use a linear dimensionality reduction to represent the label assignments in a low-dimensional space. However, by doing so, these methods actually neglect the tail labels - labels that are infrequently assigned to instances. We propose an embedding-based method that non-linearly embeds the label vectors using an stochastic approach, thereby predicting the tail labels more accurately. Moreover, the proposed method have excellent mechanisms for handling missing labels, dealing with large-scale datasets, as well as exploiting unlabeled data. With the best of our knowledge, our proposed method is the first multi-label classifier that simultaneously addresses all of the mentioned challenges. Experiments on real-world datasets show that our method outperforms stateof-the-art multi-label classifiers by a large margin, in terms of prediction performance, as well as training time.

📄 PDF Abstract BibTeX arXiv:1606.05725

Code (0)

등록된 구현이 없습니다.

Tasks

Dimensionality ReductionMissing LabelsMulti-Label ClassificationMUlTI-LABEL-ClASSIFICATION

Similar Papers 제목 키워드 기반

Neural Semi-supervised Learning for Text Classification Under Large-Scale Pretraining

2020-11-17 · Zijun Sun, Chun Fan, Xiaofei Sun, Yuxian Meng 외

The goal of semi-supervised learning is to utilize the unlabeled, in-domain dataset U to improve models trained on the labeled dataset D. Under the context of large-scale language-model (LM) pretraining, how we can make …

General ClassificationLanguage ModellingPseudo Labeltext-classification+1

VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation

2021-01-02 · ACL 2021 5 · Changhan Wang, Morgane Rivière, Ann Lee, Anne Wu 외

We introduce VoxPopuli, a large-scale multilingual corpus providing 100K hours of unlabelled speech data in 23 languages. It is the largest open data to date for unsupervised representation learning as well as semi-super…

Representation Learningspeech-recognitionSpeech Recognition

Confidence-Guided Semi-supervised Learning in Land Cover Classification

2023-05-17 · Wanli Ma, Oktay Karakus, Paul L. Rosin

Semi-supervised learning has been well developed to help reduce the cost of manual labelling by exploiting a large quantity of unlabelled data. Especially in the application of land cover classification, pixel-level manu…

ClassificationDiversityLand Cover ClassificationPseudo Label

Semi-supervised Vision Transformers at Scale

2022-08-11 · Zhaowei Cai, Avinash Ravichandran, Paolo Favaro, Manchen Wang 외

We study semi-supervised learning (SSL) for vision transformers (ViT), an under-explored topic despite the wide adoption of the ViT architectures to different tasks. To tackle this problem, we propose a new SSL pipeline,…

Inductive BiasSemi-Supervised Image Classification

Boosting Facial Expression Recognition by A Semi-Supervised Progressive Teacher

2022-05-28 · Jing Jiang, Weihong Deng

In this paper, we aim to improve the performance of in-the-wild Facial Expression Recognition (FER) by exploiting semi-supervised learning. Large-scale labeled data and deep learning methods have greatly improved the per…

Facial Expression RecognitionFacial Expression Recognition (FER)