paper-with-me

Papers

DEDPUL: Difference-of-Estimated-Densities-based Positive-Unlabeled Learning

2019-02-19 · Dmitry Ivanov

Positive-Unlabeled (PU) learning is an analog to supervised binary classification for the case when only the positive sample is clean, while the negative sample is contaminated with latent instances of positive class and hence can be considered as an unlabeled mixture. The objectives are to classify the unlabeled sample and train an unbiased PN classifier, which generally requires to identify the mixing proportions of positives and negatives first. Recently, unbiased risk estimation framework has achieved state-of-the-art performance in PU learning. This approach, however, exhibits two major bottlenecks. First, the mixing proportions are assumed to be identified, i.e. known in the domain or estimated with additional methods. Second, the approach relies on the classifier being a neural network. In this paper, we propose DEDPUL, a method that solves PU Learning without the aforementioned issues. The mechanism behind DEDPUL is to apply a computationally cheap post-processing procedure to the predictions of any classifier trained to distinguish positive and unlabeled data. Instead of assuming the proportions to be identified, DEDPUL estimates them alongside with classifying unlabeled sample. Experiments show that DEDPUL outperforms the current state-of-the-art in both proportion estimation and PU Classification.

📄 PDF Abstract BibTeX arXiv:1902.06965

Code (1)

dimonenka/DEDPUL 공식 구현 pytorch

Tasks

Binary ClassificationDensity EstimationGeneral Classification

Similar Papers 제목 키워드 기반

Meta-learning for Positive-unlabeled Classification

2024-06-06 · Atsutoshi Kumagai, Tomoharu Iwata, Yasuhiro Fujiwara

We propose a meta-learning method for positive and unlabeled (PU) classification, which improves the performance of binary classifiers obtained from only PU data in unseen target tasks. PU learning is an important proble…

ClassificationDensity Ratio EstimationInformation RetrievalMeta-Learning+1

Class-prior Estimation for Learning from Positive and Unlabeled Data

2016-11-05 · Marthinus C. du Plessis, Gang Niu, Masashi Sugiyama

We consider the problem of estimating the class prior in an unlabeled dataset. Under the assumption that an additional labeled dataset is available, the class prior can be estimated by fitting a mixture of class-wise dat…

Analysis of Learning from Positive and Unlabeled Data

2014-12-01 · NeurIPS 2014 12 · Marthinus C. Du Plessis, Gang Niu, Masashi Sugiyama

Learning a classifier from positive and unlabeled data is an important class of classification problems that are conceivable in many practical applications. In this paper, we first show that this problem can be solved by…

General ClassificationOutlier Detection

Clustering Unclustered Data: Unsupervised Binary Labeling of Two Datasets Having Different Class Balances

2013-05-01 · Marthinus Christoffel du Plessis, Masashi Sugiyama

We consider the unsupervised learning problem of assigning labels to unlabeled data. A naive approach is to use clustering methods, but this works well only when data is properly clustered and each cluster corresponds to…

ClusteringDensity Estimation

Marginal Densities, Factor Graph Duality, and High-Temperature Series Expansions

2019-01-07 · Mehdi Molkaraie

We prove that the marginal densities of a global probability mass function in a primal normal factor graph and the corresponding marginal densities in the dual normal factor graph are related via local mappings. The mapp…

Vocal Bursts Intensity Prediction