paper-with-me

Papers

Propensity-scored Probabilistic Label Trees

2021-10-20 · Marek Wydmuch, Kalina Jasinska-Kobus, Rohit Babbar, Krzysztof Dembczyński

Extreme multi-label classification (XMLC) refers to the task of tagging instances with small subsets of relevant labels coming from an extremely large set of all possible labels. Recently, XMLC has been widely applied to diverse web applications such as automatic content labeling, online advertising, or recommendation systems. In such environments, label distribution is often highly imbalanced, consisting mostly of very rare tail labels, and relevant labels can be missing. As a remedy to these problems, the propensity model has been introduced and applied within several XMLC algorithms. In this work, we focus on the problem of optimal predictions under this model for probabilistic label trees, a popular approach for XMLC problems. We introduce an inference procedure, based on the $A^*$-search algorithm, that efficiently finds the optimal solution, assuming that all probabilities and propensities are known. We demonstrate the attractiveness of this approach in a wide empirical study on popular XMLC benchmark datasets.

📄 PDF Abstract BibTeX arXiv:2110.10803

Code (1)

mwydmuch/napkinXC 공식 구현

Tasks

Extreme Multi-Label ClassificationMulti-Label ClassificationMUlTI-LABEL-ClASSIFICATIONRecommendation Systems

Similar Papers 제목 키워드 기반

Best-scored Random Forest Density Estimation

2019-05-09 · Hanyuan Hang, Hongwei Wen

This paper presents a brand new nonparametric density estimation strategy named the best-scored random forest density estimation whose effectiveness is supported by both solid theoretical analysis and significant experim…

Density Estimation

Online probabilistic label trees

2020-07-08 · Kalina Jasinska-Kobus, Marek Wydmuch, Devanathan Thiruvenkatachari, Krzysztof Dembczyński

We introduce online probabilistic label trees (OPLTs), an algorithm that trains a label tree classifier in a fully online manner without any prior knowledge about the number of training instances, their features and labe…

Few-Shot LearningMulti-class Classification

Positivity Validation Detection and Explainability via Zero Fraction Multi-Hypothesis Testing and Asymmetrically Pruned Decision Trees

2021-11-07 · Guy Wolf, Gil Shabat, Hanan Shteingart

Positivity is one of the three conditions for causal inference from observational data. The standard way to validate positivity is to analyze the distribution of propensity. However, to democratize the ability to do caus…

Causal Inference

Calibrated and Conformal Propensity Scores for Causal Effect Estimation

2023-06-01 · Shachi Deshpande, Volodymyr Kuleshov

Propensity scores are commonly used to estimate treatment effects from observational data. We argue that the probabilistic output of a learned propensity score model should be calibrated -- i.e., a predictive treatment p…

Projection-based Annotation of a Polish Dependency Treebank

2014-05-01 · LREC 2014 5 · Alina Wr{\'o}blewska, Adam Przepi{\'o}rkowski

This paper presents an approach of automatic annotation of sentences with dependency structures. The approach builds on the idea of cross-lingual dependency projection. The presented method of acquiring dependency trees …

ARCDependency ParsingMachine TranslationQuestion Answering+1