paper-with-me

홈 › Papers

Utility-Theoretic Ranking for Semi-Automated Text Classification

2015-03-02 · Giacomo Berardi, Andrea Esuli, Fabrizio Sebastiani

\emph{Semi-Automated Text Classification} (SATC) may be defined as the task of ranking a set $\mathcal{D}$ of automatically labelled textual documents in such a way that, if a human annotator validates (i.e., inspects and corrects where appropriate) the documents in a top-ranked portion of $\mathcal{D}$ with the goal of increasing the overall labelling accuracy of $\mathcal{D}$, the expected increase is maximized. An obvious SATC strategy is to rank $\mathcal{D}$ so that the documents that the classifier has labelled with the lowest confidence are top-ranked. In this work we show that this strategy is suboptimal. We develop new utility-theoretic ranking methods based on the notion of \emph{validation gain}, defined as the improvement in classification effectiveness that would derive by validating a given automatically labelled document. We also propose a new effectiveness measure for SATC-oriented ranking methods, based on the expected reduction in classification error brought about by partially validating a list generated by a given ranking method. We report the results of experiments showing that, with respect to the baseline method above, and according to the proposed measure, our utility-theoretic ranking methods can achieve substantially higher expected reductions in classification error.

📄 PDF Abstract BibTeX arXiv:1503.00491

Code (0)

등록된 구현이 없습니다.

Tasks

ClassificationGeneral Classificationtext-classificationText Classification

Similar Papers 제목 키워드 기반

Marginal-Certainty-aware Fair Ranking Algorithm

2022-12-18 · Tao Yang, Zhichao Xu, Zhenduo Wang, Anh Tran 외

Ranking systems are ubiquitous in modern Internet services, including online marketplaces, social media, and search engines. Traditionally, ranking systems only focus on how to get better relevance estimation. When relev…

Fairness

Towards computer vision technologies: Semi-automated reading of automated utility meters

2022-11-24 · Maria Spichkova, Johan van Zyl

In this report we analysed a possibility of using computer vision techniques for automated reading of utility meters. In our study, we focused on two computer vision techniques: an open-source solution Tensorflow Object …

object-detectionObject Detection

Semi-supervised Ranking Pursuit

2013-07-02 · Evgeni Tsivtsivadze, Tom Heskes

We propose a novel sparse preference learning/ranking algorithm. Our algorithm approximates the true utility function by a weighted sum of basis functions using the squared loss on pairs of data points, and is a generali…

regression

Context-aware Reranking with Utility Maximization for Recommendation

2021-10-18 · Yunjia Xi, Weiwen Liu, Xinyi Dai, Ruiming Tang 외

As a critical task for large-scale commercial recommender systems, reranking has shown the potential of improving recommendation results by uncovering mutual influence among items. Reranking rearranges items in the initi…

counterfactualGraph AttentionPositionRecommendation Systems+1

Convergence of Learning Dynamics in Information Retrieval Games

2018-06-14 · Omer Ben-Porat, Itay Rosenberg, Moshe Tennenholtz

We consider a game-theoretic model of information retrieval with strategic authors. We examine two different utility schemes: authors who aim at maximizing exposure and authors who want to maximize active selection of th…

Information RetrievalRetrieval