paper-with-me

홈 › Papers

Should I Follow the Crowd? A Probabilistic Analysis of the Effectiveness of Popularity in Recommender Systems

2019-02-06 · Rocío Cañamares, Pablo Castells

The use of IR methodology in the evaluation of recommender systems has become common practice in recent years. IR metrics have been found however to be strongly biased towards rewarding algorithms that recommend popular items –the same bias that state of the art recommendation algorithms display. Recent research has confirmed and measured such biases, and proposed methods to avoid them. The fundamental question remains open though whether popularity is really a bias we should avoid or not; whether it could be a useful and reliable signal in recommendation, or it may be unfairly rewarded by the experimental biases. We address this question at a formal level by identifying and modeling the conditions that can determine the answer, in terms of dependencies between key random variables, involving item rating, discovery and relevance. We find conditions that guarantee popularity to be effective or quite the opposite, and for the measured metric values to reflect a true effectiveness, or qualitatively deviate from it. We exemplify and confirm the theoretical findings with empirical results. We build a crowdsourced dataset devoid of the usual biases displayed by common publicly available data, in which we illustrate contradictions between the accuracy that would be measured in a common biased offline experimental setting, and the actual accuracy that can be measured with unbiased observations.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Recommendation Systems

Similar Papers 제목 키워드 기반

Improve Learning from Crowds via Generative Augmentation

2021-07-22 · Zhendong Chu, Hongning Wang

Crowdsourcing provides an efficient label collection schema for supervised machine learning. However, to control annotation cost, each instance in the crowdsourced data is typically annotated by a small number of annotat…

BIG-bench Machine LearningData Augmentation

Probabilistic Multigraph Modeling for Improving the Quality of Crowdsourced Affective Data

2017-01-04 · Jianbo Ye, Jia Li, Michelle G. Newman, Reginald B. Adams, Jr. 외

We proposed a probabilistic approach to joint modeling of participants' reliability and humans' regularity in crowdsourced affective studies. Reliability measures how likely a subject will respond to a question seriously…

Much Ado About Time: Exhaustive Annotation of Temporal Data

2016-07-25 · Gunnar A. Sigurdsson, Olga Russakovsky, Ali Farhadi, Ivan Laptev 외

Large-scale annotated datasets allow AI systems to learn from and build upon the knowledge of the crowd. Many crowdsourcing techniques have been developed for collecting image annotations. These techniques often implicit…

Runtime Analysis of Probabilistic Crowding and Restricted Tournament Selection for Bimodal Optimisation

2018-03-26 · Edgar Covantes Osuna, Dirk Sudholt

Many real optimisation problems lead to multimodal domains and so require the identification of multiple optima. Niching methods have been developed to maintain the population diversity, to investigate many peaks in para…

Diversity

Scalable Variational Gaussian Processes for Crowdsourcing: Glitch Detection in LIGO

2019-11-05 · Pablo Morales-Álvarez, Pablo Ruiz, Scott Coughlin, Rafael Molina 외

In the last years, crowdsourcing is transforming the way classification training sets are obtained. Instead of relying on a single expert annotator, crowdsourcing shares the labelling effort among a large number of colla…

Gaussian ProcessesUncertainty Quantification