paper-with-me

홈 › Papers

Time-Sensitive Bayesian Information Aggregation for Crowdsourcing Systems

2015-10-21 · Matteo Venanzi, John Guiver, Pushmeet Kohli, Nick Jennings

Crowdsourcing systems commonly face the problem of aggregating multiple judgments provided by potentially unreliable workers. In addition, several aspects of the design of efficient crowdsourcing processes, such as defining worker's bonuses, fair prices and time limits of the tasks, involve knowledge of the likely duration of the task at hand. Bringing this together, in this work we introduce a new time--sensitive Bayesian aggregation method that simultaneously estimates a task's duration and obtains reliable aggregations of crowdsourced judgments. Our method, called BCCTime, builds on the key insight that the time taken by a worker to perform a task is an important indicator of the likely quality of the produced judgment. To capture this, BCCTime uses latent variables to represent the uncertainty about the workers' completion time, the tasks' duration and the workers' accuracy. To relate the quality of a judgment to the time a worker spends on a task, our model assumes that each task is completed within a latent time window within which all workers with a propensity to genuinely attempt the labelling task (i.e., no spammers) are expected to submit their judgments. In contrast, workers with a lower propensity to valid labeling, such as spammers, bots or lazy labelers, are assumed to perform tasks considerably faster or slower than the time required by normal workers. Specifically, we use efficient message-passing Bayesian inference to learn approximate posterior probabilities of (i) the confusion matrix of each worker, (ii) the propensity to valid labeling of each worker, (iii) the unbiased duration of each task and (iv) the true label of each task. Using two real-world public datasets for entity linking tasks, we show that BCCTime produces up to 11% more accurate classifications and up to 100% more informative estimates of a task's duration compared to state-of-the-art methods.

📄 PDF Abstract BibTeX arXiv:1510.06335

Code (0)

등록된 구현이 없습니다.

Tasks

Bayesian InferenceEntity Linkingvalid

Similar Papers 제목 키워드 기반

Bayesian Crowdsourcing with Constraints

2020-12-20 · Panagiotis A. Traganitis, Georgios B. Giannakis

Crowdsourcing has emerged as a powerful paradigm for efficiently labeling large datasets and performing various learning tasks, by leveraging crowds of human annotators. When additional information is available about the…

Variational Inference

Truthful Online Preference Aggregation for LLM Fine-Tuning in Mobile Crowdsourcing

2026-05-22 · Shugang Hao, Lingjie Duan arxiv

To better serve users' demands in mobile applications (e.g., navigation), mobile crowdsourcing platforms can iteratively align large language model (LLM)-generated content (e.g., AI-generated traffic condition prediction…

Mitigating Cognitive Biases in Multi-Criteria Crowd Assessment

2024-07-10 · Shun Ito, Hisashi Kashima

Crowdsourcing is an easy, cheap, and fast way to perform large scale quality assessment; however, human judgments are often influenced by cognitive biases, which lowers their credibility. In this study, we focus on cogni…

A Bayesian Approach for Sequence Tagging with Crowds

2018-11-02 · IJCNLP 2019 11 · Edwin Simpson, Iryna Gurevych

Current methods for sequence tagging, a core task in NLP, are data hungry, which motivates the use of crowdsourcing as a cheap way to obtain labelled data. However, annotators are often unreliable and current aggregation…

Active LearningArgument Miningnamed-entity-recognitionNamed Entity Recognition+1

Eliciting Categorical Data for Optimal Aggregation

2016-12-01 · NeurIPS 2016 12 · Chien-Ju Ho, Rafael Frongillo, Yi-Ling Chen

Models for collecting and aggregating categorical data on crowdsourcing platforms typically fall into two broad categories: those assuming agents honest and consistent but with heterogeneous error rates, and those assumi…

Multiple-choice